AI method reports gains in multi-label image recognition
Preprint: PuRF reported higher mAP than comparison methods across five multi-label image benchmarks.
981–1000
Preprint: PuRF reported higher mAP than comparison methods across five multi-label image benchmarks.
A synthetic, counterfactually constructed benchmark favored TSIM in retrieval and accuracy tests, but the results do not establish performance on natural conversations.
Preprint: The policy posted stronger benchmark and pooled physical results, but each action chunk took 44 milliseconds longer than reproduced π0.5.
Preprint: A controlled evaluation found that finite reference lists could reverse which confidence rule looked better for open-ended model beliefs.
Preprint: A model of connected-vehicle services links jitter and traffic variability with lower deadline reliability and higher tail latency, especially in the cloud.
Preprint: A same-crystal comparison finds a near-quarter-cycle offset between magnetization and a strain-sensor signal, but no unique Berry phase.
Preprint: A hidden-state method reads an AI model’s reasoning signals and leads tested approaches in judging research-idea novelty.
Preprint testing found higher scores on two public benchmarks and one company dataset, while a verification gate rejected 37% of proposed changes.
Preprint: A simulated benchmark reports higher recognition scores under a privileged reference trajectory than under a stationary initial view, while an active policy recovered only part of the gap.
Preprint: PRISM reports broader feasible sampling in a single-joint test, task-specific success rates under manual and Bayesian tuning, and reported transfer to physical UR5e robots.
A preprint argues that useful citations must trace training data, retrieved datasets and individual knowledge-graph facts.
Preprint: In a study of 12 adults, spear-phishing paired with greed had the highest reported yes-rate among 25 simulated scenarios, at 75%.
Preprint: A mathematical study reports conditional curvature and Hessian estimates alongside a new concavity inequality.
An arXiv preprint compares Dyson, non-Dyson and screened calculations against selected-CI reference values.
A review proposes testing the code behind vulnerability claims before human triage, but its proposed safeguards have not yet been validated.
Preprint: COSMOS-Web observations find star formation rising slightly faster than stellar mass at low masses before flattening, a pattern that better matches the evolving galaxy mass distribution.
Preprint: a theorem-driven analysis finds expansion, boundedness or collapse as the boundary-density parameter changes.
An arXiv preprint reports an acquisition-retention trade-off in Qwen3 tests, while a smaller mathematics experiment showed a much weaker pattern.
Preprint analysis of 5,124 massive central galaxies supports a possible early-type-to-spiral pathway, but does not directly observe the transformation.
Preprint: The Transformer-based route led in the study’s Heavy OOD category, while supervised-only training performed better in Slight OOD.