AI can tidy model cards, but generated errors persist
Preprint: Two LLM workflows kept most existing information in place, but only 26 of 48 newly generated cards were judged fully correct.
1361–1380
Preprint: Two LLM workflows kept most existing information in place, but only 26 of 48 newly generated cards were judged fully correct.
Preprint: A retrieval-based language model outperformed an LLM-only baseline and was used to plan routes for prospective target phases.
A preprint reports lower first-arrival picking RMSE than a Vision Transformer on Brunswick, Halfmile and Dongbei, while its noise evidence remains qualitative.
Preprint: The study reports higher scores when a detector-agnostic restoration model processes synthetic pairs and 200 real nighttime UAV images, while real-world validation remains limited.
Preprint: A controlled 3-D simulation found semantic labels remained coherent deeper into noise, while fading-input tests mislabelled 18% of held-out novel objects as known.
This arXiv Preprint compares geometry-specific readouts with text-coordinate output across grounding benchmarks, simulated manipulation, transfer tests and physical robot trials.
A preprint reports a two-orders-of-magnitude cost advantage in SrMnO3, but the method remains a limited computational demonstration.
Preprint: A theoretical analysis finds that classical fields recover links between radiation energy, pressure and flux, but not the spectrum’s absolute normalization.
Preprint: A review compares model-based, sensory, incremental and hybrid nonlinear dynamic inversion, including a Citation II flight-test example.
An arXiv preprint reports higher post-RLVR scores and sharply lower compute costs for the mixed training route.
Preprint: Measurements and simulations follow D2+ formation through an indirect pathway with direct, roaming and delayed branches.
An arXiv preprint reports the highest displayed scores across four MAVOS-DD scenarios, while AVLips, cross-dataset and video-only comparisons are less uniform.
A statistical preprint proposes a way to classify estimates into rival possibilities while controlling the risk of assigning the wrong class.
Preprint: The method was tested on three ECG datasets under offline, continual online and independent online adaptation protocols.
An arXiv version 2 preprint dated 25 August 2026 maps 7,120 papers and finds core-task coverage centered on Modern Standard Arabic.
MediSkill-Evo led several fixed benchmark comparisons, but the results do not establish clinical safety or prospective validity.
This preprint reports lower judged error rates for compact models, but its benchmark and language-model evaluation limit the result.
An arXiv version 1 preprint reports strong results on CREMA-D and AV-MNIST, mixed results elsewhere, and no win over VGGSound’s strongest baselines.
A preprint reports lower collision rates on nuScenes and stronger benchmark scores across Bench2Drive and NAVSIM.
The arXiv preprint targets x^n−1, x^n and x^n+1 when the exponent’s prime factors all fall in a tightly defined class.