LIVE-Last scan updating-53 sources active-879 signals today-RESEARCH PVBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
Research radar

arXiv preprints, filtered for developer impact.

Concise summaries of new AI papers, ranked for how likely they are to change how production systems are built. arXiv collection remains rate-limited.

Papers tracked
2,287
matching records
Shown
10
current page
Top radar
76
http://arxiv.org/abs/2608.24848v1
Code links
0
on this page
PaperAuthorsarXiv IDCategoriesPublishedCodeRadarSummary
BrowserForge: Scaling Web Episode via Parallel Browser SandboxesFei Tang, Huawen Shen, Zhiqiong Luhttp://arxiv.org/abs/2608.24848v1cs.CLAug 25, 2026NoneRDR76Web agents that act from rendered pixels avoid the fragility and heavy token cost of reading...
StarHarness: Evolving Harnesses with Stratified Search for Enterprise EnvironmentsEsakkivel Esakkiraja, Denis Akhiyarov, Vikas Yadavhttp://arxiv.org/abs/2608.24804v1cs.AI, cs.SEAug 25, 2026NoneRDR77We present StarHarness, a framework for evolving environment-specific agent harnesses while k...
CAFE: Self-Improving Search Agents Need Co-Evolving FeedbackBoyang Liu, Senjie Jin, Peixin Wanghttp://arxiv.org/abs/2608.24794v1cs.AIAug 25, 2026NoneRDR78Outcome-supervised search agents learn when and how to retrieve evidence, but terminal reward...
Image Difference Quantification Using Autoencoder-Based Latent RepresentationsManish Sharma, Timothy Yim, Clifton Forlineshttp://arxiv.org/abs/2608.24782v1cs.CVAug 25, 2026NoneRDR84Traditional image similarity metrics such as Mean Squared Error (MSE), Peak Signal-to-Noise R...
Stochastic Estimation of Transduced Language ModelsVésteinn Snæbjarnarson, Samuel Kiegeland, Manuel de Prada Corralhttp://arxiv.org/abs/2608.27428v1cs.CLAug 27, 2026NoneRDR73Transduced language models (TLMs) compose a pretrained \emph{source} language model with a fu...
ICON Decomposition: Multivariate Concept-Level Explanations of Deep Representations for Model AuditingRoshan Prakash Rane, Marco Simnacher, Manuel Pfeufferhttp://arxiv.org/abs/2608.26083v1cs.LG, cs.AIAug 26, 2026NoneRDR80Deep neural networks often exploit spurious associations in their training data, a failure kn...
Fine-Tuning Whisper for Automatic Speech Recognition in Baniwa: A Preliminary StudyLeonardo Duart, Tiago Fonseca, Thiago Chacónhttp://arxiv.org/abs/2608.26060v1cs.CL, stat.MLAug 26, 2026NoneRDR79Automatic Speech Recognition (ASR) technologies have achieved remarkable performance in recen...
Beyond Local Surprise: Grounded Dialogue as Selective Belief Revision under Referential UncertaintyZiming Liu, Bhanu Chaitanya Jasti, Ziyang Xuhttp://arxiv.org/abs/2608.26035v1cs.CLAug 26, 2026NoneRDR74When a speaker refers to a scene that the listener cannot directly see, the listener must dec...
Do Robotic World Models Really Follow Actions? Diagnosing and Aligning Action-Conditioned Generation for Policy LearningSixiang Chen, Jiaming Liu, Jixian Wuhttp://arxiv.org/abs/2608.24885v1cs.RO, cs.CVAug 25, 2026NoneRDR79Action-conditioned world models are increasingly used as learned simulators for policy evalua...
What FID Hides: Detecting, Ranking, and Diagnosing Deviations in Generative EvaluationHao Chenhttp://arxiv.org/abs/2608.24881v1stat.ML, cs.LGAug 25, 2026NoneRDR83Generative models are commonly ranked by Fréchet Inception Distance (FID) and Kernel Inceptio...