LIVE-Last scan updating-53 sources active-864 signals today-RESEARCH PVBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
Research radar

arXiv preprints, filtered for developer impact.

Concise summaries of new AI papers, ranked for how likely they are to change how production systems are built. arXiv collection remains rate-limited.

Papers tracked
2,287
matching records
Shown
10
current page
Top radar
83
http://arxiv.org/abs/2608.27455v1
Code links
2
on this page
PaperAuthorsarXiv IDCategoriesPublishedCodeRadarSummary
CritICL: Inference-Time Weak-to-Strong Generalization from Small Language Model Failure ModesYufan Wu, Yinghui He, Zhengyi Huhttp://arxiv.org/abs/2608.27455v1cs.CLAug 27, 2026DetectedRDR83Recent advances in inference-time scaling have significantly improved the reasoning performan...
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090Kairong Luo, Jiarui Cui, Yaorui Yinhttp://arxiv.org/abs/2608.27370v1cs.CL, cs.LGAug 27, 2026DetectedRDR87Language model pretraining has become almost synonymous with prohibitive cost, placing it out...
RedEvoAgent: Automatic Red-Teaming Agent with Experience-Driven Skill EvolutionJunjie Zhang, Hui Liu, Kecheng Chenhttp://arxiv.org/abs/2608.27439v1cs.CR, cs.AIAug 27, 2026NoneRDR82LLM-based agents are increasingly deployed in product-level execution harnesses, where jailbr...
Retrieval Heads Meet Vision: Uncovering How VLMs Locate and Extract Visual InformationChanho Park, Daehyeon Choi, Jihyun Leehttp://arxiv.org/abs/2608.27417v1cs.CVAug 27, 2026NoneRDR84Vision-language models (VLMs) can locate an image region referred to by a text prompt and rou...
Reconstructing Humans and Objects in Interaction using Large Reconstruction ModelsAgniv Chatterjee, Georgios Pavlakoshttp://arxiv.org/abs/2608.27407v1cs.CVAug 27, 2026NoneRDR83Estimation of Human-Object Interactions in 3D (3D HOI) is a fundamental problem in 3D compute...
CLAP: Cross-Embodiment Video World Models are Zero-Shot Physical SimulatorsKechen Liu, Ola Shorinwahttp://arxiv.org/abs/2608.27406v1cs.RO, cs.AIAug 27, 2026NoneRDR86State-of-the-art action-conditioned video models are typically restricted to a single robot e...
Making Clinical Language Models Auditable: Concept-Guided Fine-Tuning for Robust PredictionJin Mu, Guanhua Chenhttp://arxiv.org/abs/2608.27397v1cs.CL, cs.AIAug 27, 2026NoneRDR84Clinical language models can achieve strong in-hospital accuracy yet fail under deployment sh...
LeVJEPA: Efficient & Scalable Video Pretraining without the HeuristicsLukas Kuhn, Lucas Maes, Giuseppe Serrahttp://arxiv.org/abs/2608.27395v1cs.CV, cs.AIAug 27, 2026NoneRDR87Video carries the temporal structure of the physical world, yet learning representations from...
Persona-Execution Separation: An Architecture Pattern for Evolving LLM Agents under Execution AuditYisen Xihttp://arxiv.org/abs/2608.27427v1cs.SE, cs.AIAug 27, 2026NoneRDR77Large language model (LLM) agents in governed organizations must let the persona (instruction...
RATIO: A Benchmark for Retrieval Across Typed Ideation Operations in Scientific LiteratureMaayan Sharon, Tom Hopehttp://arxiv.org/abs/2608.27394v1cs.CL, cs.IRAug 27, 2026NoneRDR84Retrieved scientific literature can serve as inspiration for both human and AI scientists. In...