LIVE-Last scan updating-53 sources active-874 signals today-RESEARCH PVBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
Research radar

arXiv preprints, filtered for developer impact.

Concise summaries of new AI papers, ranked for how likely they are to change how production systems are built. arXiv collection remains rate-limited.

Papers tracked
2,287
matching records
Shown
10
current page
Top radar
74
http://arxiv.org/abs/2608.26067v1
Code links
2
on this page
PaperAuthorsarXiv IDCategoriesPublishedCodeRadarSummary
StreamPI: Streaming Multimodal Temporal Modeling for Vision-Language-Action ModelsZhe Liu, Jinghua Hou, Yuxiang Luhttp://arxiv.org/abs/2608.26067v1cs.CVAug 26, 2026NoneRDR74Vision-Language-Action (VLA) models have demonstrated effectiveness in robot manipulation, ye...
UltraPIPS: Improving model perception in B-mode ultrasound with foundation modelsTal Grutman, Tali Ilovitshhttp://arxiv.org/abs/2608.26033v1cs.CV, eess.IVAug 26, 2026DetectedRDR80In medical imaging, it is common to use learned perceptual image patch similarity (LPIPS) to...
VISA: Agentic Self-Evolving Data Synthesis for Multimodal Instruction FollowingMin Zeng, Guanxin Tan, Libin Cenhttp://arxiv.org/abs/2608.26013v1cs.CLAug 26, 2026NoneRDR78Multimodal instruction-following models require training data that is accurate, diverse, veri...
A Self-Evolving Multi-Agent Framework Defense against LLM Jailbreak AttacksTongyan Hu, Bryan Hooihttp://arxiv.org/abs/2608.26008v1cs.CR, cs.CLAug 26, 2026NoneRDR76Large language models (LLMs) remain vulnerable to jailbreak attacks that exploit techniques s...
VoiceMem: Streaming Dual-Brain Memory for Real-Time InteractionZhifei Xie, Jiaqi Lang, Ze Anhttp://arxiv.org/abs/2608.26005v1eess.AS, cs.AIAug 26, 2026NoneRDR81Conversational systems, such as duplex speech language models (SLMs), still lack a streaming,...
AsymSpec: Context-Asymmetric Speculative Decoding for Agentic LLMsSheng Liang, Yongyue Zhang, Nathanael Brianhttp://arxiv.org/abs/2608.26004v1cs.AI, cs.CLAug 26, 2026NoneRDR78Agentic LLM pipelines face escalating inference costs as context accumulates across retrieval...
Recursive Experiential-Working Memory Evolution for Long-Horizon Agent HarnessesZhaochen Yu, Yingcheng Wu, Zhenfei Yinhttp://arxiv.org/abs/2608.24876v1cs.AI, cs.CLAug 25, 2026DetectedRDR86Recursive self-improvement (RSI) remains hard in long-horizon tasks, where growing histories...
UrbanGround: From Local Perception to Spatial Agency in a Real-Scale CityTianjie Ju, Zheng Wu, Yueqing Sunhttp://arxiv.org/abs/2608.27456v1cs.CVAug 27, 2026NoneRDR64Multimodal large language models (MLLMs) can interpret a street view, but urban agency depend...
How Language Models Organize and Structure Moral KnowledgeOrion Reblitz-Richardsonhttp://arxiv.org/abs/2608.27402v1cs.CL, cs.AIAug 27, 2026NoneRDR77How do large language models (LLMs) organize moral knowledge? Models detect moral content bro...
Successive Capacity Growth: Task-Complexity-Driven Width and Depth Expansion for Vision Transformer Encoders in JEPA World ModelsFrederik Berenzhttp://arxiv.org/abs/2608.27367v1cs.CV, cs.AIAug 27, 2026NoneRDR77Joint-Embedding Predictive Architectures (JEPAs) for world modeling typically employ fixed-si...