LIVE-Last scan updating-53 sources active-398 signals today-BENCHMARKSCorporateBench: A New Benchmark for LLM Q&A on Enterprise Data
This week

AI coding repos this week

A weekly radar of AI coding and repository automation signals that are useful to builders.

Items
7
published only
Avg RDR
79
current set
Sources
4
linked domains
Updated
Aug 29, 2026
indexable
InfrastructureGitHub trend signalAug 23, 2026

Omadia: A Self-Hostable Agentic OS for Multi-Agent AI Teams

Omadia is a self-hostable agentic operating system designed for building, running, and auditing multi-agent AI teams. It supports signed plugins, allows users to bring their own LLM keys, and emphasizes data ownership with EU/GDPR readiness.

RDR87
Benchmarksresearch signalAug 26, 2026

SWE Refactor Bench: Can Coding Agents Complete a Long-Horizon, Whole-Repository Stack Migration?

A new benchmark, SWE Refactor Bench, has been introduced to evaluate the ability of coding agents to perform complex, whole-repository software migrations. Existing benchmarks are insufficient as they do not verify if the migration actually occurred, allowing agents to pass tests by copying original code. This benchmark addresses that gap by assessing both migration completeness and behavioral correctness.

RDR82
Infrastructureofficial or source announcementAug 26, 2026

OpenAI Unveils Jalapeño Inference Chip

OpenAI has introduced Jalapeño, a custom-designed inference chip. This new hardware aims to significantly improve the speed and power efficiency of AI inference tasks.

RDR80
Research Papersresearch signalAug 29, 2026

OpenAI Report: AI Enhances Continuous Learning

A new report from OpenAI details how students and educators are leveraging ChatGPT to foster continuous learning. The findings highlight AI's role in providing educational support that transcends traditional classroom boundaries.

RDR78
Research Papersresearch signalAug 24, 2026

G-CARL: Grounded Checklist-Aligned Reward Learning for Patient-Oriented Medical Report Interpretation

Researchers have introduced G-CARL, a novel reinforcement learning framework designed for patient-oriented medical report interpretation. This framework aims to bridge the gap in existing medical vision-language tasks by addressing the dual requirements of evidence-grounded factuality and context-dependent patient communication.

RDR78
AI Toolsthird-party newsAug 25, 2026

New Hugging Face Datasets for PDE LLM Evaluation

Three new Hugging Face datasets have been released for evaluating Large Language Models (LLMs) on Partial Differential Equations (PDEs). These datasets, named bermaneh/pde-llm-eval-freegen-xmodal-qwen__qwen3-8-27b, bermaneh/pde-llm-eval-freegen-xmodal-qwen__qwen3-6-27b, and bermaneh/pde-llm-eval-freegen-xmodal-qwen__qwen3-5-27b, contain tabular and text data for free-generation PDE evaluation.

RDR68