2026-09-25

Internal Generation Record

Internal generation metadata: 406 candidate papers.

Published 2026-09-25 Target source 2026-09-23 Actual source 2026-09-23 Candidates 406 Featured 6 Tracked 20

Generation Record

This page preserves selected papers, candidate scale, and source-date metadata for traceability. The page only changes presentation, not selected papers, ordering, or counts.

Internal generation record. Fetched at 2026-09-24T23:51:17.107742+00:00. Generated at 2026-09-24T23:52:47.430690+00:00. Machine-readable details stay under data/processed and data/reports.

Selected papers

RankTakeawayTopicarXiv
1make RAG retrieval and knowledge-base QA more reliableBenchmarks and Evaluation2609.28154
2improve code generation, execution feedback, and automated repairSystems and Deployment2609.27380
3improve code generation, execution feedback, and automated repairTraining and Post-training2609.27338
4improve model reasoning, planning, and verificationCode Intelligence2609.28449
8make agents use tools and reusable skills more reliablyAgents and Tool Use2609.27277
10make RAG retrieval and knowledge-base QA more reliableSafety and Alignment2609.28335
5improve model reasoning, planning, and verificationTraining and Post-training2609.28344
6improve code generation, execution feedback, and automated repairBenchmarks and Evaluation2609.27510
7improve code generation, execution feedback, and automated repairBenchmarks and Evaluation2609.27441
9improve model reasoning, planning, and verificationBenchmarks and Evaluation2609.27173
11improve model reasoning, planning, and verificationBenchmarks and Evaluation2609.28007
12make RAG retrieval and knowledge-base QA more reliableSystems and Deployment2609.27981
13improve code generation, execution feedback, and automated repairTraining and Post-training2609.27980
14improve image generation, visual understanding, and controllable renderingBenchmarks and Evaluation2609.27656
15make RAG retrieval and knowledge-base QA more reliableVideo Generation2609.27274
16use benchmarks and evaluations to expose model weaknessesSystems and Deployment2609.27202
17make RAG retrieval and knowledge-base QA more reliableTraining and Post-training2609.28378
18make RAG retrieval and knowledge-base QA more reliableBenchmarks and Evaluation2609.28328
19make agents use tools and reusable skills more reliablyAgents and Tool Use2609.28296
20improve code generation, execution feedback, and automated repairCode Intelligence2609.28248
21make RAG retrieval and knowledge-base QA more reliableBenchmarks and Evaluation2609.28230
22improve model reasoning, planning, and verificationMultimodal Models2609.28222
23make agents use tools and reusable skills more reliablyAgents and Tool Use2609.28216
24make agents use tools and reusable skills more reliablyAgents and Tool Use2609.28197
25make agents use tools and reusable skills more reliablyAgents and Tool Use2609.28182
26improve code generation, execution feedback, and automated repairVideo Generation2609.28080