2026-08-25

Internal Generation Record

Internal generation metadata: 304 candidate papers.

Published 2026-08-25 Target source 2026-08-23 Actual source 2026-08-21 Candidates 304 Featured 6 Tracked 20 Source date fallback

Generation Record

This page preserves selected papers, candidate scale, and source-date metadata for traceability. The page only changes presentation, not selected papers, ordering, or counts.

Internal generation record. Fetched at 2026-08-24T21:39:14.407221+00:00. Generated at 2026-08-24T21:40:34.973578+00:00. Machine-readable details stay under data/processed and data/reports.

Selected papers

RankTakeawayTopicarXiv
1make agents use tools and reusable skills more reliablyAgents and Tool Use2608.21107
2make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2608.21208
3make agents use tools and reusable skills more reliablySafety and Alignment2608.21100
4make RAG retrieval and knowledge-base QA more reliableBenchmarks and Evaluation2608.21300
5make agents use tools and reusable skills more reliablyRetrieval and RAG2608.21095
9improve code generation, execution feedback, and automated repairCode Intelligence2608.21074
6make agents use tools and reusable skills more reliablyRetrieval and RAG2608.21075
7improve code generation, execution feedback, and automated repairBenchmarks and Evaluation2608.20870
8make agents use tools and reusable skills more reliablyAgents and Tool Use2608.20797
10make RAG retrieval and knowledge-base QA more reliableVideo Generation2608.21041
11make RAG retrieval and knowledge-base QA more reliableBenchmarks and Evaluation2608.20991
12improve model reasoning, planning, and verificationBenchmarks and Evaluation2608.20884
13make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2608.20851
14make agents use tools and reusable skills more reliablyAgents and Tool Use2608.20805
15improve code generation, execution feedback, and automated repairMultimodal Models2608.20748
16make agents use tools and reusable skills more reliablyCode Intelligence2608.20711
17make RAG retrieval and knowledge-base QA more reliableTraining and Post-training2608.20687
18make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2608.20664
19improve code generation, execution feedback, and automated repairCode Intelligence2608.20656
20improve code generation, execution feedback, and automated repairCode Intelligence2608.20653
21make RAG retrieval and knowledge-base QA more reliableSystems and Deployment2608.21308
22make RAG retrieval and knowledge-base QA more reliableVideo Generation2608.21290
23test temporal consistency and motion realism in video generationSafety and Alignment2608.21278
24improve model reasoning, planning, and verificationRetrieval and RAG2608.21252
25make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2608.21249
26make agents use tools and reusable skills more reliablyAgents and Tool Use2608.21209