2026-07-31

Internal Generation Record

Internal generation metadata: 339 candidate papers.

Published 2026-07-31 Target source 2026-07-29 Actual source 2026-07-29 Candidates 339 Featured 6 Tracked 20

Generation Record

This page preserves selected papers, candidate scale, and source-date metadata for traceability. The page only changes presentation, not selected papers, ordering, or counts.

Internal generation record. Fetched at 2026-07-30T22:16:23.044701+00:00. Generated at 2026-07-30T22:17:40.608292+00:00. Machine-readable details stay under data/processed and data/reports.

Selected papers

RankTakeawayTopicarXiv
1make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2607.26723
2make agents use tools and reusable skills more reliablyCode Intelligence2607.27146
3improve model reasoning, planning, and verificationTraining and Post-training2607.26654
4make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2607.26481
6make agents use tools and reusable skills more reliablyAgents and Tool Use2607.26791
8make RAG retrieval and knowledge-base QA more reliableVideo Generation2607.26511
5improve model reasoning, planning, and verificationBenchmarks and Evaluation2607.26368
7make agents use tools and reusable skills more reliablyCode Intelligence2607.26710
9make RAG retrieval and knowledge-base QA more reliableVideo Generation2607.26429
10make RAG retrieval and knowledge-base QA more reliableRobotics and Embodied AI2607.27205
11make RAG retrieval and knowledge-base QA more reliableBenchmarks and Evaluation2607.27194
12make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2607.27191
13make agents use tools and reusable skills more reliablyCode Intelligence2607.27167
14make agents use tools and reusable skills more reliablySystems and Deployment2607.27132
15identify and reduce safety, jailbreak, and alignment risksTraining and Post-training2607.26981
16make RAG retrieval and knowledge-base QA more reliableRetrieval and RAG2607.26843
17improve code generation, execution feedback, and automated repairTraining and Post-training2607.26801
18make RAG retrieval and knowledge-base QA more reliableCode Intelligence2607.26503
19improve code generation, execution feedback, and automated repairSystems and Deployment2607.26491
20strengthen multimodal understanding of charts, documents, and visual evidenceBenchmarks and Evaluation2607.27180
21make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2607.27155
22improve model reasoning, planning, and verificationBenchmarks and Evaluation2607.27145
23make RAG retrieval and knowledge-base QA more reliableBenchmarks and Evaluation2607.27136
24make agents use tools and reusable skills more reliablyBenchmarks and Evaluation2607.27084
25test temporal consistency and motion realism in video generationSafety and Alignment2607.27081
26make agents use tools and reusable skills more reliablyCode Intelligence2607.27080