Get Started
Home
Topics
Search
Library
148 papers · page 3 of 15
Research archive
Every paper covered by r*cap, newest first.
On-Policy Self-Distillation without Any Supervision
LLM Training · Aug 9
0
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers
Agents · Aug 7
0
SimWAM: A Simple World Action Model for End-to-End Autonomous Driving
Inference Optimization · Aug 7
0
Mendel Gödel Machine: Recursive Self-Improving Coding Agents via Comparative Evolution
Agents · Aug 7
0
SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries
Agents · Aug 6
0
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?
Evaluation · Aug 6
0
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning
Agents · Aug 6
0
Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning
Multimodal · Aug 6
0
SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs
LLM Training · Aug 6
0
ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation
Agents · Aug 5
0
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15