Get Started
Home
Topics
Search
Library
148 papers · page 3 of 15

Research archive

Every paper covered by r*cap, newest first.
On-Policy Self-Distillation without Any Supervision
LLM Training · Aug 9
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers
Agents · Aug 7
SimWAM: A Simple World Action Model for End-to-End Autonomous Driving
Inference Optimization · Aug 7
Mendel Gödel Machine: Recursive Self-Improving Coding Agents via Comparative Evolution
Agents · Aug 7
SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries
Agents · Aug 6
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?
Evaluation · Aug 6
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning
Agents · Aug 6
Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning
Multimodal · Aug 6
SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs
LLM Training · Aug 6
ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation
Agents · Aug 5
123456789101112131415