Get Started
Topic · 272 recaps
Research Paper
Academic and industry research papers covering various fields, including scientific discoveries, market analysis, and technical innovations.
Play all
...
Posts
Questions
Home
Topics
Search
Library
Sort
Newest
$Φ$-Bench: Can Large Language Models Engineer the Infrastructure That Powers Them?
Agents · Sep 9
0
Show-Harness: Just a VLM Agent Can Play Robots
Agents · Sep 9 · 10:04
0
Agent Memory Controls Must Follow Consequences, Not Labels
Agents · Sep 8 · 10:35
0
SyncWorld: Visual Calibration Enables World Models as Zero-Shot Simulators
Reinforcement Learning · Sep 8
0
SchemeArena: Factorized Stress Testing of Scheming in LLM Agents
Agents · Sep 8
0
AgentGrad: Intervention-guided Prompt Optimization for Multi Agent Systems
Agents · Sep 8
0
Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation
LLM Training · Sep 8 · 7:52
0
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness
Agents · Sep 8
0
Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation
Diffusion · Sep 8
0
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
Agents · Sep 8
0