Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends
papers

Transformer Transformer: A Unified Model for Motion-Conditioned Robot Co-design

Fine-Grained Food Image Understanding via Target-Aware Data Alignment






























Verification of Provers and Solvers
The Model in the Middle: Toward AI-Native Real-Time Communication
DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
MemoHarness: Agent Harnesses That Learn from Experience
Nanbeige4.2-3B: Unlocking Agentic Capabilities in a Compact Model
RETHINKING CLASSIFIER-FREE GUIDANCE IN ON-POLICY DIFFUSION DISTILLATION
STATEACT: PROGRAM STATE, BEFORE PIXELS, FOR LONG-HORIZON COMPUTER-USE AGENTS
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents
KIMI K3: OPEN FRONTIER INTELLIGENCE
MUSIC-JEPA: LEARNING A WORLD MODEL OF SOUND FROM ACTION
LLMs can’t jump
LLMs Get Lost in Evolving User Intent
SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales
Improved Distribution Matching Distillation for Fast Image Synthesis
THREE-BODY SCATTERING FOR GENERATIVE MODELING
Scaling Native Multimodal Pre-Training From Scratch
DataPrep-Bench: Benchmarking LLMs as Training Data Preparators
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills
Better Harnesses, Smaller Models: Building 90% Cheaper Agents via Automated Harness Adaptation
From Memory to Skills: Evidence-Grounded Co-Evolution Governance for Long-Horizon LLM Agents
Context-weighted Discrete Flow Matching
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation
Show, Don’t Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text
VISUAL CONTRASTIVE SELF-DISTILLATION
ReferTrack: Referring Then Tracking for Embodied Visual Tracking
K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs