Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends
papers

VISUAL CONTRASTIVE SELF-DISTILLATION

ReferTrack: Referring Then Tracking for Embodied Visual Tracking






























K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs
AREX: Towards a Recursively Self-Improving Agent for Deep Research
Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization
Scaling Laws for HyperNetwork-Based Knowledge Injection in Large Language Models
An Exam for Active Observers
Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking
SELF GRADIENT FORCING: NATIVE LONG VIDEO EX-TRAPOLATION
SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD
Vera: A Layered Difusion Model for Content-Preserving Video Editing
Two-Level Meta-Rubrics for Evaluating Open-Ended Generation: GAMUT, a Benchmark for Factual Completeness
Automated Discovery Has No Universally Superior Harness
Towards a Science of Scaling Agent Systems
AlayaWorld: Interactive Long-Horizon World Modeling - Full Technical Report
Mage-Flow: An Eficient Native-Resolution Foundation Model for Image Generation and Editing
DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines
Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers
Generative World Renderer at the Speed of Play
Infinite Interactive World Rollout on a Single Desktop GPU
UniMoMo: Unified Generative Modeling of 3D Molecules for De Novo Binder Design
Measuring Reward-Seeking via Contrastive Belief Updates
LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks
Apple-π: A Benchmark for Evaluating Video Generation Models on Physical Law Grounding
HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune
DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment
EvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary World
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs
Understanding Reasoning from Pretraining to Post-Training
Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops
Loop the Loopies!