Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends
papers

LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing

Agentic Abstention: Do Agents Know When to Stop Instead of Act?






























EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents
SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs
MultiHashFormer: Hash-based Generative Language Models
Qwen-Image-2.0-RL Technical Report
Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots
PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation
OpenTME: An Open Dataset of AI-powered H&E Tumor Microenvironment Profiles from TCGA
FlashAttention-4: Algorithm and Kernel Pipelining Co-Design for Asymmetric Hardware Scaling
DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation
ViQ: Text-Aligned Visual Quantized Representations at Any Resolution
The Verification Horizon: No Silver Bullet for Coding Agent Rewards
Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation
OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning
In-Context World Modeling for Robotic Control
DanceOPD: On-Policy Generative Field Distillation
Autodata: An agentic data scientist to create high quality synthetic data
Improved Large Language Diffusion Models
How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations
RoboAtlas: Contextual Active SLAM
Learning Robot Visual Navigation in Crowds via Intention-Aware Scene Representations
Deep Reinforcement Learning-Enhanced Event-Triggered Data-Driven Predictive Control for a 3D Cable-Driven Soft Robotic Arm
Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining
Every Nonnegative Integer Is a Sum of a Triangular, a Pentagonal, and a Heptagonal Number
Loop Engineering: The Anthropic Playbook for Designing Systems That Prompt Your Agents
Small LLMs: Pruning vs. Training from Scratch
OpenThoughts-Agent: Data Recipes for Agentic Models
LingxiDiagBench: A Multi-Agent Framework for Benchmarking LLMs in Chinese Psychiatric Consultation and Diagnosis
AOHP: An Open-Source OS-Level Agent Harness for Personalized, Efficient and Secure Interaction
MemGUI-Agent: An End-to-End Long-Horizon Mobile GUI Agent with Proactive Context Management