Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends
papers

TTPO: TEST-TIME POLICY OPTIMIZATION

URBANGROUND: FROM LOCAL PERCEPTION TO SPATIAL AGENCY IN A REAL-SCALE CITY






























PAWBENCH: HOW FAR ARE WE FROM PROBABILISTICALLY ALIGNED WORLD MODELING?
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models
WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation
VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
FrontierChallenge: Evaluating Scientific Workflow Completion
VGI-Bench: Probing Visual Intelligence in Video Generation Models
VOICEMEM: STREAMING DUAL-BRAIN MEMORY FOR REAL-TIME INTERACTION
LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus
GLOBE: A High-quality English Corpus with Global Accents for Zero-shot Speaker Adaptive Text-to-Speech
From Passive Response to Proactive Correction: Enhancing LLM Robustness Against Input Fact Perturbations
A General-Purpose Molecular Foundation Model Transfers Across Diverse Olfactory Tasks
Embedding NDRE Trajectories into Contrastive Learning for Label-Free, Physiology-Aware Crop-Stress Staging and DSS Outputs
Efficient Estimation of High Information Projections using Nearest Neighbours
Scaling Harness Intelligence via Just-in-Time Harness Evolution
GaussianDream++: Efficient 3D Gaussian World Modeling for Robotic Manipulation
Recursive Experiential–Working Memory Evolution for Long-Horizon Agent Harnesses
CYBERFACTORY: SCALING CYBER SECURITY CAPABILITIES WITH INSTANCES FROM THE WILD
On-Policy Self-Distillation in Diffusion Models
AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces
WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report
GigaBrain-0.7: Scaling Embodied Foundation Models to Emergent Capabilities with a Three-System Architecture
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs
RISE: Adaptive Imagination for World Action Models
MobilePA-Bench: Benchmarking Mobile Planner Agents on Complex Real-World Tasks
TLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming
Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision
EchoWM: Open and Enterable Omnimodal World Models
Apodex 1.1: Scaling Agentic Intelligence for Complex Work
EviRank: Structured Relevance Evidence for Multimodal Image Re-ranking
Every Coin Has Two Sides: On the Dual Nature of Generalization in On-Policy Distillation of Large Language Models