Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends

LMEnt: A Suite for Analyzing Knowledge in Language Models from Pretraining Data to Representations

Open Data Synthesis For Deep Research































LMEnt: A Suite for Analyzing Knowledge in Language Models from Pretraining Data to Representations

Open Data Synthesis For Deep Research






























Robix: A Unified Model for Robot Interaction, Reasoning and Planning
FusionProt: Fusing Sequence and Structural Information for Unified Protein Representation Learning
LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence
epiGPTope: A machine learning-based epitope generator and classifier
GenCompositor: Generative Video Compositing with Diffusion Transformer
DCPO: Dynamic Clipping Policy Optimization
Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
Baichuan-M2: Scaling Medical Capability with Large Verifier System
VerlTool: Towards Holistic Agentic Reinforcement Learning with Tool Use
ELV-Halluc: Benchmarking Semantic Aggregation Hallucinations in Long Video Understanding
AlphaEarth Foundations: An embedding field model for accurate and efficient global mapping from sparse label data
AetherCode: Evaluating LLMs' Ability to Win In Premier Programming Competitions
TileLang: A Composable Tiled Programming Model for AI Systems
DeepSeek-R1 Thoughtology: Let's think about LLM Reasoning
Multi-Ontology Integration with Dual-Axis Propagation for Medical Concept Representation
Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture
SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion
olmOCR: Unlocking Trillions of Tokens in PDFs with Vision Language Models
How Can Input Reformulation Improve Tool Usage Accuracy in a Complex Dynamic Environment? A Study on τ-bench
UI-Level Evaluation of ALLaM 34B: Measuring an Arabic-Centric LLM via HUMAIN Chat
From reactive to cognitive: brain-inspired spatial intelligence for embodied agents
No Label Left Behind: A Unified Surface Defect Detection Model for all Supervision Regimes
T2R-bench: A Benchmark for Generating Article-Level Reports from Real World Industrial Tables
PVPO: Pre-Estimated Value-Based Policy Optimization for Agentic Reasoning
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
UQ: Assessing Language Models on Unsolved Questions
CARJAN: Agent-Based Generation and Simulation of Traffic Scenarios with AJAN
TiKMiX: Take Data Influence into Dynamic Mixture for Language Model Pre-training
TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis
Droplet3D: Commonsense Priors from Videos Facilitate 3D Generation
Robix: A Unified Model for Robot Interaction, Reasoning and Planning
FusionProt: Fusing Sequence and Structural Information for Unified Protein Representation Learning
LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence
epiGPTope: A machine learning-based epitope generator and classifier
GenCompositor: Generative Video Compositing with Diffusion Transformer
DCPO: Dynamic Clipping Policy Optimization
Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
Baichuan-M2: Scaling Medical Capability with Large Verifier System
VerlTool: Towards Holistic Agentic Reinforcement Learning with Tool Use
ELV-Halluc: Benchmarking Semantic Aggregation Hallucinations in Long Video Understanding
AlphaEarth Foundations: An embedding field model for accurate and efficient global mapping from sparse label data
AetherCode: Evaluating LLMs' Ability to Win In Premier Programming Competitions
TileLang: A Composable Tiled Programming Model for AI Systems
DeepSeek-R1 Thoughtology: Let's think about LLM Reasoning
Multi-Ontology Integration with Dual-Axis Propagation for Medical Concept Representation
Automated Clinical Problem Detection from SOAP Notes using a Collaborative Multi-Agent LLM Architecture
SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion
olmOCR: Unlocking Trillions of Tokens in PDFs with Vision Language Models
How Can Input Reformulation Improve Tool Usage Accuracy in a Complex Dynamic Environment? A Study on τ-bench
UI-Level Evaluation of ALLaM 34B: Measuring an Arabic-Centric LLM via HUMAIN Chat
From reactive to cognitive: brain-inspired spatial intelligence for embodied agents
No Label Left Behind: A Unified Surface Defect Detection Model for all Supervision Regimes
T2R-bench: A Benchmark for Generating Article-Level Reports from Real World Industrial Tables
PVPO: Pre-Estimated Value-Based Policy Optimization for Agentic Reasoning
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
UQ: Assessing Language Models on Unsolved Questions
CARJAN: Agent-Based Generation and Simulation of Traffic Scenarios with AJAN
TiKMiX: Take Data Influence into Dynamic Mixture for Language Model Pre-training
TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis
Droplet3D: Commonsense Priors from Videos Facilitate 3D Generation