Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends

Deep Learning in Remote Sensing: A Review

A Regression Approach to Speech Enhancement Based on Deep Neural Networks

Deep Neural Networks for Acoustic Modeling in Speech Recognition

RoboTTT: Context Scaling for Robot Policies

SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

Efficient Estimation of Word Representations in Vector Space

Depth Map Prediction from a Single Image using a Multi-Scale Deep Network

TabNet: Attentive Interpretable Tabular Learning

AudioPaLM: A Large Language Model That Can Speak and Listen

SQuAD: 100,000+ Questions for Machine Comprehension of Text

DeepPose: Human Pose Estimation via Deep Neural Networks

Self-Improvements in Modern Agentic Systems: A Survey

Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning

SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference

MetaView: Monocular Novel View Synthesis with Scale-Aware Implicit Geometry Priors

PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails

KnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and Skill

OvisOCR2 Technical Report

Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation

Harness Handbook: Making Evolving Agent Harnesses Readable, Navigable, and Editable

Qwen-Music Technical Report

Spectral Rewiring for Exploration, Purification, and Model Merging

Rethinking the Evaluation of Harness Evolution for Agents

Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning

Bridging the Gap Between Latent and Explicit Reasoning with Looped Transformers

Towards Autonomous and Auditable Medical Imaging Model Development

MUSCRIPTOR: AN OPEN MODEL FOR MULTI-INSTRUMENT MUSIC TRANSCRIPTION

Principled Analysis of Deep Reinforcement Learning Evaluation and Design Paradigms

Know Before Fix: QA-Driven Repository Knowledge Acquisition for Software Issue Resolution

Blind-Spots-Bench: Evaluating Blind Spots in Multimodal Models

Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation

The Role of Rigor in Artificial Intelligence

Deep Learning in Remote Sensing: A Review

A Regression Approach to Speech Enhancement Based on Deep Neural Networks

Deep Neural Networks for Acoustic Modeling in Speech Recognition

RoboTTT: Context Scaling for Robot Policies

SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

Efficient Estimation of Word Representations in Vector Space

Depth Map Prediction from a Single Image using a Multi-Scale Deep Network

TabNet: Attentive Interpretable Tabular Learning

AudioPaLM: A Large Language Model That Can Speak and Listen

SQuAD: 100,000+ Questions for Machine Comprehension of Text

DeepPose: Human Pose Estimation via Deep Neural Networks

Self-Improvements in Modern Agentic Systems: A Survey

Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning

SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference

MetaView: Monocular Novel View Synthesis with Scale-Aware Implicit Geometry Priors

PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails

KnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and Skill

OvisOCR2 Technical Report

Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation

Harness Handbook: Making Evolving Agent Harnesses Readable, Navigable, and Editable

Qwen-Music Technical Report

Spectral Rewiring for Exploration, Purification, and Model Merging

Rethinking the Evaluation of Harness Evolution for Agents

Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning

Bridging the Gap Between Latent and Explicit Reasoning with Looped Transformers

Towards Autonomous and Auditable Medical Imaging Model Development

MUSCRIPTOR: AN OPEN MODEL FOR MULTI-INSTRUMENT MUSIC TRANSCRIPTION

Principled Analysis of Deep Reinforcement Learning Evaluation and Design Paradigms

Know Before Fix: QA-Driven Repository Knowledge Acquisition for Software Issue Resolution

Blind-Spots-Bench: Evaluating Blind Spots in Multimodal Models

Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation

The Role of Rigor in Artificial Intelligence