Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends

Igniting Creative Writing in Small Language Models: LLM-as-a-Judge versus Multi-Agent Refined Rewards

TMUAD: Enhancing Logical Capabilities in Unified Anomaly Detection Models with a Text Memory Bank































Igniting Creative Writing in Small Language Models: LLM-as-a-Judge versus Multi-Agent Refined Rewards

TMUAD: Enhancing Logical Capabilities in Unified Anomaly Detection Models with a Text Memory Bank






























Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?
AWorld: Orchestrating the Training Recipe for Agentic AI
MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers
rStar2-Agent: Agentic Reasoning Technical Report
Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning
MobileCLIP2: Improving Multi-Modal Reinforced Training
AI-AI Esthetic Collaboration with Explicit Semiotic Awareness and Emergent Grammar Development
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
Predicting the Order of Upcoming Tokens Improves Language Modeling
MIDAS: Multimodal Interactive Digital-human Synthesis via Real-time Autoregressive Video Generation
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
Self-Rewarding Vision-Language Model via Reasoning Decomposition
Beyond Transcription: Mechanistic Interpretability in ASR
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning
WebSight: A Vision-First Architecture for Robust Web Agents
UltraMemV2: Memory Networks Scaling to 120B Parameters with Superior Long-Context Learning
Hermes 4 Technical Report
OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation
VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space
CMPhysBench: A Benchmark for Evaluating Large Language Models in Condensed Matter Physics
TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling
Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
Understanding Tool-Integrated Reasoning
Spacer: Towards Engineered Scientific Inspiration
Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling
VibeVoice Technical Report
MMTok: Multimodal Coverage Maximization for Efficient Inference of VLMs
MV-RAG: Retrieval Augmented Multiview Diffusion
Connecting metal-organic framework synthesis to applications using multimodal machine learning
Model Context Protocols in Adaptive Transport Systems: A Survey
Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?
AWorld: Orchestrating the Training Recipe for Agentic AI
MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers
rStar2-Agent: Agentic Reasoning Technical Report
Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning
MobileCLIP2: Improving Multi-Modal Reinforced Training
AI-AI Esthetic Collaboration with Explicit Semiotic Awareness and Emergent Grammar Development
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
Predicting the Order of Upcoming Tokens Improves Language Modeling
MIDAS: Multimodal Interactive Digital-human Synthesis via Real-time Autoregressive Video Generation
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
Self-Rewarding Vision-Language Model via Reasoning Decomposition
Beyond Transcription: Mechanistic Interpretability in ASR
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning
WebSight: A Vision-First Architecture for Robust Web Agents
UltraMemV2: Memory Networks Scaling to 120B Parameters with Superior Long-Context Learning
Hermes 4 Technical Report
OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation
VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space
CMPhysBench: A Benchmark for Evaluating Large Language Models in Condensed Matter Physics
TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling
Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
Understanding Tool-Integrated Reasoning
Spacer: Towards Engineered Scientific Inspiration
Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling
VibeVoice Technical Report
MMTok: Multimodal Coverage Maximization for Efficient Inference of VLMs
MV-RAG: Retrieval Augmented Multiview Diffusion
Connecting metal-organic framework synthesis to applications using multimodal machine learning
Model Context Protocols in Adaptive Transport Systems: A Survey