Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends
papers

ManiScope: LLM-Assisted Visual Analytics of Cryptocurrency Manipulation Risk

Event-based Neural Decoding for Neuroprosthetic Motor Control






























papers

ManiScope: LLM-Assisted Visual Analytics of Cryptocurrency Manipulation Risk

Event-based Neural Decoding for Neuroprosthetic Motor Control






























Unlocking Every Expert in Domain-Specific Training
EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments
ARDY: Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation
PithTrain: A Compact and Agent-Native MoE Training System
Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models
KronQ: LLM Quantization via Kronecker-Factored Hessian
Trust Region Policy Distillation
Video Generation Models are General-Purpose Vision Learners
Scalable Visual Pretraining for Language Intelligence
Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL
Atomic Task Graph: A Unified Framework for Agentic Planning and Execution
LongE2V: Long-Horizon Event-based Video Reconstruction, Prediction, and Frame Interpolation with Video Diffusion Models
UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks
Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation
Why Can’t I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition
Video-Oasis: Rethinking Evaluation of Video Understanding
Vidu S1: A Real-Time Interactive Video Generation Model
Measuring the Gap Between Human and LLM Research Ideas
The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agentic AI
Infinite Worlds with Versatile Interactions
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
LAME M-VLA: DUAL LATENT MEMORY IN VISION-LANGUAGE-ACTION MODELS FOR ROBOTIC MANIPULATION
Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning
Parallelized Autoregressive Decoding for Omni-Modal Dense Video Captioning
Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory
Vision as Unified Multimodal Generation
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling
Unlocking Every Expert in Domain-Specific Training
EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments
ARDY: Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation
PithTrain: A Compact and Agent-Native MoE Training System
Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models
KronQ: LLM Quantization via Kronecker-Factored Hessian
Trust Region Policy Distillation
Video Generation Models are General-Purpose Vision Learners
Scalable Visual Pretraining for Language Intelligence
Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL
Atomic Task Graph: A Unified Framework for Agentic Planning and Execution
LongE2V: Long-Horizon Event-based Video Reconstruction, Prediction, and Frame Interpolation with Video Diffusion Models
UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks
Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation
Why Can’t I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition
Video-Oasis: Rethinking Evaluation of Video Understanding
Vidu S1: A Real-Time Interactive Video Generation Model
Measuring the Gap Between Human and LLM Research Ideas
The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agentic AI
Infinite Worlds with Versatile Interactions
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
LAME M-VLA: DUAL LATENT MEMORY IN VISION-LANGUAGE-ACTION MODELS FOR ROBOTIC MANIPULATION
Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning
Parallelized Autoregressive Decoding for Omni-Modal Dense Video Captioning
Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory
Vision as Unified Multimodal Generation
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling