Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends
papers

LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks

Apple-π: A Benchmark for Evaluating Video Generation Models on Physical Law Grounding






























papers

LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks

Apple-π: A Benchmark for Evaluating Video Generation Models on Physical Law Grounding






























HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune
DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment
EvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary World
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs
Understanding Reasoning from Pretraining to Post-Training
Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops
Loop the Loopies!
On-Policy Delta Distillation
Cura 1T: A Healthcare-Specialized LLM Trained Through a Human-Gated Self-Evolution Loop
From Human-Centric to Agentic Code Review: The Impact of Different Generations of Generative AI Technology on Review Quality
RecGPT-V3 Technical Report
xHC: Expanded Hyper-Connections
Towards Predictive, Aligned, and Scalable Robot Learning
Flow Matching in Feature Space for Stochastic World Modeling
Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit
TRACE: TURN-LEVEL REWARD ASSIGNMENT VIA CREDIT ESTIMATION FOR LONG-HORIZON AGENTS
KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation
BadWAM: When World-Action Models Dream Right but Act Wrong
SearchOS-V1 : Towards Robust Open-Domain Information-Seeking Agent Collaboration
SEED: SELF-EVOLVING ON-POLICY DISTILLATION FOR AGENTIC REINFORCEMENT LEARNING
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget
Deep Learning in Remote Sensing: A Review
A Regression Approach to Speech Enhancement Based on Deep Neural Networks
Deep Neural Networks for Acoustic Modeling in Speech Recognition
RoboTTT: Context Scaling for Robot Policies
SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
Efficient Estimation of Word Representations in Vector Space
Depth Map Prediction from a Single Image using a Multi-Scale Deep Network
HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune
DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment
EvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary World
TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs
Understanding Reasoning from Pretraining to Post-Training
Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops
Loop the Loopies!
On-Policy Delta Distillation
Cura 1T: A Healthcare-Specialized LLM Trained Through a Human-Gated Self-Evolution Loop
From Human-Centric to Agentic Code Review: The Impact of Different Generations of Generative AI Technology on Review Quality
RecGPT-V3 Technical Report
xHC: Expanded Hyper-Connections
Towards Predictive, Aligned, and Scalable Robot Learning
Flow Matching in Feature Space for Stochastic World Modeling
Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit
TRACE: TURN-LEVEL REWARD ASSIGNMENT VIA CREDIT ESTIMATION FOR LONG-HORIZON AGENTS
KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation
BadWAM: When World-Action Models Dream Right but Act Wrong
SearchOS-V1 : Towards Robust Open-Domain Information-Seeking Agent Collaboration
SEED: SELF-EVOLVING ON-POLICY DISTILLATION FOR AGENTIC REINFORCEMENT LEARNING
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget
Deep Learning in Remote Sensing: A Review
A Regression Approach to Speech Enhancement Based on Deep Neural Networks
Deep Neural Networks for Acoustic Modeling in Speech Recognition
RoboTTT: Context Scaling for Robot Policies
SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
Efficient Estimation of Word Representations in Vector Space
Depth Map Prediction from a Single Image using a Multi-Scale Deep Network