Command Palette
Search for a command to run...
Papers
Daily updated cutting-edge AI research papers to help you keep up with the latest AI trends
papers

ScienceIDE: Turning World’s Scientific Codebase into Agent Learnable Environments

ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks






























Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening
Agora: Git as Shared Memory for Collective AutoResearch
Confidence Comes from Experience: Experiential Confidence Estimation from Reasoning to Agents
LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence
ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents
AI for Games in the Foundation Model Era
StepAudio 3 Music Technical Report
StepAudio 3 Realtime Technical Report
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement
Continual Learning Mechanisms Compose for Long-Horizon Memorization
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression
Atria Dawn: The Dawn of Agentic Superintelligence On the Evolving Roles of Human–AI Collaboration
PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models
Dream-RSI: Recursive Self-Improvement through Evolving Worlds
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation
AI in Science: Early Insights
Exact Feasibility Certification and Optimal Responsibility Allocation for Multi-Robot CBF Safety Filters
CROSS-BLOCK CONDITIONING IN DEEP BOLTZMANN MACHINES FOR STATISTICAL DATA FUSION
The Complexity of Weak Partition Connectivity in Hedgegraphs
Accelerated Local Algorithms for Personalized and Regularized PageRank
Geometric Signatures of Conceptual Reorganization: A Counterfactual Embedding Framework for Detecting Scientific Revolutions
RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments
Breaking the Token Ceiling: Distilling Smaller, Stronger Byte Models
COBRA-SKILLS: CONTEXTUAL BANDIT-GUIDED EVOLUTION FOR AGENT SKILL OPTIMIZATION
SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation
DataFlex-RL: An Evaluation Platform for RLVR Data Policies
Breaking the Vision–Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models