-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 86 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 236 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 162 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
Collections
Discover the best community collections!
Collections including paper arxiv:2608.05139
-
Writing in the Margins: Better Inference Pattern for Long Context Retrieval
Paper • 2408.14906 • Published • 144 -
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
Paper • 2410.10819 • Published • 7 -
LLMtimesMapReduce: Simplified Long-Sequence Processing using Large Language Models
Paper • 2410.09342 • Published • 37 -
PDFTriage: Question Answering over Long, Structured Documents
Paper • 2309.08872 • Published • 55
-
ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning
Paper • 2604.24300 • Published • 67 -
Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis
Paper • 2604.24198 • Published • 23 -
KernelBench-X: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels
Paper • 2605.04956 • Published • 7 -
Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
Paper • 2605.13062 • Published • 33
-
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources
Paper • 2606.29538 • Published • 144 -
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation
Paper • 2608.02287 • Published • 31 -
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
Paper • 2608.05139 • Published • 27 -
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure
Paper • 2608.11079 • Published • 16
-
Agent Learning via Early Experience
Paper • 2510.08558 • Published • 256 -
Learning on the Job: An Experience-Driven Self-Evolving Agent for Long-Horizon Tasks
Paper • 2510.08002 • Published • 24 -
Self-Improving LLM Agents at Test-Time
Paper • 2510.07841 • Published • 10 -
The Denario project: Deep knowledge AI agents for scientific discovery
Paper • 2510.26887 • Published • 8
-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 86 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 236 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 162 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
-
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources
Paper • 2606.29538 • Published • 144 -
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation
Paper • 2608.02287 • Published • 31 -
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
Paper • 2608.05139 • Published • 27 -
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure
Paper • 2608.11079 • Published • 16
-
Writing in the Margins: Better Inference Pattern for Long Context Retrieval
Paper • 2408.14906 • Published • 144 -
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads
Paper • 2410.10819 • Published • 7 -
LLMtimesMapReduce: Simplified Long-Sequence Processing using Large Language Models
Paper • 2410.09342 • Published • 37 -
PDFTriage: Question Answering over Long, Structured Documents
Paper • 2309.08872 • Published • 55
-
ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning
Paper • 2604.24300 • Published • 67 -
Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis
Paper • 2604.24198 • Published • 23 -
KernelBench-X: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels
Paper • 2605.04956 • Published • 7 -
Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
Paper • 2605.13062 • Published • 33
-
Agent Learning via Early Experience
Paper • 2510.08558 • Published • 256 -
Learning on the Job: An Experience-Driven Self-Evolving Agent for Long-Horizon Tasks
Paper • 2510.08002 • Published • 24 -
Self-Improving LLM Agents at Test-Time
Paper • 2510.07841 • Published • 10 -
The Denario project: Deep knowledge AI agents for scientific discovery
Paper • 2510.26887 • Published • 8