-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 86 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 236 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 162 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
Collections
Discover the best community collections!
Collections including paper arxiv:2607.05147
-
Dockerless: Environment-Free Program Verifier for Coding Agents
Paper • 2606.28436 • Published • 85 -
Program-as-Weights: A Programming Paradigm for Fuzzy Functions
Paper • 2607.02512 • Published • 307 -
EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments
Paper • 2607.02440 • Published • 48 -
Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks
Paper • 2606.29082 • Published • 42
-
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
Paper • 2512.02556 • Published • 273 -
DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
Paper • 2606.19348 • Published • 46 -
DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation
Paper • 2607.05147 • Published • 50 -
DualPath: Breaking the Storage Bandwidth Bottleneck in Agentic LLM Inference
Paper • 2602.21548 • Published • 57
-
VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training
Paper • 2602.10693 • Published • 39 -
Flash-KMeans: Fast and Memory-Efficient Exact K-Means
Paper • 2603.09229 • Published • 85 -
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
Paper • 2603.11076 • Published • 5 -
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning
Paper • 2603.21065 • Published • 80
-
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
Paper • 2402.04291 • Published • 51 -
OneBit: Towards Extremely Low-bit Large Language Models
Paper • 2402.11295 • Published • 24 -
A Survey on Transformer Compression
Paper • 2402.05964 • Published • 1 -
Towards Next-Level Post-Training Quantization of Hyper-Scale Transformers
Paper • 2402.08958 • Published • 4
-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 86 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 236 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 162 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
-
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
Paper • 2512.02556 • Published • 273 -
DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
Paper • 2606.19348 • Published • 46 -
DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation
Paper • 2607.05147 • Published • 50 -
DualPath: Breaking the Storage Bandwidth Bottleneck in Agentic LLM Inference
Paper • 2602.21548 • Published • 57
-
Dockerless: Environment-Free Program Verifier for Coding Agents
Paper • 2606.28436 • Published • 85 -
Program-as-Weights: A Programming Paradigm for Fuzzy Functions
Paper • 2607.02512 • Published • 307 -
EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments
Paper • 2607.02440 • Published • 48 -
Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks
Paper • 2606.29082 • Published • 42
-
VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training
Paper • 2602.10693 • Published • 39 -
Flash-KMeans: Fast and Memory-Efficient Exact K-Means
Paper • 2603.09229 • Published • 85 -
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
Paper • 2603.11076 • Published • 5 -
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning
Paper • 2603.21065 • Published • 80
-
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
Paper • 2402.04291 • Published • 51 -
OneBit: Towards Extremely Low-bit Large Language Models
Paper • 2402.11295 • Published • 24 -
A Survey on Transformer Compression
Paper • 2402.05964 • Published • 1 -
Towards Next-Level Post-Training Quantization of Hyper-Scale Transformers
Paper • 2402.08958 • Published • 4