| 2024/09 | Expediting and Elevating Large Language Model Reasoning via Hidden Chain-of-Thought Decoding |  | - |
| 2024/09 | Uncovering Latent Chain of Thought Vectors in Language Models |  | - |
| 2024/10 | Understanding Reasoning in Chain-of-Thought from the Hopfieldian View |  | - |
| 2024/10 | Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation |  | Github |
| 2024/11 | Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding |  | Github |
| 2024/12 | Training Large Language Models to Reason in a Continuous Latent Space |  | Github |
| 2024/12 | Compressed Chain of Thought: Efficient Reasoning Through Dense Representations |  | - |
| 2024/12 | Deliberation in Latent Space via Differentiable Cache Augmentation |  | - |
| 2025/01 | Latent-space adversarial training with post-aware calibration for defending large language models against jailbreak attacks |  | Github |
| 2025/01 | LF-Steering: Latent Feature Activation Steering for Enhancing Semantic Consistency in Large Language Models |  | - |
| 2025/02 | Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning |  | - |
| 2025/02 | Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization |  | - |
| 2025/02 | Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach |  | Github |
| 2025/02 | LLM Pretraining with Continuous Concepts |  | Github |
| 2025/02 | SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs |  | Github |
| 2025/02 | Human Preferences in Large Language Model Latent Space: A Technical Analysis on the Reliability of Synthetic Data in Voting Outcome Prediction |  | - |
| 2025/02 | Reasoning with Latent Thoughts: On the Power of Looped Transformers |  | - |
| 2025/02 | Beyond Words: A Latent Memory Approach to Internal Reasoning in LLMs |  | - |
| 2025/02 | CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation |  | Github |
| 2025/03 | Reasoning to Learn from Latent Thoughts |  | Github |
| 2025/03 | Think Before Recommend: Unleashing the Latent Reasoning Power for Sequential Recommendation |  | - |
| 2025/03 | MoLAE: Mixture of Latent Experts for Parameter-Efficient Language Models |  | - |
| 2025/04 | Beyond Chains of Thought: Benchmarking Latent-Space Reasoning Abilities in Large Language Models | - | - |
| 2025/04 | Efficient Pretraining Length Scaling |  | - |
| 2025/05 | SoftCoT++: Test-Time Scaling with Soft Chain-of-Thought Reasoning |  | Github |
| 2025/05 | Reasoning by Superposition: A Theoretical Perspective on Chain of Continuous Thought |  | Github |
| 2025/05 | Enhancing Latent Computation in Transformers with Latent Tokens |  | - |
| 2025/05 | Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space |  | Github |
| 2025/05 | Internal Chain-of-Thought: Empirical Evidence for Layer-wise Subtask Scheduling in LLMs |  | Github |
| 2025/05 | Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space |  | - |
| 2025/05 | Think Silently, Think Fast: Dynamic Latent Compression of LLM Reasoning Chains |  | Github |
| 2025/05 | LARES: Latent Reasoning for Sequential Recommendation |  | - |
| 2025/05 | Hybrid Latent Reasoning via Reinforcement Learning |  | Github |
| 2025/05 | System-1.5 Reasoning: Traversal in Language and Latent Spaces with Dynamic Shortcuts |  | - |
| 2025/05 | Reinforced Latent Reasoning for LLM-based Recommendation |  | Github |
| 2025/05 | Continuous Chain of Thought Enables Parallel Exploration and Reasoning |  | Github |
| 2025/05 | Soft Reasoning: Navigating Solution Spaces in Large Language Models through Controlled Embedding Exploration |  | Github |
| 2025/06 | Efficient Post-Training Refinement of Latent Reasoning in Large Language Models |  | Github |
| 2025/06 | DART: Distilling Autoregressive Reasoning to Silent Thought |  | - |
| 2025/06 | Parallel Continuous Chain-of-Thought with Jacobi Iteration |  | Github |
| 2025/07 | Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer |  | Github |
| 2025/07 | CTRLS: Chain-of-Thought Reasoning via Latent State Transition |  | - |
| 2025/07 | Geometry of Knowledge Allows Extending Diversity Boundaries of Large Language Models |  | - |
| 2025/08 | Bridging Search and Recommendation through Latent Cross Reasoning |  | - |
| 2025/08 | LatentPrompt: Optimizing Promts in Latent Space |  | - |
| 2025/08 | Latent Fusion Jailbreak: Blending Harmful and Harmless Representations to Elicit Unsafe LLM Outputs |  | - |
| 2025/09 | Decoding in Latent Spaces for Efficient Inference in LLM-based Recommendation |  | - |
| 2025/09 | LTA-thinker: Latent Thought-Augmented Training Framework for Large Language Models on Complex Reasoning |  | Github |
| 2025/09 | The Transfer Neurons Hypothesis: An Underlying Mechanism for Language Latent Space Transitions in Multilingual LLMs |  | - |
| 2025/09 | LatentGuard: Controllable Latent Steering for Robust Refusal of Attacks and Reliable Response Generation |  | - |
| 2025/09 | SIM-CoT: Supervised Implicit Chain-of-Thought |  | Github |
| 2025/09 | PonderLM-2: Pretraining LLM with Latent Thoughts in Continuous Space |  | Github |
| 2025/09 | Fast Thinking for Large Language Models |  | - |
| 2025/09 | Learning to Ponder: Adaptive Reasoning in Latent Space |  | - |
| 2025/09 | Identity Bridge: Enabling Implicit Reasoning via Shared Latent Memory |  | - |
| 2025/09 | MemGen: Weaving Generative Latent Memory for Self-Evolving Agents |  | Github |
| 2025/09 | LatentEvolve: Self-Evolving Test-Time Scaling in Latent Space |  | Github |
| 2025/09 | MARCOS: Deep Thinking by Markov Chain of Continuous Thoughts |  | - |
| 2025/09 | A Formal Comparison Between Chain of Thought and Latent Thought | - | - |
| 2025/09 | Latent Thinking Optimization: Your Latent Reasoning Language Model Secretly Encodes Reward Signals in Its Latent Thoughts |  | - |
| 2025/10 | Thoughtbubbles: an Unsupervised Method for Parallel Thinking in Latent Space |  | Github |
| 2025/10 | Analyzing Latent Concepts in Code Language Models |  | - |
| 2025/10 | Exploring System 1 and 2 communication for latent reasoning in LLMs |  | - |
| 2025/10 | KaVa: Latent Reasoning via Compressed KV-Cache Distillation |  | - |
| 2025/10 | Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimization |  | Github |
| 2025/10 | LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning |  | Github |
| 2025/10 | SwiReasoning: Switch-Thinking in Latent and Explicit for Pareto-Superior Reasoning LLMs |  | Github |
| 2025/10 | Encode, Think, Decode: Scaling test-time reasoning with recursive latent thoughts |  | - |
| 2025/10 | Parallel Test-Time Scaling for Latent Reasoning Models |  | Github |
| 2025/10 | LatentBreak: Jailbreaking Large Language Models through Latent Space Feedback |  | - |
| 2025/10 | Kelp: A Streaming Safeguard for Large Models via Latent Dynamics-Guided Risk Detection |  | Github |
| 2025/10 | Tracing the Traces: Latent Temporal Signals for Efficient and Accurate Reasoning |  | - |
| 2025/10 | Unlocking Out-of-Distribution Generalization in Transformers via Recursive Latent Space Reasoning |  | Github |
| 2025/10 | Language Models are Injective and Hence Invertible |  | - |
| 2025/10 | LLM Latent Reasoning as Chain of Superposition |  | Github |
| 2025/10 | ActivationReasoning: Logical Reasoning in Latent Activation Spaces |  | - |
| 2025/10 | Emotions Where Art Thou: Understanding and Characterizing the Emotional Latent Space of Large Language Models |  | - |
| 2025/10 | SALS: Sparse Attention in Latent Space for KV cache Compression |  | - |
| 2025/10 | SemCoT: Accelerating Chain-of-Thought Reasoning through Semantically-Aligned Implicit Tokens |  | Github |
| 2025/10 | Scaling Latent Reasoning via Looped Language Models |  | - |
| 2025/10 | Cache-to-Cache: Direct Semantic Communication Between Large Language Model |  | Github |
| 2025/10 | Thought Communication in Multiagent Collaboration |  | - |
| 2025/11 | SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization |  | Github |
| 2025/11 | Think Consistently, Reason Efficiently: Energy-Based Calibration for Implicit Chain-of-Thought |  | - |
| 2025/11 | Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models |  | Github |
| 2025/11 | SpiralThinker: Latent Reasoning through an Iterative Process with Text-Latent Interleaving |  | - |
| 2025/11 | Enabling Agents to Communicate Entirely in Latent Space |  | - |
| 2025/11 | Improving Latent Reasoning in LLMs via Soft Concept Mixing |  | - |
| 2025/11 | Your Latent Reasoning is Secretly Policy Improvement Operator | - | - |
| 2025/11 | CLaRa: Bridging Retrieval and Generation with Continuous Latent Reasoning |  | Github |
| 2025/11 | Learning When to Stop: Adaptive Latent Reasoning via Reinforcement Learning |  | Github |
| 2025/11 | Visualizing LLM Latent Space Geometry Through Dimensionality Reduction |  | Github |
| 2025/11 | Polarity-Aware Probing for Quantifying Latent Alignment in Language Models |  | Github |
| 2025/11 | Latent Collaboration in Multi-Agent Systems |  | Github |
| 2025/11 | Next-Latent Prediction Transformers Learn Compact World Models |  | Github |
| 2025/12 | Latent Debate: A Surrogate Framework for Interpreting LLM Thinking |  | Github |
| 2025/12 | Lightweight Latent Reasoning for Narrative Tasks |  | - |
| 2025/12 | Think-While-Generating: On-the-Fly Reasoning for Personalized Long-Form Generation |  | - |
| 2025/12 | ReLaX: Reasoning with Latent Exploration for Large Reasoning Models |  | - |
| 2025/12 | Reinforcement Learning for Latent-Space Thinking in LLMs |  | Github |
| 2025/12 | Reasoning Palette: Modulating Reasoning via Latent Contextualization for Controllable Exploration for (V)LMs |  | - |
| 2025/12 | JEPA-Reasoner: Decoupling Latent Reasoning from Token Generation |  | - |
| 2025/12 | Do Latent Tokens Think? A Causal and Adversarial Analysis of Chain-of-Continuous-Thought |  | - |
| 2025/12 | iCLP: Large Language Model Reasoning with Implicit Cognition Latent Planning |  | Github |
| 2025/12 | Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space |  | - |
| 2025/12 | Learning Evolving Latent Strategies for Multi-Agent Language Systems without Model Fine-Tuning | - | - |
| 2026/01 | Parallel Latent Reasoning for Sequential Recommendation |  | - |
| 2026/01 | Latent Space Communication via K-V Cache Alignment |  | - |
| 2026/01 | Layer-Order Inversion: Rethinking Latent Multi-Hop Reasoning in Large Language Models |  | Github |
| 2026/01 | FlashMem: Distilling Intrinsic Latent Memory via Computation Reuse |  | - |
| 2026/01 | IIB-LPO: Latent Policy Optimization via Iterative Information Bottleneck |  | Github |
| 2026/01 | Breaking Model Lock-in: Cost-Efficient Zero-Shot LLM Routing via a Universal Latent Space |  | Github |
| 2026/01 | Silence the Judge: Reinforcement Learning with Self-Verifier via Latent Geometric Clustering |  | - |
| 2026/01 | Reasoning Beyond Chain-of-Thought: A Latent Computational Mode in Large Language Models |  | - |
| 2026/01 | RISER: Orchestrating Latent Reasoning Skills for Adaptive Activation Steering |  | Github |
| 2026/01 | GeoSteer: Faithful Chain-of-Thought Steering via Latent Manifold Gradients |  | - |
| 2026/01 | Reasoning While Recommending: Entropy-Guided Latent Reasoning in Generative Re-ranking Models |  | - |
| 2026/01 | Latent-Space Contrastive Reinforcement Learning for Stable and Efficient LLM Reasoning |  | - |
| 2026/01 | UniCog: Uncovering Cognitive Abilities of LLMs through Latent Mind Space Analysis |  | Github |
| 2026/01 | S2GR: Stepwise Semantic-Guided Reasoning in Latent Space for Generative Recommendation |  | - |
| 2026/01 | The Geometric Reasoner: Manifold-Informed Latent Foresight Search for Long-Context Reasoning |  | - |
| 2026/01 | PILOT: Planning via Internalized Latent Optimization Trajectories for Large Language Models |  | - |
| 2026/01 | Beyond Imitation: Reinforcement Learning for Active Latent Planning |  | Github |
| 2026/01 | Latent Adversarial Regularization for Offline Preference Optimization |  | - |
| 2026/01 | Latent Chain-of-Thought as Planning: Decoupling Reasoning from Verbalization |  | Github |
| 2026/01 | Depth-Recurrent Attention Mixtures: Giving Latent Reasoning the Attention it Deserves |  | - |
| 2026/01 | From Logits to Latents: Contrastive Representation Shaping for LLM Unlearning | - | - |
| 2026/01 | ReGuLaR: Variational Latent Reasoning Guided by Rendered Chain-of-Thought |  | Github |
| 2026/02 | G-MemLLM: Gated Latent Memory Augmentation for Long-Context Reasoning in Large Language Models |  | - |
| 2026/02 | Do Latent-CoT Models Think Step-by-Step? A Mechanistic Study on Sequential Reasoning Tasks |  | Github |
| 2026/02 | Capabilities and Fundamental Limits of Latent Chain-of-Thought | - | - |
| 2026/02 | Restoring Exploration after Post-Training: Latent Exploration Decoding for Large Reasoning Models |  | Github |
| 2026/02 | No Global Plan in Chain-of-Thought: Uncover the Latent Planning Horizon of LLMs | - | Github |
| 2026/02 | CoLT: Reasoning with Chain of Latent Tool Calls |  | - |
| 2026/02 | Internalizing LLM Reasoning via Discovery and Replay of Latent Actions |  | Github |
| 2026/02 | Inference-Time Rethinking with Latent Thought Vectors for Math Reasoning |  | - |
| 2026/02 | LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning |  | Github |
| 2026/02 | DeltaKV: Residual-Based KV Cache Compression via Long-Range Similarity |  | Github |
| 2026/02 | Pretraining with Token-Level Adaptive Latent Chain-of-Thought |  | - |
| 2026/02 | Latent Reasoning with Supervised Thinking States |  | - |
| 2026/02 | Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure |  | - |
| 2026/02 | Next Concept Prediction in Discrete Latent Space Leads to Stronger Language Models |  | Github |
| 2026/02 | Talking with the Latents -- how to convert your LLM into an astronomer |  | - |
| 2026/02 | Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens |  | Github |
| 2026/02 | Prioritize the Process, Not Just the Outcome: Rewarding Latent Thought Trajectories Improves Reasoning in Looped Language Models |  | - |
| 2026/02 | LoopFormer: Elastic-Depth Looped Transformers for Latent Reasoning via Shortcut Modulation |  | Github |
| 2026/02 | Jailbreaking Leaves a Trace: Understanding and Detecting Jailbreak Attacks from Internal Representations of Large Language Models |  | - |
| 2026/02 | Native Reasoning Models: Training Language Models to Reason on Unverifiable Data |  | - |
| 2026/02 | ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces |  | - |
| 2026/02 | SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion |  | - |
| 2026/02 | GTS: Inference-Time Scaling of Latent Reasoning with a Learnable Gaussian Thought Sampler |  | - |
| 2026/02 | Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation |  | - |
| 2026/02 | Inner Loop Inference for Pretrained Transformers: Unlocking Latent Capabilities Without Training | - | Github |
| 2026/02 | LatentMem: Customizing Latent Memory for Multi-Agent Systems |  | Github |
| 2026/02 | Agent Primitives: Reusable Latent Building Blocks for Multi-Agent Systems |  | - |
| 2026/03 | LaSER: Internalizing Explicit Reasoning into Latent Space for Dense Retrieval |  | Github |
| 2026/03 | Multi-Head Low-Rank Attention | - | Github |
| 2026/03 | AdaPonderLM: Gated Pondering Language Models with Token-Wise Adaptive Depth |  | - |
| 2026/03 | PonderLM-3: Adaptive Token-Wise Pondering with Differentiable Masking |  | - |
| 2026/03 | When Shallow Wins: Silent Failures and the Depth-Accuracy Paradox in Latent Reasoning | - | Github |
| 2026/03 | β-REASONER: LLM REASONING VIA TEST-TIMEGRADIENT DESCENT IN LATENT SPACE | - | - |
| 2026/03 | SPOT: Span-level Pause-of-Thought for Efficient and Interpretable Latent Reasoning in Large Language Models |  | - |
| 2026/03 | NextMem: Towards Latent Factual Memory for LLM-based Agents |  | Github |
| 2026/03 | Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations |  | - |
| 2026/03 | LoopRPT: Reinforcement Pre-Training for Looped Language Models |  | - |