论文就是你所需要的。
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| MiniMax Sparse Attention | 2026 | MiniMax | ./papers/00179-MSA.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model | 2025 | Hugging Face | ./papers/00159-SmolLM2.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Training-Free Group Relative Policy Optimization | 2025 | Tencent | ./papers/00108-Training-Free-GRPO.pdf |
| HY-MT1.5 Technical Report | 2025 | Tencent | ./papers/00152-HY-MT1.5.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding | 2026 | Xiaomi | ./papers/00162-ThinkOmni.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| LongCat-Flash Technical Report | 2025 | Meituan LongCat Team | ./papers/00093-LongCat-Flash.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B | 2025 | Sina Weibo | ./papers/00137-VibeThinker-1.5B.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| FlowSearch: Advancing deep research with dynamic structured knowledge flow | 2025 | Shanghai Artificial Intelligence Laboratory | ./papers/00142-FlowSearch.pdf |
| Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement | 2026 | Shanghai Artificial Intelligence Laboratory | ./papers/00188-Harness-of-Harness.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| YaRN: Efficient Context Window Extension of Large Language Models | 2023 | EleutherAI | ./papers/00034-YaRN.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Memento: Fine-tuning LLM Agents without Fine-tuning LLMs | 2025 | Huawei Noah’s Ark Lab | ./papers/00089-Memento.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies | 2024 | Tsinghua University | ./papers/00003-MiniCPM.pdf |
| ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools | 2024 | Tsinghua University | ./papers/00005-ChatGLM.pdf |
| LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs | 2024 | Tsinghua University | ./papers/00038-LongWriter.pdf |
| ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search | 2024 | Tsinghua University | ./papers/00043-ReST-MCTS.pdf |
| GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models | 2025 | Tsinghua University | ./papers/00113-GLM-4.5.pdf |
| GLM-5: from Vibe Coding to Agentic Engineering | 2026 | Tsinghua University | ./papers/00157-GLM-5.pdf |
| AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning | 2025 | Tsinghua University | ./papers/00160-AReaL.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning | 2025 | Carnegie Mellon University | ./papers/00053-MRT.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Efficient Memory Management for Large Language Model Serving with PagedAttention | 2023 | University of California, Berkeley | ./papers/00086-vllm.pdf |
| SGLang: Efficient Execution of Structured Language Model Programs | 2024 | University of California, Berkeley | ./papers/00164-SGLang.pdf |
| Ring Attention with Blockwise Transformers for Near-Infinite Context | 2023 | University of California, Berkeley | ./papers/00087-RingAttention.pdf |
| Blockwise Parallel Transformer for Large Context Models | 2023 | University of California, Berkeley | ./papers/00088-BPT.pdf |
| RL Grokking Recipe: How Does RL Unlock and Transfer New Algorithms in LLMs? | 2025 | University of California, Berkeley | ./papers/00115-grokking.pdf |
| SimpleMem: Efficient Lifelong Memory for LLM Agents | 2026 | University of California, Berkeley | ./papers/00145-SimpleMem.pdf |
| From Model Scaling to System Scaling: Scaling the Harness in Agentic AI | 2026 | University of California, Berkeley | ./papers/00182-CheetahClaws.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Striped Attention: Faster Ring Attention for Causal Transformers | 2023 | MIT | ./papers/00096-Striped-Attention.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| In-the-Flow Agentic System Optimization for Effective Planning and Tool Use | 2025 | Stanford University | ./papers/00125-in-the-flow.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Skill-Targeted Adaptive Training | 2025 | Princeton University | ./papers/00119-STAT.pdf |
| OpenClaw-RL: Train Any Agent Simply by Talking | 2026 | Princeton University | ./papers/00171-OpenClaw-RL.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Understanding R1-Zero-Like Training: A Critical Perspective | 2025 | National University of Singapore | ./papers/00048-understand-r1-zero.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| EvoSyn: Generalizable Evolutionary Data Synthesis for Verifiable Learning | 2025 | Fudan University | ./papers/00120-EvoSyn.pdf |
| AgentPRM: Process Reward Models for LLM Agents via Step-Wise Promise and Progress | 2025 | Fudan University | ./papers/00135-AgentPRM.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| LIMO: Less is More for Reasoning | 2025 | Shanghai Jiao Tong University | ./papers/00051-LIMO.pdf |
| Context Engineering 2.0: The Context of Context Engineering | 2025 | Shanghai Jiao Tong University | ./papers/00131-Context-Engineering-2.0.pdf |
| Repo0: Design-Driven Zero-to-All Code Generation | 2026 | Shanghai Jiao Tong University | ./papers/00184-Repo0.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| A Comprehensive Survey on Long Context Language Modeling | 2025 | Nanjing University | ./papers/00058-LCLM.pdf |
| DART: Diffusion-Inspired Speculative Decoding for Fast LLM Inference | 2026 | Nanjing University | ./papers/00175-DART.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning | 2025 | University of Illinois Urbana-Champaign | ./papers/00060-DeepRetrieval.pdf |
| Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning | 2025 | University of Illinois Urbana-Champaign | ./papers/00061-Search-R1.pdf |
| Executable Code Actions Elicit Better LLM Agents | 2024 | University of Illinois Urbana-Champaign | ./papers/00166-CodeAct.pdf |
| CoRNStack: High-Quality Contrastive Data for Better Code Retrieval and Reranking | 2024 | University of Illinois Urbana-Champaign | ./papers/00170-CoRNStack.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Dr. RTL: Autonomous Agentic RTL Optimization through Tool-Grounded Self-Improvement | 2026 | Hong Kong University of Science and Technology | ./papers/00169-Dr.RTL.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| AssertMiner: Module-Level Spec Generation and Assertion Mining using Static Analysis Guided LLMs | 2025 | Institute of Computing Technology, Chinese Academy of Sciences | ./papers/00173-AssertMiner.pdf |
| AssertMiner-pro: Enhanced module-level spec generation and assertion mining with LLM guided by top-down hierarchical strategies | 2026 | Institute of Computing Technology, Chinese Academy of Sciences | ./papers/00177-AssertMiner-pro.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| NoTB: Oracle-Free Triage of LLM-Generated RTL via Cross-Model Formal Consensus | 2026 | Columbia University | ./papers/00183-NoTB.pdf |
151 commits
TeX
88.6%
BibTeX Style
11.3%
论文就是你所需要的。
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| MiniMax Sparse Attention | 2026 | MiniMax | ./papers/00179-MSA.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model | 2025 | Hugging Face | ./papers/00159-SmolLM2.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Training-Free Group Relative Policy Optimization | 2025 | Tencent | ./papers/00108-Training-Free-GRPO.pdf |
| HY-MT1.5 Technical Report | 2025 | Tencent | ./papers/00152-HY-MT1.5.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding | 2026 | Xiaomi | ./papers/00162-ThinkOmni.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| LongCat-Flash Technical Report | 2025 | Meituan LongCat Team | ./papers/00093-LongCat-Flash.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B | 2025 | Sina Weibo | ./papers/00137-VibeThinker-1.5B.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| FlowSearch: Advancing deep research with dynamic structured knowledge flow | 2025 | Shanghai Artificial Intelligence Laboratory | ./papers/00142-FlowSearch.pdf |
| Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement | 2026 | Shanghai Artificial Intelligence Laboratory | ./papers/00188-Harness-of-Harness.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| YaRN: Efficient Context Window Extension of Large Language Models | 2023 | EleutherAI | ./papers/00034-YaRN.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Memento: Fine-tuning LLM Agents without Fine-tuning LLMs | 2025 | Huawei Noah’s Ark Lab | ./papers/00089-Memento.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies | 2024 | Tsinghua University | ./papers/00003-MiniCPM.pdf |
| ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools | 2024 | Tsinghua University | ./papers/00005-ChatGLM.pdf |
| LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs | 2024 | Tsinghua University | ./papers/00038-LongWriter.pdf |
| ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search | 2024 | Tsinghua University | ./papers/00043-ReST-MCTS.pdf |
| GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models | 2025 | Tsinghua University | ./papers/00113-GLM-4.5.pdf |
| GLM-5: from Vibe Coding to Agentic Engineering | 2026 | Tsinghua University | ./papers/00157-GLM-5.pdf |
| AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning | 2025 | Tsinghua University | ./papers/00160-AReaL.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning | 2025 | Carnegie Mellon University | ./papers/00053-MRT.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Efficient Memory Management for Large Language Model Serving with PagedAttention | 2023 | University of California, Berkeley | ./papers/00086-vllm.pdf |
| SGLang: Efficient Execution of Structured Language Model Programs | 2024 | University of California, Berkeley | ./papers/00164-SGLang.pdf |
| Ring Attention with Blockwise Transformers for Near-Infinite Context | 2023 | University of California, Berkeley | ./papers/00087-RingAttention.pdf |
| Blockwise Parallel Transformer for Large Context Models | 2023 | University of California, Berkeley | ./papers/00088-BPT.pdf |
| RL Grokking Recipe: How Does RL Unlock and Transfer New Algorithms in LLMs? | 2025 | University of California, Berkeley | ./papers/00115-grokking.pdf |
| SimpleMem: Efficient Lifelong Memory for LLM Agents | 2026 | University of California, Berkeley | ./papers/00145-SimpleMem.pdf |
| From Model Scaling to System Scaling: Scaling the Harness in Agentic AI | 2026 | University of California, Berkeley | ./papers/00182-CheetahClaws.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Striped Attention: Faster Ring Attention for Causal Transformers | 2023 | MIT | ./papers/00096-Striped-Attention.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| In-the-Flow Agentic System Optimization for Effective Planning and Tool Use | 2025 | Stanford University | ./papers/00125-in-the-flow.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Skill-Targeted Adaptive Training | 2025 | Princeton University | ./papers/00119-STAT.pdf |
| OpenClaw-RL: Train Any Agent Simply by Talking | 2026 | Princeton University | ./papers/00171-OpenClaw-RL.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Understanding R1-Zero-Like Training: A Critical Perspective | 2025 | National University of Singapore | ./papers/00048-understand-r1-zero.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| EvoSyn: Generalizable Evolutionary Data Synthesis for Verifiable Learning | 2025 | Fudan University | ./papers/00120-EvoSyn.pdf |
| AgentPRM: Process Reward Models for LLM Agents via Step-Wise Promise and Progress | 2025 | Fudan University | ./papers/00135-AgentPRM.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| LIMO: Less is More for Reasoning | 2025 | Shanghai Jiao Tong University | ./papers/00051-LIMO.pdf |
| Context Engineering 2.0: The Context of Context Engineering | 2025 | Shanghai Jiao Tong University | ./papers/00131-Context-Engineering-2.0.pdf |
| Repo0: Design-Driven Zero-to-All Code Generation | 2026 | Shanghai Jiao Tong University | ./papers/00184-Repo0.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| A Comprehensive Survey on Long Context Language Modeling | 2025 | Nanjing University | ./papers/00058-LCLM.pdf |
| DART: Diffusion-Inspired Speculative Decoding for Fast LLM Inference | 2026 | Nanjing University | ./papers/00175-DART.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning | 2025 | University of Illinois Urbana-Champaign | ./papers/00060-DeepRetrieval.pdf |
| Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning | 2025 | University of Illinois Urbana-Champaign | ./papers/00061-Search-R1.pdf |
| Executable Code Actions Elicit Better LLM Agents | 2024 | University of Illinois Urbana-Champaign | ./papers/00166-CodeAct.pdf |
| CoRNStack: High-Quality Contrastive Data for Better Code Retrieval and Reranking | 2024 | University of Illinois Urbana-Champaign | ./papers/00170-CoRNStack.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| Dr. RTL: Autonomous Agentic RTL Optimization through Tool-Grounded Self-Improvement | 2026 | Hong Kong University of Science and Technology | ./papers/00169-Dr.RTL.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| AssertMiner: Module-Level Spec Generation and Assertion Mining using Static Analysis Guided LLMs | 2025 | Institute of Computing Technology, Chinese Academy of Sciences | ./papers/00173-AssertMiner.pdf |
| AssertMiner-pro: Enhanced module-level spec generation and assertion mining with LLM guided by top-down hierarchical strategies | 2026 | Institute of Computing Technology, Chinese Academy of Sciences | ./papers/00177-AssertMiner-pro.pdf |
| 论文 | 年份 | 论文单位 | 笔记地址 |
|---|---|---|---|
| NoTB: Oracle-Free Triage of LLM-Generated RTL via Cross-Model Formal Consensus | 2026 | Columbia University | ./papers/00183-NoTB.pdf |
151 commits
TeX
88.6%
BibTeX Style
11.3%