yanfeng98/paper-is-all-you-need

论文就是你所需要的。

TeX

2

151 commits

updated Sep 14, 2026

See the code

README

paper-is-all-you-need

论文就是你所需要的。

论文

OpenAI

DeepSeek-AI

KimiTeam

MiniMax

论文年份论文单位笔记地址
MiniMax Sparse Attention2026MiniMax./papers/00179-MSA.pdf

Microsoft

论文年份论文单位笔记地址
ZeRO: Memory Optimizations Toward Training Trillion Parameter Models2020Microsoft./papers/00002-ZeRO.pdf
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone2024Microsoft./papers/00004-Phi-3.pdf
DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales2023Microsoft./papers/00013-DeepSpeed-Chat.pdf
ZeroQuant(4+2): Redefining LLMs Quantization with a New FP6-Centric Strategy for Diverse Generative Tasks2023Microsoft./papers/00014-ZeroQuant(4+2).pdf
DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale2022Microsoft./papers/00015-DeepSpeed-Inference.pdf
DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models2023Microsoft./papers/00016-DeepSpeed-Ulysses.pdf
ZeRO-Offload: Democratizing Billion-Scale Model Training2021Microsoft./papers/00017-ZeRO-Offload.pdf
1-bit Adam: Communication Efficient Large-Scale Training with Adam's Convergence Speed2021Microsoft./papers/00018-1-bit-Adam.pdf
ZeRO-Infinity: Breaking the GPU Memory Wall for Extreme Scale Deep Learning2021Microsoft./papers/00019-ZeRO-Infinity.pdf
1-bit LAMB: Communication Efficient Large-Scale Large-Batch Training with LAMB's Convergence Speed2021Microsoft./papers/00020-1-bit-LAMB.pdf
The Stability-Efficiency Dilemma: Investigating Sequence Length Warmup for Training GPT Models2022Microsoft./papers/00021-Sequence-Length-Warmup.pdf
Maximizing Communication Efficiency for Large-scale Training via 0/1 Adam2022Microsoft./papers/00023-0-1-Adam.pdf
DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale2022Microsoft./papers/00023-DeepSpeed-MoE.pdf
Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model2022Microsoft./papers/00024-Megatron-Turing.pdf
Extreme Compression for Pre-trained Transformers Made Simple and Efficient2022Microsoft./papers/00025-XTC.pdf
ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers2022Microsoft./papers/00026-ZeroQuant.pdf
DeepSpeed Data Efficiency: Improving Deep Learning Model Quality and Training Efficiency via Efficient Data Sampling and Routing2024Microsoft./papers/00027-DeepSpeed-Data-Efficiency.pdf
Understanding INT4 Quantization for Transformer Models: Latency Speedup, Composability, and Failure Cases2023Microsoft./papers/00029-INT4-Quantization.pdf
ZeroQuant-FP: A Leap Forward in LLMs Post-Training W4A8 Quantization Using Floating-Point Formats2023Microsoft./papers/00030-ZeroQuant-FP.pdf
ZeRO++: Extremely Efficient Collective Communication for Giant Model Training2023Microsoft./papers/00031-ZeRO++.pdf
ZeroQuant-HERO: Hardware-Enhanced Robust Optimized Post-Training Quantization Framework for W8A8 Transformers2023Microsoft./papers/00032-ZeroQuant-HERO.pdf
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking2025Microsoft./papers/00044-rStar-Math.pdf
From Local to Global: A Graph RAG Approach to Query-Focused Summarization2024Microsoft./papers/00046-GraphRAG.pdf
Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning2025Microsoft./papers/00050-Logic-RL.pdf
LongRoPE2: Near-Lossless LLM Context Window Scaling2025Microsoft./papers/00063-LongRoPE2.pdf
BitNet b1.58 2B4T Technical Report2025Microsoft./papers/00064-BitNet.pdf
1-bit AI Infra: Part 1.1, Fast and Lossless BitNet b1.58 Inference on CPUs2024Microsoft./papers/00065-1-bit.pdf
Bitnet.cpp: Efficient Edge Inference for Ternary LLMs2025Microsoft./papers/00066-Bitnet.cpp.pdf
MiniLLM: Knowledge Distillation of Large Language Models2023Microsoft./papers/00075-MiniLLM.pdf
rStar2-Agent: Agentic Reasoning Technical Report2025Microsoft./papers/00092-rStar.pdf
RPG: A Repository Planning Graph for Unified and Scalable Codebase Generation2025Microsoft./papers/00111-RPG.pdf
Agent Lightning: Train ANY AI Agents with Reinforcement Learning2025Microsoft./papers/00124-agent-lightning.pdf
X-Coder: Advancing Competitive Programming with Fully Synthetic Tasks, Solutions, and Tests2026Microsoft./papers/00161-X-Coder.pdf

Google

Meta

DeepMind

NVIDIA

Hugging Face

Alibaba Group

论文年份论文单位笔记地址
AlphaMath Almost Zero: process Supervision without process2024Alibaba Group./papers/00007-AlphaMath.pdf
Qwen2.5-Coder Technical Report2024Alibaba Group./papers/00033-Qwen2.5-Coder.pdf
Qwen3 Technical Report2025Alibaba Group./papers/00069-Qwen3_Technical_Report.pdf
Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models2025Alibaba Group./papers/00073-Qwen3-Embedding.pdf
WebWalker: Benchmarking LLMs in Web Traversal2025Alibaba Group./papers/00097-WebWalker.pdf
WebDancer: Towards Autonomous Information Seeking Agency2025Alibaba Group./papers/00098-WebDancer.pdf
WebSailor: Navigating Super-human Reasoning for Web Agent2025Alibaba Group./papers/00082-WebSailor.pdf
WebShaper: Agentically Data Synthesizing via Information-Seeking Formalization2025Alibaba Group./papers/00099-WebShaper.pdf
WebWatcher: Breaking New Frontier of Vision-Language Deep Research Agent2025Alibaba Group./papers/00100-WebWatcher.pdf
WebResearcher: Unleashing unbounded reasoning capability in Long-Horizon Agents2025Alibaba Group./papers/00101-WebResearcher.pdf
ReSum: Unlocking Long-Horizon Search Intelligence via Context Summarization2025Alibaba Group./papers/00102-ReSum.pdf
WebWeaver: Structuring Web-Scale Evidence with Dynamic Outlines for Open-Ended Deep Research2025Alibaba Group./papers/00103-WebWeaver.pdf
WebSailor-V2: Bridging the Chasm to Proprietary Agents via Synthetic Data and Scalable Reinforcement Learning2025Alibaba Group./papers/00104-WebSailor-V2.pdf
Scaling Agents via Continual Pre-training2025Alibaba Group./papers/00105-AgentFounder.pdf
Towards General Agentic Intelligence via Environment Scaling2025Alibaba Group./papers/00106-AgentScaler.pdf
On-Policy RL Meets Off-Policy Experts: Harmonizing Supervised Fine-Tuning and Reinforcement Learning via Dynamic Weighting2025Alibaba Group./papers/00094-CHORD.pdf
Group Sequence Policy Optimization2025Alibaba Group./papers/00107-GSPO.pdf
Tree Search for LLM Agent Reinforcement Learning2025Alibaba Group./papers/00127-Tree-GRPO.pdf
Search Self-play: Pushing the Frontier of Agent Capability without Supervision2025Alibaba Group./papers/00129-SSP.pdf
Soft Adaptive Policy Optimization2025Alibaba Group./papers/00134-SAPO.pdf
AgentEvolver: Towards Efficient Self-Evolving Agent System2025Alibaba Group./papers/00138-AgentEvolver.pdf
ArenaRL: Scaling RL for Open-Ended Agents via Tournament-based Relative Ranking2026Alibaba Group./papers/00143-ArenaRL.pdf
AMAP Agentic Planning Technical Report2025Alibaba Group./papers/00151-AMAP.pdf
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning2026Alibaba Group./papers/00153-DASD.pdf
Let It Flow: Agentic Crafting on Rock and Roll, Building the ROME Model within an Open Agentic Learning Ecosystem2025Alibaba Group./papers/00165-ALE.pdf
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver2026Alibaba Group./papers/00167-SkillClaw.pdf
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks2026Alibaba Group./papers/00185-LongHorizon-Harness.pdf
Context as an Environment: Programmatic Context Management for Long-Horizon Agents2026Alibaba Group./papers/00186-Scroll.pdf

ByteDance

Tencent

Xiaomi

OPPO AI Agent Team

Meituan LongCat Team

论文年份论文单位笔记地址
LongCat-Flash Technical Report2025Meituan LongCat Team./papers/00093-LongCat-Flash.pdf

Xiaohongshu

Sina Weibo

Baidu

Shanghai Artificial Intelligence Laboratory

EleutherAI

Huawei Noah’s Ark Lab

论文年份论文单位笔记地址
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs2025Huawei Noah’s Ark Lab./papers/00089-Memento.pdf

Tsinghua University

Peking University

Carnegie Mellon University

论文年份论文单位笔记地址
Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning2025Carnegie Mellon University./papers/00053-MRT.pdf

University of California, Berkeley

MIT

Stanford University

Princeton University

National University of Singapore

论文年份论文单位笔记地址
Understanding R1-Zero-Like Training: A Critical Perspective2025National University of Singapore./papers/00048-understand-r1-zero.pdf

Fudan University

Shanghai Jiao Tong University

Nanjing University

Renmin University of China

University of Illinois Urbana-Champaign

Hong Kong University of Science and Technology

论文年份论文单位笔记地址
Dr. RTL: Autonomous Agentic RTL Optimization through Tool-Grounded Self-Improvement2026Hong Kong University of Science and Technology./papers/00169-Dr.RTL.pdf

Institute of Computing Technology, Chinese Academy of Sciences

Columbia University

其他

论文年份论文单位笔记地址
LightRAG: Simple and Fast Retrieval-Augmented Generation2024Beijing University of Posts and Telecommunications./papers/00047-LightRAG.pdf
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models2024The Ohio State University./papers/00071-HippoRAG.pdf
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models2025The Ohio State University./papers/00072-HippoRAG2.pdf
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model2024University of Waterloo./papers/00011-Chinese-Tiny-LLM.pdf
OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models2024University College London./papers/00041-OpenR.pdf
Retrieval-Augmented Generation with Graphs (GraphRAG)2025Michigan State University./papers/00052-GraphRAG-review.pdf
RoFormer: Enhanced Transformer with Rotary Position Embedding2021Zhuiyi Technology./papers/00009-RoPE.pdf
Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions2022Allen Institute for AI./papers/00079-IRCoT.pdf
Memory as Action: Autonomous Context Curation for Long-Horizon Agentic Tasks2025Beijing Jiaotong University./papers/00110-MemAct.pdf
Scientific Algorithm Discovery by Augmenting AlphaEvolve with Deep Research2025IBM Research./papers/00112-DeepEvolve.pdf
Liger-Kernel: Efficient Triton Kernels for LLM Training2025LinkedIn./papers/00130-Liger-Kernel.pdf
HyperRAG: Reasoning N-ary Facts over Hypergraphs for Retrieval Augmented Generation2026National Yang Ming Chiao Tung University./papers/00163-HyperRAG.pdf
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems2026VILA Lab./papers/00180-Dive-into-Claude-Code.pdf
Miles v0.1: Production-Level Post-Training2026Miles Team./papers/00187-Miles.pdf
MechMem-RTL: Reusing Verified Mechanism Memories for LLM-Based RTL Repair2026Hunan University./papers/00190-MechMem-RTL.pdf

书籍

  1. ./books/00001-Claude-Code-Harness-Engineering-Book.pdf

工具

  1. Overleaf

论文集

  1. huggingface daily papers
  2. arxiv

Contributors

yanfeng98

151 commits

yanfeng98/paper-is-all-you-need

论文就是你所需要的。

TeX

2

151 commits

updated Sep 14, 2026

See the code

README

paper-is-all-you-need

论文就是你所需要的。

论文

OpenAI

DeepSeek-AI

KimiTeam

MiniMax

论文年份论文单位笔记地址
MiniMax Sparse Attention2026MiniMax./papers/00179-MSA.pdf

Microsoft

论文年份论文单位笔记地址
ZeRO: Memory Optimizations Toward Training Trillion Parameter Models2020Microsoft./papers/00002-ZeRO.pdf
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone2024Microsoft./papers/00004-Phi-3.pdf
DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales2023Microsoft./papers/00013-DeepSpeed-Chat.pdf
ZeroQuant(4+2): Redefining LLMs Quantization with a New FP6-Centric Strategy for Diverse Generative Tasks2023Microsoft./papers/00014-ZeroQuant(4+2).pdf
DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale2022Microsoft./papers/00015-DeepSpeed-Inference.pdf
DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models2023Microsoft./papers/00016-DeepSpeed-Ulysses.pdf
ZeRO-Offload: Democratizing Billion-Scale Model Training2021Microsoft./papers/00017-ZeRO-Offload.pdf
1-bit Adam: Communication Efficient Large-Scale Training with Adam's Convergence Speed2021Microsoft./papers/00018-1-bit-Adam.pdf
ZeRO-Infinity: Breaking the GPU Memory Wall for Extreme Scale Deep Learning2021Microsoft./papers/00019-ZeRO-Infinity.pdf
1-bit LAMB: Communication Efficient Large-Scale Large-Batch Training with LAMB's Convergence Speed2021Microsoft./papers/00020-1-bit-LAMB.pdf
The Stability-Efficiency Dilemma: Investigating Sequence Length Warmup for Training GPT Models2022Microsoft./papers/00021-Sequence-Length-Warmup.pdf
Maximizing Communication Efficiency for Large-scale Training via 0/1 Adam2022Microsoft./papers/00023-0-1-Adam.pdf
DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale2022Microsoft./papers/00023-DeepSpeed-MoE.pdf
Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model2022Microsoft./papers/00024-Megatron-Turing.pdf
Extreme Compression for Pre-trained Transformers Made Simple and Efficient2022Microsoft./papers/00025-XTC.pdf
ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers2022Microsoft./papers/00026-ZeroQuant.pdf
DeepSpeed Data Efficiency: Improving Deep Learning Model Quality and Training Efficiency via Efficient Data Sampling and Routing2024Microsoft./papers/00027-DeepSpeed-Data-Efficiency.pdf
Understanding INT4 Quantization for Transformer Models: Latency Speedup, Composability, and Failure Cases2023Microsoft./papers/00029-INT4-Quantization.pdf
ZeroQuant-FP: A Leap Forward in LLMs Post-Training W4A8 Quantization Using Floating-Point Formats2023Microsoft./papers/00030-ZeroQuant-FP.pdf
ZeRO++: Extremely Efficient Collective Communication for Giant Model Training2023Microsoft./papers/00031-ZeRO++.pdf
ZeroQuant-HERO: Hardware-Enhanced Robust Optimized Post-Training Quantization Framework for W8A8 Transformers2023Microsoft./papers/00032-ZeroQuant-HERO.pdf
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking2025Microsoft./papers/00044-rStar-Math.pdf
From Local to Global: A Graph RAG Approach to Query-Focused Summarization2024Microsoft./papers/00046-GraphRAG.pdf
Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning2025Microsoft./papers/00050-Logic-RL.pdf
LongRoPE2: Near-Lossless LLM Context Window Scaling2025Microsoft./papers/00063-LongRoPE2.pdf
BitNet b1.58 2B4T Technical Report2025Microsoft./papers/00064-BitNet.pdf
1-bit AI Infra: Part 1.1, Fast and Lossless BitNet b1.58 Inference on CPUs2024Microsoft./papers/00065-1-bit.pdf
Bitnet.cpp: Efficient Edge Inference for Ternary LLMs2025Microsoft./papers/00066-Bitnet.cpp.pdf
MiniLLM: Knowledge Distillation of Large Language Models2023Microsoft./papers/00075-MiniLLM.pdf
rStar2-Agent: Agentic Reasoning Technical Report2025Microsoft./papers/00092-rStar.pdf
RPG: A Repository Planning Graph for Unified and Scalable Codebase Generation2025Microsoft./papers/00111-RPG.pdf
Agent Lightning: Train ANY AI Agents with Reinforcement Learning2025Microsoft./papers/00124-agent-lightning.pdf
X-Coder: Advancing Competitive Programming with Fully Synthetic Tasks, Solutions, and Tests2026Microsoft./papers/00161-X-Coder.pdf

Google

Meta

DeepMind

NVIDIA

Hugging Face

Alibaba Group

论文年份论文单位笔记地址
AlphaMath Almost Zero: process Supervision without process2024Alibaba Group./papers/00007-AlphaMath.pdf
Qwen2.5-Coder Technical Report2024Alibaba Group./papers/00033-Qwen2.5-Coder.pdf
Qwen3 Technical Report2025Alibaba Group./papers/00069-Qwen3_Technical_Report.pdf
Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models2025Alibaba Group./papers/00073-Qwen3-Embedding.pdf
WebWalker: Benchmarking LLMs in Web Traversal2025Alibaba Group./papers/00097-WebWalker.pdf
WebDancer: Towards Autonomous Information Seeking Agency2025Alibaba Group./papers/00098-WebDancer.pdf
WebSailor: Navigating Super-human Reasoning for Web Agent2025Alibaba Group./papers/00082-WebSailor.pdf
WebShaper: Agentically Data Synthesizing via Information-Seeking Formalization2025Alibaba Group./papers/00099-WebShaper.pdf
WebWatcher: Breaking New Frontier of Vision-Language Deep Research Agent2025Alibaba Group./papers/00100-WebWatcher.pdf
WebResearcher: Unleashing unbounded reasoning capability in Long-Horizon Agents2025Alibaba Group./papers/00101-WebResearcher.pdf
ReSum: Unlocking Long-Horizon Search Intelligence via Context Summarization2025Alibaba Group./papers/00102-ReSum.pdf
WebWeaver: Structuring Web-Scale Evidence with Dynamic Outlines for Open-Ended Deep Research2025Alibaba Group./papers/00103-WebWeaver.pdf
WebSailor-V2: Bridging the Chasm to Proprietary Agents via Synthetic Data and Scalable Reinforcement Learning2025Alibaba Group./papers/00104-WebSailor-V2.pdf
Scaling Agents via Continual Pre-training2025Alibaba Group./papers/00105-AgentFounder.pdf
Towards General Agentic Intelligence via Environment Scaling2025Alibaba Group./papers/00106-AgentScaler.pdf
On-Policy RL Meets Off-Policy Experts: Harmonizing Supervised Fine-Tuning and Reinforcement Learning via Dynamic Weighting2025Alibaba Group./papers/00094-CHORD.pdf
Group Sequence Policy Optimization2025Alibaba Group./papers/00107-GSPO.pdf
Tree Search for LLM Agent Reinforcement Learning2025Alibaba Group./papers/00127-Tree-GRPO.pdf
Search Self-play: Pushing the Frontier of Agent Capability without Supervision2025Alibaba Group./papers/00129-SSP.pdf
Soft Adaptive Policy Optimization2025Alibaba Group./papers/00134-SAPO.pdf
AgentEvolver: Towards Efficient Self-Evolving Agent System2025Alibaba Group./papers/00138-AgentEvolver.pdf
ArenaRL: Scaling RL for Open-Ended Agents via Tournament-based Relative Ranking2026Alibaba Group./papers/00143-ArenaRL.pdf
AMAP Agentic Planning Technical Report2025Alibaba Group./papers/00151-AMAP.pdf
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning2026Alibaba Group./papers/00153-DASD.pdf
Let It Flow: Agentic Crafting on Rock and Roll, Building the ROME Model within an Open Agentic Learning Ecosystem2025Alibaba Group./papers/00165-ALE.pdf
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver2026Alibaba Group./papers/00167-SkillClaw.pdf
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks2026Alibaba Group./papers/00185-LongHorizon-Harness.pdf
Context as an Environment: Programmatic Context Management for Long-Horizon Agents2026Alibaba Group./papers/00186-Scroll.pdf

ByteDance

Tencent

Xiaomi

OPPO AI Agent Team

Meituan LongCat Team

论文年份论文单位笔记地址
LongCat-Flash Technical Report2025Meituan LongCat Team./papers/00093-LongCat-Flash.pdf

Xiaohongshu

Sina Weibo

Baidu

Shanghai Artificial Intelligence Laboratory

EleutherAI

Huawei Noah’s Ark Lab

论文年份论文单位笔记地址
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs2025Huawei Noah’s Ark Lab./papers/00089-Memento.pdf

Tsinghua University

Peking University

Carnegie Mellon University

论文年份论文单位笔记地址
Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning2025Carnegie Mellon University./papers/00053-MRT.pdf

University of California, Berkeley

MIT

Stanford University

Princeton University

National University of Singapore

论文年份论文单位笔记地址
Understanding R1-Zero-Like Training: A Critical Perspective2025National University of Singapore./papers/00048-understand-r1-zero.pdf

Fudan University

Shanghai Jiao Tong University

Nanjing University

Renmin University of China

University of Illinois Urbana-Champaign

Hong Kong University of Science and Technology

论文年份论文单位笔记地址
Dr. RTL: Autonomous Agentic RTL Optimization through Tool-Grounded Self-Improvement2026Hong Kong University of Science and Technology./papers/00169-Dr.RTL.pdf

Institute of Computing Technology, Chinese Academy of Sciences

Columbia University

其他

论文年份论文单位笔记地址
LightRAG: Simple and Fast Retrieval-Augmented Generation2024Beijing University of Posts and Telecommunications./papers/00047-LightRAG.pdf
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models2024The Ohio State University./papers/00071-HippoRAG.pdf
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models2025The Ohio State University./papers/00072-HippoRAG2.pdf
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model2024University of Waterloo./papers/00011-Chinese-Tiny-LLM.pdf
OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models2024University College London./papers/00041-OpenR.pdf
Retrieval-Augmented Generation with Graphs (GraphRAG)2025Michigan State University./papers/00052-GraphRAG-review.pdf
RoFormer: Enhanced Transformer with Rotary Position Embedding2021Zhuiyi Technology./papers/00009-RoPE.pdf
Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions2022Allen Institute for AI./papers/00079-IRCoT.pdf
Memory as Action: Autonomous Context Curation for Long-Horizon Agentic Tasks2025Beijing Jiaotong University./papers/00110-MemAct.pdf
Scientific Algorithm Discovery by Augmenting AlphaEvolve with Deep Research2025IBM Research./papers/00112-DeepEvolve.pdf
Liger-Kernel: Efficient Triton Kernels for LLM Training2025LinkedIn./papers/00130-Liger-Kernel.pdf
HyperRAG: Reasoning N-ary Facts over Hypergraphs for Retrieval Augmented Generation2026National Yang Ming Chiao Tung University./papers/00163-HyperRAG.pdf
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems2026VILA Lab./papers/00180-Dive-into-Claude-Code.pdf
Miles v0.1: Production-Level Post-Training2026Miles Team./papers/00187-Miles.pdf
MechMem-RTL: Reusing Verified Mechanism Memories for LLM-Based RTL Repair2026Hunan University./papers/00190-MechMem-RTL.pdf

书籍

  1. ./books/00001-Claude-Code-Harness-Engineering-Book.pdf

工具

  1. Overleaf

论文集

  1. huggingface daily papers
  2. arxiv

Contributors

yanfeng98

151 commits

Languages

TeX

88.6%

BibTeX Style

11.3%