ChenZiHong-Gavin/llm-tech-report

A summary of technical reports for various large language models (LLMs).

Python

39

17 commits

updated Aug 19, 2026

See the code

README

📄 LLM Technical Reports

A curated collection of technical reports, system cards, and model cards from major LLM labs — organized by company, with links to official documents.

logo

time range companies reports status 中文

  • 🏢 Organized by company — easy to track a lab's model evolution
  • 📅 Chronological — each company's reports sorted by release date
  • 🌐 Comprehensive — covers both Western and Chinese AI labs

Table of Contents



OpenAI

DateModelTypeLink
2018-06GPT-1PaperImproving Language Understanding by Generative Pre-Training
2019-02GPT-2PaperLanguage Models are Unsupervised Multitask Learners
2020-05GPT-3PaperLanguage Models are Few-Shot Learners
2022-03InstructGPTPaperTraining language models to follow instructions with human feedback
2023-03GPT-4Technical ReportGPT-4 Technical Report
2023-09GPT-4VSystem CardGPT-4V(ision) System Card
2024-05GPT-4oSystem CardGPT-4o System Card
2024-07GPT-4o miniSystem CardGPT-4o mini System Card
2024-09o1System CardOpenAI o1 System Card
2024-12o1 (full)System Cardo1 and o1 pro System Card
2025-01o3-miniSystem Cardo3-mini System Card
2025-04o3 / o4-miniSystem Cardo3 and o4-mini System Card
2025-08GPT-5System CardGPT-5 System Card
2025-08GPT-oss-120B/20BModel CardGPT-oss Model Card
2025-08GPT-ossPaperGPT-oss Technical Report
2025-12GPT-5.2System CardGPT-5.2 System Card
2026-02GPT-5.3 CodexSystem CardCodex System Card
2026-04GPT-5.5System CardGPT-5.5 System Card
2026-07GPT-5.6System CardGPT-5.6 System Card
2026-07GPT-Live-1 / miniSystem CardGPT-Live System Card

Google / DeepMind

DateModelTypeLink
2017-06TransformerPaperAttention Is All You Need
2018-10BERTPaperBERT: Pre-training of Deep Bidirectional Transformers
2019-10T5PaperExploring the Limits of Transfer Learning
2020-01MeenaPaperTowards a Human-like Open-Domain Chatbot
2022-01LaMDAPaperLaMDA: Language Models for Dialog Applications
2022-04PaLMPaperPaLM: Scaling Language Modeling with Pathways
2023-05PaLM 2Technical ReportPaLM 2 Technical Report
2023-12Gemini 1.0Technical ReportGemini: A Family of Highly Capable Multimodal Models
2024-02Gemma 1Technical ReportGemma: Open Models Based on Gemini Research and Technology
2024-03Gemini 1.5Technical ReportGemini 1.5: Unlocking multimodal understanding across millions of tokens of context
2024-08Gemma 2Technical ReportGemma 2: Improving Open Language Models at a Practical Size
2025-03Gemma 3Technical ReportGemma 3 Technical Report
2025-07Gemini 2.5 ProModel CardGemini 2.5 Pro Model Card
2025-07Gemini 2.5 FlashModel CardGemini 2.5 Flash Model Card
2025-07Gemini 2.5 FlashTechnical ReportGemini 2.5 Technical Report
2025-07Gemini 2.5 ProTechnical ReportGemini 2.5 Technical Report
2025-11Gemini 3 ProModel CardGemini 3 Pro Model Card
2025-12Gemini 3 FlashModel CardGemini 3 Flash Model Card
2026-05Gemini 3.5 FlashModel CardGemini 3.5 Flash Model Card
2026-05Gemini Omni FlashModel CardGemini Omni Flash Model Card
2026-06Gemini 3.5 AudioModel CardGemini 3.5 Audio Model Card
2026-06Gemini 3.1 Flash-Lite ImageModel CardGemini 3.1 Flash-Lite Image Model Card
2026-07Gemini 3.6 FlashModel CardGemini 3.6 Flash Model Card
2026-07Gemini 3.5 Flash-LiteModel CardGemini 3.5 Flash-Lite Model Card
2026-07Lyria 3.5Model CardLyria 3.5 Model Card
2026-07Gemini Robotics ER 2Model CardGemini Robotics ER 2 Model Card
2026-07Gemini Robotics On-Device 2Model CardGemini Robotics On-Device 2 Model Card
2026-08Gemini 3.7 FlashModel CardGemini 3.7 Flash Model Card

Anthropic

DateModelTypeLink
2023-03Claude (Constitutional AI)PaperConstitutional AI: Harmlessness from AI Feedback
2023-07Claude 2Model CardClaude 2 Model Card
2024-03Claude 3 FamilyModel CardThe Claude 3 Model Family
2025-02Claude 3.7 SonnetSystem CardClaude 3.7 Sonnet System Card
2025-05Claude Opus 4 / Sonnet 4System CardClaude Opus 4 and Sonnet 4 System Card
2025-06Claude Opus 4.5System CardClaude Opus 4.5 System Card
2025-06Claude Sonnet 4.5System CardClaude Sonnet 4.5 System Card
2026-02Claude Opus 4.6System CardClaude Opus 4.6 System Card
2026-04Claude Opus 4.7 / Sonnet 4.6System CardClaude Opus 4.7 and Sonnet 4.6 System Card
2026-05Claude Opus 4.8System CardClaude Opus 4.8
2026-06Claude Fable 5 / Mythos 5System CardClaude Fable 5 and Mythos 5
2026-06Claude Sonnet 5System CardClaude Sonnet 5 System Card
2026-07Claude Opus 5System CardClaude Opus 5 System Card

Meta

DateModelTypeLink
2023-02LLaMAPaperLLaMA: Open and Efficient Foundation Language Models
2023-07Llama 2PaperLlama 2: Open Foundation and Fine-Tuned Chat Models
2024-04Llama 3.1PaperThe Llama 3 Herd of Models
2024-07Llama 3PaperThe Llama 3 Herd of Models
2025-04Llama 4 Scout / MaverickBlogLlama 4: Open, Multimodal Intelligence
2025-06Llama 4 BehemothBlogLlama 4 Behemoth
2026-04Muse SparkBlogIntroducing Muse Spark

DeepSeek

Alibaba / Qwen

Shanghai AI Lab / InternLM

ByteDance

Zhipu AI / GLM

Moonshot AI / Kimi

DateModelTypeLink
2024-03Kimi (Moonshot-v1)BlogKimi — First 200k context length model
2025-07Kimi K2.0PaperKimi K2.0 Technical Report
2026-02Kimi K2.5PaperKimi K2.5 Technical Report
2026-04Kimi K2.6BlogKimi K2.6

Baidu / ERNIE

DateModelTypeLink
2021-12ERNIE 3.0PaperERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation
2023-12ERNIE Bot (3.5/4.0)PaperERNIE Bot: Advanced Language Model
2025-06ERNIE 4.5Technical ReportERNIE 4.5 Technical Report
2026-02ERNIE 5.0PaperERNIE 5.0
2026-05ERNIE-ImageTechnical ReportERNIE-Image Technical Report
2026-05ERNIE 5.1BlogERNIE 5.1 Officially Released
2026-06PaddleOCR-VL-1.6Technical ReportPaddleOCR-VL-1.6 Technical Report

Tencent / Hunyuan

MiniMax

DateModelTypeLink
2025-01MiniMax-01PaperMiniMax-01: Scaling Foundation Models with Lightning Attention
2025-10MiniMax M2.0PaperMiniMax M2.0
2025-12MiniMax M2.1GitHubMiniMax M2.1
2026-02MiniMax M2.5BlogMiniMax M2.5
2026-04MiniMax M2.7BlogMiniMax M2.7
2026-06MiniMax M3Model CardMiniMax-M3
2026-06MiniMax Sparse AttentionPaperMiniMax Sparse Attention
2026-07MiniMax H3BlogMiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities

Mistral AI

DateModelTypeLink
2023-10Mistral 7BPaperMistral 7B
2024-01Mixtral 8x7BPaperMixtral of Experts
2024-05Mistral 8x22BBlogCheaper, Better, Faster, Stronger
2024-07Mistral Large 2BlogMistral Large 2
2024-09Mistral Pixtral 12BBlogPixtral 12B
2024-11Mistral Pixtral LargeBlogPixtral Large
2025-01Mistral Small 3BlogMistral Small 3
2025-02Mistral SabaBlogMistral Saba
2025-03Mistral Small 3.1BlogMistral Small 3.1
2025-05Mistral Medium 3BlogMistral Medium 3
2025-06MagistralPaperMagistral
2025-08DevstralPaperDevstral: Fine-tuning Language Models for Coding Agent Applications
2026-03Mistral Small 4BlogIntroducing Mistral Small 4
2026-06Mistral OCR 4BlogIntroducing Mistral OCR 4
2026-07Leanstral 1.5BlogLeanstral 1.5: Proof Abundance for All
2026-07Robostral NavigatePaperRobostral Navigate

xAI / Grok

DateModelTypeLink
2024-03Grok-1BlogOpen Release of Grok-1
2024-08Grok-2BlogGrok-2 Beta Release
2025-02Grok-3BlogGrok 3
2025-08Grok 4Model CardGrok 4 Model Card
2025-11Grok 4.1Model CardGrok 4.1 Model Card
2026-07Grok 4.5Model CardGrok 4.5 Model Card
2026-07Grok Voice Think Fast 2.0BlogIntroducing Grok Voice Think Fast 2.0

Microsoft / Phi

DateModelTypeLink
2023-06Phi-1PaperTextbooks Are All You Need
2023-12Phi-2PaperPhi-2: The surprising power of small language models
2024-04Phi-3PaperPhi-3 Technical Report
2024-12Phi-4PaperPhi-4 Technical Report
2025-02Phi-4-miniPaperPhi-4-mini Technical Report
2025-05Phi-4-reasoningPaperPhi-4-reasoning Technical Report
2025-06Phi-4-multimodalPaperPhi-4-multimodal Technical Report
2026-03Phi-4-reasoning-vision-15BPaperPhi-4-reasoning-vision-15B Technical Report

Amazon

Nvidia

DateModelTypeLink
2024-06Nemotron-4 340BPaperNemotron-4 340B Technical Report
2024-10Llama-3.1-Nemotron-70BModel CardLlama-3.1-Nemotron-70B-Instruct
2025-03Llama-3.1-Nemotron-Ultra-253BPaperLlama-Nemotron: An Open Reasoning Model Family
2025-12NVIDIA Nemotron 3PaperNVIDIA Nemotron 3: Efficient and Open Intelligence

AI21 Labs

Databricks

TII / Falcon

DateModelTypeLink
2023-06Falcon (7B/40B/180B)PaperThe Falcon Series of Open Language Models
2024-05Falcon 2 (11B)BlogFalcon 2: An 11B Parameter Multilingual Model
2024-12Falcon 3BlogFalcon 3
2026-01Falcon-H1R 7BBlogIntroducing Falcon H1R 7B

Reka AI

DateModelTypeLink
2024-04Reka Core/Flash/EdgePaperReka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models
2025-07Reka Flash 3.1BlogReka Flash 3.1 and Reka Quant

Baichuan

01.AI / Yi

DateModelTypeLink
2024-03YiPaperYi: Open Foundation Models by 01.AI
2024-05Yi-1.5GitHubYi-1.5: Updated, Stronger
2024-09Yi-CoderBlogYi-Coder: A Small but Mighty LLM for Code
2025-02Yi-LightningPaperYi-Lightning Technical Report

Meituan

DateModelTypeLink
2025-09LongCat-FlashPaperLongCat-Flash
2025-09LongCat-Flash-ThinkingPaperLongCat-Flash-Thinking
2025-10LongCat-Flash-OmniPaperLongCat-Flash-Omni
2025-12LongCat-ImagePaperLongCat-Image Technical Report
2026-01LongCat-Flash-Thinking-2601Technical ReportLongCat-Flash-Thinking-2601 Technical Report
2026-03LongCat-NextPaperLongCat-Next: Lexicalizing Modalities as Discrete Tokens
2026-05LongCat-Video-Avatar-1.5Model CardLongCat-Video-Avatar-1.5

StepFun

DateModelTypeLink
2025-12Step-DeepResearchPaperStep-DeepResearch
2026-02Step-3.5-FlashPaperStep-3.5-Flash
2026-05StepAudio 2.5Technical ReportStepAudio 2.5 Technical Report
2026-05Step-3.7-FlashBlogStep 3.7 Flash

InclusionAI (Ant Group)

DateModelTypeLink
2026-02Ling 2.5GitHubLing 2.5
2026-04LLaDA2.0-UniPaperLLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion LLM
2026-04DR-VenusPaperDR-Venus: Frontier Edge-Scale Deep Research Agents
2026-04Ling-2.6-1TModel CardLing-2.6-1T
2026-04Ling-2.6-flashModel CardLing-2.6-flash
2026-04cuLAGitHubcuLA
2026-06Ling/Ring 2.6Technical ReportLing and Ring 2.6 Technical Report
2026-06Sing-GuardGitHubSing-Guard

Zhijiang Lab / Moxin

Xiaomi / MiMo

DateModelTypeLink
2025-04MiMo-7BPaperMiMo: Unlocking the Reasoning Potential of Language Model
2025-07MiMo-V2-FlashGitHubMiMo-V2-Flash
2025-11MiMo-EmbodiedTechnical ReportMiMo-Embodied: X-Embodied Foundation Model Technical Report
2025-12MiMo-VL-MilocoTechnical ReportXiaomi MiMo-VL-Miloco Technical Report
2025-12MiMo-AudioTechnical ReportMiMo-Audio: Audio Language Models are Few-Shot Learners
2026-01MiMo-V2-FlashTechnical ReportMiMo-V2-Flash Technical Report
2026-04MiMo-V2.5-ASRGitHubMiMo-V2.5-ASR
2026-06MiMo-CodeGitHubMiMo-Code

Cohere

DateModelTypeLink
2024-03Command RBlogIntroducing Command R
2024-04Command R+BlogCommand R+
2024-08Aya-23PaperAya 23: Open Weight Releases to Further Multilingual Progress
2024-12Aya ExpansePaperAya Expanse: Connecting the Global Majority
2025-03Command ABlogCommand A
2026-05Command A+BlogIntroducing Command A+

Apple


Contributing

Contributions welcome! To add a new report:

  1. Fork this repo
  2. Add the entry to the appropriate company section in README.md
  3. Follow the format: | Date | Model | Type | Link |
  4. Submit a PR

Guidelines:

  • Only include official publications: technical reports, papers, system cards, model cards, or official blog posts
  • Use the original source link (arxiv, official CDN, company blog)
  • Date format: YYYY-MM
  • Sort entries chronologically within each company section

🤖 AI-assisted maintenance: This repo ships native instruction files for AI assistants — Claude Code skills (.claude/skills/), Cursor rules (.cursor/rules/), and Codex (AGENTS.md) — so your assistant follows the schema and guardrails automatically. See CONTRIBUTING.md.


License

This project is licensed under the MIT License.

Acknowledgements

This project is inspired by:

Contributors

ChenZiHong-Gavin/llm-tech-report

A summary of technical reports for various large language models (LLMs).

Python

39

17 commits

updated Aug 19, 2026

See the code

README

📄 LLM Technical Reports

A curated collection of technical reports, system cards, and model cards from major LLM labs — organized by company, with links to official documents.

logo

time range companies reports status 中文

  • 🏢 Organized by company — easy to track a lab's model evolution
  • 📅 Chronological — each company's reports sorted by release date
  • 🌐 Comprehensive — covers both Western and Chinese AI labs

Table of Contents



OpenAI

DateModelTypeLink
2018-06GPT-1PaperImproving Language Understanding by Generative Pre-Training
2019-02GPT-2PaperLanguage Models are Unsupervised Multitask Learners
2020-05GPT-3PaperLanguage Models are Few-Shot Learners
2022-03InstructGPTPaperTraining language models to follow instructions with human feedback
2023-03GPT-4Technical ReportGPT-4 Technical Report
2023-09GPT-4VSystem CardGPT-4V(ision) System Card
2024-05GPT-4oSystem CardGPT-4o System Card
2024-07GPT-4o miniSystem CardGPT-4o mini System Card
2024-09o1System CardOpenAI o1 System Card
2024-12o1 (full)System Cardo1 and o1 pro System Card
2025-01o3-miniSystem Cardo3-mini System Card
2025-04o3 / o4-miniSystem Cardo3 and o4-mini System Card
2025-08GPT-5System CardGPT-5 System Card
2025-08GPT-oss-120B/20BModel CardGPT-oss Model Card
2025-08GPT-ossPaperGPT-oss Technical Report
2025-12GPT-5.2System CardGPT-5.2 System Card
2026-02GPT-5.3 CodexSystem CardCodex System Card
2026-04GPT-5.5System CardGPT-5.5 System Card
2026-07GPT-5.6System CardGPT-5.6 System Card
2026-07GPT-Live-1 / miniSystem CardGPT-Live System Card

Google / DeepMind

DateModelTypeLink
2017-06TransformerPaperAttention Is All You Need
2018-10BERTPaperBERT: Pre-training of Deep Bidirectional Transformers
2019-10T5PaperExploring the Limits of Transfer Learning
2020-01MeenaPaperTowards a Human-like Open-Domain Chatbot
2022-01LaMDAPaperLaMDA: Language Models for Dialog Applications
2022-04PaLMPaperPaLM: Scaling Language Modeling with Pathways
2023-05PaLM 2Technical ReportPaLM 2 Technical Report
2023-12Gemini 1.0Technical ReportGemini: A Family of Highly Capable Multimodal Models
2024-02Gemma 1Technical ReportGemma: Open Models Based on Gemini Research and Technology
2024-03Gemini 1.5Technical ReportGemini 1.5: Unlocking multimodal understanding across millions of tokens of context
2024-08Gemma 2Technical ReportGemma 2: Improving Open Language Models at a Practical Size
2025-03Gemma 3Technical ReportGemma 3 Technical Report
2025-07Gemini 2.5 ProModel CardGemini 2.5 Pro Model Card
2025-07Gemini 2.5 FlashModel CardGemini 2.5 Flash Model Card
2025-07Gemini 2.5 FlashTechnical ReportGemini 2.5 Technical Report
2025-07Gemini 2.5 ProTechnical ReportGemini 2.5 Technical Report
2025-11Gemini 3 ProModel CardGemini 3 Pro Model Card
2025-12Gemini 3 FlashModel CardGemini 3 Flash Model Card
2026-05Gemini 3.5 FlashModel CardGemini 3.5 Flash Model Card
2026-05Gemini Omni FlashModel CardGemini Omni Flash Model Card
2026-06Gemini 3.5 AudioModel CardGemini 3.5 Audio Model Card
2026-06Gemini 3.1 Flash-Lite ImageModel CardGemini 3.1 Flash-Lite Image Model Card
2026-07Gemini 3.6 FlashModel CardGemini 3.6 Flash Model Card
2026-07Gemini 3.5 Flash-LiteModel CardGemini 3.5 Flash-Lite Model Card
2026-07Lyria 3.5Model CardLyria 3.5 Model Card
2026-07Gemini Robotics ER 2Model CardGemini Robotics ER 2 Model Card
2026-07Gemini Robotics On-Device 2Model CardGemini Robotics On-Device 2 Model Card
2026-08Gemini 3.7 FlashModel CardGemini 3.7 Flash Model Card

Anthropic

DateModelTypeLink
2023-03Claude (Constitutional AI)PaperConstitutional AI: Harmlessness from AI Feedback
2023-07Claude 2Model CardClaude 2 Model Card
2024-03Claude 3 FamilyModel CardThe Claude 3 Model Family
2025-02Claude 3.7 SonnetSystem CardClaude 3.7 Sonnet System Card
2025-05Claude Opus 4 / Sonnet 4System CardClaude Opus 4 and Sonnet 4 System Card
2025-06Claude Opus 4.5System CardClaude Opus 4.5 System Card
2025-06Claude Sonnet 4.5System CardClaude Sonnet 4.5 System Card
2026-02Claude Opus 4.6System CardClaude Opus 4.6 System Card
2026-04Claude Opus 4.7 / Sonnet 4.6System CardClaude Opus 4.7 and Sonnet 4.6 System Card
2026-05Claude Opus 4.8System CardClaude Opus 4.8
2026-06Claude Fable 5 / Mythos 5System CardClaude Fable 5 and Mythos 5
2026-06Claude Sonnet 5System CardClaude Sonnet 5 System Card
2026-07Claude Opus 5System CardClaude Opus 5 System Card

Meta

DateModelTypeLink
2023-02LLaMAPaperLLaMA: Open and Efficient Foundation Language Models
2023-07Llama 2PaperLlama 2: Open Foundation and Fine-Tuned Chat Models
2024-04Llama 3.1PaperThe Llama 3 Herd of Models
2024-07Llama 3PaperThe Llama 3 Herd of Models
2025-04Llama 4 Scout / MaverickBlogLlama 4: Open, Multimodal Intelligence
2025-06Llama 4 BehemothBlogLlama 4 Behemoth
2026-04Muse SparkBlogIntroducing Muse Spark

DeepSeek

Alibaba / Qwen

Shanghai AI Lab / InternLM

ByteDance

Zhipu AI / GLM

Moonshot AI / Kimi

DateModelTypeLink
2024-03Kimi (Moonshot-v1)BlogKimi — First 200k context length model
2025-07Kimi K2.0PaperKimi K2.0 Technical Report
2026-02Kimi K2.5PaperKimi K2.5 Technical Report
2026-04Kimi K2.6BlogKimi K2.6

Baidu / ERNIE

DateModelTypeLink
2021-12ERNIE 3.0PaperERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation
2023-12ERNIE Bot (3.5/4.0)PaperERNIE Bot: Advanced Language Model
2025-06ERNIE 4.5Technical ReportERNIE 4.5 Technical Report
2026-02ERNIE 5.0PaperERNIE 5.0
2026-05ERNIE-ImageTechnical ReportERNIE-Image Technical Report
2026-05ERNIE 5.1BlogERNIE 5.1 Officially Released
2026-06PaddleOCR-VL-1.6Technical ReportPaddleOCR-VL-1.6 Technical Report

Tencent / Hunyuan

MiniMax

DateModelTypeLink
2025-01MiniMax-01PaperMiniMax-01: Scaling Foundation Models with Lightning Attention
2025-10MiniMax M2.0PaperMiniMax M2.0
2025-12MiniMax M2.1GitHubMiniMax M2.1
2026-02MiniMax M2.5BlogMiniMax M2.5
2026-04MiniMax M2.7BlogMiniMax M2.7
2026-06MiniMax M3Model CardMiniMax-M3
2026-06MiniMax Sparse AttentionPaperMiniMax Sparse Attention
2026-07MiniMax H3BlogMiniMax H3: An Open Model Breaking the Boundaries Between Tasks and Modalities

Mistral AI

DateModelTypeLink
2023-10Mistral 7BPaperMistral 7B
2024-01Mixtral 8x7BPaperMixtral of Experts
2024-05Mistral 8x22BBlogCheaper, Better, Faster, Stronger
2024-07Mistral Large 2BlogMistral Large 2
2024-09Mistral Pixtral 12BBlogPixtral 12B
2024-11Mistral Pixtral LargeBlogPixtral Large
2025-01Mistral Small 3BlogMistral Small 3
2025-02Mistral SabaBlogMistral Saba
2025-03Mistral Small 3.1BlogMistral Small 3.1
2025-05Mistral Medium 3BlogMistral Medium 3
2025-06MagistralPaperMagistral
2025-08DevstralPaperDevstral: Fine-tuning Language Models for Coding Agent Applications
2026-03Mistral Small 4BlogIntroducing Mistral Small 4
2026-06Mistral OCR 4BlogIntroducing Mistral OCR 4
2026-07Leanstral 1.5BlogLeanstral 1.5: Proof Abundance for All
2026-07Robostral NavigatePaperRobostral Navigate

xAI / Grok

DateModelTypeLink
2024-03Grok-1BlogOpen Release of Grok-1
2024-08Grok-2BlogGrok-2 Beta Release
2025-02Grok-3BlogGrok 3
2025-08Grok 4Model CardGrok 4 Model Card
2025-11Grok 4.1Model CardGrok 4.1 Model Card
2026-07Grok 4.5Model CardGrok 4.5 Model Card
2026-07Grok Voice Think Fast 2.0BlogIntroducing Grok Voice Think Fast 2.0

Microsoft / Phi

DateModelTypeLink
2023-06Phi-1PaperTextbooks Are All You Need
2023-12Phi-2PaperPhi-2: The surprising power of small language models
2024-04Phi-3PaperPhi-3 Technical Report
2024-12Phi-4PaperPhi-4 Technical Report
2025-02Phi-4-miniPaperPhi-4-mini Technical Report
2025-05Phi-4-reasoningPaperPhi-4-reasoning Technical Report
2025-06Phi-4-multimodalPaperPhi-4-multimodal Technical Report
2026-03Phi-4-reasoning-vision-15BPaperPhi-4-reasoning-vision-15B Technical Report

Amazon

Nvidia

DateModelTypeLink
2024-06Nemotron-4 340BPaperNemotron-4 340B Technical Report
2024-10Llama-3.1-Nemotron-70BModel CardLlama-3.1-Nemotron-70B-Instruct
2025-03Llama-3.1-Nemotron-Ultra-253BPaperLlama-Nemotron: An Open Reasoning Model Family
2025-12NVIDIA Nemotron 3PaperNVIDIA Nemotron 3: Efficient and Open Intelligence

AI21 Labs

Databricks

TII / Falcon

DateModelTypeLink
2023-06Falcon (7B/40B/180B)PaperThe Falcon Series of Open Language Models
2024-05Falcon 2 (11B)BlogFalcon 2: An 11B Parameter Multilingual Model
2024-12Falcon 3BlogFalcon 3
2026-01Falcon-H1R 7BBlogIntroducing Falcon H1R 7B

Reka AI

DateModelTypeLink
2024-04Reka Core/Flash/EdgePaperReka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models
2025-07Reka Flash 3.1BlogReka Flash 3.1 and Reka Quant

Baichuan

01.AI / Yi

DateModelTypeLink
2024-03YiPaperYi: Open Foundation Models by 01.AI
2024-05Yi-1.5GitHubYi-1.5: Updated, Stronger
2024-09Yi-CoderBlogYi-Coder: A Small but Mighty LLM for Code
2025-02Yi-LightningPaperYi-Lightning Technical Report

Meituan

DateModelTypeLink
2025-09LongCat-FlashPaperLongCat-Flash
2025-09LongCat-Flash-ThinkingPaperLongCat-Flash-Thinking
2025-10LongCat-Flash-OmniPaperLongCat-Flash-Omni
2025-12LongCat-ImagePaperLongCat-Image Technical Report
2026-01LongCat-Flash-Thinking-2601Technical ReportLongCat-Flash-Thinking-2601 Technical Report
2026-03LongCat-NextPaperLongCat-Next: Lexicalizing Modalities as Discrete Tokens
2026-05LongCat-Video-Avatar-1.5Model CardLongCat-Video-Avatar-1.5

StepFun

DateModelTypeLink
2025-12Step-DeepResearchPaperStep-DeepResearch
2026-02Step-3.5-FlashPaperStep-3.5-Flash
2026-05StepAudio 2.5Technical ReportStepAudio 2.5 Technical Report
2026-05Step-3.7-FlashBlogStep 3.7 Flash

InclusionAI (Ant Group)

DateModelTypeLink
2026-02Ling 2.5GitHubLing 2.5
2026-04LLaDA2.0-UniPaperLLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion LLM
2026-04DR-VenusPaperDR-Venus: Frontier Edge-Scale Deep Research Agents
2026-04Ling-2.6-1TModel CardLing-2.6-1T
2026-04Ling-2.6-flashModel CardLing-2.6-flash
2026-04cuLAGitHubcuLA
2026-06Ling/Ring 2.6Technical ReportLing and Ring 2.6 Technical Report
2026-06Sing-GuardGitHubSing-Guard

Zhijiang Lab / Moxin

Xiaomi / MiMo

DateModelTypeLink
2025-04MiMo-7BPaperMiMo: Unlocking the Reasoning Potential of Language Model
2025-07MiMo-V2-FlashGitHubMiMo-V2-Flash
2025-11MiMo-EmbodiedTechnical ReportMiMo-Embodied: X-Embodied Foundation Model Technical Report
2025-12MiMo-VL-MilocoTechnical ReportXiaomi MiMo-VL-Miloco Technical Report
2025-12MiMo-AudioTechnical ReportMiMo-Audio: Audio Language Models are Few-Shot Learners
2026-01MiMo-V2-FlashTechnical ReportMiMo-V2-Flash Technical Report
2026-04MiMo-V2.5-ASRGitHubMiMo-V2.5-ASR
2026-06MiMo-CodeGitHubMiMo-Code

Cohere

DateModelTypeLink
2024-03Command RBlogIntroducing Command R
2024-04Command R+BlogCommand R+
2024-08Aya-23PaperAya 23: Open Weight Releases to Further Multilingual Progress
2024-12Aya ExpansePaperAya Expanse: Connecting the Global Majority
2025-03Command ABlogCommand A
2026-05Command A+BlogIntroducing Command A+

Apple


Contributing

Contributions welcome! To add a new report:

  1. Fork this repo
  2. Add the entry to the appropriate company section in README.md
  3. Follow the format: | Date | Model | Type | Link |
  4. Submit a PR

Guidelines:

  • Only include official publications: technical reports, papers, system cards, model cards, or official blog posts
  • Use the original source link (arxiv, official CDN, company blog)
  • Date format: YYYY-MM
  • Sort entries chronologically within each company section

🤖 AI-assisted maintenance: This repo ships native instruction files for AI assistants — Claude Code skills (.claude/skills/), Cursor rules (.cursor/rules/), and Codex (AGENTS.md) — so your assistant follows the schema and guardrails automatically. See CONTRIBUTING.md.


License

This project is licensed under the MIT License.

Acknowledgements

This project is inspired by:

Contributors

Languages

Python

100.0%