Cluster 436191
387 repos
leeguooooo/AgentParty, wshobson/agents, yyjeqhc/webcodex, ctxrs/ctx, ferrislucas/Circus-Chief, Ut8v/apronagents
Cluster 436304
346 repos
wx-chevalier/Awesome-Lists, wx-chevalier/Awesome-Lists-and-CheatSheets, szabgab/awesome-lists, andrew/ultimate-awesome, zupcode-com/awesome-free-services-for-your-next-startup-or-saas, dictcp/awesome-git
Cluster 448828
238 repos
fishaudio/s1-mini, CohereForAI/aya-23-8B, fishaudio/openaudio-s1-mini, fishaudio/fish-speech-1.5, CohereForAI/aya-expanse-8b, BlinkDL/rwkv-6-world
Cluster 436924
204 repos
Aratako/T5Gemma-TTS, Aratako/Irodori-TTS, dots-studio/dots.tts-base, rednote-hilab/dots.tts-base, loudreader/loudkit, fluxions/vui
Cluster 435996
178 repos
jiahaog/nativefier, nativefier/nativefier, atom/atom, PowerShell/PowerShell, jlevy/the-art-of-command-line, brave/brave-browser
Cluster 437443
178 repos
AutomatosX/AX-Qwen3.6-35B-A3B-MLX-AXQ-4bit, AutomatosX/AX-Qwen3.8-27B-MLX-AXQ-4bit, AutomatosX/AX-Qwen3.6-27B-MLX-AXQ-4bit, AutomatosX/AX-Holo-3.1-35B-A3B-MLX-AXQ-MXFP4, AutomatosX/AX-Ministral-3-14B-Instruct-2512-MLX-AXQ-4bit, AutomatosX/AX-Qwen3.8-27B-MLX-AXQ-6bit
Cluster 436173
170 repos
TeamAmaze/AmazeFileManager, microg/GmsCore, nextcloud/android, The412Banner/Bannerlator, LawnchairLauncher/lawnchair, cinit/QAuxiliary
Cluster 436185
156 repos
Bin-Huang/chatbox, chatboxai/chatbox, FranciscoMoretti/sparka, feiskyer/kube-copilot, hwchase17/langchain, codecentric/c4-genai-suite
Cluster 436704
148 repos
stevearc/oil.nvim, stevearc/conform.nvim, stevearc/dressing.nvim, stevearc/overseer.nvim, tamago324/lir.nvim, charm-and-friends/freeze.nvim
Cluster 439696
141 repos
facebook/wav2vec2-conformer-rel-pos-large-960h-ft, facebook/s2t-small-librispeech-asr, jonatasgrosman/wav2vec2-large-xlsr-53-chinese-zh-cn, facebook/wav2vec2-xlsr-53-espeak-cv-ft, jonatasgrosman/wav2vec2-large-xlsr-53-arabic, ccoreilly/wav2vec2-large-100k-voxpopuli-catala
Cluster 436090
140 repos
MaikuB/flutter_appauth, miguelpruivo/plugins_flutter_file_picker, koel/player, Notsfsssf/pixez-flutter, ncosgray/cuppa_mobile, skiptools/skip
Cluster 441583
138 repos
BAAI/bge-small-en-v1.5, llmrails/ember-v1, BAAI/bge-large-zh-v1.5, BAAI/bge-large-en-v1.5, BAAI/bge-base-zh-v1.5, PandaLLMCommunity/panda-index-large-en
Cluster 435982
133 repos
xo/usql, uptrace/bun, pressly/goose, amacneil/dbmate, flyxl/datazen, EdurtIO/incubator-dbm
Cluster 439001
120 repos
CompVis/stable-diffusion-v1-4, runwayml/stable-diffusion-inpainting, naclbit/trinart_stable_diffusion_v2, digiplay/Photon_v1, Linaqruf/anything-v3.0, gsdf/Counterfeit-V2.5
Cluster 437426
116 repos
m-bain/whisperX, QuintinShaw/openasr, MahmoudAshraf97/whisper-diarization, efeslab/LiteASR, Macoron/whisper.unity, istupakov/onnx-asr
Cluster 448806
114 repos
google-bert/bert-base-chinese, hfl/chinese-roberta-wwm-ext, hfl/chinese-bert-wwm-ext, hfl/chinese-macbert-base, hfl/chinese-macbert-large, Geotrend/bert-base-ru-cased
Cluster 436915
109 repos
Tencent-Hunyuan/HunyuanVideo-1.5, AILab-CVC/VideoCrafter, Saravutw/Omni-videos-custom, zai-org/CogVideoX1.5-5B, THUDM/CogVideoX1.5-5B, Wan-AI/Wan2.1-T2V-1.3B-Diffusers
Cluster 441460
90 repos
pugant/Qwen3.8-Flash-Next-Q4_0_ROCMFP4_STRIX_LEAN-GGUF, pugant/Qwen3.8-Flash-Next-ROCMFP4_STRIX_LEAN-GGUF, philtheriver/Qwen3.6-27B-ROCmFPX, philtheriver/Qwopus3.6-27B-v2-MTP-ROCmFPX, pugant/Qwen3.6-35B-A3B-MTP-Q6_0_ROCMFPX, pugant/grug-35b-v2-ROCmFP4-STRIX_LEAN
Cluster 448673
89 repos
black-forest-labs/FLUX.2-dev, black-forest-labs/FLUX.2-klein-base-9b-nvfp4, black-forest-labs/FLUX.2-klein-9b-nvfp4, black-forest-labs/FLUX.2-klein-9b-kv, black-forest-labs/FLUX.2-klein-9B, black-forest-labs/FLUX.2-klein-base-9B
Cluster 440326
82 repos
Salesforce/moirai-1.0-R-base, Salesforce/moirai-1.0-R-small, Salesforce/moirai-1.0-R-large, Salesforce/moirai-2.0-R-small, ibm-granite/granite-timeseries-ttm-r2, time-series-foundation-models/Lag-Llama
Cluster 436118
81 repos
golang/vscode-go, mrmlnc/vscode-postcss-sorting, swift-server/vscode-swift, swiftlang/vscode-swift, semgrep/semgrep-vscode, fwcd/vscode-kotlin
Cluster 451952
78 repos
mlboydaisuke/MiniCPM5-1B-CoreAI, mlboydaisuke/Qwen3.6-35B-A3B-CoreAI, mlboydaisuke/Nanbeige4.1-3B-CoreAI, mlboydaisuke/GLM-4.7-Flash-CoreAI, mlboydaisuke/MiniCPM-V-4.6-CoreAI, mlboydaisuke/LFM2.5-1.2B-CoreAI
Cluster 436436
78 repos
praetorian-inc/noseyparker, KeygraphHQ/shannon, bearer/bearer, securego/gosec, Hackmanit/Web-Cache-Vulnerability-Scanner, zizmorcore/zizmor
Cluster 449610
78 repos
joelito/legal-maltese-roberta-base, joelito/legal-polish-roberta-base, joelito/legal-hungarian-roberta-base, joelito/legal-slovak-roberta-base, joelito/legal-slovenian-roberta-base, joelito/legal-romanian-roberta-base
Cluster 439293
77 repos
2toINF/X-VLA-libero-goal-peft, 2toINF/X-VLA-libero-long-peft, 2toINF/X-VLA-libero-object-peft, 2toINF/X-VLA-simpler-widowx-peft, 2toINF/X-VLA-libero-spatial-peft, sarulab-speech/sidon_raw_weight
PyTorch Model Hub & Image Processing
75 repos
Libraries and tools for publishing, sharing, and managing PyTorch models through standardized hub integrations, with emphasis on serialization formats like safetensors and the model_hub_mixin ecosystem. The cluster includes image processing and segmentation models, alongside multimodal encoders and image-text architectures. These repositories provide infrastructure for model versioning, distribution, and interoperability across the Hugging Face Hub and related platforms.
Cluster 436059
73 repos
open-telemetry/opentelemetry-collector, odigos-io/odigos, monoscope-tech/monoscope, krzko/otelgen, uptrace/uptrace, SigNoz/signoz
Cluster 448841
72 repos
TheBloke/Llama-2-7B-Chat-GGML, TheBloke/Llama-2-13B-Chat-GGML, TheBloke/Llama-2-70B-fp16, uukuguy/speechless-llama2-13b, uukuguy/speechless-llama2-hermes-orca-platypus-wizardlm-13b, TheBloke/Speechless-Llama2-13B-GGML
Cluster 436509
71 repos
asyncapi/website, rody-huancas/docuventus, xitanggg/open-resume, haydenbleasel/next-forge, vercel/next-forge, amelioro/ameliorate
T5 Text-to-Text Transformers
70 repos
Pre-trained and fine-tuned variants of the T5 (Text-to-Text Transfer Transformer) model family, including base, small, and extra-large configurations optimized for various domains like biomedical text (Pubmed/PMC), Chinese language tasks, and general text generation. These repositories provide model weights, inference endpoints, and implementations compatible with the Hugging Face Transformers library for tasks spanning machine translation, summarization, question answering, and other sequence-to-sequence problems.
Code Generation & LLM Inference
68 repos
Large language models fine-tuned and optimized for code generation tasks, along with infrastructure for deploying and serving text-generation models at scale. The cluster centers on specialized code LLMs (Wizardcoder, Speechless variants, and similar models) designed to improve upon base models like CodeLlama through instruction-tuning and alignment techniques. Repositories here cover model weights, inference servers, and deployment patterns compatible with standard text-generation endpoints.
Cluster 436049
67 repos
charmbracelet/bubbles, ratatui-org/ratatui, ratatui/ratatui, rhysd/tui-textarea, lrstanley/bubblezone, Textualize/textual
Cluster 436102
66 repos
eduardofuncao/squix, slick/slick, launix-de/memcp, snowflakedb/snowflake-connector-net, h2database/h2database, cwida/duckdb
Cluster 453276
63 repos
open-gigaai/GigaBrain-0.7-3.5B-Base, RLWRLD/RLDX-1-PT, open-gigaai/GigaBrain-0.7-SampleData, RLWRLD/RLDX-1-PT-IMG, RLWRLD/RLDX-1-FT-LIBERO, RLWRLD/RLDX-1-FT-ROBOCASA
Cluster 435993
63 repos
coq-community/paramcoq, mattam82/Coq-Equations, rocq-community/aac-tactics, coq-community/aac-tactics, rocq-prover/equations, coq-community/bignums
Tree-sitter Parser Generators
62 repos
Parser generators and language bindings built on tree-sitter, a parsing toolkit for building fast, incremental parsers for programming languages. The cluster consists primarily of tree-sitter grammar definitions and bindings for major languages (JavaScript, C, Rust, Java, Ruby, OCaml, Julia, and others), along with supporting tools and libraries. Developers exploring this area will find grammar specifications, language-specific implementations, and infrastructure for building syntax-aware tools like editors, linters, and code analyzers.
ESLint plugins and rule ecosystems
58 repos
ESLint rule plugins and tooling for extending JavaScript/TypeScript linting with specialized rules across testing, imports, promises, and code quality. This cluster centers on the ESLint plugin architecture—both general-purpose plugins like eslint-plugin-import-x and domain-specific ones for Jest, promises, and other concerns—alongside shared infrastructure for building and maintaining ESLint rules. Contributors and maintainers in this space work on both the plugins themselves and the meta-tooling that makes plugin development sustainable.
Retrieval-Augmented Generation (RAG) Systems
57 repos
Libraries, frameworks, and implementations for building retrieval-augmented generation systems that combine large language models with external knowledge retrieval. The cluster spans foundational RAG architectures (MiniRAG, LightRAG, R2R), semantic search approaches (SearchLM, SubgraphRAG), and agent-based variants that integrate retrieval with agentic AI workflows. Most repos are Python-based tools and frameworks for developers implementing RAG pipelines at various scales of complexity.
Cluster 435981
57 repos
qltysh/qlty, realm/SwiftLint, PyCQA/pylint, crystal-ameba/ameba, analysis-tools-dev/static-analysis, detekt/detekt
Cluster 436380
56 repos
mozilla/rust, rust-lang/rust, rust-lang/rustup.rs, rust-lang-nursery/rustup.rs, graydon/rust, esp-rs/rust
npm CLI and Package Management
53 repos
JavaScript tooling for npm's command-line interface and package registry operations. This cluster covers the core utilities that power npm's functionality—from package fetching and registry communication to dependency resolution and package metadata handling. Developers here will find libraries for interacting with npm registries, managing package dependencies, handling cryptographic package integrity (SSRI), and orchestrating the CLI workflows that make npm one of JavaScript's essential tools.
iOS/macOS Native Development
53 repos
Libraries, frameworks, and tools for building native applications on Apple platforms using Swift and Objective-C. The cluster emphasizes database solutions (notably Realm), SDK integrations, and runtime utilities for iOS and macOS development. Beyond the database-centric principal repos, it encompasses a broad range of supporting libraries for features like voice communication, error tracking, and container management that developers commonly integrate into native Apple apps.
Cluster 449609
52 repos
shai-msy/per-classifier, cardiffnlp/twitter-roberta-base-2021-124m-topic-multi, cardiffnlp/twitter-roberta-base-2021-124m-hate, cardiffnlp/twitter-roberta-base-2021-124m-offensive, cardiffnlp/twitter-roberta-base-2021-124m-topic-single, cardiffnlp/twitter-roberta-base-2021-124m-sentiment
C# / .NET Libraries and Templates
50 repos
Libraries, SDKs, and project templates for the .NET ecosystem, with primary focus on Azure Functions, Web Jobs, WPF desktop applications, and generative AI integrations. The cluster spans cloud platform tooling, game engine bindings (Unity), source code linking utilities, and various development templates for C#-based projects. While most repos are C# libraries and frameworks, a smaller set of PowerShell and shell tooling supports infrastructure automation and developer workflows.
Text Generation & Inference Optimization
49 repos
Libraries, models, and frameworks for efficient large language model inference and text generation at scale. The cluster centers on techniques like speculative decoding and optimized tensor operations (via safetensors), with a focus on making transformer-based models faster and more practical to deploy. Primary repos include Qwen-series models optimized for inference performance, alongside text-generation-inference frameworks and related transformer utilities for production applications.
Frontend Build Tools & Plugin Ecosystems
49 repos
TypeScript-based tooling and plugin systems for modern web development, particularly focused on build optimization, bundling, and developer experience. The cluster centers on plugin architectures (plugma, unplugin ecosystem) and framework integrations for React, Vue, and Angular, alongside tooling like Vite and Storybook that enable rapid frontend development. Repositories here address common problems in frontend tooling: code splitting, icon management, console debugging enhancements, and PWA capabilities.
Scala data processing and streaming
49 repos
Distributed computing frameworks and data pipeline tools built primarily in Scala and Java. The cluster centers on Apache Spark for large-scale data processing, Akka for actor-based concurrency and reactive systems, and Akka HTTP for building reactive web services. These projects enable building scalable data workflows, streaming applications, and real-time processing systems, with supporting libraries like Alpakka for reactive integrations and Linkis for data engine orchestration.
Webpack loaders and plugins
48 repos
JavaScript bundler extensions and middleware for webpack, including specialized loaders for processing various file types (HTML, CoffeeScript) and plugins for optimizing bundle size through compression and asset transformation. These repositories represent the webpack plugin ecosystem that extends the core bundler's capabilities for web performance and build-time asset handling.
Large Language Models and NLP
48 repos
Open-source large language models, training frameworks, and natural language processing tools built primarily in Python. The cluster centers on accessible implementations and variants of LLMs—including models like Baichuan, MOSS, and camel—alongside infrastructure for fine-tuning, inference, and dialogue systems. Developers here will find both pretrained models ready for deployment and the underlying frameworks for building and adapting language models to specific tasks.
GGML and LLaMA C++ Inference
47 repos
Fast, efficient large language model inference through GGML (a tensor library optimized for inference on consumer hardware) and LLaMA.cpp, a popular C++ implementation enabling quantized model execution on CPU and GPU. This cluster contains mostly C++ implementations, quantization tools, and framework bindings focused on making LLM inference practical and portable across platforms without heavy dependencies.
Cluster 441320
47 repos
cy0307/d-dreamerv3-world-model, cy0307/b-nerf-from-scratch, cy0307/b-nerfstudio-nerfacto, cy0307/d-splatam-slam, cy0307/ag-habitat-navigation, cy0307/b-gaussian-splatting-3d
LLM Agents & Agentic AI
47 repos
Libraries, frameworks, and tools for building autonomous agents powered by large language models. This cluster covers agent orchestration, multi-agent systems, production-ready deployment patterns, and agent-as-a-service architectures. The repos span Python-heavy implementations alongside TypeScript and Rust alternatives, reflecting both rapid prototyping and performance-critical use cases in the emerging agentic AI space.
Open-source LLM Applications & Chat Tools
47 repos
Self-hosted and open-source applications built around large language models, primarily chat interfaces and AI-powered tools. The cluster is dominated by TypeScript-based web applications and Python backends for deploying, managing, and interacting with LLMs, with a strong emphasis on analytics, customization, and local deployment patterns. Common themes include chat frontends, prompt management, model integration, and community-driven alternatives to commercial AI services.
OpenAPI and REST API tooling
46 repos
Libraries, parsers, and code generators for OpenAPI/Swagger specifications across TypeScript, JavaScript, and Go. This cluster encompasses tools for validating and parsing OpenAPI documents (swagger-parser, apidom), generating client and server code from specs (swagger-codegen), and providing interactive documentation and exploration interfaces (swagger-ui). Developers use these tools to build, document, and maintain REST APIs with specification-driven workflows.
Apache Commons Java Utilities
46 repos
A collection of reusable Java libraries and components maintained under the Apache Commons project, providing foundational utilities for common programming tasks. The cluster includes widely-used modules for command-line argument parsing (CLI), object pooling (Pool), XML/object manipulation (Digester, JXPath), and process execution (Exec), along with supporting infrastructure for Maven-based builds and dependency management. These libraries form the backbone of many Java applications and frameworks, offering stable, battle-tested solutions for recurring development challenges.
Container orchestration and Kubernetes tooling
46 repos
Go-based tools and utilities for building, deploying, and managing containerized applications on Kubernetes and container runtimes. The cluster spans runtime implementations (podman, gvisor), orchestration and cluster management (kops, kompose), and build/deployment automation (ko). These projects address the full lifecycle of cloud-native applications, from container image creation through production orchestration.
Markdown parsing and rendering
44 repos
Libraries and tools for parsing, processing, and rendering Markdown across multiple languages. The cluster centers on CommonMark and GFM (GitHub Flavored Markdown) implementations, with a heavy concentration of JavaScript/TypeScript projects alongside Rust implementations like pulldown-cmark and marked. Developers exploring this area will find parsers, AST processors, linters, and ecosystem tooling for handling Markdown in web development, documentation generation, and content processing.
Deep Learning Computer Vision & YOLO
44 repos
Deep learning frameworks and models for real-time object detection and computer vision tasks, centered on PyTorch implementations. The cluster contains numerous YOLO (You Only Look Once) variants and related detection architectures, alongside foundational vision models like DINO. Repositories span from core model implementations and training pipelines to inference optimization and practical applications, making this a comprehensive resource for practitioners building vision systems with modern deep learning approaches.
Protocol Buffers tooling and runtimes
44 repos
Libraries, code generators, and runtime implementations for Protocol Buffers serialization across Go, TypeScript, Java, and other languages. This cluster covers the ecosystem around protoc compilation, protobuf wire format handling, and language-specific marshaling/unmarshaling — from core runtime libraries like hyperpb-go and ProtoBuf.js to higher-level tools like buf that improve the protobuf development workflow.
Proxy and VPN protocol implementations
44 repos
Open-source implementations and clients for proxy and VPN protocols including Shadowsocks, V2Ray, Xray, Trojan, and VMess. The cluster is dominated by Go server implementations (v2ray-core, Xray-core, Marzban) alongside mobile and desktop clients in various languages, plus shell-based deployment and configuration tools. This area spans both protocol development and practical tooling for building censorship-resistant proxy networks and privacy-focused tunneling infrastructure.
Computational Pathology & Medical Image Analysis
43 repos
Deep learning approaches for analyzing histopathology and medical images, with a focus on vision transformers, feature extraction, and self-supervised learning methods. This cluster centers on PyTorch-based tools for processing gigapixel pathology slides and extracting interpretable features from medical imagery, with notable work on foundation models like UNI and TITAN that serve as pretrained backbones for downstream pathology tasks. Repositories span implementations of attention-based multiple instance learning, feature aggregation pipelines, and domain-specific vision architectures optimized for the unique computational constraints of whole-slide image analysis.
Multimodal Vision-Language Models
43 repos
Large-scale transformer-based models that combine vision and language capabilities for tasks like image understanding, visual question answering, and cross-modal feature extraction. The cluster centers on the InternVL family of models—ranging from compact 2B parameters to large 76B variants—which demonstrate how to build efficient multimodal systems across different scale points. Repositories here focus on model architectures, safetensor checkpoints, multilingual support, and practical implementations of vision-language integration using modern transformer frameworks.
Multimodal Vision-Language Models
42 repos
Libraries and model implementations for vision-language tasks, particularly image captioning and image-to-text generation using transformer architectures. The cluster centers on Japanese language variants of LLaVA models and related multimodal AI systems that combine visual understanding with text generation. While the language composition appears sparse in the metadata, the consistent presence of image-captioning, transformers, and image-text-to-text topics across repositories indicates a focused area around building and deploying models that process both images and text.
BitTorrent and P2P file sharing
42 repos
JavaScript and Go implementations of BitTorrent protocol, peer discovery, and distributed file-sharing systems. The cluster centers on WebTorrent (a JavaScript-based torrent client for browsers and Node.js) and related protocol implementations, alongside tracker systems and P2P networking infrastructure. Developers working here build low-level protocol handlers, distributed content discovery mechanisms, and client libraries for decentralized file distribution.
JVM Ecosystem & Cross-Language SDKs
42 repos
Libraries and frameworks for building on the Java Virtual Machine, with substantial support for polyglot development across Java, C#, Python, JavaScript, and Rust. The cluster includes core JVM tools (Avro for data serialization, Selenium for browser automation), language-specific SDKs and code generation frameworks, and copilot/AI-assisted development tooling. Projects span data processing, testing infrastructure, and developer productivity, unified by a focus on interoperability and SDK patterns that bridge multiple ecosystems.
NLP Tasks, Datasets & Benchmarks
42 repos
Libraries, datasets, and benchmarking frameworks for natural language processing tasks including text classification, semantic similarity, sentence embeddings, and structured NLP pipelines. The cluster spans from low-level linguistic tools (like spaCy and AllenNLP for parsing and annotation) to high-level task frameworks (like PromptSource for prompt-based learning) and multilingual resources (Thai sentence vectors, cross-lingual benchmarks). Most repos are Python-based tools and Jupyter notebooks exploring NLP methods empirically.
BERT and Transformer-based NLP
41 repos
Libraries, models, and applications built on BERT and transformer architectures for natural language processing tasks. The cluster spans multiple language implementations and use cases—from foundational model variants like ModernBERT and language-specific adaptations (KoBERT, KcBERT, turkish-sentiment-analysis-with-bert) to sentiment analysis and text classification pipelines. Most repos are Python-based with Jupyter notebooks for experimentation, reflecting the practical, applied nature of this work across production and research contexts.
Docker Images and Container Packaging
41 repos
Docker container images and Dockerfile-based packaging for applications and services. The cluster centers on creating, building, and distributing containerized versions of popular software like web servers (httpd), caching layers (memcached), content management systems (WordPress), and lightweight embedded systems (busybox), with heavy use of shell scripting for build automation and image configuration. This represents the practical tooling and patterns for containerizing existing applications rather than container orchestration or runtime systems.
WebAssembly Virtual Machines & JavaScript Integration
41 repos
JavaScript engines and virtual machines compiled to or implemented in WebAssembly, enabling high-performance JavaScript execution in browser and non-browser environments. The cluster spans runtime implementations, type systems, promise integration, exception handling, and threading support for WASM-based VMs. Central repos like gc, function-references, and exception-handling represent core language features and proposals, while the broader collection explores how JavaScript semantics map onto WebAssembly's execution model.
Cluster 436559
40 repos
getsentry/sentry-python, getsentry/sentry-cli, getsentry/sentry-go, getsentry/sentry-electron, getsentry/sentry-php, getsentry/sentry-laravel
AWS Lambda serverless utilities
40 repos
Libraries and tools for building and operating AWS Lambda functions across multiple languages. The cluster centers on the AWS Lambda Powertools suite (Python, TypeScript, .NET), which provide observability, middleware, and utility functions for serverless applications, alongside related serverless framework plugins and Lambda-specific tooling. These repos help developers standardize logging, tracing, and operational patterns across serverless workloads.
Cluster 436077
39 repos
vlang-io/V, vlang/v, nature-lang/nature, blade-lang/blade, NicoNex/tau, carbon-language/carbon-lang
Quantized Language Models and Inference
39 repos
Optimized implementations and quantized variants of large language models, primarily focused on the GGUF format for efficient inference and deployment. The cluster centers on model compression techniques (particularly 4-bit and 8-bit quantization), multilingual model variants, and infrastructure for serving conversational AI systems with reduced computational footprint. Repositories here span model distributions, quantization tools, and frameworks that make LLM inference practical on resource-constrained hardware.
Music Audio and Synthesis Tools
39 repos
Libraries, datasets, and applications for music generation, audio synthesis, and instrument modeling. The cluster spans music information retrieval (datasets and scoring systems), synthesizer design and audio processing, and streaming music platforms. Most repositories are language-agnostic tools and datasets rather than language-specific libraries, with Python and C++ implementations supporting both research and production use cases.
Cluster 436947
39 repos
Shubhamsaboo/awesome-llm-apps, patchy631/ai-engineering-hub, getzep/graphiti, Sumanth077/Hands-On-AI-Engineering, NirDiamant/GenAI_Agents, LazyAGI/LazyLLM
Kubernetes Core API and Control Plane
38 repos
Go libraries and servers for Kubernetes' API layer, including the API server, client libraries, custom resource definitions, and kubectl command-line interface. This cluster covers the fundamental control plane components and client tooling that enable interaction with Kubernetes clusters, from low-level API protocol handling to higher-level cluster management interfaces. The dominance of Go and kubernetes-related topics reflects the mature, tightly-integrated nature of these core upstream projects.
Gradle Build System and Plugins
38 repos
Java and Groovy tooling centered on the Gradle build automation system, including plugins for build optimization, caching, profiling, and native compilation. The cluster covers both plugin development patterns and practical guides for improving build performance in JVM projects, with secondary support for cloud-native and Python-related build scenarios.
Language Server Protocol implementations
38 repos
Language servers that enable IDE features like autocompletion, diagnostics, and refactoring across different programming languages by implementing the Language Server Protocol standard. This cluster contains LSP servers for TypeScript, Elixir, Elm, Erlang, AWK, F#, and other languages, spanning implementations in TypeScript, Rust, Go, and functional languages. Repositories here range from foundational LSP server libraries to complete language-specific implementations used by editor extensions and IDEs.
WebAssembly Compilation & Tooling
38 repos
Toolchains, compilers, and runtime systems for compiling to and executing WebAssembly across multiple languages and platforms. The cluster centers on core infrastructure like wasm-tools and emscripten (C/C++ to WASM compilation), alongside language-specific tooling for Go, Rust, and other ecosystems. Repositories here cover compiler backends, bytecode manipulation, optimization, and WASM runtime support for both browser and server-side execution.
GPU-Accelerated ML & CUDA Kernels
37 repos
Libraries and frameworks for machine learning acceleration on NVIDIA GPUs, centered on CUDA kernel optimization, collective communication primitives, and low-level GPU compute abstractions. Includes foundational tools like NCCL (distributed GPU communication), CUTLASS (GPU tensor operations), and cuML (GPU-accelerated ML algorithms), alongside operator implementations and GPU infrastructure components. Most repos are production Python and C++ code targeting high-performance ML workloads and system-level GPU programming.
Web markup formatting and beautification
37 repos
Libraries and tools for parsing, formatting, and optimizing HTML, CSS, and JavaScript code. This cluster centers on code beautifiers, minifiers, and formatters—including js-beautify and minify as core examples—alongside web frameworks and component libraries like Bootstrap and Bulma that heavily depend on well-structured markup and styling. Most repos are JavaScript-based utilities for developer tooling and frontend build pipelines, with educational resources on web standards and markup best practices.
Diffusion Models & Image Generation
37 repos
Deep learning libraries and applications for training, fine-tuning, and extending diffusion-based generative models, particularly Stable Diffusion and similar architectures. The cluster spans practical tools for model optimization (like SimpleTuner for efficient training), creative applications combining diffusion with other modalities (face synthesis, video editing, multi-view generation), and frameworks for experimenting with diffusion variants and mixtures. Most repositories are Python-based implementations suitable for both research and production use.
GitHub Actions for Google Cloud
37 repos
TypeScript-based GitHub Actions workflows and utilities for deploying to and managing Google Cloud Platform services. The cluster covers common GCP integration patterns including cloud storage uploads, Cloud Run and Cloud Functions deployments, authentication, secrets management, and GKE cluster access. These are production-ready action libraries that enable CI/CD pipelines to interact with GCP resources directly from GitHub workflows.
Math, Code, and Reasoning in LLMs
36 repos
Systems and techniques for enhancing large language models' capabilities in mathematical reasoning, code generation, and logical problem-solving. This cluster covers instruction-tuning datasets (like MathCodeInstruct), retrieval-augmented approaches, and specialized model architectures designed to improve transformer-based text generation for technical domains. Projects here address the challenge of making LLMs more reliable and accurate when handling symbolic reasoning, mathematical proofs, and code synthesis tasks.
LLM Serving and Inference Optimization
36 repos
Python libraries and frameworks focused on deploying, serving, and optimizing large language models in production. The cluster centers on high-performance inference engines, batch processing, memory optimization, and hardware acceleration (particularly AMD support). vLLM, the dominant repository by internal connectivity, exemplifies the core work here—providing efficient LLM serving infrastructure with features like paged attention and continuous batching. Repositories in this cluster address the practical engineering challenges of running LLMs at scale, from quantization and pruning to distributed inference and dynamic scheduling.
OCaml metaprogramming and PPX
36 repos
OCaml code generation and compile-time metaprogramming via PPX (preprocessor extensions) and related macro systems. The cluster centers on libraries like ppx_sexp_conv, ppx_expect, ppx_fields_conv, and ppx_inline_test that use OCaml's PPX framework to automatically generate boilerplate code, enable expressive testing patterns, and reduce manual serialization logic. Developers exploring this area will find infrastructure for building domain-specific language extensions, testing frameworks, and productivity tools within the OCaml ecosystem.
Flutter app development libraries
36 repos
Libraries and utilities for building Flutter applications, with a focus on caching, state management, and localization. The cluster includes image caching solutions (flutter_cached_network_image, flutter_cache_manager), state management tools (Redux integration), translation/localization frameworks (Flutter-Translate), and other Dart-based mobile development utilities. Most repos are pure Dart implementations, with some C++ and Rust components for performance-critical operations.
Privacy-focused infrastructure and reverse proxies
36 repos
Tools and services for deploying privacy-preserving infrastructure, VPN solutions, and reverse proxies with automatic HTTPS. The cluster spans VPN installers and servers (WireGuard, OpenVPN, IPSec), privacy-respecting frontends for public services (like Libreddit), and proxy/gateway solutions—primarily written in shell scripts and Go. These are practical DevOps and self-hosting utilities for individuals and small teams prioritizing privacy and security.
Terminal and CLI color libraries
36 repos
Libraries and utilities for adding colored text and ANSI styling to command-line interfaces and terminal applications. This cluster spans multiple languages (primarily JavaScript and TypeScript, with Go implementations) and focuses on making it easy for developers to output formatted, colorized text in CLI tools and scripts. Central projects like chalk, colorette, and yoctocolors represent different approaches to terminal color abstraction, from feature-rich to minimal, while the broader cluster includes related terminal formatting and console output tooling.
Reproducible ML Research & Open Science
36 repos
Reproducible implementations and open-source artifacts from machine learning research papers, primarily from ICML 2026 and related venues. This cluster contains code reproductions, experimental implementations, and collaborative research tooling focused on making academic ML work verifiable and accessible. The repositories emphasize reproducibility infrastructure, agent collaboration frameworks, and static analysis tools for validating research claims across diverse algorithmic topics.