ONNX-powered LLM + LoRA inferencing for multi-accelerator setups
C#
0
6 commits
updated Sep 12, 2026
Leosang-lx/LLM-Inference
Research and develop distributed LLM inference
johnmai-dev/ANE-LM
LLM inference on Apple Neural Engine (ANE)
144
kouhxp/sherpa-onnx-diarization-models
CodeBySonu95/Sherpa-onnx-models
onnx-community/LFM2-VL-1.6B-ONNX
1
TJU-NSL/awesome-papers
37
efeslab/Nanoflow
A throughput-oriented high-performance serving framework for LLMs
975
onnx-community/LFM2-VL-3B-ONNX
99.4%