On-Device LLM Inference & Apple Silicon

39 repos

Running large language models efficiently on personal devices and Apple Silicon hardware, with emphasis on optimized inference, memory constraints, and native integration. The cluster centers on practical implementations of quantized and specialized model variants (including multimodal vision-language models) designed for macOS and iOS, along with Swift-native tooling to deploy and interact with these models locally without cloud dependencies.

Swift · 2
Python · 1
macos ·6,817
llm ·6,814
apple-silicon ·6,775
swift ·6,405
on-device ·6,404
homebrew ·6,360
foundationmodels ·6,360
cli ·6,360
macos-26 ·6,360
apple-intelligence ·6,360