I’m building the fastest local inference engine for Apple Silicon (r/LLMDevs)
Hi everyone! I’m passionate about making local models accessible to everyone. I’m trying to make an inference engine optimized across the stack for consumer MacBooks. Currently it is the fastest way to run LFM, Qwen3.5, and the new Clef Flash decision model locally if you are on Mac. Would greatly…