Perplexity open-sourced their Mac inference server for Qwen 3.6

Reddit r/LocalLLaMA Tools

Summary

Perplexity has open-sourced a Mac inference server optimized for the Qwen 3.6 model to achieve best performance on Apple Silicon.

https://preview.redd.it/6h4xc5o8l6nh1.png?width=1158&format=png&auto=webp&s=b65074b6baaa1faa2347e5259229c8ba803bcd4b Here is link to repo: https://github.com/perplexityai/pplx-garden/tree/main/lily It's optimized for just one model to get best perf on apple silicon
Original Article

Similar Articles

Run Qwen3.8 27B locally: real numbers from my Mac Studio

Hacker News Top

The article provides real-world performance benchmarks for running the Qwen3.8 27B AI model locally on a Mac Studio, comparing it to its predecessor and discussing hardware requirements and quantization effects.