Gemma 4 26B A4B running on iPhone 17 Pro via model paging

Reddit r/LocalLLaMA Models

Summary

Google's Gemma 4 26B A4B model can run on iPhone 17 Pro using model paging, enabling powerful on-device AI.

No content available
Original Article

Similar Articles

Gemma 4 + LiteRT-LM on mobile: much better memory/perf than my llama.cpp setup

Reddit r/LocalLLaMA

A user shares a hands-on comparison of running Gemma 4 with LiteRT-LM on mobile devices versus their previous llama.cpp setup, noting significantly better memory usage (1.5-2 GB vs 4-5 GB) and faster inference (2-4 seconds vs 7-10 seconds) on smartphones like Samsung S25 Ultra and iPhone 13 Pro Max.