mobile-deployment

Tag

Cards List
#mobile-deployment

Llama-Mobile: Efficient 2.7-Bit Quantization of VLMs

Hugging Face Daily Papers · 2026-08-21 Cached

This paper presents a framework for quantizing vision-language models to 2.7 bits per parameter, enabling efficient mobile deployment by compressing the Llama 3.2 11B Vision Instruct model to 3.7 GB while preserving performance on visual QA tasks.

0 favorites 0 likes
#mobile-deployment

1-Bit Bonsai Image 4B Image Generation for Local Devices

Hacker News Top · 2026-05-31 Cached

PrismML releases Bonsai Image 4B, a family of compact image generation models using 1-bit and ternary weights, enabling high-quality diffusion inference on local devices like laptops and iPhones with significantly reduced memory footprint.

0 favorites 0 likes
← Back to home

Submit Feedback