mac-optimization

Tag

Cards List
#mac-optimization

Qwen3.8-Flash-Next (95.5 GiB) on a 64GB Mac at ~27 tok/s, checkpoint + fork

Reddit r/LocalLLaMA ↗ · 2026-09-18

The author optimized the Qwen3.8-Flash-Next model to run on a 64GB Mac using expert streaming and other techniques, achieving ~27 tok/s by publishing a checkpoint and a llama.cpp fork.

0 favorites 0 likes
#mac-optimization

Qwen3.8-Flash-Next optimised for Macs

Reddit r/LocalLLaMA ↗ · 2026-08-30

The article details custom optimizations for running the Qwen3.8-Flash-Next AI model on Mac M1 Max hardware, including SSD streaming, custom quantizations, and a sparse attention mechanism to improve performance.

0 favorites 0 likes
#mac-optimization

@Lonely__MH: Unleashed! The uncensored version of Qwen3.8-27B with safety restrictions removed is here! Kudos to the community for the speed! Deeply optimized for Mac M chips! I see everyone discussing the DGX Spark deployment for ling-3.0-flash, and many people's first reaction is that the compute power is too expensive to buy. Since cloud costs are high...

X AI KOLs Timeline ↗ · 2026-08-18 Cached

Qwen3.8-27B uncensored version released, optimized for Mac M chips, supports local deployment, retains multimodal capabilities and safety research features, with simplified installation steps.

0 favorites 0 likes
#mac-optimization

TRELLIS.2 image-to-3D now runs on Mac (Apple Silicon) - no NVIDIA GPU needed

Reddit r/LocalLLaMA ↗ · 2026-04-20

A developer ported Microsoft's TRELLIS.2 image-to-3D model to run on Apple Silicon Macs by replacing CUDA-only dependencies with PyTorch MPS equivalents, enabling offline 3D mesh generation without requiring NVIDIA GPUs.

0 favorites 0 likes
← Back to home

Submit Feedback