mlx-serve

Tag

Cards List
#mlx-serve

Qwen3.8-Flash-Next on MLX-serve, 1m context is released!

Reddit r/LocalLLaMA · 18h ago

The article announces the release of Qwen3.8-Flash-Next on MLX-serve, supporting 1 million token context with efficient performance on M5 Max hardware using quantized weights.

0 favorites 0 likes
#mlx-serve

@ddalcu: What a crazy last 4 days for local AI... insane... https://github.com/ddalcu/mlx-serve/releases/tag/v26.8.2… @liquidai …

X AI KOLs Timeline · 2026-08-04 Cached

A developer updates MLX-Serve, a fast local inference server for Apple Silicon, to support recent models like LiquidAI 2.6B, MiniMax H3 video generation, and DeepSeek V4 Flash, with AntLing 3.0-flash coming soon.

0 favorites 0 likes
← Back to home

Submit Feedback