minimax-m3

Tag

Cards List
#minimax-m3

Vision Support for Minimax-M3 has been merged into llama.cpp

Reddit r/LocalLLaMA ↗ · 2026-07-26 Cached

Vision support for the Minimax-M3 model has been merged into the llama.cpp project, enabling multimodal inference for this model locally.

0 favorites 0 likes
#minimax-m3

Minimax M3 support with MSA has been merged into llama.cpp

Reddit r/LocalLLaMA ↗ · 2026-07-26 Cached

Minimax M3 support with MSA has been merged into llama.cpp, enabling inference for the Minimax M3 model using the MSA architecture.

0 favorites 0 likes
#minimax-m3

8-16 MI50s Minimax M3 @19 tps TG (peak)

Reddit r/LocalLLaMA ↗ · 2026-06-21

Reports a peak throughput of 19 tokens per second for the Minimax M3 model running on 8-16 MI50 GPUs.

0 favorites 0 likes
#minimax-m3

@browser_use: Open-weights models have officially caught up We tried GLM 5.2 in BrowserCode > Near Opus-level score > Cheapest model …

X AI KOLs Following ↗ · 2026-06-19 Cached

Open-weights models have caught up with proprietary ones, with GLM 5.2 achieving near Opus-level scores in browser agent tasks at low cost. Other models like Minimax M3 and Kimi k2.7 also show notable improvements.

0 favorites 0 likes
#minimax-m3

@dealignai: MiniMax m3, made for 128gb Mac’s Thank you to @hornsby_andrew for preparing the pruning calibration dataset and doing e…

X AI KOLs Timeline ↗ · 2026-06-18 Cached

A pruned and quantized version of MiniMax-M3 (MiniMax-M3-Medium-JANG_2L) optimized to run on 128GB Macs using vMLX, featuring 32% expert pruning and JANG_2L mixed-precision quantization to fit within ~105 GB.

0 favorites 0 likes
#minimax-m3

PM tried M3's 1M context on a real Q3 brief: where it held, where it broke

Reddit r/AI_Agents ↗ · 2026-06-17

A product manager shares hands-on testing of Minimax M3's 1M context window on a real Q3 strategic brief, noting strong source attribution up to ~200K tokens but synthesis degradation beyond that.

0 favorites 0 likes
#minimax-m3

@atomic_chat_hq: Open-weight MiniMax M3 filled out a US customs form from a driver's license photo For this test we deployed MiniMax M3 …

X AI KOLs Timeline ↗ · 2026-06-15 Cached

A test of the open-weight MiniMax M3 model using MLX-VLM on a Mac Studio shows it can autonomously fill out a US customs form from a driver's license photo and a scanned document, using tool calls for fields, checkboxes, and signature.

0 favorites 0 likes
#minimax-m3

@0xSero: Minimax-M3 running on 4x RTX Pro 6000s - 800k context - 4x concurrency at 250k - 70-120 tok/s - 2000 tok/s prefill no c…

X AI KOLs Following ↗ · 2026-06-14 Cached

Minimax-M3 is demonstrated running on 4x RTX Pro 6000 GPUs with 800k context, achieving 70-120 tok/s inference and 2000 tok/s prefill at 4x concurrency using 376GB VRAM in mxfp4 format.

0 favorites 0 likes
#minimax-m3

Kimi K2.6 vs Minimax M3: 5x the cost for worse results? I ran the tests.

Reddit r/AI_Agents ↗ · 2026-06-12

A hands-on comparison of Kimi K2.6 and Minimax M3 in real agent workflows shows M3 costs roughly 5x less while delivering nearly identical quality, making it more cost-effective for production systems.

0 favorites 0 likes
#minimax-m3

Unsloth Minimax M3 GGUF

Reddit r/LocalLLaMA ↗ · 2026-06-12

Unsloth is uploading a GGUF quantized version of the MiniMax M3 model to Hugging Face.

0 favorites 0 likes
#minimax-m3

unsloth/MiniMax-M3-GGUF

Hugging Face Models Trending ↗ · 2026-06-12 Cached

Unsloth releases a GGUF quantized version of the MiniMax-M3 multimodal model, enabling image-text-to-text tasks with support for Transformers, llama.cpp, vLLM, and other inference engines.

0 favorites 0 likes
#minimax-m3

As we know Minimax M3 is just going to be open sourced in few days and because of that I was surfing on internet searching for its scores and I found out pretty interesting results. Is Minimax M3 really that good in agentic stuff and in coding? Is it better than older gpt models?

Reddit r/LocalLLaMA ↗ · 2026-06-11

A user inquires about the upcoming open-source Minimax M3 model's performance in agentic tasks and coding, asking how it compares to older GPT models like GPT 5.2.

0 favorites 0 likes
#minimax-m3

MiniMax promises M3 weights after 1M-context model launch (2 minute read)

TLDR AI ↗ · 2026-06-03 Cached

MiniMax released M3, a model with a 1M-token context window and native multimodal input, via API. The company promises open-weight release and a technical report within 10 days.

0 favorites 0 likes
#minimax-m3

@RyanLeeMiniMax: MiniMax-M3 will by arrive on HuggingFace openweight at next week!

X AI KOLs Following ↗ · 2026-06-01 Cached

MiniMax announced MiniMax-M3, an open-weights model combining frontier coding and agentic capabilities with sparse attention scaling to 1M context, set to arrive on HuggingFace next week.

0 favorites 0 likes
← Back to home

Submit Feedback