@ai_xiaomu: Measured data: Running qwen3.8 (averaging about 65G per machine) on dual Spark machines does not affect ComfyUI's normal functioning, but it gets quite hot.
Summary
Tests indicate that dual Spark machines can run qwen3.8 and ComfyUI simultaneously, but the devices heat up severely.
View Cached Full Text
Cached at: 08/28/26, 07:48 AM
Actual test data: Running Qwen3.8 (average ~65G memory usage per machine) on two Spark machines simultaneously does not affect the normal operation of ComfyUI, though the heat generation is somewhat severe.
Huang Xiaomu (@ai_xiaomu): ok, Qwen 3.8 Flash is running, connecting to DeepSeek harness to try out the effect.
Similar Articles
@zhixianio: Finished testing, feeling quite surprised, not sure if I'm using it wrong. Feel free to provide counterexamples. Here are my results: On M5 Max, pitting this community fine-tuned gemma-4-12B-coder (llama.cpp) against my daily driver Qwen3.6-35B-…
The user tested the community fine-tuned gemma-4-12B-coder against Qwen3.6-35B-A3B MoE on three programming tasks, finding that gemma performed poorly on complex stateful programs, while Qwen 35B remained robust.
Qwen 3.6 27B is a BEAST
A developer reports that the new 27B Qwen 3.6 model runs excellently on a 24GB VRAM laptop, passing all PySpark/Python data-transformation benchmarks and eliminating the need for cloud subscriptions.
@AI_Community_9: Unsloth Desktop takes the experience of Qwen3.8 27B to new heights. Out-of-the-box + cross-platform, the core is optimizing VRAM for fine-tuning and inference to the extreme (VRAM saved by 70%, speed increased by 2 times). It packs 'training, inference, and Agent toolchain' into a free, open-source desktop application, …
Unsloth Desktop is a free, open-source desktop application that integrates training, inference, and Agent toolchain through extreme optimization of VRAM and speed, significantly improving the user experience of AI models like Qwen3.8 27B.
@Lonely__MH: Unleashed! The uncensored version of Qwen3.8-27B with safety restrictions removed is here! Kudos to the community for the speed! Deeply optimized for Mac M chips! I see everyone discussing the DGX Spark deployment for ling-3.0-flash, and many people's first reaction is that the compute power is too expensive to buy. Since cloud costs are high...
Qwen3.8-27B uncensored version released, optimized for Mac M chips, supports local deployment, retains multimodal capabilities and safety research features, with simplified installation steps.
Running Qwen 3.6 35B MoE (Q4_K_M) on a Zeus (Xiaomi 12 Pro, 12GB RAM)
Demonstrates running the Qwen 3.6 35B MoE model in Q4_K_M quantization on a Xiaomi 12 Pro with 12GB RAM, showing local AI inference on a mobile device.