Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis
Summary
An experiment ran the Qwen 3.8 27B model for 63 hours on a RTX 3090 to attempt solving the Riemann hypothesis, demonstrating autonomous reasoning, self-correction, and no hallucination.
Similar Articles
Qwen 3.8 27B Running LIVE on a RTX 5090 to solve an Open Math Problem - Covering Design C(25,15,5)
A live experiment running the Qwen 3.8 27B model on an RTX 5090 to solve a covering design math problem, demonstrating the potential of open-source AI on consumer hardware for scientific innovation.
Wow! Qwen 3.6:35b-a3b on a 3090... pretty amazing.
A user shares impressive results running a quantized Qwen 3.6:35b-a3b model on a used RTX 3090, achieving 160 tokens per second output after fitting the model into VRAM, and demonstrates vision capabilities with a 75-second video processing time.
Qwen 3.6 27B Speculative Decoding Bench: Pushing ~100 TPS on a single RTX 3090
A detailed benchmark comparing speculative decoding engines for Qwen 3.6 27B on a single RTX 3090, showing ik_llama achieving ~100 tokens per second in code generation. Results include decode TPS, TTFT, VRAM usage, and context degradation across 5 engine variants.
The bear can dance: Qwen 3.8 27B on one 3090 for 3 weeks
An experiment demonstrated that a quantized Qwen 3.8 27B model, running locally on a single RTX 3090 GPU, autonomously pursued optimizing CUDA inference for over three weeks, producing functional kernels and benchmarks while maintaining coherent long-term goal-following.
@ItsmeAjayKV: Achievement Unlocked: Running Qwen3.6-27b dense Thanks to the RTX 3090, now I can do this. Running @Alibaba_Qwen Qwen 3…
User benchmarks Qwen3.6-27B on an RTX 3090 using llama.cpp, achieving 35 tok/s generation and 1247 tok/s prompt processing.