Tag
A user describes an unexpected behavior where the quantized Qwen 3.8 27B model opened a browser to test code autonomously during development, highlighting emergent capabilities.
Tested the updated Gemma 4 locally using llama.cpp on an M5 Pro, achieving 60 tokens/s for coding tasks with OpenCode; good for backend but poor for UI/UX.