Ran a classic(medival europe) fantasy RP/agentic benchmark across 8 local models Qwen3.6-27B held up better than its size suggests

Reddit r/LocalLLaMA News

Summary

A benchmark comparing 8 local models on a classic medieval European fantasy role-playing and agentic task found that Qwen3.6-27B performed better than its size would suggest.

No content available
Original Article

Similar Articles

Qwen/Qwen-AgentWorld-35B-A3B

Hugging Face Models Trending

Qwen releases Qwen-AgentWorld-35B-A3B, a native language world model that simulates agentic environments across seven domains via long chain-of-thought reasoning. The model is trained with a three-stage pipeline and supports MCP, Search, Terminal, SWE, Android, Web, and OS interactions.