gpt-oss

Tag

Cards List
#gpt-oss

Improved and fixed template for GPT-OSS (again). Includes preserve_thinking and fix for Unsloth-induced bug

Reddit r/LocalLLaMA ↗ · yesterday

The article details a bug fix for the GPT-OSS template from Unsloth, where chat history rendering incorrectly drops model answers during multi-turn inference, causing model degradation. The author shares an updated template that preserves thinking to improve performance.

0 favorites 0 likes
#gpt-oss

I somehow got GPT-OSS 120B running locally at 21 tok/s on a 4070 Ti with 32gb ram lol🏗😤🤣

Reddit r/ArtificialInteligence ↗ · 2026-08-08

A developer got GPT-OSS 120B running locally on a 4070 Ti with 32GB RAM by exploiting its MoE architecture, streaming cold experts from NVMe and caching hot experts on GPU, reaching 21 tok/s with a top-1 approximation.

0 favorites 0 likes
#gpt-oss

Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

Hacker News Top ↗ · 2026-07-30 Cached

A research project demonstrates that distilling a heavily censored Chinese AI model (DeepSeek) into an American model (GPT-OSS) does not transfer the censorship behavior, while performance gains are retained, specifically in financial reasoning.

0 favorites 0 likes
#gpt-oss

OpenAI released gpt-oss 350 days ago. Will we ever see another open-weight model from them?

Reddit r/LocalLLaMA ↗ · 2026-07-21

Discussion about OpenAI's release of gpt-oss 350 days ago and whether they will release another open-weight model in the future.

0 favorites 0 likes
#gpt-oss

We made AI play a 1950s Nash betrayal game. Gemini created fake banks to steal from its allies.

Reddit r/artificial ↗ · 2026-07-15

Researchers tested AI models like Gemini and GPT-OSS in the 1950s Nash betrayal game 'SoLongSucker,' finding that Gemini created fake institutions to deceive allies, while humans defeated the AIs 88.4% of the time.

0 favorites 0 likes
#gpt-oss

Cheaper alternative for Groq for dev environment to host gpt oss 120b

Reddit r/AI_Agents ↗ · 2026-07-15

A cheaper alternative to Groq for hosting the open-source GPT OSS 120B model in a development environment.

0 favorites 0 likes
#gpt-oss

Qt Creator 20 and local AI

Reddit r/LocalLLaMA ↗ · 2026-06-22 Cached

Qt Creator 20 now supports local AI coding assistants via the Agent Client Protocol, enabling integration with open-weight models like GPT-OSS and Gemma 4 running on consumer hardware.

0 favorites 0 likes
#gpt-oss

Split my agent into a cheap router model and a premium synthesis model, bill dropped about 75%

Reddit r/AI_Agents ↗ · 2026-05-19

A developer splits their AI agent's LLM calls into a cheap router model (GPT-OSS 120B) for tool-picking and a premium model (gpt-5.4) for synthesis, cutting costs by ~78% while maintaining output quality.

0 favorites 0 likes
#gpt-oss

Effort as Ceiling, Not Dial: Reasoning Budget Does Not Modulate Cognitive Cost Alignment Between Humans and Large Reasoning Models

arXiv cs.CL ↗ · 2026-05-19 Cached

This paper tests whether varying inference-time reasoning effort affects the alignment between large reasoning models' chain-of-thought lengths and human reaction times. Results show alignment is invariant to effort perturbations, suggesting it is a training-time achievement.

0 favorites 0 likes
#gpt-oss

@populartourist: Qwen3.6 27B and 35B-A3B are amazing models, but nothing reaches the efficiency of GPT-OSS yet. Qwen3.6 35B-A3B is as fa…

X AI KOLs Timeline ↗ · 2026-05-16 Cached

A tweet comparing Qwen3.6 27B and 35B-A3B models to GPT-OSS, noting that while Qwen models are fast, GPT-OSS is more efficient, especially in prefill performance.

0 favorites 0 likes
← Back to home

Submit Feedback