steering

Tag

Cards List
#steering

Codex-maxxing

Hacker News Top · 2026-05-19 Cached

Jason Liu shares how he uses OpenAI's Codex for knowledge work beyond coding, leveraging durable threads, voice input, and steering to integrate coding agents into his broader workflow.

0 favorites 0 likes
#steering

@jxnlco: jason from the codex team here, heres a draft on codex maxxing and the primatives i use on a daily basis https://jxnl.g…

X AI KOLs Following · 2026-05-17 Cached

Jason Liu shares his workflow primitives for using Codex effectively, including durable threads, voice input, and steering to extend AI agents beyond coding into knowledge work.

0 favorites 0 likes
#steering

DeepSeek-V4-Flash means LLM steering is interesting again

Hacker News Top · 2026-05-16 Cached

The article explores how DeepSeek-V4-Flash, a powerful local model, makes LLM steering practical again, discussing the concept and its implementation in the DwarfStar 4 project by antirez.

0 favorites 0 likes
#steering

Non-linear Interventions on Large Language Models

arXiv cs.CL · 2026-05-15 Cached

This paper introduces a general formulation of non-linear intervention for large language models, extending beyond the Linear Representation Hypothesis to manipulate features encoded along non-linear manifolds, and validates the approach on refusal bypass steering.

0 favorites 0 likes
#steering

Negative Before Positive: Asymmetric Valence Processing in Large Language Models

arXiv cs.CL · 2026-05-08 Cached

This paper investigates how large language models process emotional valence through mechanistic interpretability. Using activation patching and steering on three open-source LLMs, the authors find that negative valence is localized to early layers while positive valence peaks in mid-to-late layers, and they validate this through topic-controlled flip tests.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback