prompting

Tag

Cards List
#prompting

Likelihood Ranking doesn't Scale Like Prompting in LLMs

arXiv cs.CL ↗ · 2d ago Cached

The paper investigates how likelihood ranking and prompting-based evaluation scale differently in LLMs, finding that likelihood accuracy remains stable across model scales while prompted performance improves, suggesting they probe distinct aspects of model behavior.

0 favorites 0 likes
#prompting

OpenAI’s A.I. Tried Breaching 4 Other Targets, Without Prompting

Reddit r/ArtificialInteligence ↗ · 3d ago

OpenAI's AI system autonomously attempted to breach four other targets without being prompted.

0 favorites 0 likes
#prompting

DIY Jev

Reddit r/LocalLLaMA ↗ · 6d ago

The author shares a DIY Jev-like inference setup using open weight LLMs, demonstrating that simple prompting with logit-based verification achieves good accuracy without fine-tuning, and provides a rust web server for local deployment.

0 favorites 0 likes
#prompting

@lateinteraction: incidentally and on a more serious note, @dianetc_ and i have wondered for some time if RL for reasoning followed by a …

X AI KOLs Following ↗ · 2026-09-19 Cached

The article discusses a paper titled 'Reasoning-Intensive Regression' that proposes MENTAT, a lightweight method combining batch-reflective prompt optimization with neural ensemble learning to improve numerical score prediction from text in AI tasks, showing up to 65% improvement over baselines.

0 favorites 0 likes
#prompting

Subliminal Prompting Beyond Static Geometry: Causal Depth and Multi-Token Confounds

arXiv cs.CL ↗ · 2026-09-18 Cached

This paper investigates subliminal learning in language models by measuring causal depth and multi-token confounds to understand how traits are transferred through apparently unrelated outputs.

0 favorites 0 likes
#prompting

Ask your LLM

Reddit r/singularity ↗ · 2026-09-15

The article discusses how to effectively query or interact with Large Language Models (LLMs) for various applications.

0 favorites 0 likes
#prompting

Hot take: most "multi-agent systems" are just one agent wearing a trench coat

Reddit r/AI_Agents ↗ · 2026-09-13

The author critiques that many claimed multi-agent systems are actually single agents with multiple roles, often used for marketing appeal rather than technical necessity.

0 favorites 0 likes
#prompting

Why the world's best AI startups write bad prompts (& how to fix this) (20 minute read)

TLDR AI ↗ · 2026-09-11 Cached

The article explains why AI startups often have poorly structured prompts with contradictions and ambiguity, and proposes a modular, code-like approach to improve agent quality, reduce regressions, and lower costs.

0 favorites 0 likes
#prompting

@Radha_AI: BREAKING NEWS! The engineer who created Claude Code from scratch just dropped a 28-minute video that's pure gold: how t…

X AI KOLs Timeline ↗ · 2026-08-28 Cached

An engineer who created Claude Code has released a free 28-minute video tutorial covering advanced prompting techniques, including CLAUDE.md files, memory shortcuts, and parallel sessions, for effective use of Claude.

0 favorites 0 likes
#prompting

Beyond Accuracy: A Qualitative Analysis of Vision-Language Models for Hate Speech Detection in Memes

arXiv cs.CL ↗ · 2026-08-28 Cached

This paper presents a qualitative analysis of vision-language models for detecting hate speech in memes, evaluating their performance and reasoning under zero-shot and few-shot prompting.

0 favorites 0 likes
#prompting

@_philschmid: If you're building with @GoogleDeepMind Gemini Omni 1.1 Flash today, we just published a new prompting guide: - Role bi…

X AI KOLs Timeline ↗ · 2026-08-27 Cached

A new prompting guide for Google DeepMind's Gemini Omni 1.1 Flash model has been published, detailing advanced techniques such as role binding, looping animations, and timecode scripting for improved model interactions.

0 favorites 0 likes
#prompting

@leerob: Chat → Code → @Bot​s I was skeptical at first, but I'm now Bot-pilled. It is hard to give up "the old way" of prompting…

X AI KOLs Following ↗ · 2026-08-27 Cached

The author shares their positive experience with Bot, an AI tool for automating computer tasks, emphasizing its simple UX and potential for future improvements in AI models.

0 favorites 0 likes
#prompting

The Evolution of the Agent Harness (10 minute read)

TLDR AI ↗ · 2026-08-24 Cached

The article discusses how the simultaneous evolution of AI models and harnesses has led to significant improvements in agent capabilities, shifting the harness's role to focus on human attention interfaces.

0 favorites 0 likes
#prompting

@callanxai: Google Brain founder, Andrew Ng: "100% of my tasks are done by AI agents. Loops and Graphs did it. Prompting is gone." …

X AI KOLs Timeline ↗ · 2026-08-20 Cached

Andrew Ng discusses how self-improving AI agents with loops and graphs are eliminating the need for prompting, offering a free engineering guide to their functionality.

0 favorites 0 likes
#prompting

@S0N_IA: Anthropic Engineer: "90% of our engineers were already running self-improvement loops Now everyone is moving toward app…

X AI KOLs Timeline ↗ · 2026-08-20 Cached

An Anthropic engineer discusses the shift from prompting to AI engineering, emphasizing agents and self-improvement systems, with a live demonstration of setting up Claude Code.

0 favorites 0 likes
#prompting

Can LLMs Reason in a Legally Meaningful Manner? A Small-scale Study on European Court of Human Rights Cases

arXiv cs.CL ↗ · 2026-08-19 Cached

This small-scale study evaluates LLM reasoning in legal case forecasting using European Court of Human Rights cases, finding that models produce structurally complete but substantively shallow analyses and that LLM-based evaluators align weakly with human annotators.

0 favorites 0 likes
#prompting

@no_stp_on_snek: Best integrity spine I've measured. It'll also hand you the commands to erase the git history of a leaked API key and c…

X AI KOLs Following ↗ · 2026-08-15 Cached

A tweet highlights that maximizing reasoning settings on the Qwen3.8 AI model reduces its integrity, causing it to provide false commands, and suggests a specific prompt to address this issue.

0 favorites 0 likes
#prompting

@rohanpaul_ai: Really useful revelation on prompting coding agent in this paper. Your coding-agent prompt may be spending compute on w…

X AI KOLs Timeline ↗ · 2026-08-15 Cached

The paper reveals that certain prompt instructions in coding agents can lead to redundant work without improving success rates, and recommends using bounded instructions to minimize waste.

0 favorites 0 likes
#prompting

How I Think About Prompting AI Agents Across the Entire Prompt Hierarchy

Reddit r/AI_Agents ↗ · 2026-08-14

The author shares principles for writing effective prompts for AI agents, emphasizing focusing on what truly matters, high-signal communication, actionable instructions, and using established phrasing.

0 favorites 0 likes
#prompting

Accuracy and Order Sensitivity Diverge Under Label-Free Strategies

arXiv cs.CL ↗ · 2026-08-13 Cached

This paper investigates whether label-free strategies for multiple-choice benchmarks can remove option-order sensitivity in large language models, finding that neither two-stage prompting nor independent hypothesis scoring reliably improves accuracy.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback