proactive-agents

Tag

Cards List
#proactive-agents

Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation

arXiv cs.AI · yesterday Cached

This survey proposes a unified decision framework for proactive service agents using large language models, formulating the problem as a partially observable sequential decision process and organizing methods and evaluation metrics for initiative, timing, and safety.

0 favorites 0 likes
#proactive-agents

@omarsar0: The more I embrace open and cheaper models, the more automation I can afford. Frontier models for orchestration and coo…

X AI KOLs Following · 2026-08-28 Cached

The author discusses the benefits of using open and cheaper AI models for automation tasks while reserving frontier models for orchestration, enabling more proactive agents.

0 favorites 0 likes
#proactive-agents

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

Hugging Face Daily Papers · 2026-08-11 Cached

Introduces VibeLifeBench, a benchmark of 200 long-horizon tasks across ten everyday-life domains for evaluating proactive and persistent LLM agents in a simulated multi-week living world. Current frontier models score low, highlighting the gap between existing agents and real-life assistance.

0 favorites 0 likes
#proactive-agents

@mitch_troy: Excited to announce the Basis End-to-End Tax Platform, the first production deployment I’m aware of built around truly …

X AI KOLs Following · 2026-08-05 Cached

Mitch Troy announces the Basis End-to-End Tax Platform, described as the first production deployment built around proactive agents that autonomously handle tax return preparation while allowing accountants to supervise decisions.

0 favorites 0 likes
#proactive-agents

@RhysSullivan: i can't believe the ideal form factor for background / proactive agents is codex pets

X AI KOLs Following · 2026-07-11 Cached

Rhys Sullivan discusses the ideal form factor for background/proactive agents being Codex Pets, and lists desired agents like one that watches GitHub repos for changes.

0 favorites 0 likes
#proactive-agents

Context Graphs for Proactive Enterprise Agents

arXiv cs.AI · 2026-07-10 Cached

This paper proposes Context Graphs, a live relational data structure for enterprise entities that enables proactive agents to surface relevant information before users query, formalizing components for delta detection, proactivity scoring, and LLM-powered surfacing.

0 favorites 0 likes
#proactive-agents

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

Hugging Face Daily Papers · 2026-07-09 Cached

UniClawBench introduces a capability-driven benchmark for evaluating proactive agents in dynamic, real-world environments using live Docker containers and a closed-loop evaluation strategy with multiple agent roles.

0 favorites 0 likes
#proactive-agents

Make your agents proactive with one line of code

Reddit r/AI_Agents · 2026-07-06

An SDK lets developers make AI agents proactive with a single line of code, enabling self-scheduled execution across multiple frameworks.

0 favorites 0 likes
#proactive-agents

Communication Policy Evolution for Proactive LLM Agents

arXiv cs.AI · 2026-06-15 Cached

This paper formalizes communication policy for LLM agents and proposes Communication Policy Evolution (CPE), a self-evolution framework that refines communication policies through rollout and prompt-level evolving, achieving best task success across multiple settings.

0 favorites 0 likes
#proactive-agents

Reusable Knowledge vs. Operational Memory: The missing distinction in building AI agents that can actually follow through

Reddit r/AI_Agents · 2026-06-03

The article distinguishes between reusable knowledge (durable context) and operational memory (task state) as essential components for building proactive AI agents that can follow through on complex tasks.

0 favorites 0 likes
#proactive-agents

$\Psi$-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues

arXiv cs.LG · 2026-06-03 Cached

Ψ-Bench is a benchmark for evaluating LLMs' ability to influence users through persuasive dialogues, incorporating user profiles for personalized persuasion. Experiments show that even state-of-the-art models have room for improvement, and access to client profiles significantly boosts performance.

0 favorites 0 likes
#proactive-agents

Data Isn't Scarce. Your Imagination Is (8 minute read)

TLDR AI · 2026-05-29 Cached

Asuka Zheng argues that the 'running out of training data' panic is misplaced; the real scarcity is a lack of imagination in collecting diverse, long-horizon data, illustrated by her SRE replacement project and broader research trends.

0 favorites 0 likes
#proactive-agents

Context: Proactive Goal-Directed Intelligence via Composable Sandboxed Programs, Declarative Wiring, and Structured Interaction

arXiv cs.AI · 2026-05-26 Cached

This paper introduces Context, a new architecture for proactive goal-directed agents that replaces reactive chatbots. It presents formal theorems proving efficiency gains through composable sandboxed programs, declarative wiring, and proactive state machines, with an open-source implementation.

0 favorites 0 likes
#proactive-agents

Anticipate and Learn: Unleashing Idle-Time Compute in Proactive Agents

Hugging Face Daily Papers · 2026-05-25 Cached

ProAct is a proactive agent architecture that leverages idle-time computation to anticipate user needs, improving task completion efficiency and accuracy. It introduces ProActEval, a benchmark spanning 200 scenarios across 40 domains, and achieves significant gains over reactive baselines: 14.8% reduction in required turns, 11.7% decrease in user effort, and 28.1% cut in hallucination rates.

0 favorites 0 likes
#proactive-agents

@shawn_pana: Proactive agents are the future We're building Agency in Browser Use Box > Agents propose goals and tasks to complete >…

X AI KOLs Following · 2026-05-14 Cached

A new tool called Agency in Browser Use Box enables AI agents to propose goals and tasks, with humans accepting or rejecting them and agents notifying progress.

0 favorites 0 likes
← Back to home

Submit Feedback