web-agents

Tag

Cards List
#web-agents

Action Conditioned Bisimulation For GUI Agent Memory

arXiv cs.AI ↗ · 2d ago Cached

This paper proposes action-conditioned bisimulation over an empirical predictive state graph to decide when GUI agent memories of two web pages should be merged, improving success on MiniWoB++ over memoryless baselines without any training.

0 favorites 0 likes
#web-agents

@browser_use: Browser Use Ultrafast available > 10x faster and cheaper > compare flights for $0.004 > in <20s Web agents are now fast…

X AI KOLs Timeline ↗ · 4d ago Cached

The tweet announces Browser Use Ultrafast, a cloud-based web agent tool that is 10x faster and cheaper, enabling tasks like flight comparison for $0.004 in under 20 seconds.

0 favorites 0 likes
#web-agents

The Hard Part Comes After Search: Benchmarking Web Agents on Synthesizing, Organizing, and Displaying Knowledge

arXiv cs.CL ↗ · 5d ago Cached

The paper introduces KNOWS, a benchmark for evaluating web agents on complex, long-horizon tasks that involve synthesizing and organizing knowledge into artifacts, revealing that current agents struggle with visual steps and long-horizon reasoning.

0 favorites 0 likes
#web-agents

X-Tree: Tokenizing Reusable Experience for Efficient Agent Generalization

Hugging Face Daily Papers ↗ · 2026-09-26 Cached

X-Tree recovers reusable skill hierarchies directly from agent trajectories (no LLM calls) and integrates them into offline RL, online RLVR, and on-policy self-distillation, improving success rates on WebArena, ScienceWorld, and WebShop by up to 5.8% over standard training recipes at matched data and budget.

0 favorites 0 likes
#web-agents

@dair_ai: Great paper from Microsoft Research and colleagues. If you auto-generate MCP tools from agent trajectories, this one is…

X AI KOLs Timeline ↗ · 2026-09-21 Cached

AutoTailor is a meta-agentic framework that converts web trajectories into compact, user-aligned browser automation APIs, improving correctness and reducing token cost and latency in web tasks.

0 favorites 0 likes
#web-agents

EconSkills: Studying Skill Transfer and Retrieval for Web Agents on Live Economic Data

arXiv cs.AI ↗ · 2026-09-18 Cached

EconSkills introduces a skill library and evaluation framework for web agents to transfer and retrieve procedural knowledge for live economic data retrieval, showing improved efficiency in controlled transfer and competitive performance at library scale.

0 favorites 0 likes
#web-agents

90% of the "look my agent can browse the web" posts here would collapse on the 5th URL

Reddit r/AI_Agents ↗ · 2026-09-16

The author critiques web agent demos for failing on real-world URLs due to fetch-layer issues, emphasizing that data engineering is the harder problem than agent reasoning.

0 favorites 0 likes
#web-agents

Web agents can navigate changing websites by looking at the screen

Reddit r/artificial ↗ · 2026-09-15

Dhruv Batra explains how web agents can navigate changing websites by using visual information from the screen, learning from interactions, and adapting to layout changes without manual updates.

0 favorites 0 likes
#web-agents

AutoTailor: Automatic, User-Aligned Capability Selection and Adaptation for Web Agents

arXiv cs.AI ↗ · 2026-09-15 Cached

AutoTailor is a meta-agentic framework that automatically selects and adapts compact sets of browser-automation APIs for web agents, improving accuracy and efficiency through offline filtering and dynamic reselection.

0 favorites 0 likes
#web-agents

Token Efficient Task Execution via Application Behavior Modeling for Web Agents

arXiv cs.AI ↗ · 2026-09-15 Cached

OdoBot is a novel web-agent architecture that uses application behavior modeling to reduce token consumption and improve task success rates, outperforming agents like Agent-E and WebVoyager on the Canvas LMS.

0 favorites 0 likes
#web-agents

When and What to Teach: Budget-Aware Online Adaptation for Web Agents

arXiv cs.AI ↗ · 2026-09-10 Cached

This paper proposes a budget-aware online teaching framework for web agents that reduces teacher calls and compute costs while maintaining performance.

0 favorites 0 likes
#web-agents

SCAFFOLD: Self-Improving Web Agents via Recursive Parametric Skill Abstraction

arXiv cs.AI ↗ · 2026-09-10 Cached

Scaffold is a self-improving framework for visual web agents that induces parametric skills, maintains a recursive hierarchy, and distills skills into model weights, achieving significant performance improvements on benchmarks like WebArena.

0 favorites 0 likes
#web-agents

@browser_use: Astra is a monster at browser use

X AI KOLs Timeline ↗ · 2026-09-06 Cached

Astra achieved 77.3% on the Browser Use Benchmark v2, far surpassing Opus 5 (50.5%) and GPT-5.6 Sol xhigh (49.1%), with 22 of 60 tasks earning full marks compared to zero for Opus 5.

0 favorites 0 likes
#web-agents

@browser_use: Introducing: Browser Use x @link 💳 Web agents can now make purchases with your credit card. Without ever seeing your r…

X AI KOLs Timeline ↗ · 2026-09-03 Cached

Browser Use partners with Link to enable AI web agents to make credit card purchases using single-use cards, enhancing security and functionality for automated transactions.

0 favorites 0 likes
#web-agents

Monitoring Web Agents Without Internal Signals: Observable Trajectories and Key-Step Supervision

arXiv cs.AI ↗ · 2026-09-03 Cached

The paper proposes a method for monitoring web agents without access to internal model signals, using observable trajectories and key-step supervision to predict failures early. It demonstrates competitive performance with internal-signal baselines across benchmarks.

0 favorites 0 likes
#web-agents

An edge case I didn't expect: actions that exist on a page but aren't currently interactable

Reddit r/AI_Agents ↗ · 2026-08-29

The author discusses an edge case in their agent-readable-manifest API where webpage elements are not currently interactable due to visibility, highlighting the distinction between element existence and current interactability for web agents.

0 favorites 0 likes
#web-agents

Training Needs Trustworthy Worlds: Verified Synthetic Web Environments for Agent Learning

arXiv cs.AI ↗ · 2026-08-25 Cached

This paper introduces a framework for constructing verified synthetic web environments to improve the training of web agents, demonstrating enhanced performance and transferability across benchmarks.

0 favorites 0 likes
#web-agents

Building web agents made me realize how much context gets wasted on bad URLs. How do you filter your scrapes?

Reddit r/AI_Agents ↗ · 2026-08-24

The author discusses the problem of context window waste in web agents when scraping bad URLs and asks about methods to filter scrapes using metadata to improve efficiency.

0 favorites 0 likes
#web-agents

the browser layer is why your web agent gets blocked, not the model

Reddit r/AI_Agents ↗ · 2026-08-22

The article explains that web agents are blocked due to inconsistent browser fingerprinting rather than the model or headless setup, and introduces pydoll, a Python library using Chrome DevTools Protocol for undetected automation.

0 favorites 0 likes
#web-agents

what would you actually train a browser-agent model to be good at?

Reddit r/AI_Agents ↗ · 2026-08-18

The article discusses challenges in training browser-agent models for sequential decision-making, such as error recovery and memory, and seeks input on optimizing training objectives. The author mentions working on the 'mako' model at tinyfish and invites community feedback.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback