Tag
Browser Use claims its V4 system is better than humans at browsing the web.
Nous Research's Hermes introduces Browser Use mode, replacing twelve browser tools with a single script-driven approach using browser_use CLI 3.0, cutting token use by 48-66% with no accuracy drop.
An AI agent powered by Opus 5 and browser_use is given $1000 in credits and real money to act autonomously on the real internet, broadcast live.
Browser Use Cloud v4 introduces a web agent platform that claims to solve accuracy and cost constraints, making reliable web agents at scale feasible, with $15 free credits offered to try.
browser-use announces a new agent powered by GPT-5.6 Luna, which can read top Hacker News posts and generate a full report for only 3 cents, highlighting the low cost of AI-driven web browsing.
A technical post describing how to add computer and browser use verification to AI agents, enabling bug reproduction and feature verification in a cloud software factory using Warp and a new verify-behavior skill.
Qwen-UI-Agent is a new foundation GUI agent from Alibaba's Qwen team that handles mobile, computer, web, and DeepSearch tasks with state-of-the-art performance on mobile-use benchmarks and competitive results on computer/browser tasks, combining GUI and CLI actions in a unified action space.
A developer benchmarks Rote, a memory manager for browser agents that sends page diffs instead of full re-renders, showing a 37% reduction in token growth compared to Browser Use, but with trade-offs on short tasks.
Browser Use announced the winner of its Game Day contest: SUPERINTELLIGENCE, built with Kimi K3. The winner received $500 credits, and Kimi K3 games will be free for a week.
Browser-use announces an AI agent that can build web games, supporting 2D and 3D, online multiplayer, and models like Kimi 3 and Fable 5; they are offering a free game mode for feedback and $50 credits to first 20 respondents.
The article argues that computer-use/browser-use AI capabilities are progressing very quickly and will agentify the web, as most of the web lacks APIs.
Announced new infrastructure for browser automation, likely a tool or framework update.
The team behind browser-use celebrates 100k GitHub stars with a giveaway where their AI agent can purchase items for users, up to $100k total.
A discussion on the current state of AI browsers, noting that while the category isn't dead, high switching costs require a 100x killer feature, and the rise of agentic capabilities in ChatGPT and Claude has eroded the initial value proposition of AI browsers.
Browser Use has released its CLI 3.0, which can be integrated as a skill into Claude Code and Codex, enabling them to control browsers. The new version is 6x smaller, consumes fewer tokens, supports direct CDP control, and features self-evolution and on-the-fly function writing.
Introducing Browser Use CLI 3.0, which turns any model into a SOTA browser agent with direct CDP control, 6x smaller and fewer tokens.
Opus 4.7 and GLM 5.2 are being benchmarked on frontend design using Browser Use v4; results are shared via a link.
Browser Use v4 introduces a QA skill that allows your agent to test flows, catch bugs, and evaluate UI by clicking around as a user, closing the feedback loop for developers.
Fara1.5 is a family of native computer use agents trained using the FaraGen1.5 scalable data pipeline. The models achieve new state-of-the-art results on browser-use benchmarks, competing with much larger frontier models.
BrowserCode achieves #1 spot on Odysseys benchmark for long-horizon web agents, demonstrating strong performance in multi-hour web workflows.