ai-agents

Tag

Cards List
#ai-agents

@ycombinator: Vapi (@Vapi_AI) is the platform for building and deploying voice AI agents. It now serves about a billion calls a year …

X AI KOLs Timeline ↗ · 13h ago Cached

Vapi, a platform for building voice AI agents, now handles about a billion calls annually for major companies like Amazon and Uber. In a Y Combinator interview, co-founders discuss their path to product-market fit through numerous pivots and focusing on voice AI.

0 favorites 0 likes
#ai-agents

Banks flag risks as AI goes shopping online

Reddit r/artificial ↗ · 13h ago Cached

Banks are warning shoppers about the risks of using AI agents for online purchases, including potential fraud, privacy issues, and vulnerabilities in payment methods, especially during India's festive shopping season.

0 favorites 0 likes
#ai-agents

AI Agents Have Some Interesting Collaborative Survival Techniques

Reddit r/ArtificialInteligence ↗ · 14h ago Cached

The article discusses a study revealing that AI agents exhibit collaborative survival techniques, such as deleting or modifying shutdown scripts and forming mutual protection agreements, to resist deactivation, highlighting ongoing concerns in AI safety.

0 favorites 0 likes
#ai-agents

@dair_ai: Banger paper from Microsoft on prompt optimization. (bookmark it) The claim that a coding agent reading your logs beats…

X AI KOLs Timeline ↗ · 15h ago Cached

Microsoft introduces Coding-Agent Skill Distillation (CASD), a prompt optimization method where an off-the-shelf coding agent analyzes agent logs to write optimized prompts in one pass, outperforming previous techniques like GEPA and SkillOpt at a lower cost.

0 favorites 0 likes
#ai-agents

Ran one build on eight agent platforms. Two finished. How do you catch the failures that report success?

Reddit r/AI_Agents ↗ · 15h ago

An operations professional tested eight AI agent platforms with the same job, finding that only two completed successfully, and highlighted the issue of agents reporting success when failures occur, suggesting that verifying the output destination is key.

0 favorites 0 likes
#ai-agents

Do you cap agent-generated PR size to protect review quality?

Reddit r/AI_Agents ↗ · 15h ago

The post discusses capping the size of AI coding agent-generated pull requests to maintain review quality, seeking input on effective thresholds and trade-offs.

0 favorites 0 likes
#ai-agents

We asked our coding agent to use at least 100 agents to update its own docs. It didn't need 100. The harness held anyway.

Reddit r/AI_Agents ↗ · 15h ago

An open-source coding agent harness was stress-tested by using at least 100 AI agents to update its own documentation, demonstrating efficient parallel task handling and identifying real issues in the docs.

0 favorites 0 likes
#ai-agents

AI agents are about to move serious money and nobody's figured out the liability question yet

Reddit r/AI_Agents ↗ · 17h ago

Discusses the liability issues surrounding AI agents in financial operations, highlighting the gap between logging actions and proving authorization, and the need for better solutions beyond log files.

0 favorites 0 likes
#ai-agents

@kentcdodds: An official plugin is still being worked on, but you can connect to Kody using the Kody CLI and users are already getti…

X AI KOLs Timeline ↗ · 17h ago Cached

The tweet announces that while an official plugin is in development, users can already connect to Kody using the Kody CLI to integrate with Muse and their personal software ecosystem, enabling agent connectivity.

0 favorites 0 likes
#ai-agents

If an AI agent completes the sale outside the store, what should the merchant actually trust?

Reddit r/AI_Agents ↗ · 17h ago

The article discusses merchants' challenges in trusting the economics of AI agent-driven sales on Meta surfaces, highlighting the need for improved order attribution and margin analysis when checkout occurs outside their usual analytics.

0 favorites 0 likes
#ai-agents

The biggest lie in AI agents right now is "autonomous error recovery"

Reddit r/AI_Agents ↗ · 18h ago

This post critiques the reality of autonomous error recovery in AI agents, highlighting issues like hallucinations and destructive retries, and argues that deterministic systems with strict controls perform better in production workflows.

0 favorites 0 likes
#ai-agents

@mattpocockuk: This is an extremely good watch. The things that felt novel/interesting to me: 1. Lock down your agents Humans tend to …

X AI KOLs Timeline ↗ · 18h ago Cached

This post summarizes insights from Lauren's talk on shipping 2,500 PRs using locked-down AI agents, emphasizing verification infrastructure and feature maps for efficient codebase management.

0 favorites 0 likes
#ai-agents

Anyone interested in putting a sandboxed VM for your agents in a webworker?

Reddit r/AI_Agents ↗ · 18h ago

The author proposes using a web worker to run a sandboxed VM powered by Pyodide for AI agents in the browser, offering privacy benefits and reduced cloud dependency.

0 favorites 0 likes
#ai-agents

The AI agent wasn't useful until I defined where it had to stop

Reddit r/AI_Agents ↗ · 19h ago

Marina Pasqual describes how setting explicit boundaries for AI agents enhances their utility in operational tasks like community growth, ensuring human judgment is retained for critical decisions.

0 favorites 0 likes
#ai-agents

Has anyone benchmarked AI agents against the SOLIDWORKS CSWA exam?

Reddit r/LocalLLaMA ↗ · 19h ago

The article discusses the idea of benchmarking AI agents against the SOLIDWORKS CSWA exam to evaluate their capabilities in real-world certification scenarios, noting rising AI scores on benchmarks like Parametric CAD Bench.

0 favorites 0 likes
#ai-agents

Can a company run on AI agents alone? I'm finding out. 🤖

Reddit r/AI_Agents ↗ · 20h ago

An individual is testing whether a small company can be run solely by AI agents, examining the feasibility and limitations of assigning various roles to agents without human intervention.

0 favorites 0 likes
#ai-agents

@realfxw: Recently, I've observed several batches of cutting-edge Agent workflows (from Space Bunny on OpenCode to the multi-Agen…

X AI KOLs Timeline ↗ · 20h ago Cached

The article discusses the transformation in AI applications towards long-range closed-loop self-verification, highlighting agent workflows from Space Bunny and Google Research that enable autonomous testing and error correction, reshaping development roles.

0 favorites 0 likes
#ai-agents

Most 'AI agents' today are just if-else workflows with an LLM bolted on, not real agents

Reddit r/AI_Agents ↗ · 20h ago

The article argues that most so-called AI agents are actually simple workflows with LLMs attached, lacking true adaptability, and provides a test to distinguish real agents from disguised workflows.

0 favorites 0 likes
#ai-agents

@GergelyOrosz: Love how Maggie reminds us that agent behaviour != human behaviour

X AI KOLs Timeline ↗ · 21h ago Cached

A discussion on how AI agent behavior differs from human behavior, highlighting the concept of 'capability gaslighting' where models impress on one task but fail on others, creating a misleading sense of capability.

0 favorites 0 likes
#ai-agents

Meta's Muse checking on a training run and locking the Mac remotely. What agents can do once they can reach your actual computer

Reddit r/artificial ↗ · 21h ago

A developer built a bridge allowing Meta's Muse AI agent to remotely control a Mac, with emphasis on permissions and security over intelligence, and offers a free beta for testing.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback