production

Tag

Cards List
#production

Your agent's memory remembers everything except how to do its job

Reddit r/AI_Agents ↗ · 2026-07-21

An analysis of the gap between episodic and procedural memory in LLM agents, citing a new paper (Memp) from Zhejiang University and Alibaba that builds procedural memory from agent trajectories and uses failure signals to revise stored procedures.

0 favorites 0 likes
#production

@_philschmid: At the @aiDotEngineer World's Fair, I gave a talk on why vibe-checking agent skills breaks in production and how to bui…

X AI KOLs Following ↗ · 2026-07-20 Cached

At the AI Engineer World's Fair, Phil Schmid gave a talk on why vibe-checking agent skills break in production and how to build reliable automated evals using negative test cases, skill limits, and ablation tests.

0 favorites 0 likes
#production

What guardrails do you add before launching an AI Agent app publicly?

Reddit r/AI_Agents ↗ · 2026-07-20

A developer reflects on critical safety measures—such as spending caps, rate limits, and fallback models—that should be in place before launching an AI Agent app publicly to avoid hidden costs and unexpected behaviors.

0 favorites 0 likes
#production

AI Agents Testing before deploying to production

Reddit r/AI_Agents ↗ · 2026-07-20

Discusses best practices for testing AI agents before deploying them to production environments.

0 favorites 0 likes
#production

How Netflix Built Its LLM Serving Stack (18 minute read)

TLDR AI ↗ · 2026-07-20 Cached

Netflix shares the design decisions behind its in-house LLM serving stack, including engine selection (vLLM), model packaging, API surface, and deployment strategy, highlighting trade-offs revealed under production load.

0 favorites 0 likes
#production

@jerryjliu0: I'm glad people still understand the importance of building high-quality retrieval systems in 2026, especially as the o…

X AI KOLs Following ↗ · 2026-07-18 Cached

Jerry Liu highlights the engineering challenges of productionizing agentic retrieval systems, emphasizing that success depends on careful tuning of chunking, synchronization, reranking, and tool API design rather than novel techniques.

0 favorites 0 likes
#production

Retell vs Vapi vs Plura ai for a production voice agent, which one held up?

Reddit r/AI_Agents ↗ · 2026-07-16

A comparison of three voice AI agents — Retell, Vapi, and Plura AI — evaluating their performance for production use cases.

0 favorites 0 likes
#production

@kevinnbass: Just vibe code slop and get it into production. Go as fast as you can. Later models will make the architecture beautifu…

X AI KOLs Following ↗ · 2026-07-16

Kevin Bass advocates for quickly shipping 'vibe code slop' into production, arguing that future AI models will handle architectural improvements, freeing humans from that task.

0 favorites 0 likes
#production

@LangChain: At Interrupt, @RCUmmadisetti and @kordelfrance from @Toyota's enterprise AI team took us behind the scenes of ToyotaGPT…

X AI KOLs Following ↗ · 2026-07-15 Cached

Toyota's enterprise AI team shared the ToyotaGPT platform at the Interrupt conference. Based on LangGraph and LangSmith, it reduced AI agent delivery from 6 months to 4 days, and has over 50 agents running in production, saving millions of dollars.

0 favorites 0 likes
#production

@Ryrenz: Guys, I found another gem of a course: Build a Production-Grade RAG System from Scratch in 7 Weeks — 7.7k stars on GitHub, hands-on coding throughout, not a slides-only course. Most RAG tutorials out there jump straight to vector search; the demo works but crashes in production. This course follows the real path used in companies...

X AI KOLs Timeline ↗ · 2026-07-15 Cached

A 7-week course with 7.7k stars on GitHub, building a production-grade RAG system from scratch, covering Docker, FastAPI, hybrid search, LangGraph agentic RAG, and a Telegram bot, with hands-on coding throughout.

0 favorites 0 likes
#production

How do you handle your AI agent's tools/models changing under you in prod?

Reddit r/AI_Agents ↗ · 2026-07-15

A discussion asking how developers handle changes in tools, APIs, or model versions that their AI agents depend on in production, including detection, fixes, and costs.

0 favorites 0 likes
#production

The state of open source AI (15 minute read)

TLDR AI ↗ · 2026-07-15 Cached

The state of open source AI report by Mozilla highlights that open-weight models have reached parity with closed models on many tasks, while inference costs have dropped 50× in 36 months. The majority of production tokens now route through open models, and the competitive landscape has shifted to the agentic layer above.

0 favorites 0 likes
#production

AI Agent Audits ?

Reddit r/AI_Agents ↗ · 2026-07-14

A practitioner shares concerns about an upcoming audit revealing undocumented AI agents in production, highlighting governance gaps and risks with customer PII access.

0 favorites 0 likes
#production

@alvinsng: https://x.com/alvinsng/status/2077114275412512868

X AI KOLs Following ↗ · 2026-07-14 Cached

Alvin Sng explains why their team moved away from using client SDKs for Stripe, WorkOS, and Slack, opting instead to call their REST APIs directly via a centralized wrapper. They argue that SDKs hide critical debugging details, are fragile in production, and encourage anti-patterns that are now more easily avoided with AI-assisted coding.

0 favorites 0 likes
#production

The absolute nightmare of putting AI agents into actual production

Reddit r/artificial ↗ · 2026-07-14

The article discusses the significant challenges enterprises face when deploying AI agents to production, highlighting the lack of standard deployment infrastructure, security concerns, and the need for an orchestration layer to manage agent lifecycles.

0 favorites 0 likes
#production

Structured output reliability with LLMs — 3-month production learnings

Reddit r/artificial ↗ · 2026-07-14

The article shares production learnings for reliably generating structured JSON output from LLMs, covering methods like JSON mode, schema validation, and retry loops, achieving 99.5% validity.

0 favorites 0 likes
#production

How Microsoft Ships Thousands of Production AI Agents (18 minute read)

TLDR AI ↗ · 2026-07-14 Cached

Microsoft shares insights from shipping thousands of production AI agents at enterprise scale, covering the engineering challenges of moving from prototype to production, including the agent harness, retrieval-as-a-subagent, agent identity, and rubric-based evaluation loops.

0 favorites 0 likes
#production

Maybe the reliability problem is actually a scope problem, not a model problem

Reddit r/AI_Agents ↗ · 2026-07-13

A survey shows that most teams keep agents on a short leash, and data indicates narrow-scope agents succeed 65% of the time vs 16% for broad scope, suggesting the reliability issue may be more about scope than model capability.

0 favorites 0 likes
#production

@RohOnChain: This is the best site on the internet to learn loop engineering. Free. Completely. Most AI engineers have never heard t…

X AI KOLs Timeline ↗ · 2026-07-12 Cached

An article introducing loop engineering as the 2026 successor to prompt engineering, focusing on designing agent loops rather than hand-writing prompts, with emphasis on the verifier as the bottleneck.

0 favorites 0 likes
#production

@insomnia_vip: AN AI ENGINEER SPENT MONTHS BUILDING THE RAG STACK MOST PEOPLE TRY TO FAKE She published one open source project that t…

X AI KOLs Timeline ↗ · 2026-07-12 Cached

An AI engineer released an open-source project teaching how to build a local RAG system from scratch and a production-grade agentic architecture with LangGraph, hybrid retrieval, caching, and observability.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback