production-deployment

Tag

Cards List
#production-deployment

I evaluated 7 production agent runtimes against 7 criteria: strengths, trade-offs, and who each one is actually for

Reddit r/AI_Agents ↗ · 2026-06-11

An evaluation of seven production agent runtimes (Cloudflare Agents, AWS Bedrock AgentCore, Google AX, Anthropic Claude Managed Agents, kagent, Vercel Open Agents, and Agyn) against seven criteria including self-hostability, multi-vendor support, isolation, and credential security, highlighting trade-offs and best-fit use cases.

0 favorites 0 likes
#production-deployment

How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions

arXiv cs.AI ↗ · 2026-06-09 Cached

This paper presents a deployment-focused study comparing LoRA fine-tuning of 24 model variants (270M–8B parameters) for merchant information extraction from financial transaction strings. The authors find that smaller models like Qwen 3.5 4B achieve 96.6% F1, within 0.35 points of the 8B baseline, while offering significant reductions in latency and cost.

0 favorites 0 likes
#production-deployment

@sydneyrunkle: here's a quick overview of a) what is deepagents b) what makes deepagents good at complex tasks c) how to easily take o…

X AI KOLs Following ↗ · 2026-06-01 Cached

DeepAgents is a customizable AI agent framework designed for complex real-world tasks. It features execution environments, context management, delegation, and human-in-the-loop capabilities, and offers a hosted version for production-level deployment.

0 favorites 0 likes
#production-deployment

why AI agent pilots feel amazing but production deployment turns into a mess

Reddit r/AI_Agents ↗ · 2026-05-31

The author shares experiences moving AI agent systems from sandbox to production, highlighting how human roles become ambiguous and teams disengage when agents execute tasks, leading to operational failures.

0 favorites 0 likes
#production-deployment

@nini_incrypto_: Want to learn AI system design? Just look at the real-world experience of top-tier companies! This amazing repository on GitHub aggregates over 500 real GenAI deployment cases from more than 130 big companies. It doesn't teach basic textbook theory but specifically breaks down the technical decisions of top teams in real production environments: 1. Uber: …

X AI KOLs Timeline ↗ · 2026-05-29

A repository on GitHub aggregates over 500 real GenAI deployment cases from more than 130 big companies, breaking down top teams' technical decisions in production environments, such as Uber's real-time traffic scheduling across multiple model providers.

0 favorites 0 likes
#production-deployment

how to scale AI agents in production workflows when the underlying business process is broken?

Reddit r/AI_Agents ↗ · 2026-05-26

A practitioner shares challenges scaling multi-agent AI systems in production, including dealing with shadow workflows (undocumented Slack threads and spreadsheets), context loss across different systems (ERP to CRM), and cross-departmental ownership issues. They seek advice from others who have navigated these real-world problems.

0 favorites 0 likes
#production-deployment

Switched our agent stack from Dify to OpenAgent. Here's why we made the call.

Reddit r/AI_Agents ↗ · 2026-05-26

A developer explains why their team switched from Dify and Langflow to OpenAgent for production agent workflows, highlighting OpenAgent's simpler architecture, direct REST/SSE endpoints, built-in prompt versioning, and native Atlas Cloud integration.

0 favorites 0 likes
#production-deployment

@DanKornas: Agent demos are easy. The production stack is the messy part. Awesome Production Agentic Systems is a curated GitHub li…

X AI KOLs Timeline ↗ · 2026-05-25 Cached

A curated GitHub list of open-source libraries for deploying, monitoring, scaling, and securing production agentic systems, organizing the ecosystem into practical sections.

0 favorites 0 likes
#production-deployment

The accountability gap in AI agent deployments is growing faster than the capability gap and nobody's talking about it

Reddit r/ArtificialInteligence ↗ · 2026-05-25

The article highlights the growing accountability gap in AI agent deployments, where audit trails are insufficient, and argues for infrastructure-level execution governance with verifiable records. It mentions W3's solution using Proof of Compute on Avalanche.

0 favorites 0 likes
#production-deployment

Structure-Guided Entity Resolution: Fine-Tuning LLMs for Robust Name Matching in Complex Linguistic Contexts

arXiv cs.CL ↗ · 2026-05-25 Cached

This paper presents Structure-Guided Entity Resolution (SGER), a framework that fine-tunes LLMs through curriculum learning for robust person name matching in linguistically diverse contexts, achieving 99.02% accuracy on Indian identity data and deployed at Dream11.

0 favorites 0 likes
#production-deployment

@no_stp_on_snek: got it here if ya want to try it out:

X AI KOLs Following ↗ · 2026-05-23 Cached

A fork of llama.cpp integrating TurboQuant+ for advanced KV-cache and weight quantization, with cross-backend kernel support (Apple Silicon, NVIDIA CUDA, AMD ROCm, Vulkan) and used in production by LocalAI, Chronara, and AtomicChat.

0 favorites 0 likes
#production-deployment

How are you actually predicting AI costs before they hit your invoice?

Reddit r/AI_Agents ↗ · 2026-05-20

A developer shares the hidden cost variables that cause AI bills to exceed estimates, including reasoning model chain-of-thought tokens, multimodal per-image charges, and function calling system tokens, and asks the community how they predict costs upfront.

0 favorites 0 likes
#production-deployment

Operationalizing Document AI: A Microservice Architecture for OCR and LLM Pipelines in Production

arXiv cs.AI ↗ · 2026-05-20 Cached

This paper presents a microservice architecture for production document AI pipelines that combine classification, OCR, and LLM extraction, sharing design decisions and batch profiling insights that reveal OCR, not LLM parsing, dominates latency.

0 favorites 0 likes
#production-deployment

Single-model AI image detection failed in production. Here’s what 6 models in ensemble actually look like

Reddit r/artificial ↗ · 2026-05-18

A developer shares practical lessons from moving from a single AI image detection model to an ensemble of six models plus non-ML signals in production, highlighting the roles each model plays and the value of disagreement signals. The post also asks the community about retraining cadence and model retirement strategies.

0 favorites 0 likes
#production-deployment

AI memory demos show week one , Production is a month six problem lol

Reddit r/AI_Agents ↗ · 2026-05-18

The article discusses the gap between initial AI memory demos and long-term production challenges, where memory degrades due to contradictions, drift, and outdated preferences, and benchmarks fail to capture these issues.

0 favorites 0 likes
#production-deployment

we gave an AI autonomy over real business decisions with real money for eight months. the thing we learned that surprised us most was not about capability.

Reddit r/ArtificialInteligence ↗ · 2026-05-17

After eight months of real-world deployment, PayWithLocus found that the hardest problem for their autonomous AI system is not capability but confidence: the AI executes confidently wrong decisions in novel situations, highlighting a metacognitive gap that current architectures don't address.

0 favorites 0 likes
#production-deployment

Case Study: Dogfooding a Facebook Agent Before Deploying It to a Realtor

Reddit r/AI_Agents ↗ · 2026-05-17

A real estate firm built an AI agent for Facebook page management, dogfooding it for 10 days on their own page. They share lessons on drift detection, policy-gated runtimes, and API stability, highlighting that production behavior reveals orchestration and integration failures beyond model intelligence.

0 favorites 0 likes
#production-deployment

How much payment authority are people giving their agents in production?

Reddit r/AI_Agents ↗ · 2026-05-14

The article outlines three levels of payment authority given to AI agents in production: query/recommend, limited caps with human review, and broader authority in specific domains, noting that most deployments are still in the first two stages.

0 favorites 0 likes
#production-deployment

The approval queue is the architecture: how I built an autonomous Claude Code agent that runs a real product

Reddit r/AI_Agents ↗ · 2026-05-12

The author details the architecture of 'Aiden,' an autonomous Claude Code agent managing a product called Delegate, emphasizing a human-in-the-loop approval queue system to ensure safety and efficiency in production.

0 favorites 0 likes
#production-deployment

Stop building AI agents.

Reddit r/AI_Agents ↗ · 2026-05-11

The author argues that most founders requesting AI agents actually need straightforward automations with minimal LLM integration, citing production failures, compliance hurdles, and higher ROI from simpler workflows. The piece provides a practical decision framework to help builders and founders prioritize reliable automations over complex, unpredictable agents.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback