Tag
A developer introduces TMDD/TTMDD, a pattern where code measures its own time and tokens to distinguish a working AI agent from a hung one.
agentglass is an open-source local dashboard that provides real-time monitoring, cost tracking, and fleet management for multiple AI coding agents like Claude Code, Codex, and Gemini.
Dashy is an open-source, customizable homepage dashboard for homelabs with pre-built widgets, real-time status checks, and multi-user auth, all running in a single Docker container.
Due to recent escalations with the Iran Conflict, we are reintroducing a Custom Timeline to monitor the situation.
Screenpipe is an open-source, local-first tool that records your work and converts it into searchable memory and SOPs for AI agents, enabling personalized automation.
DrDroid's Alert Grouping feature helps reduce alert noise, improving monitoring efficiency.
A discussion asking how developers handle changes in tools, APIs, or model versions that their AI agents depend on in production, including detection, fixes, and costs.
The paper examines how length penalties applied during chain-of-thought reasoning can reduce the ability to monitor the reasoning process, raising concerns for interpretability and alignment.
Discusses the phenomenon where AI agents appear to succeed at tasks but later reveal failures, highlighting challenges in agent evaluation and monitoring.
A developer details Observation-Oriented Programming (OOP), a paradigm for organizing a swarm of terminal AI agents around watching observability data rather than task execution. The system uses tmux, cron, bash, and an append-only log file, with agents acting as observers that correlate signals and materialize actions only after interpretation.
The author details overhauling their homelab setup, moving from a Zimaboard with NixOS to a new Debian-based server with Dokploy, Jellyfin, Immich, and other self-hosted services, including improvements to home networking with fiber.
A discussion on the lack of processes for retiring AI agents, focusing on how to decide when to shut down an agent, track usage, and who should make the kill call.
reaction is a daemon that monitors program outputs for repeated patterns and triggers predefined actions, useful for automation and alerting.
A tool to detect discrepancies between an AI agent's claimed actions and its actual behavior.
Un utilisateur a créé un SaaS personnel (Sillage et Mogwaï) avec Claude Code et Screenpipe pour surveiller 14 flux numériques et un agent proactif qui l'interrompt seulement dans 94 situations précises, illustrant une application avancée de l'IA pour la mémoire et la productivité personnelle.
A curated list of open source libraries for deploying, monitoring, versioning, scaling, and securing production machine learning systems.
An open-source companion to the LLM Engineer's Handbook that provides a complete blueprint for building production-ready LLM systems, covering synthetic data generation, training (including DPO), RAG, deployment on AWS, evaluation, and monitoring.
Trovis is a plugin for OpenClaw that records all agent actions, costs, and deviations, presenting them in plain English for better observability.
A newly installed monitoring system on the seafloor of the Amsterdam-Saint Paul Plateau captured a sudden burst of crustal spreading, revealing that most extension occurred in a short time window and some events lacked seismic activity.
Routebase is a tool to detect API drift before it affects customers.