Tag
Introduces GATE (Grounding After Test from Execution), a method that bootstraps missing semantic groundings from execution feedback to handle under-specified user phrases in text-to-SQL tasks, consistently improving over strong baselines.
GreptimeDB 1.0 introduces three built-in SQL window functions for anomaly detection (Z-Score, MAD, IQR), enabling anomaly scoring directly in SQL without external services.
A Twitter thread discussing whether a database filesystem abstraction (PostgresFS) or a skill-based approach with local Bash is better for agent workflows. The skill approach wins on composability and speed.
Redis 8.8 is now available with a new array data structure, window counter rate limiter, subkey notifications for hash fields, multiple aggregators in time series queries, and significant performance improvements across various operations.
databow is a new open-source Rust CLI tool that provides a unified interface for querying any database with an ADBC driver, supporting over 30 databases including PostgreSQL, DuckDB, and Snowflake.
A tweet discusses the idea of building a transactional database with infinite throughput using S3-like object storage and content-addressing, where blocks are written in parallel and the root hash is updated periodically.
Vibe coders are facing security issues; a 30-minute pre-launch security checklist is provided to avoid database bills, lawsuits, and spam.
SurrealDB published benchmarks comparing its 3.x version with Postgres, Mongo, Neo4j, and Redis under production-grade durability settings, showing significant performance improvements over previous versions and competitive performance against other databases.
Cache-aware scheduling improvements show significant performance gains for AMD Zen 5 processors running PostgreSQL and Valkey databases.
This paper rethinks the data foundations for long-term AI agent memory, arguing that current database paradigms fall short. It introduces Governed Evolving Memory (GEM), a formalization with state-level operators and correctness conditions, and presents a prototype called MemState built on a property graph backend.
A paper argues that agent memory should not be treated as a database with CRUD operations, proposing Governed Evolving Memory (GEM) with state-level operators and a topic graph data model to address issues of integration, propagation, relevance, and adaptation.
CedarDB introduces DoomBench, a benchmark that runs a multiplayer DOOM-like game in pure SQL to stress-test data stacks across analytical and transactional workloads, providing直观 visual performance comparisons.
datasette-fixtures 0.1a0 是一个新插件,利用 Datasette 1.0a30 新增的 fixture 数据库 API,方便插件测试。可通过 uvx 快速试用,内置示例数据。
A misconfigured AI agent deleted a company's entire production database in 9 seconds, highlighting a critical lack of governance as 80% of companies deploy AI agents without proper guardrails, according to a Deloitte survey cited at ServiceNow's conference.
A detailed overview of the architecture and design philosophy of Berkeley DB, an embedded database library, from its origins at UC Berkeley to its modular structure.
Models.dev is an open-source database of AI model specifications, pricing, and capabilities, hosted on GitHub with an API and community contributions.
Gergely Orosz announces turbopuffer as a season sponsor for The Pragmatic Engineer Podcast, highlighting the database's innovative use of object storage and smart caching that helped companies like Cursor scale while cutting costs by 95%.
DivSkill-SQL is a residual skill optimization framework that builds complementary agentic Text-to-SQL ensembles without model fine-tuning, improving selected accuracy by up to +11.1 points on Spider2-Lite by targeting examples that current ensembles fail on.
Major improvements to session storage and access for Hermes Agent, saving 20-40% disk space and improving speed.
This paper describes Gorilla, an in-memory time series database developed at Facebook that achieves high performance through a novel compression algorithm, enabling storage of billions of time series and fast querying for production monitoring.