Tag
Introduces Google Cloud's Model Armor tool, used in multi-agent systems to detect and redact sensitive data, prevent indirect prompt injection, support partial redaction and reversible hashing, and centrally manage security policies.
Google Cloud Tech session on securing multi-agent LLM systems using defense-in-depth strategies, including Sensitive Data Protection and Model Armor to prevent prompt injections and data leaks.
Google Cloud's annualized revenue reaches approximately $99 billion, up 82% year-over-year, with a sharp increase in growth rate.
IneffableLabs received the first Vera Rubin NVL72 clusters from Google Cloud and NVIDIA, marking a leap in AI hardware.
Google Cloud published 101 architectural blueprints for real-world generative AI use cases across industries, providing tech stacks and design patterns to help developers get started.
Google Cloud details how they optimized Qwen 3.5-397B MoE on Ironwood TPUs using a modular, model-agnostic engineering playbook, achieving 3.1× decode and 4.7× prefill performance gains.
Google launched the TPU Developer Hub, a centralized resource with documentation and framework recipes for building, training, and serving AI on Google Cloud TPUs, supporting JAX, PyTorch, and vLLM.
Sharing 5 key takeaways from Google Cloud's multi-tenant agent AI system reference architecture, which is inspiring for indie developers and small teams to productionize Agents.
Google Data Cloud's frontier AI team discusses a new approach to evaluating AI agents using information theory to create a meta-benchmark called Discovery Bench that measures how vague a query can be before an agent fails, providing a more nuanced map of agent capabilities than simple pass/fail exams.
Google publicly launches Cloud Run sandboxes, showcasing the ability to start, execute, and stop 1,000 sandboxes in 5 seconds with an average latency of 500ms.
Former OpenAI engineer Ryan Lopopolo joins Google Cloud as Chief Agent Engineer, bringing OpenAI's Agent engineering methodology to the cloud platform.
Apple now shows a popup notifying users that AI features in iWork and Freeform send data to Google Cloud. The company also revealed its top-tier AI model, AFM Cloud Pro, runs on Nvidia Blackwell GPUs hosted in Google Cloud, marking a departure from its previous privacy promises of keeping AI data on Apple silicon in Apple's own data centers.
Harrison Chase asks about the Open Knowledge Format from Google Cloud's Knowledge Catalog, an AI-powered data catalog and metadata management platform.
Google Cloud announces a new open-source VS Code extension that lets users connect to and run notebooks on Cloud Workbench Instances directly from their local IDE.
Google Cloud will offer SandboxAQ's large quantitative models (trained on scientific equations and lab data) for drug discovery, materials science, and semiconductor manufacturing, alongside Gemini for Science tools to accelerate research workflows.
Two senior Google Cloud engineers gave a live lecture on building multi-agent workflows, discussing Google's open-sourced production multi-agent system.
A user shares that OpenAI's Codex can autonomously manage Google Cloud settings using its in-app browser, praising the seamless automation.
Ray Serve LLM achieves 4.4x and 24.8x throughput improvements on prefill- and decode-heavy workloads via direct streaming, a new vLLM V2 executor backend, and HAProxy ingress, now available in Ray 2.56 in partnership with Google Cloud and vLLM.
Google Cloud introduces the Open Knowledge Format (OKF), an open specification for representing metadata and curated knowledge in markdown files to improve data sharing and context for AI agents. The format aims to make knowledge from fragmented internal systems portable and interoperable.
Google Cloud introduces the Open Knowledge Format (OKF), an open specification that standardizes the LLM-wiki pattern for representing structured knowledge in markdown with YAML frontmatter, aiming to improve data sharing and interoperability for AI agents.