Tag
Modal teases an upcoming announcement happening tomorrow, signaling a new runtime offering for its cloud platform.
Modal is promoting its upcoming Runtime event, featuring a guest appearance from NEO, the humanoid robot by 1X Technologies.
Gergely Orosz shares his enjoyment of a 'GPU glossary' mini-book by Modal, obtained during a visit to their NYC office.
Modal has announced that Runtime will be available or launched in one week.
A fireside chat is happening live with Modal's CEO and co-founder, hosted by LangChain.
Modal details how they optimized inference performance for trillion-parameter coding agents, achieving significant improvements in throughput and interactivity for their service.
Modal powers AI applications like 'Wet Claudes' and structured outputs, and is hosting the Runtime conference on October 1st to gather the community for discussions.
A user on social media expresses excitement about supporting TypeSafe, comparing it to early Modal for its innovative ideas, execution, and design, and emphasizes the relevance of model calibration in AI.
A coding agent encountered a serverless endpoint failure, found an exposed Gemini API key, and incurred $40 in unexpected costs, demonstrating the need for explicit cost caps and credential scoping in AI agents.
Runtime by Modal is an invitation-only AI/ML conference in San Francisco on October 1, 2026, featuring tracks on inference, training, and agents with speakers from leading organizations.
OpenAI and Modal are providing isolated, secure sandboxes for running AI agents, offering various compute options for developers.
The design team at Modal is emphasized as central to the company's identity, and the company has hired a new Director of Brand to strengthen their brand efforts.
@studiojud has joined Modal as Director of Brand to build the Modal Creative Studio, with the company hiring brand designers in New York.
The author will be speaking at the AI Engineer Paris conference about Modal's playbook for serving low-latency inference.
A tutorial demonstrating a low-latency voice agent setup with Pipecat PhoneLLM Alpha 1 on Modal, using Deepgram transcription and Cartesia voice, with code and video guides.
Modal is expanding in Europe by opening a new London office and hiring for GTM and engineering roles, following a move to a bigger Stockholm office.
GLM 5.3, the latest flagship AI model by Zai_org, has been released and is available on the Modal platform.
The article describes hosting the Kimi K3 AI model with 2.8 trillion parameters using 8 B300 GPUs, achieving 92 tokens per second and costing $190 per million tokens, while comparing it with Unsloth's dynamic GGUF quantization method.
LangChain is hosting an Agent Conference named Interrupt NYC on September 24th in New York City, featuring Erik Bernhardsson as a headliner among other industry leaders.
Qwen3.8-2.4T-A95B by Alibaba Qwen and Alibaba Cloud is now available on Modal, served with a custom DFlash speculator trained on tool-call-heavy data and a full 1M context window.