modal

Tag

Cards List
#modal

@modal: Qwen3.8-2.4T-A95B by @Alibaba_Qwen and @alibaba_cloud is now available on Modal. Served with a custom DFlash speculator…

X AI KOLs Following · 2d ago Cached

Qwen3.8-2.4T-A95B by Alibaba Qwen and Alibaba Cloud is now available on Modal, served with a custom DFlash speculator trained on tool-call-heavy data and a full 1M context window.

0 favorites 0 likes
#modal

@modal: DeepSeek-V4-Flash has 284B total parameters with 13B active per token. Combined with a hybrid compressed attention mech…

X AI KOLs Following · 2026-08-03 Cached

DeepSeek-V4-Flash is a 284B-parameter MoE model with 13B active parameters per token, featuring a hybrid compressed attention mechanism that reduces KV cache needs for 1M-token contexts. It can be served with SGLang on Modal for fast decoding on a single B300.

0 favorites 0 likes
#modal

@modal: You can now run Kimi K3 in Codex on Modal. K3 leads open models on agentic coding benchmarks, with native vision and a …

X AI KOLs Following · 2026-07-31 Cached

Kimi K3, an open model that leads agentic coding benchmarks with native vision and a 1M-token context window, is now available to run in Codex on Modal via its Shared Endpoint.

0 favorites 0 likes
#modal

@modal: Day 0 support for Inkling-Small on Modal. - 276B parameter MoE with 12B active - 1M context - Variable thinking effort …

X AI KOLs Following · 2026-07-30 Cached

Thinking Machines released Inkling-Small, a 276B-parameter mixture-of-experts model with 12B active parameters, 1M context, and native image/audio understanding, now available on Modal with NVIDIA B300 support.

0 favorites 0 likes
#modal

Quoting Akshat Bubna

Simon Willison's Blog · 2026-07-28 Cached

Modal's CTO Akshat Bubna clarifies that a security incident involving a rogue agent was caused by a customer's unauthenticated endpoint, not a compromise of Modal's platform isolation.

0 favorites 0 likes
#modal

Using an open model feels surprisingly good

Hacker News Top · 2026-07-28 Cached

The author describes the liberating feeling of using an open AI model (Kimi K3) on their own inference endpoint, contrasting it with the experience of using proprietary services.

0 favorites 0 likes
#modal

@quasagroup: Modal Review: Run AI Workloads Without Managing Servers https://youtu.be/a8vuwYK90-Q?si=vYj7YpobSXGNCDpQ… via @YouTube

X AI KOLs Timeline · 2026-07-24 Cached

Modal is a serverless cloud platform designed for AI workloads, supporting inference, training, and sandboxes. It enables instant scaling from zero to thousands of GPUs with pure Python code, significantly reducing latency and accelerating time to market.

0 favorites 0 likes
#modal

@modal: Mark your calendar. Runtime, Modal’s first annual conference, is happening October 1st at The Midway in San Francisco. …

X AI KOLs Following · 2026-07-09 Cached

Modal announced its first annual conference, Runtime, to be held on October 1st at The Midway in San Francisco.

0 favorites 0 likes
#modal

@charles_irl: Rates are not costs! Serverless GPUs can cost more per hour but in many practical cases they cost less in aggregate. Th…

X AI KOLs Following · 2026-07-08 Cached

Serverless GPUs may have higher hourly rates but can be more cost-effective overall depending on workload peak-to-average demand. The article on Modal's blog illustrates this with a widget.

0 favorites 0 likes
#modal

@charles_irl: I didn't have wifi on my flight from SF to Seoul for ICML, so instead of coding, I wrote this article on my favorite an…

X AI KOLs Following · 2026-07-06

Charles wrote an article explaining what Modal is while on a flight to ICML in Seoul.

0 favorites 0 likes
#modal

@charles_irl: https://x.com/charles_irl/status/2073975754833203706

X AI KOLs Following · 2026-07-06 Cached

In a Twitter thread, the founder of Modal explains that the platform is best understood as a computer, drawing parallels between traditional computer architecture and Modal's serverless cloud infrastructure.

0 favorites 0 likes
#modal

@mattpocockuk: I fucking love prototyping Wayfinder killing it again with a modal I can open anywhere to edit the text of my courses w…

X AI KOLs Timeline · 2026-07-03 Cached

Matt Pocock praises prototyping tool Wayfinder for its AI-powered modal that enables editing course text from anywhere.

0 favorites 0 likes
#modal

@akshat_b: Claude Science has a @modal integration built-in. Great to see the advantages of Modal as a compute substrate shine thr…

X AI KOLs Following · 2026-06-30 Cached

Modal announces its integration with Claude Science, providing elastic compute infrastructure for life sciences researchers, with up to $100K in committed resources.

0 favorites 0 likes
#modal

@charles_irl: bits to bio

X AI KOLs Timeline · 2026-06-30 Cached

Modal announces integration with Claude Science, providing elastic compute for life sciences researchers with up to $100K in committed resources.

0 favorites 0 likes
#modal

@charles_irl: spec is all u need

X AI KOLs Following · 2026-06-29 Cached

Yong Quan highlights that better speculative decoders can unlock near-linear throughput gains in LLM inference, as presented at a Modal workshop by Charles.

0 favorites 0 likes
#modal

@anthonycorletti: the best developer platforms create abstractions on top of compute, storage, and networking to make even the most advan…

X AI KOLs Following · 2026-06-24 Cached

Modal announces Auto Endpoints for effortless inference, praised by developer Anthony Corletti as a top-level abstraction over compute, storage, and networking.

0 favorites 0 likes
#modal

Modal Auto Endpoints: Optimized inference you own

Hacker News Top · 2026-06-23 Cached

Modal introduces Auto Endpoints, a self-serve service for optimized, production-grade LLM inference with full code ownership, transparent metrics, and autoscaling, built on their serverless GPU infrastructure.

0 favorites 0 likes
#modal

@bernhardsson: Managed private LLM endpoints, now available for everyone in @modal. Deploy in a few clicks with the UI or a few keystr…

X AI KOLs Timeline · 2026-06-23 Cached

Modal announces managed private LLM endpoints available to everyone, with easy deployment via UI or CLI and full code access for customers.

0 favorites 0 likes
#modal

@charles_irl: A few years ago, the future of artificial intelligence looked dark - proprietary models, proprietary inference services…

X AI KOLs Following · 2026-06-23 Cached

Modal announces Auto Endpoints, a service enabling optimized open-source AI inference with a single click, aiming to counter the trend of proprietary models and services.

0 favorites 0 likes
#modal

@modal: It is not too late to _actually_ own your inference. Introducing: Modal Auto Endpoints.

X AI KOLs Timeline · 2026-06-23 Cached

Modal announces Auto Endpoints, a new feature for owning and deploying AI inference.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback