serverless-gpu

Tag

Cards List
#serverless-gpu

GPUHedge: Hedging serverless GPU providers improves cold start p95 latency from 117s to 30s [P]

Reddit r/MachineLearning · 2026-07-13

GPUHedge is an open-source tool that uses speculative execution to hedge between serverless GPU providers, reducing cold start p95 latency from 117s to 30s.

0 favorites 0 likes
#serverless-gpu

@charles_irl: Rates are not costs! Serverless GPUs can cost more per hour but in many practical cases they cost less in aggregate. Th…

X AI KOLs Following · 2026-07-08 Cached

Serverless GPUs may have higher hourly rates but can be more cost-effective overall depending on workload peak-to-average demand. The article on Modal's blog illustrates this with a widget.

0 favorites 0 likes
#serverless-gpu

@modal: New replicas of @vllm_project and @sgl_project servers start up 3-10x faster on Modal. Read the article to learn how --…

X AI KOLs Following · 2026-05-12 Cached

Modal has announced that replicas of vLLM and SGLang servers now start up 3-10x faster, leveraging improvements in GPU health management and CUDA context checkpointing.

0 favorites 0 likes
← Back to home

Submit Feedback