@asterailabs: Introducing Aster Inference -- The world's fastest inference API created by AI research agents We serve the world's fas…

X AI KOLs Timeline Products

Summary

Aster Labs launches Aster Inference, claiming the world's fastest inference API using AI research agents, with benchmark speeds for models like OpenAI's gpt-oss-120b and GLM 5.2.

Introducing Aster Inference -- The world's fastest inference API created by AI research agents We serve the world's fastest inference on GPU: - OpenAI's gpt-oss-120b @ 644 tps - http://Z.ai's GLM 5.2 @ 281 tps At Aster, we're automating open-ended research, and we use inference optimization as a task to benchmark our agents against. We're creating a product out of the inference discoveries made from our system. As our agents discover more, we plan to further improve our inference product and ship new, SOTA AI products.
Original Article
View Cached Full Text

Cached at: 07/16/26, 02:19 PM

Introducing Aster Inference – The world’s fastest inference API created by AI research agents

We serve the world’s fastest inference on GPU:

  • OpenAI’s gpt-oss-120b @ 644 tps
  • http://Z.ai’s GLM 5.2 @ 281 tps

At Aster, we’re automating open-ended research, and we use inference optimization as a task to benchmark our agents against. We’re creating a product out of the inference discoveries made from our system.

As our agents discover more, we plan to further improve our inference product and ship new, SOTA AI products.

Similar Articles