Tag
NVIDIA BioNeMo Inference Runtime accelerates biomolecular structure-prediction models using optimized kernels, CUDA Graphs, and Ray for high throughput while maintaining PyTorch workflow integration.
mere.run is a local-first inference runtime for Apple Silicon and headless Linux that provides a single CLI for text, image, video, music, 3D, and more without requiring Python.
Embodied.cpp is a portable C++ inference runtime that enables efficient deployment of vision-language-action and world-action models across heterogeneous edge devices and robots through modular execution layers and optimized inference.
Conifer launches as an open-source local AI runtime and IDE, offering native inference on Mac, Linux, and Windows with its own Metal engine and a sandboxed local agent called Typhoon.