ml-inference

Tag

Cards List
#ml-inference

Vendor-agnostic ML inference on production edge devices [R]

Reddit r/MachineLearning · 2026-07-29

Describes using ncnn's Vulkan backend for vendor-agnostic ML inference on production edge devices, achieving 10x speedup over CPU ONNX for face detection and embedding models.

0 favorites 0 likes
#ml-inference

@robertnishihara: I learned recently that @onepot_ai can synthesize and deliver custom molecules in 5 days, which is incredibly fast. Nor…

X AI KOLs Following · 2026-05-29 Cached

Onepot AI can synthesize and deliver custom molecules in just 5 days by combining robotic synthesis with large-scale ML inference on Anyscale, dramatically accelerating drug discovery.

0 favorites 0 likes
#ml-inference

Gemma 4 MTP vs DFlash on 1x H100: dense vs MoE results

Reddit r/LocalLLaMA · 2026-05-12

This benchmark compares Gemma 4's Multi-Token Prediction (MTP) and z-lab's DFlash speculative decoding methods on a single H100 GPU, showing MTP faster for dense models and DFlash faster for MoE models.

0 favorites 0 likes
← Back to home

Submit Feedback