openvino

Tag

Cards List
#openvino

Inside Our Distributed LLM Inference Research for Intel PCs

Reddit r/artificial · 2026-08-20 Cached

The article presents research on distributed LLM inference for Intel PC fleets, focusing on pipeline-parallel sharded inference using OpenVINO with performance optimizations for heterogeneous hardware.

0 favorites 0 likes
#openvino

@googlegemma: Gemma 4 E2B goes super fast on Intel AI PCs thanks to LiteRT NPU support on OpenVINO! 1.3x faster prefill performance o…

X AI KOLs Timeline · 2026-06-16 Cached

Gemma 4 E2B achieves 1.3x faster prefill and 2.8x better performance-per-watt on Intel AI PCs using OpenVINO with LiteRT NPU support, enabling efficient background LLM tasks.

0 favorites 0 likes
← Back to home

Submit Feedback