intel-arc

Tag

Cards List
#intel-arc

Ling-3.0 (BailingMoE3) lands in llama.cpp mainline - Quick benchmarks on Intel Arc B580

Reddit r/LocalLLaMA · yesterday

Ling-3.0 (BailingMoE3) is now officially supported in llama.cpp, with benchmarks on Intel Arc B580 showing efficient local inference including 128K context in 12GB VRAM.

0 favorites 0 likes
#intel-arc

Show-off Saturday: Intel Arc B140 build.

Reddit r/LocalLLaMA · 4d ago

A showcase of a personal local AI inference build featuring Intel Arc B140 GPUs and custom hardware, running llama.cpp with SYCL back-end on Ubuntu.

0 favorites 0 likes
#intel-arc

Intel Arc Pro b70 box & 3D model

Reddit r/LocalLLaMA · 2026-07-21

Intel reveals the Arc Pro B70 graphics card box and a 3D model of the card, indicating a forthcoming product launch.

0 favorites 0 likes
#intel-arc

MSI Claw 8 EX AI+ Review: Great Power, Shocking Price

Wired · 2026-07-15 Cached

The MSI Claw 8 EX AI+ is a powerful handheld gaming PC with Intel's Arc GPU, excellent ergonomics, and a premium design, but its high price and lack of OLED display are drawbacks.

0 favorites 0 likes
#intel-arc

@hotschmoe: After reading this post, I decided to get nvfp4 running on my Intel arc b70s just to see, after 12 hours it's running a…

X AI KOLs Following · 2026-07-04 Cached

A user successfully ran nvfp4 quantization on Intel Arc B70s GPUs, achieving nearly double speed and higher accuracy compared to their best int4 configuration, challenging hardware-specific format assumptions.

0 favorites 0 likes
#intel-arc

Tip: use this llama.cpp PR to improve PP on Intel ARC

Reddit r/LocalLLaMA · 2026-07-02

A llama.cpp PR significantly improves prompt processing speed on Intel ARC GPUs, with benchmark showing speed increase from 245t/s to 462t/s on a B580. The improvement currently works for F16 KV quantization, with plans to support other quants.

0 favorites 0 likes
#intel-arc

sycl : port multi-column MMVQ from CUDA backend (~45% speculative decoding speedup on Intel Arc) by masonmilby · Pull Request #21845 · ggml-org/llama.cpp

Reddit r/LocalLLaMA · 2026-06-05 Cached

A pull request for llama.cpp ports multi-column MMVQ from CUDA to SYCL, achieving approximately 45% speculative decoding speedup on Intel Arc GPUs.

0 favorites 0 likes
#intel-arc

@TeksEdge: Solved! Qwen3.6-27B-FP8 is now running on Intel Arc Pro B70! LocalMaxxing shows a working 4× Arc Pro B70 32GB run at ~5…

X AI KOLs Following · 2026-05-15 Cached

Qwen3.6-27B-FP8 model is now running on Intel Arc Pro B70 GPUs at ~50 tok/s with a vLLM bug fix, marking a significant milestone for Intel GPU local AI inference.

0 favorites 0 likes
#intel-arc

Nvidia RTX 3090 vs Intel Arc Pro B70 llama.cpp Benchmarks

Reddit r/LocalLLaMA · 2026-04-23

Community benchmark shows Intel Arc Pro B70 averages ~71% slower prompt processing and ~54% slower token generation than RTX 3090 under llama.cpp, with SYCL backend sometimes beating Vulkan on the same card.

0 favorites 0 likes
← Back to home

Submit Feedback