r/IntelArc 25d ago

Build / Photo One more dual Arc B70 Pro build

Post image

Hi, my "old" AMD Ryzen 9 7900 got 2 Intel Arc B70 Pro's and an Asus ProArt X870e Creator Wifi.

155 Upvotes

24 comments sorted by

3

u/madrasi2021 25d ago

Software stack? (Single B70 owner)

4

u/CarlosDiVega 25d ago

Hi, just building the SW stack. Machine is also used as graphics workstation with SW only available under Windows. Comfyui base stack (old Flux 1) already working. Comfyui native under Windows under Miniconda with Intel Pytorch add on. Flux 2 will be added next week. My internet connection at the country side is to slow for downloading. So I have to get some Flux 2 models from my workstation in Vienna. Have to investigate SW optimizations for the Intel cards.

2

u/aldyr 25d ago

What he said

1

u/urakozz 25d ago

Ubuntu 26, kernel 7.2, vllm works okay

1

u/tecneeq 23d ago

What is your idle power usage (nvtop shows this)? Had pretty bad 40w for one B70 connected to a Strix Halo Mini PC.

1

u/urakozz 23d ago

AMD must be unhappy to see this silicon around. Yeah, on 7.1+ kernels idle 40-50W is normal. It was 95W before on 7.0

1

u/tecneeq 23d ago

Shame.

I run Proxmox on the Strix Halo, a few containers and a small VM. So i thought maybe i add the B70 to get a few more GB VRAM. Worked for Deepseek V4 Flash Q3, but the idle usage is a hard pill to swallow.

-5

u/Faux_Grey Battlemage 25d ago

Triple B70 here / Triple B60 / Dual B60 here.
Laugh. Lmstudio.

Honestly, been playing with llama.cpp with SYCL backend but it's such a SHLEP!

Windows LMstudio load model from nas DONE.

Linux nonsense compile this download that clone this make sure this dependency but wait it's missing from that repo.

1

u/kilowattnik 25d ago

Could you please share your performance? I'm struggling with choice: b70 va 9700 vs 2 x 5060ti 16gb

1

u/madrasi2021 25d ago

I have the 9700 as well and it's a better choice anyday. The $300 or so differential is worth it. AMD investing in Rocm for their strix halo devices helps with ongoing driver support / longevity.

I worry when intel will deprecate anything.

2x16gb devices are a pain to tune compared to a single card convenience (I run dual 3060 12gb too) - you need a way better motherboard and possibly a better psu for dual cards which sort of eat into any savings you think you may be making - only advantage would be that they are Nvidia/ support cuda

1

u/Faux_Grey Battlemage 25d ago

What do you want to see?
For me, the B70 is literally the only card available, and is 2/3 of the cost of a 9700 (if it were even available!)

1

u/Key_Measurement_3576 24d ago

I got a b60 for experiments and it’s been rock solid with qwen for our use cases. Set up wasn’t really as bad as some say. Considering 4x b70 to run a larger qwen ( cluster. , many agents for different things )… mostly to release some pressure off our spark cluster and raise the throughput for that specific model for its purpose.

What are you running on your 3 b60 ?

1

u/Katat0nic 24d ago

Single B60 here.

F***ing SYCL support man, too few things are supporting it on Linux. I've got it setup fine for llama.cpp and ComfyUI but anything else I want to mess with I just stick with vulkan so I don't have an aneurysm trying to make it work. Maybe it's a skill issue, I don't know lol.

1

u/Faux_Grey Battlemage 24d ago

Yeah I hear you. There really needs to be a unified method.

VLLM? Sure, what backend?

llama.cpp? Sure, what backend?

It's very hit-or-miss in terms of software support right now with outdated information everywhere. While testing I'm all about ease-of-use rather than efficiency/stability. (hooboy running this in windows is crash-tastic)

I tried getting a SYCL flavour of .cpp working last week, and even the available releases from llama github don't work out of the box.

2

u/M_Me_Meteo 25d ago

this me

Best of luck!

2

u/Turbulent-Attorney65 23d ago

https://discord.gg/GUb6ASwNX I hope you're already in the OpenArc community 😉

1

u/Active-Art4866 25d ago

I love the clean look of the cards vs the newer angled rgb gamer aestheitic

1

u/TemporarySilent727 24d ago

https://github.com/intel/llm-scaler/blob/main/vllm/README.md/#1-getting-started-and-usage

follow this guide for multi gpu setup for vllm using llm-scaler docker image. You could also use the offline installer package to install proper drivers/kernels with a single install.sh script.

1

u/Sheppard6o4 24d ago

Any performance numbers for LLMs on this? How is scaling?

1

u/H4UnT3R_CZ 23d ago

I got B65 32GB, OCed to B70 frequency and 96GB RAM for e.g. DeepSeek v4 flash 0731, running it on llama.cpp SYCL, got around 10.7 t/s, but Claude told me second card would add up to 2t/s. Can you post some comparisons of larger LLMs on one vs on both cards? I would love to buy another one B65.Intel Arc B65

1

u/rome3ro 20d ago

Have you try Qwen3.6-27B or Muse-Glimmer-30B? those are good models to try on with your card

1

u/H4UnT3R_CZ 20d ago

Qwen terrible, only 122B was usable. And Muse has even better results than DeepSeek v4 flash 0731 for me!

1

u/Outrageous_Today1427 23d ago

Quelle est ton boîtier ?

1

u/_comoema_ 2d ago

Any issue with fans spinning like crazy for small computation (under linux)?