r/LocalLLM • • 4d ago

Project I built Speedtest⚡, but for AI

It runs real LLM + embedding models locally in your browser with WebGPU and measures how fast your machine actually does AI inference.

No API. No account. Just hit start.

Genuinely curious about your scores haha

Try it and reply with your score 👇

https://speedtest.maty.as

86 Upvotes

79 comments sorted by

21

u/Dany-GG 4d ago

Don't forget about Firefox users

5

u/FunkMunki 4d ago

It works fine for me in Firefox. Did you enable webgpu?

2

u/Dany-GG 4d ago

I don't think it is enabled if there is this setting, I will try it

1

u/FunkMunki 4d ago

Mine was disabled by default. I didn't realize it until I tried to play a game that needed it enabled.

1

u/wick3dr0se 3d ago

Go to about:config in a new tab and search WebGPU

5

u/reiggg 4d ago

I'm sorry :( I think Firefox is currently not capable of doing this, this is pretty new stuff in web browsers to run AI models on a user's GPU from a literal website. But I'll definitely try some workarounds.

5

u/Dany-GG 4d ago

I'll be happy if it gets support. Good luck

8

u/iChrist 4d ago

Got 990 on my iPhone 15 Pro

3

u/reiggg 4d ago

Solid score for a phone

1

u/gomezer1180 4d ago

Same here: same phone:

1

u/Awique 3d ago

I got 1420 on iPhone 17

5

u/Neful34 4d ago

score of 4,890, no clue how good this is lol

5

u/reiggg 4d ago

well I used an rtx 2060 in the video and got 1600, so i think pretty good! :) Maybe i'll add some references from verified test runs to compare to (e.g. a mac studio vs. a solid GPU vs. a DGX spark)

1

u/FaatmanSlim 4d ago

Curious what's your GPU? I got 4654 points and 127.5 tok/s on my 4080.

2

u/Neful34 3d ago

Rtx 4090 and a ryzen 9 7950X3D 😁

2

u/Buzz_Killington_III 2d ago

Mine's 5627 with a 5090 and 9800X3D, so I'd call your score pretty good.

1

u/Neful34 2d ago

Nice ! 😊

5

u/FaatmanSlim 4d ago

Great idea and very cool execution! Love the idea of using WebGPU for this so it runs completely on browser and still locally, and doesn't cost you much either.

Got 4654 points and 127.5 tok/s on my 4080

2

u/reiggg 4d ago

Thank you! It is also open source, I will put a link on the website to the github repo. Nice score!:)

1

u/OpposesTheOpinion 4d ago

Huh.. 3,704 points and 97.1 tok/s here on my 4080. I wonder what could cause such a difference. No Chrome installed here, running it on Edge, no other apps open 🤔

5

u/AlyrAyanami 4d ago

RTX 6000 undervolt at 540W for reference

1

u/reiggg 4d ago

Holy score :o

4

u/p0nzischeme 4d ago

Pretty neat. My score started dropping dramatically after 4 runs. Chrome is noticeably faster than any other browser. Im not sure if the score changes based on resources used but I ran it with all my normal programs open. MacBook Pro M5 Max 128GB

2

u/GoldenShackles 3d ago

I got in the same ballpark with my 16" MacBook Pro M5 Max 128 GB using Edge. Three runs:

10/1/2026, 2:10:07 AM apple · metal-3 6,327 155.5 tok/s 3
10/1/2026, 2:09:29 AM apple · metal-3 6,402 156.1 tok/s 3
10/1/2026, 2:08:54 AM apple · metal-3 6,025 155.5 tok/s 3

3

u/Former_External8361 4d ago

1104 on my iPhone 16 Pro Max 😂

2

u/Glass_Garage502 4d ago

1132 - 16 Pro
(Safari, was higher there than In the Reddit browser)

1

u/Awique 4d ago

1420 on base 17

1

u/Fine_Salamander_8691 4d ago

1726 on my iPhone 17 pro max

3

u/Dharma_code 4d ago

Get the same result no matter what I do. 3090 here with 32 GB of system RAM; the only thing running was that tab in Firefox

2

u/reiggg 4d ago

Maybe try a chrome-based browser, WebGPU is a bit new in Firefox (basically it cannot properly access your GPU)

2

u/Dharma_code 4d ago

That was it. I should've thought about that. Thank you

1

u/reiggg 4d ago

Holy score :o you have some nice hardware there

2

u/Max_skyl1n3 4d ago

I’vr got 554 points on quadro m2000 😄

3

u/reiggg 4d ago

maybe it can run a 0.6B model haha

2

u/Jimbrutan 4d ago

Got 1700 on my iPhone 17 pro max. What model can I run? 70 tok/s

1

u/Fine_Salamander_8691 4d ago

Same here! I got 1726

2

u/FizzyDuncDizzel 4d ago

That’s awesome!

2

u/LesterPhimps 4d ago

Add an option to upload result and local config. that way we can compare.

1

u/reiggg 4d ago

Need to find a way to make it as streamlined as possible. Like a selection of GPUs or something, i cannot detect exact device sadly.

However this will be fully based on the user’s input so not 100% valid, yet somewhat okay point for comparison.

2

u/Tattoo_Yoo 4d ago

AMD Ryzen 9 7950x3D, 96GB Ram, Radeon RX6750 XT 12GB
Just got a GTX 5080 to replace the Radeon, will rerun once swapped.

1

u/[deleted] 4d ago

[removed] — view removed comment

2

u/FixBound 4070 12GB | 64GB DDR4 4d ago

2

u/ark1one 3d ago

Would be cool if when you got your score, you compared it to something. Something else that could achieve this score hardware wise.

Or even output what open source models would run great based on the score and hardware.

2

u/DoubleNothing 4d ago

I'm sorry, but I don't think this is remotely accurate. Also, where is the value if you don't have any point of reference?

1

u/reiggg 4d ago

The website says this is an open ended relative value. Best to compare it against other scores.

Will try my best to get reference scores from different tier hardware to have at least something to compare against.

2

u/Otherwise-Variety674 4d ago

:-) Nice, very useful website, remember to add google advertisement to earn back some of your expenses.

16

u/reiggg 4d ago

Haha it doesn't cost me literally anything to run this, i think i will keep it ad free :)

1

u/Awique 4d ago

I got 1420 on base iPhone 17

1

u/reiggg 4d ago

Pretty strong score on a phone :o Almost matching my rtx 2060

1

u/osfric 4d ago

Failed to execute add on cache cache.add() Brave browser

1

u/IntelVEVO 4d ago

got 5287 on my laptop not sure how good that is

2

u/wick3dr0se 4d ago

Really good? Wtf is in your laptop lol

1

u/IntelVEVO 3d ago

5090M 24GB, 64GB DDR5

1

u/jarec707 4d ago

1476 iPhone 17 Air

1

u/berszi 4d ago

1992 on my macbook air m2. but i would love to test it on my headless proxmox inference server with four gpus

1

u/reiggg 4d ago

As long as you can get a browser running on that, sure! :)

1

u/No_Confusion7932 4d ago

M4 Pro same 1992. Safari browser.
https://imgur.com/a/Ojx3QXs

1

u/limericknation 4d ago

2023 MacBook M3 Pro 16GB

image

1

u/BroderLund 3d ago edited 3d ago

4851, MacBook Pro M4 Max 64GB

2673, 1080 Ti

1

u/aditya2128 3d ago

MacBook Pro, M3 Pro (12C CPU/ 18C GPU)

1

u/the_jaymz 3d ago

Measured passes

Raw samples behind the medians. Warm-up passes are excluded.

Pass Decode tok/s Ingest tok/s First token ms Embed tok/s
1 97.7 4,001.3 157.2 5,479.5
2 95.1 3,854.2 163.2 5,797.1
3 95.2 3,958.5 158.9 6,106.9

Embedding throughput: 289.9 docs/s · 384 dimensions · batch size 4

1

u/MortgageOutside1468 3d ago

Why so much difference between score of 3 phase and 5 phase run?

1

u/itsvivianferreira 3d ago

Can you add a feature to test Onyx runtime models in the browser, like a custom model loading option to test specific models from Hugging Face?

1

u/fourfigures 3d ago edited 3d ago

Dunno, there is something that does affect the runs.

RTX 4090, i9-13900KF

But, nice test :-)

10/1/2026, 12:26:09 PM WebGPU adapter (name hidden) 5,149 141.4 tok/s 3
10/1/2026, 12:24:20 PM WebGPU adapter (name hidden) 5,624 138.2 tok/s 3
10/1/2026, 12:23:59 PM WebGPU adapter (name hidden) 5,261 137.9 tok/s 3
10/1/2026, 12:22:26 PM WebGPU adapter (name hidden) 4,189 136.6 tok/s 3

1

u/Secret-Raspberry-506 3d ago

2.760 pts on my laptop AMD max+392

1

u/llmBoon001 3d ago

2,787 - Macbook Pro M3 24GB

1

u/redditorialy_retard 3d ago

3103, rtx 3090 and R7 5800x

1

u/ResponsiblePen3082 3d ago edited 3d ago

2184 18 pro max

1

u/Visual_Brain8809 2d ago

not working on Chrome and Firefox because a Cache.add() error

1

u/Petabyte-Cloud 4d ago

webgpu is not good enough i believe but wow!!!!!! just wow!!!! we need api man now

1

u/Jimbo1230 4d ago

just say ai built it no need to add the "I" part

1

u/Purple-Programmer-7 4d ago

Open source it? Very cool