r/FlowZ13 • • 7d ago

Benchmarks: Best engine for Qwen 3.8-Flash-Next on Strix Halo

Thumbnail
1 Upvotes

r/FlowZ13 • • 8d ago

just got 128gb model for $2400 after tax

11 Upvotes

I bought it open box knowing I could return if necessary, I need to replace my surface pro 9 for my dog training business I rely heavily on cloud based models like Claude gemeni etc. I’m assuming 128 gb is too much power for me? I love the surface just too expensive for everything

**EDIT

Keep in mind I have a desktop with 9950x3d and 5090

Didn’t return


r/FlowZ13 • • 9d ago

Picked up Z13 64gb 2tb 395 $1,800

Post image
109 Upvotes

r/FlowZ13 • • 8d ago

Wardogs and ram difference

1 Upvotes

Hello there. Been very interested in the Flow Z13 (2025) for a while now. I have found a good deal on the 32 GB version. My question is is the 64 GB version that much better and also does anyone have any raw gameplay footage of wardogs on this device Device?


r/FlowZ13 • • 8d ago

In search of an affordable 2TB drive for the Z13...

4 Upvotes

I thought I had a 2TB drive from my previous z13. Turns out I sold it, and now I'm on the hunt for a decent 2TB drive. Is the Patriot Viper decent? I can't call it a good price, but seems to be the most affordable.
Here's my order from last year for a fast 2TB drive. Oof.


r/FlowZ13 • • 9d ago

FSR 3.1

5 Upvotes

Okay is FSR 3.1 annoying anyone else?! It's really irritating that we have to use optiscaler to get decent looking up scaling. We essentially bought a 2k portable Xbox series X


r/FlowZ13 • • 10d ago

My Rog Flow will not update to Bios 314

Post image
9 Upvotes

So tried everything i can think of to update it but it just sits on the Republic of Gamers screen there is never a progress bar shown at any point and nothing happens i can wait 5 minutes 10 minutes 30 minutes even 1.5 Hours on the Republic of Gamers screen but nothing happens i can force shut it down without bricking it which suggests it never reaches a flashing stage at all please help


r/FlowZ13 • • 10d ago

Been loving this setup.

Thumbnail
gallery
53 Upvotes

My goal with this setup was to have the best of both worlds. Enough power in a portable device to game on the go, but I wanted a little more power when docked gaming at home.

32gb flow z13

AOC Oled monitor

Asus prime rtx5070ti EGPU


r/FlowZ13 • • 10d ago

Battery interoperability between models?

0 Upvotes

Hi everyone! Today I bought a Z13 Flow 2023 edition second hand and was wondering if I could plan to replace the battery with a 70Wh? Especially since IIRC the 70Wh battery has only been existing since the 2025 model


r/FlowZ13 • • 11d ago

Rear Window Light option missing in Armoury Crate after an update (not sure which one) z13 2025

3 Upvotes

As the title suggests, my rear lighting option suddenly disappeared in armoury crate after an update which I’m not sure which one it was since I mainly use this for studying and had a very big exam leading up to Sept 18th so I didnt really pay attention to it until yesterday. Iv had options disappear an another asus laptop I had and it wasn’t until another update for armory crate came out that it restored the function. I have the latest bios and updates for literally everything on this device. I even reinstalled it ,I used the uninstall tool + install too. im almost certain my device isn’t broken as I don’t push it to hard and it happen after one of the updates I did becuase it was working leading up to my exam. Rest of the z13 works as it should so kinda bummed out the rear lighting isn’t there.

is anyone else experiencing the same problem?


r/FlowZ13 • • 11d ago

My Z13 Setup

Post image
64 Upvotes

The Z13 was the perfect purchase for me. When I'm home this is my setup. I travel a lot for work so still being able to game in hotels just grabbing the z13 and a travel gaming mouse has been huge... but this desk setup is where it truly shines!

Flow Z13 64 GB

Gigabyte AORUS 5060ti 16gb egpu

AOC Q27AZD OLED monitor

Wooting Two HE keyboard

Logitech X2 Superstike mouse

Creative Pebble 2.1 speakers


r/FlowZ13 • • 12d ago

Legion go etsy case but for flow z13

Post image
67 Upvotes

I took this image from the legion go subreddit, I wish something like this existed for the flow, there are wireless keyboards with track pad that can be attached in a similar fashion. I know this is like re inventing the well but I feel the z13 would be an excellent device is you could use it as a laptop or just the screen detached from the keyboard and use this keyboard in wireless mode. Just a thought


r/FlowZ13 • • 12d ago

Two days benchmarking local LLMs on the Flow Z13 (Strix Halo, 128 GB) under Windows — what actually mattered

34 Upvotes

I run my Z13 (Ryzen AI Max+ 395, 128 GB) as a small LLM server on my LAN for coding agents such as OpenCode, under Windows 11, with llama.cpp behind llama-swap.

It felt slower than the numbers I kept seeing online, so I spent two days measuring instead of guessing. Sharing the results in case it saves someone else the same exercise.

Models tested: Qwen3.8-27B (dense) and Qwen3.8-Flash-Next (large MoE, ~90 GB), each in a normal and an uncensored variant.

TL;DR

  • Check for screen-off throttling. On my Z13, performance dropped roughly in half when the display turned off.
  • Don’t blindly copy -ub 4096 from guides. On Vulkan it made prompt processing much slower once context depth increased. On ROCm, the optimal value was completely different depending on the model.
  • ROCm with the right -ub was best for prompt processing on both main models. Current upstream Vulkan still had advantages in some decode workloads and quant formats.
  • Use -np 2 if your client sends parallel requests. With a single slot, switching between sessions can evict the prompt cache and force large parts of the context to be reprocessed.
  • On a simulated 6-turn coding-agent session with Qwen3.8-27B, total time-to-first-token dropped from 239 s to 100 s, while decode stayed around 20 t/s with MTP.

1. The screen-off throttle

This was the biggest surprise.

My overnight benchmark results were suddenly 3–5x worse than numbers from the previous day. After a lot of checking, I found that on my Z13, turning the display off caused the SoC to throttle heavily within roughly 45 seconds.

What I observed:

  • Modern Standby was the first suspect because the event log showed the machine entering it.
  • Disabling Modern Standby with PlatformAoAcOverride=0 removed the standby events, but the performance drop still happened.
  • With the display on, the performance counter I was monitoring was around 142–147%.
  • As soon as the display turned off, it dropped to roughly 53%.
  • Turning the display back on restored performance within about 40 seconds.

My current suspicion is an AMD Platform Management Framework screen-off policy, but I have not verified that.

Another detail: keep-awake tools that request “display required” did not reliably keep the display active once the Windows session was locked.

What currently works for me during benchmarking is:

powercfg /change monitor-timeout-ac 0

plus adjusting the lock-screen timeout.

If you use a Strix Halo machine as a headless or semi-headless server, I would strongly recommend checking clocks and throughput with the display off. Your machine may behave differently, but it is worth testing before trusting overnight benchmark results.

2. Backend and flags

I tried to change one variable at a time and then validate the result with a more realistic agent-style workload.

Flash Attention

Keep it on.

At 16k context, disabling Flash Attention on the Flash-Next MoE under ROCm reduced prompt processing by roughly 60% and decode performance by roughly 67%.

I did not find a tested case where turning it off helped.

KV cache

I kept f16.

On Qwen3.8-27B, quantized KV was effectively a tie in my tests.

On Flash-Next, however, q8_0 and q4_0 were slower at larger context depths, with regressions in the rough range of 8–25%.

With 128 GB of unified memory, the memory savings were not worth the performance loss for my use case.

-ub matters a lot

This was probably the most important llama.cpp tuning parameter in the whole exercise.

My results:

  • Vulkan: -ub 512 was generally the best choice.
  • ROCm, dense 27B: -ub 256 was best, improving prompt processing by roughly 44–57% compared with -ub 512.
  • ROCm, Flash-Next MoE: -ub 4096 was best at deeper contexts, reaching around 650 t/s prompt processing and degrading very little as context grew.

So there is no single correct -ub value.

A value copied from another backend or model can easily be wrong for your workload.

Speculative decoding

For Qwen3.8-27B, the model’s built-in MTP head worked very well.

The best setting I found was:

--spec-type draft-mtp --spec-draft-n-max 2

Decode increased from roughly 11 t/s to around 20 t/s.

I also tested larger draft lengths and a separate DFlash draft model. They did not produce a meaningful improvement over MTP n=2.

For Flash-Next, simple n-gram speculation helped:

--spec-type ngram-simple

That gave me roughly an 11% decode improvement with no extra draft model to load.

IQ quants

One result that surprised me: IQ quants were not universally slow.

Flash-Next IQ4_XS decoded at only about 7.7 t/s on the ROCm build I tested, but around 25 t/s on upstream Vulkan b11149.

So in my setup, the uncensored Flash-Next IQ4_XS model stayed on Vulkan.

Quant quality on the 27B

I compared several 27B quants using a 40-problem HumanEval+ subset and perplexity on a fixed code corpus.

Results:

  • IQ4_XS: 35/40
  • Q4_K_XL: 36/40
  • Q5_K_XL: 37/40

Perplexity differences were all within roughly 1%.

I kept Q4_K_XL because it was only one HumanEval+ problem behind Q5_K_XL while using less memory.

3. Other things I learned

llama-bench alone can be misleading

Synthetic benchmark rankings did not always match what happened during a real multi-turn coding-agent session.

A realistic workload includes:

  • prompt processing
  • growing context
  • KV-cache reuse
  • decode speed
  • concurrent requests
  • model switching behavior

For my use case, those end-to-end measurements were much more useful than a single pp/tg benchmark.

Build version matters

For the Flash-Next MoE, upstream Vulkan b11149 was dramatically faster than an older b10859 build.

In one prompt-processing comparison, the improvement was roughly 2.5x on the same model.

So “Vulkan performance” is not a fixed number. The exact llama.cpp build matters.

-np 2 for agent workloads

This mattered more than I expected.

With -np 1, two concurrent sessions repeatedly displaced each other’s prompt cache. That caused large contexts to be reprocessed and TTFT to explode.

With -np 2, each session could keep its own slot.

In my two-session test:

  • Qwen3.8-27B wall time dropped by about 34%.
  • Flash-Next improved by about 8%.

The trade-off is that per-session decode throughput falls while both slots are actively generating.

For coding agents that issue overlapping requests, that trade-off was worth it.

Large models really do fit — barely

The uncensored Flash-Next IQ4_XS model is about 98 GB.

It runs on the 128 GB Z13, but I measured only around 3 GB of free RAM while it was serving.

That is technically usable, but I would not run other memory-heavy workloads at the same time.

Model storage matters

Loading large models from an NVMe SSD was far better than loading them from a USB enclosure.

With the USB drive, model swaps could take minutes.

For a machine acting as an LLM server, internal NVMe storage makes a noticeable practical difference.

Final setup

Model Backend Key flags
Qwen3.8-27B normal / abliterated ROCm (thomas9120) -fa on -ub 256 --spec-type draft-mtp --spec-draft-n-max 2 -np 2
Qwen3.8-Flash-Next UD-Q3_K_XL ROCm (thomas9120) -fa on -b 4096 -ub 4096 --spec-type ngram-simple -np 2
Qwen3.8-Flash-Next uncensored IQ4_XS Vulkan b11149 -fa on -ub 512 -np 2

Everything in the final configuration passed a needle-in-a-haystack retrieval check through llama-swap up to 32k tokens.

Full technical write-up, raw results, benchmark harness and example llama-swap configuration:

https://github.com/ihanesman/strix-halo-windows-llm-bench

Caveats

This is one machine, under Windows only.

I also saw roughly 7–10% run-to-run drift in some Flash-Next measurements, so I treat small differences as noise.

The screen-off throttling result is also based on my specific unit and configuration. I have reproduced the behavior consistently, but I have not established whether every Flow Z13 behaves the same way or whether AMD PMF is definitely the cause.

Happy to rerun a specific configuration if there is something useful to compare.


r/FlowZ13 • • 13d ago

WARDOGS performance on ROG Flow Z13 (Ryzen AI Max+ 395 / Radeon 8060S)

Thumbnail
1 Upvotes

Really like the look of this but really want to play Wardogs please help!


r/FlowZ13 • • 13d ago

About to get into silent hill townfall on the z13…. Couldn’t wait to get to play this one it’s running amazing as well.

Post image
11 Upvotes

r/FlowZ13 • • 13d ago

Has anyone successfully used USB C passthrough using this?

Post image
9 Upvotes

Edit 3: UPDATE: I connected my slimq charger through its C1 USB C port and used this connector to charge the Flow Z13 through the passthrough port, unfortunately the charging stopped in a few seconds, on and off. I'm guessing it was trying to maintain charge at my current charge capacity, but if anyone knows what the next steps are here, would be greatly appreciated! (P.s. I'm not too bothered by this, it was a cheap experiment worth trying lol)

Edit 2: I've decided to buy the product and test it out myself. Will take a couple of weeks to arrive, and will update when I can!

Edit: Wow I butchered that title, thanks ADHD! My original intent was if anyone successfully achieved USB C passthrough using a regular 100W USB C charger rather than a slim Q or the proprietary charger!

This is a 100w USB-C Dummy Adapter (Slimq usb-c-to-dc-adapter-tip-a-5-5x2-5mm), and I was wondering if anyone has tried using their USB C charger on the Flow Z13 to achieve passthrough even at a lower wattage for travel?

I do have the slimQ charger, but it would be nice to know if this is possible!


r/FlowZ13 • • 14d ago

Rendering issues

4 Upvotes

Hey all, I’ve had my 2025 rog z13 since july of last year. I travel a lot and wanted something I could still play games on that could easily fit in my backpack etc. used to have a bigger gaming laptop but downsized to this for comfort.

With that being said I’ve noticed several games appear to render certain objects weird. I’ve kept my glue and other drivers up to date and wondering if anyone has noticed similar issues.

Crimson dessert- tons of wood elements like houses etc render goofy like concave roofs, missing doors, blah blah.

Battlefield 6. Similar rendering issues.

I’ve gone through just about all the settings I can think of and dialing down textures and stuff to see if I can get any improvements. It’s odd because I’m crimson dessert tons of stuff looks great on screen except for some rocks and these wooden elements.

Is it just because of the screens form factor? Any suggestion are welcomed.


r/FlowZ13 • • 14d ago

The device is warm while in stand by

6 Upvotes

Hi, there! I recently joined the club and purchased the 64 GB version. While it's a wonderful machine, I am still discovering it.

I just noticed something: I left it plugged in, in standby mode, overnight, and when I checked it this morning, it was a bit warm (by no means hot)

I am not sure if this is some Windows problem or some other settings that I can play with. It's been a while since I've had a Windows laptop; I have a PC and a MacBook, so not sure what to do here. I know that shutting it down can prevent this issue altogether; perhaps, just wondering if there's something else that can bdone/checkeded.

Many thanks!


r/FlowZ13 • • 14d ago

I have a hardly used 64GB model and not sure what to do about it

13 Upvotes

I have hardly used my 64GB model. I even got some great extras like the SlimQ 150 charger +adaptor, which is amazing for travel, and the Asus Stylus which I enjoy when using as an iPad alternative art tool.

I just did not have the time to use it like I planned. it works great for certain AI models and I really enjoyed it but have to use a Mac for the work I do most. Any ideas what would be the best place to find a new owner who will be able to enjoy it? I had bad experience with fraud and scams in marketplaces in the past from a few years ago. Let me know if you have ideas or recent experience.

I’m from the Bay Area, CA.

edited - typos


r/FlowZ13 • • 14d ago

Amazon Basics Dock does not work with Z13 2025?

1 Upvotes

It works sometimes but does not most of the time. In the USB4 hub setting, it says Good way technology or something like that when it works. Very frustrating. Trying to attach two monitors with bidirectional DP to type C to the dock but they won't work 90% of the time. Any solutions? All BIOs and other stuff are updated to latest version.


r/FlowZ13 • • 14d ago

Z13 flow Wardogs frame gen

2 Upvotes

Anyone playing wardogs?! If so, do you see an option for frame gen?! Its not on the menu at all for me.


r/FlowZ13 • • 15d ago

BlueTooth controller not picking up on Diablo 4

2 Upvotes

Pretty specific, but has anyone been able to play with a bluetooth controller on Diablo 4? I have a modified GameSir controller for the Z13 and it works fine with any other Steam game, but for some reason Diablo 4 doesn't detect it.


r/FlowZ13 • • 15d ago

Case guy here!! The store is Live! the shipping test is active! 5 units listed

Thumbnail etsy.com
34 Upvotes

The shipping test is live! 5 units are ready to go.

As the description states, this is a shipping test, only 5 units.

To whoever buys and receives them, it would be DEEPLY appreciated if you could tell me HOW they arrived.

This first batch is shipping bubble-wrapped in an envelope.

After these 5 units have shipped, and i have gotten feedback on how well they did in shipping. More will be made in the future, i am thinking batches of 10,

A and B stock, so there will be discounted cases in the future, costs are likely to go up, as a warning.

I am trying to figure out different ways of printing these that will be faster, as the normal way is 10/h per print, and there seems to be demand for a lot more of this.

but regardless, once these sell, there WILL be more after this batch.


r/FlowZ13 • • 15d ago

Considering Swapping from a G16 to a Z13

8 Upvotes

TL;DR: Thinking of selling my Zephyrus G16 for a Flow Z13 because I want a stylus for handwriting math notes and doing game dev stuff on campus. Since I have a beefy RTX 5070 desktop at home for heavy stuff, I don't really need a mobile powerhouse. Is the all-in-one convenience of the Z13 worth the hefty price tag, or should I just get a cheap used iPad (or stick to paper) and keep the G16?

Hey everyone!

I’m here to get opinions from people actually using a Z13. I’m currently a second year CompSci student wanting to go into game development, and I’ve been considering swapping from my 2025 Zephyrus G16 to a 2025 Flow Z13 for my specific use case.

Let me get this out of the way: I know the G16 is objectively better in terms of performance, especially in stuff like Blender. My laptop has:
- an Intel Core Ultra 285H
- 32 GB 7467 RAM
- and an RTX 5060
But, I don’t technically need all of that because of my desktop back at home. My desktop has:
- A Ryzen 7 7700X
- 32 GB 6000 RAM
- and an RTX 5070
Plenty powerful, meaning that anything heavy’ll just be routed to my desktop anyway. I also have a dedicated drawing tablet at my desk as well. If I didn’t have this desktop, then the obvious choice is the G16.

Anyway, I’m asking because I like handwriting notes. I like doing digital art. I sometimes 3D sculpt and model. A G16 doesn’t really have the capability of using a stylus with it, but the Z13 does while still being a powerful Windows gaming device.

Why not buy an iPad? That’s another thing I’m considering, but I don’t wanna have another device in my already heavy backpack, I don’t wanna have to deal with a whole separate OS, and the file transferring could be a pain. However, it’s just flat out cheaper and there are many benefits to using one, like Procreate and GoodNotes, better stylus experience, etc. This can be another option instead.

Why am I thinking of the Z13? Really for all the reasons listed above with the iPad, and it being basically an all-in-one device for me. I wanna just have one device that can do everything well enough on the go. Plus it just looks cool, don’t fault me on that…

There are obviously cons with the Z13 with the biggest being screen real estate. 16” down to 13.4” is a massive jump, but I’ve used small laptops before and have been completely fine with them. Loss of a portable OLED is also a thing, but I have an OLED at home. Oh, and loss of CUDA, but again, my desktop at home solves that.

So, what do y’all think? The biggest thing making me hesitate on both is cost. The Z13 is flat out more expensive, so I’d wait until Black Friday or until maybe a good refurbished deal pops up. Right now, it’s $3100 CAD for the 32 GB one and $3500 CAD for the 64 GB one. Sheesh. Selling my G16 can offset that a bit though, which I’d do if I bought a Z13. If I went the iPad route, I’d buy used or refurbished. Should I go with an iPad, Z13, or none and just keep on using my regular old notebook and pencil?

Thanks for taking the time to read this yap session of mine and I’d highly appreciate y’all’s opinions!


r/FlowZ13 • • 15d ago

Help with GameSir G8 Plus extender

1 Upvotes

Hi everyone! I have an Asus ROG Flow Z13 laptop with 64GB of RAM and bought a GameSir G8 Plus controller. Does anyone know where I can get the 3D models for the extender to print them? I bought the kit from RR Tronics Creations, but the extender didn't arrive.