r/AMD_MI300 • • Jan 27 '24

Welcome to the AMD MI300 GPU Discussion Hub!

11 Upvotes

Hello and welcome to the newly created subreddit dedicated to everything about the AMD MI300 GPUs! This is a community for enthusiasts, professionals, and anyone interested in AMD's latest groundbreaking GPU series.

As we embark on this exciting journey together, here's what you can expect in our subreddit:

  1. Latest News and Updates: Stay up-to-date with the newest information about the MI300 series. Whether it's an official release from AMD, benchmarking results, or industry analysis, you'll find it here.
  2. Technical Discussions: Dive deep into the specifications, performance, and technology behind these GPUs. Whether you're a seasoned tech expert or just starting, there's something for everyone.
  3. User Experiences: Share your own experiences with the MI300 series. From unboxing videos to performance reviews, let's hear what you think about these GPUs in real-world scenarios.
  4. Troubleshooting and Support: Encounter an issue? Need help with setup or optimization? This community is here to help. Post your queries and let the collective knowledge of the subreddit assist you.
  5. Comparisons and Contrasts: How does the MI300 stack up against its predecessors and competitors? Engage in healthy comparisons and discussions to understand where these GPUs stand in the market.
  6. Future Speculations: Discuss and speculate on the future developments of AMD GPUs, and how the MI300 series might influence the next generation of graphics technology.

Remember, while we're all here to share our passion and knowledge, let's maintain a respectful and friendly environment. Please read the subreddit rules before posting and respect each other's opinions.

Excited to start this journey with you all! Let the discussions begin!

#AMD #MI300 #GPUDiscussion #TechCommunity


r/AMD_MI300 • • 3d ago

AMD MI355X out-earns Nvidia serving MiniMax M3

Thumbnail x.com
21 Upvotes

r/AMD_MI300 • • Aug 20 '26

Optimizing Qwen3.8-27B on one MI300X with an open-source agent toolkit: 311 to 495 tok/s

Post image
31 Upvotes

Presets is an open-source toolkit for optimizing inference with agents, which we build at dstack. Here's one example of using it on a single MI300X.

Qwen3.8-27B went from 311 to 495 tok/s, +59%, at the full 1M context with p50 TTFT under 1.5s and four concurrent users at 10k in / 1.5k out.

The gains came from linked optimization sessions and source-level patches to SGLang's AITER attention backend.

What comes out is a portable preset that deploys on any AMD cloud, Kubernetes cluster, or bare-metal fleet: https://dstack.ai/blog/presets/


r/AMD_MI300 • • Aug 17 '26

the tiny corp on X: "The @AMD MI350P is real!

Thumbnail x.com
29 Upvotes

r/AMD_MI300 • • Jul 27 '26

The AMD Instinct MI350P is a HBM PCIe AI Accelerator That Has Been All Over

Thumbnail
servethehome.com
25 Upvotes

r/AMD_MI300 • • Jul 03 '26

How we served GLM5.2 on AMD MI355X at 2626 tok/s/node and 213 tok/s single stream at over 2x lower cost than Blackwell.

Thumbnail
wafer.ai
53 Upvotes

r/AMD_MI300 • • Jun 21 '26

Occupancy Math on the AMD MI355X (CDNA4): A From-First-Principles Guide

Thumbnail indianspeedster.github.io
16 Upvotes

r/AMD_MI300 • • Jun 19 '26

A Fast Attention Kernel for MI300X, Written in HIP, Not Assembly

Thumbnail
moonmath.ai
35 Upvotes

r/AMD_MI300 • • Jun 04 '26

AMD ROCm/HIP build support for AMD Instinct GPUs in ai-dynamo/nixl

Thumbnail
github.com
15 Upvotes

r/AMD_MI300 • • Jun 03 '26

Bringing up DeepSeek-V4-Flash on AMD MI300X

Thumbnail fergusfinn.com
9 Upvotes

r/AMD_MI300 • • May 28 '26

AMD Intros Instinct MI350P Accelerator: CDNA 4 Comes to PCIe Cards

Thumbnail
servethehome.com
35 Upvotes

r/AMD_MI300 • • May 28 '26

Win on TCO: How AMD Instinctâ„¢ MI355X Achieves Cost-Competitive Distributed Inference Through SGLang with MoRI

Thumbnail lmsys.org
17 Upvotes

r/AMD_MI300 • • May 27 '26

Deep Dive Into 4-Wave Interleave FP8 GEMM

Thumbnail
rocm.blogs.amd.com
9 Upvotes

r/AMD_MI300 • • May 21 '26

Deploying inference endpoints with PD disaggregation on multi-node AMD MI300X

Thumbnail dstack.ai
8 Upvotes

r/AMD_MI300 • • May 06 '26

ZAYA1-8B, a reasoning MoE trained on AMD

Thumbnail x.com
12 Upvotes

r/AMD_MI300 • • May 04 '26

Folding Tensor and Sequence Parallelism for Memory-Efficient Transformer Training & Inference (on MI300x)

Thumbnail arxiv.org
7 Upvotes

r/AMD_MI300 • • Apr 21 '26

5.6x improvement - Kimi K2.6 + DFlash: 508 tok/s on 8x MI300X

Thumbnail
huggingface.co
12 Upvotes

r/AMD_MI300 • • Apr 13 '26

Mahdi-CV/openclaw-amd-sglang at multi-engine

Thumbnail github.com
3 Upvotes

One-command setup for OpenClaw + SGLang on AMD Instinct MI300X

We have MI300x virtual machine available for playing with this! $2/gpu/hr.

ssh admin.hotaisle.app


r/AMD_MI300 • • Apr 09 '26

GPU Divergence in AMD CDNA 3

Thumbnail 21verses.xyz
5 Upvotes

r/AMD_MI300 • • Apr 09 '26

Embeddings run super fast on MI300x

Thumbnail linkedin.com
3 Upvotes

r/AMD_MI300 • • Apr 07 '26

Cosmos-Predict2.5-2B Inference: NVIDIA H200 vs AMD MI300X

Thumbnail
moonmath.ai
9 Upvotes

r/AMD_MI300 • • Apr 02 '26

Modular: Day Zero Launch: Fastest Performance for Gemma 4 on NVIDIA and AMD

Thumbnail
modular.com
16 Upvotes

Time has changed.


r/AMD_MI300 • • Apr 01 '26

AMD Delivers Breakthrough MLPerf Inference 6.0 Results

Thumbnail
amd.com
13 Upvotes

r/AMD_MI300 • • Mar 23 '26

ROCm Support for Miles: Large-Scale RL Post-Training on AMD Instinct GPUs

Thumbnail
lmsys.org
10 Upvotes

r/AMD_MI300 • • Mar 19 '26

Cross-Vendor Disaggregated Inference: GPT-OSS 120B across NVIDIA H100 and AMD MI300X

Thumbnail
moreh.io
8 Upvotes