Apps & Tools
Qwen3.8-Flash-Next 125B on AMD R9700
Full local setup and benchmark of the 125B Qwen3.8-Flash-Next model on AMD hardware.
Open source
Qwen3.8-Flash-Next 125B on AMD R9700 media is blocked
Allow external media to connect to the provider and play this content.
Description
A technical demonstration and guide for running the 125-billion-parameter Qwen3.8-Flash-Next Mixture-of-Experts model locally on a single AMD Radeon AI PRO R9700 GPU. The setup utilizes ROCm and llama.cpp to achieve local inference without NVIDIA hardware or cloud-based APIs, proving that flagship-scale models can run on non-CUDA consumer hardware.
Descriptions, tags, and model credits may be AI-generated or inferred from public sources and can be incomplete or wrong. Learn more.