Posted inAMD Distributed Inference
Running DeepSeek V4 Flash Q2 on Strix Halo 128GB
AIfinitee belives local compute as an alternative to cloud and that the most capable open models should be runnable on hardware you actually own. This post is the recipe for…




