touisteur 8 hours ago

As much as I enjoy these articles and for AMD to write more light technical articles, it really feels constrained, even strained, to be unable to cite the equivalent terms from the precursor here (NVIDIA). Another batch of jargon for very similar architectures and programming models... HIP and ROCm have actually made amazing strides in making CUDA developers' porting work easy, and I know playing catchup to a (monopolist) moving target you have no power over is bad... but I feel this is part of the thousand paper cuts.

broadsidepicnic 8 hours ago

So when can I buy MI350s for my homelab?

  • touisteur 8 hours ago

    Hopefully the MI350P is available soon, at last a standard PCIe SKU, if a bit too-much previous-generation and gimped compared to the MI350X

    • lostmsu 8 hours ago

      MI350 happens to not fit DeepSeek V4 Flash (144GB vs 180+ needed) so it will be dead on arrival.

      • boroboro4 7 hours ago

        For production loads? No one will run DS without expert parallelism so it’s not important for model to fit on one gpu.

        • smallerize 5 hours ago

          Do you mean people will run multiple GPUs, or that streaming a small % of experts from disk won't completely kill performance?

      • tpurves 4 hours ago

        This thing has 50% more memory than GTX PRO 6000, which fetch 18k-25k+ each right now on newegg, and generally sold out. If nvdia can still sell out 96GB PCI cards (on a 2 year old architecture) at those prices, how is an AMD equivalent with better specs dead on arrival?

        • lostmsu 1 hour ago

          Market takes time to catch up. When released Qwen 3.6 had a lot of sense and it did fit, but frontier since went too far ahead. Now you'd want DS V4 Flash from yesterday, or you are too far behind.

          Unless MI350P will be much cheaper than 6000 Pro, which is unlikely given its VRAM, it will lose to it, because you'd need 2 of either.