I'll note on that Ryzen AI Halo? It's Memory bandwidth is 256Gbs, so it's operating at a level on par with Framework's Max 395+ Desktop.
That being said, for local AI compute, your options for Frontier Models is to either network multiple devices like this together through stuff like exolab's software, or to get something like a Threadripper/EPYC, or Intel equiv. to get together a bunch of old Server GPUs like V100s from Nvidia or the AMD options with a couple of adapter cards if you can't find a central server unit that the SXM interface expects.
At that point, we're less getting a model together for AI Waifu Coom sessions, and building out a complicated stack for Homelab tier of AI compute, which means we're either exploring some niche use cases for AI, or building a bespoke coding assistant, or such.