Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Do they really need HF for that?

Dgx Spark and Strix Halo have very close specs and deliver similar performance. If nvidia makes their stack be more efficient with for example 50% more tockens on similar hardware specs agaist competitors, they don't need HF.

My take is that Chinese Labs, though slightly behind on the frontier (due to compute constraints) are on a trajectory to surpass Western labs (this is me speculating, reasons are better ecosystem creation on China's part and potentially better/more data environment). Qwen-3.5-122b was the king in it's category and noone came up with something better, even though many tried like poolside with laguna. Similar with 35b and 27b param models. I think poolside and HF acquisitions show us that nvidia really wants to have competitive models on the prosumer (~100-150b param size) and likely at the 300-500b as well. Together with a hardware to run them that's a good market to be in. And as the recently rumored Xiaomi AI cube shows us (together with gorgon/medusa halo and mac studios), this is a market segment that will have competition.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: