My biggest frustration with the frontier AI companies isn't what they're announcing, but that the announced-thing that exists ~6 months later is severely nerfed to reduce compute spend. It doesn't resemble the demo in any way. For example, this was what the 4o voice capability sounded like in 2024(!) <a href="https://www.youtube.com/watch?v=vgYi3Wr7v_g" rel="nofollow">https://www.youtube.com/watch?v=vgYi3Wr7v_g. What exists today pales in comparison.
Pretty much. Always-on doesn't scale as well as JIT access to a massive array of GPUs, because always-on means it's always-using-memory - so you're at least going to paying the cost of a minimum chips VPS for every Dot right?
Maybe people are ok with that, but I feel like it would be nicer to just sell a cheap SBC like a raspberry pi that you can plug in (and unplug!) and just pay for the tokens used instead of having your data stored offsite and paying cloud prices.
mvkel · · focus · HN ↗
jasongi · · focus · HN ↗
Maybe people are ok with that, but I feel like it would be nicer to just sell a cheap SBC like a raspberry pi that you can plug in (and unplug!) and just pay for the tokens used instead of having your data stored offsite and paying cloud prices.