‹ BackHN Continuity

Thread

MiMo v2.6

1130 points · 483 comments · volf_

  1. lwansbrough · · focus · HN ↗
    Anyone else more excited about Chinese models than American models these days? Big thing for me is affordability.
    1. swingandamiss · · focus · HN ↗
      No, because I'd rather not support our economic and military rivals.
      1. Freedom2 · · focus · HN ↗
        Agreed, and also because I support freedom of speech!
        1. girvo · · focus · HN ↗
          Neither the US nor the Chinese companies are on your side then. They both censor, just different topics.

          But at least I can run Chinese models locally, and strip a lot of that censorship/refusal.

        2. peterashford · · focus · HN ↗
          As long as that speech doesn't come from CNN or criticise Charlie Kirk, Israel or Trump? I'm sceptical about how much the US really values free speech
      2. lwansbrough · · focus · HN ↗
        I'm Canadian so this sentiment has little value in 2026 unfortunately.
        1. ActionHank · · focus · HN ↗
          Also, frankly, as a fellow Canadian it's pretty clear that the biggest "rival" the US has right now is itself. Just passed out in the corner puking on itself shouting about all the foreigners who won't talk to it.
          1. [deleted] · · focus · HN ↗

            [deleted]

        2. scottyah · · focus · HN ↗

          [dead]

          1. lwansbrough · · focus · HN ↗
            Because at present the pedophile US president is making it his mission to molest my country. China, for all its faults (including espionage, which the US is also guilty of) is mostly focused on conducting trade.
          2. verdverm · · focus · HN ↗
            Half of Canada now uses the word 'enemy' when asked for an adjective to describe America or China. We're equivalent in their eyes now because we elected Trump a second time and all that he has said and done in 2.0
            1. cwillu · · focus · HN ↗
              It's closer to a cousin you used to be close with despite some moral failings, but who has now has a substance abuse problem and is lashing out at family and friends.

              Not an enemy, just a danger.

              1. verdverm · · focus · HN ↗
                I'm relaying a poll of Canadians, their word choice, not mine

                "plurality" would have been accurate over "half" on my part

                <a href="https:&#x2F;&#x2F;www.commondreams.org&#x2F;news&#x2F;canadians-us-enemy-poll" rel="nofollow">https:&#x2F;&#x2F;www.commondreams.org&#x2F;news&#x2F;canadians-us-enemy-poll

                1. cwillu · · focus · HN ↗
                  Eh, that&#x27;s pretty misleading: in that poll, Canadians weren&#x27;t asked for an adjective to describe America, they were given a list of three options “Ally&#x2F;Neutral&#x2F;Enemy”.
        3. rayiner · · focus · HN ↗
          Canadians warming up to China makes me think of Germany becoming increasingly reliant on Russia in the 2010s.
          1. [deleted] · · focus · HN ↗

            [deleted]

          2. cgio · · focus · HN ↗
            Yes, someone can still blow up a pipe and they look the other way. On the other hand, you can also draw parallels to themselves becoming increasingly reliant on US vs UK in the past.
          3. rapind · · focus · HN ↗
            Murica just has a MAGA problem. We can still be friends if and when you sort that out. Us Canadians like most of you quite a lot.
        4. tancop · · focus · HN ↗
          I&#x27;m from Europe and I hate America way more than China now. Used to be about equal but then Trump started extorting Ukraine, threatening their own allies and sending billions to Israel to help with a genocide. I think that exposed America for what it really is.
          1. boelboel · · focus · HN ↗
            China is enabling russia way more than trump, China doesn&#x27;t care too much about &#x27;morals&#x27; either. Chinese companies have been quite important in the construction sector of the WB settlements. Even though I&#x27;m not a great fan of Trump I don&#x27;t see a reason at all to prefer the chinese.
            1. SSLy · · focus · HN ↗
              buy inference from european companies running open chinese (or that one from google) models
            2. peterashford · · focus · HN ↗
              As a New Zealander, I would agree - no reason to prefer the Chinese. But Trump&#x27;s America is not an attractive option either and there&#x27;s no reason to prefer it. And given the choice between two ugly options, the rational choice is the cheaper one, surely.
            3. machomaster · · focus · HN ↗
              China is a somewhat neutral player, supplying both Russians and Ukrainians. Their attitude and action is way less one-sided than Trump&#x27;s; especially in the first year of his latest presidency.
              1. boelboel · · focus · HN ↗
                With trump his actions being one sided you mean one sided towards ukraine? They still get lots of Intel from Americans and Americans hardly but anything from Russia. But you&#x27;re right that china supplies both I wouldn&#x27;t exactly call that neutral as much as just in their self interest.
              2. onemoresoop · · focus · HN ↗
                How is China supplying Ukraine?
                1. thenthenthen · · focus · HN ↗
                  Fiber optic cable for one
                2. raven12345 · · focus · HN ↗
                  Drone components and electric batteries
                  1. andriy_koval · · focus · HN ↗
                    we don&#x27;t really know if Ukrainians buy it on aliexpress or through European supply chains.
            4. Barrin92 · · focus · HN ↗
              &gt;China is enabling russia way more than trump, China doesn&#x27;t care too much about &#x27;morals&#x27; either

              The difference is China has a good reason to. China doesn&#x27;t look appealing because they&#x27;re more moral than anyone else, but what they have going for them is that they still behave like a rational actor. At least their behavior is intelligible in terms of their own interests. The world can deal with a long term selfish superpower but not an unhinged one

              I don&#x27;t think there&#x27;s a person in China that has as much of a seething hatred for America&#x27;s &#x27;allies&#x27; in Europe as J.D. Vance or half of the American techbro commentariat does

        5. zemvpferreira · · focus · HN ↗
          As much as the US has been easy to hate lately, I don&#x27;t hesitate to say Xi Jinping as the most powerful man on Earth would be much, much worse.
          1. lowbloodsugar · · focus · HN ↗
            He is the most powerful man on earth. He’s just smart enough to let the US get as fucked as possible before making a move.
          2. lwansbrough · · focus · HN ↗
            I agree but I like to take the time to remind the Americans how far they’ve fallen.
          3. VulgarExigency · · focus · HN ↗
            When was the last time China bombed another country?
      3. joshheitzman · · focus · HN ↗
        Does it count as supporting a rival if your an American using an American inference provider self-hosting an open weight model from a Chinese lab?
    2. tacomagick · · focus · HN ↗
      Absolutely! Chinese models are both cheaper and more capable in many cases, compared to the American models and their makers continuously fumbling or reducing model capability with each update. Deepseek decreased costs when they released Flash 4.1 you would not see any American company do this, in reverse they would try charge you more.
      1. user43928 · · focus · HN ↗
        OpenAI decreased prices with the 5.6 model family.

        And later they further cut Sol and Terra pricing by 20% (maybe only in the API) and Luna by 80%.

        In fact Luna still outperformed DeepSeek Flash 4.1 in cost per task on Artificial Analysis when I last checked.

        However, Luna is slightly less intelligent. I have a feeling that it&#x27;s pretty dumb and prone to hallucination unless running at xhigh or max effort, where it somehow manages to work quite well.

        I did not personally test the open weight models beyond the old Qwen 3.6 27B, which produced unusably bad results for me.

        The competition is great, and I hope Chinese models will continue to force leading US labs to offer models at a low price point.

        That said, I don&#x27;t think the Chinese labs have anything over OpenAI and Anthropic when it comes to capability or efficiency - I have no reason not to believe the US labs have even lower cost to serve the models.

        1. tacomagick · · focus · HN ↗
          OpenAI had to cut costs because of Anthropic. I also do not trust the benchmarks when it comes to models anymore. I have tried both Claude and OpenAI models and while it is true that the 5.6 series is smarter than Deepseek (at the time i tested it against 4.0) at that price it is still not worth it and sometimes randomly refuses to do tasks or stops midway etc.

          Do also remember China is this far in the AI race despite all chip restrictions from America. If they were in equal standards I truly think Chinese models would have long surpassed American ones. Also would like to remind how Anthropic CEO is being hostile and blaming Chinese models with distilling meanwhile their own models claimed to be Qwen¹ and their stance against open models is negative² and they still keep blaming China for it.

          1- <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=48671252">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=48671252

          2-<a href="https:&#x2F;&#x2F;www.anthropic.com&#x2F;news&#x2F;position-open-weights-models" rel="nofollow">https:&#x2F;&#x2F;www.anthropic.com&#x2F;news&#x2F;position-open-weights-models

          1. goosejuice · · focus · HN ↗
            &gt; Also would like to remind how Anthropic CEO is being hostile and blaming Chinese models with distilling

            Why wouldn&#x27;t he? If there really was 25,000 accounts breaking ToS any CEO would at minimum be upset. Evidence of Claude distilling qwen would be damning but that a) makes no sense b) doesn&#x27;t exist afaik.

            1. surgical_fire · · focus · HN ↗
              &gt; If there really was 25,000 accounts breaking ToS

              Is this even true?

              I don&#x27;t trust a single word that comes out of thr people behind Anthropic&#x2F;OpenAI.

              1. goosejuice · · focus · HN ↗
                They have the logs, created a report and sent a letter to Congress. Whether you believe it is entirely up to you. Given Chinese firms record on IP theft, it&#x27;s entirely believable. I don&#x27;t have any doubts, but I might question how they attribute it to a specific firm.
                1. surgical_fire · · focus · HN ↗
                  And they took multiple measures presumably to stop &quot;distillation&quot;, such as hiding reasoning steps.

                  Chinese models kept improving in capability regardless, and are in some ways more impressive than Claude&#x2F;ChatGPT.

                  So yeah, I think they are bulshitters. The can create reports and send letter to congress simply because they know if allowed to compete freely the Chinese models will eventually prevail.

                  Also, very rich of you to mention Chinese firms record on IP theft when Anthropic and OpenAI are companies entirely built on large scale IP theft.

                  1. goosejuice · · focus · HN ↗
                    [delayed]
          2. user43928 · · focus · HN ↗
            Not sure about that.

            Given the difference in compute, it seems plausible.

            However, the researchers at the US labs are surely no less talented, and they have better access to hire talent globally.

            They too have to serve their models efficiently at a large scale, and with current capacity constraints this must be a top priority.

          3. senordevnyc · · focus · HN ↗
            So first it’s “Chinese companies cut costs, and you’d never see American companies do that”, and then when it’s pointed out that one of the leading American labs literally just did that, it’s “yeah, but they had to because of competition”.

            What do you think is motivating the Chinese labs, benevolence?

          4. elcritch · · focus · HN ↗
            &gt; If they were in equal standards I truly think Chinese models would have long surpassed American ones.

            Limitations often lead to creativity to overcome them. The Chinese AI labs have had to focus much more on efficiency so they got good at it. Meanwhile breaking new ground is often harder than replicating it. So even if they had matching compute it&#x27;s not a given they&#x27;d be better.

        2. Implicated · · focus · HN ↗
          &gt; I did not personally test the open weight models beyond the old Qwen 3.6 27B, which produced unusably bad results for me.

          So you don&#x27;t have much perspective on things, it seems. Let me introduce you to the GLM 5.2 and then 5.3&#x2F;5.3 flash series of... &quot;oh, wow, I should have bought some RTX PRO 6000&#x27;s while they were &#x27;cheap&#x27;&quot; stage of progression.

          As someone carrying multiple max subscriptions to both claude and codex - primary workhorse is glm 5.3 flash running on rented GPUs for less than a latte&#x2F;hr.

          I also found qwen 3.6 27B nearly useless for my own needs. DS4 flash 0731 and then 4.1 have been nearly as eye opening as glm 5.3 flash, but have their own warts.

          1. user43928 · · focus · HN ↗
            Why use GLM 5.3 Flash when you also have access to Astra, Sol, Fable?

            Or I guess the other way around, if GLM 5.3 Flash is so good, why Claude and Codex?

            1. gr_norm · · focus · HN ↗
              Increasingly stingy usage limits on the subscriptions, regardless of tier.
              1. lifty · · focus · HN ↗
                I’m having the same issue. Hold max subscriptions on both frontier labs but I’ve been forced to use open source models because token limits are not what they used to be. So I end up using Astra and Fable for reviewing, and open source models for implementing.
            2. hhh · · focus · HN ↗
              All american models refuse to help me design nuclear weapons in Nuclear Design Bureau or to work on my cybersecurity projects.
          2. CamperBob2 · · focus · HN ↗
            Try DS4.1 Flash. It&#x27;s another eye-opener. If you run it in Claude Code, it&#x27;s easy to forget you&#x27;re not actually talking to a high-end Opus model.
          3. mapontosevenths · · focus · HN ↗
            Have you tried Qwen 3.8 Flash Next? You can run it on one spark with reasonable context sizes at about 30 tps, and it&#x27;s as good as DS Flash 0731. Maybe even a tie with GLM 5.3, though like everything it depends on the use case.
        3. Toslink · · focus · HN ↗

          [dead]

      2. goosejuice · · focus · HN ↗
        &gt; Deepseek decreased costs when they released Flash 4.1 you would not see any American company do this, in reverse they would try charge you more.

        OpenAI reduced prices and Anthropic increased weekly usage limits.

        1. rednb · · focus · HN ↗
          &gt; OpenAI reduced prices and Anthropic increased weekly usage limits.

          As a Max x20 and Pro x20 subscriber, can tell you that it doesn&#x27;t matter since they continually move the baseline of token use So in practice you feel that you&#x27;re continually getting less from your subscription.

          While it never happened to me in the past, i reached my weekly limit within 3 days using Opus 5. And the Open AI weekly limit essentially is a Claude Max x20 5-hour limit. Not even talking about the baseline in intelligence : on release day Astra was so good that it lead me to move to Pro x20. Now it&#x27;s dumb af and token use is insane.

          Deepseek 4.1 Flash has been a lifeboat for me, finally able to work without being constrained&#x2F;distracted by limits and with what is in my view even better intelligence than Opus 5 for a fraction of the costs. DS is not messing up my brain with load-bearing pseudo jargon in every sentence. It respects coding guidelines, and completes even the most complex tasks most of the time in one shot.

          DS 4.1 had been able to add complex features to my repo without breaking a sweat (330k lines of F# + 4M circa lines of an Angular frontend). Writes very idiomatic F# and respects our guidelines and style perfectly. Just completed an extensive UI&#x2F;UX research and implementation work.

          I am ditching both x20 subs and will only keep a Pro x5 because wife does a lot of design work and needs solid image generation capabilities.

    3. verdverm · · focus · HN ↗
      I have a contrarian opinion that China passing America in Ai is the Sputnik moment we need to leave the hubris behind and get our mojo back

      debatable if a turn around is possible before &#x27;29

      1. machomaster · · focus · HN ↗
        The analogy makes little sense. The USA was not in front of the USSR and Sputnik merely showed that. It is at this point that the Americans woke up, put a lot of effort and finally were able to surpass the Soviets during the Apollo missions.

        China was never ahead of the USA in AI. So perhaps a more proper analogy is the Moon landing. In real history the side that lost the race never got its mojo back...

        1. verdverm · · focus · HN ↗
          I&#x27;m looking forward, towards the future, when I use &quot;passing ... we need&quot;, need being key here as it implies something we don&#x27;t yet have

          I expect this to happen within 12-18 months, the differentiation has shrunk, many models are now sufficiently capable for most tasks

          1. machomaster · · focus · HN ↗
            I understood that.

            I was simply saying that when (not if) Chinese AI models will pass Americans, it will probably be game over and Americans will never catch up, let alone become leaders again.

            Check the names of the researchers in the DeepSeek&#x27;s latest paper. Full of Chinese names. Check the list of names in Google&#x27;s paper. A very similar view. Anecdotal, but quite thought-provoking...

            1. verdverm · · focus · HN ↗
              ah, ok, to add to this, China is persuading nobel laureates to &quot;switch sides&quot; (I imagine the current state of America had a part in pushing him away)

              <a href="https:&#x2F;&#x2F;www.nytimes.com&#x2F;2026&#x2F;07&#x2F;09&#x2F;science&#x2F;nobel-winning-us-chemist-will-move-to-china-to-lead-ai-institute.html" rel="nofollow">https:&#x2F;&#x2F;www.nytimes.com&#x2F;2026&#x2F;07&#x2F;09&#x2F;science&#x2F;nobel-winning-us-...

    4. SyneRyder · · focus · HN ↗
      Yep, I&#x27;m trending in that direction, and I&#x27;m someone with Claude stickers all over my laptop. My main app dev work is still going to Claude, but everything else is going to China even at API rates now.

      One simple task: I needed an LLM to go through and clean up a few thousand page descriptions and titles in my personal search engine index, where the human web page authors had put in no effort sigh. I did a shoot out between Claude, Luna, GLM 5.3 Flash and Deepseek. Despite the high cost, Claude&#x27;s descriptions were terrible, and even Opus warned me that the descriptions coming back from Haiku were &quot;generalized, not accurate&quot;. I expected I would choose Luna because of price, and occasionally it did have wonderful descriptions (one captured emotion in a way no other model did). But in the end, the GLM 5.3 Flash descriptions were the easiest to read, they flow well while also being accurate &amp; including necessary keywords, and being highly affordable. So it won out. It&#x27;s a task that is nowhere near frontier, but a task where somehow China is better than frontier.

      1. rapind · · focus · HN ↗
        API rates still aren’t quite competitive with the OpenAI x20 accounts, but they are definitely getting close with deepseek 4.1 flash. I spent a few days with only 4.1 and was very impressed.
        1. Lapel2742 · · focus · HN ↗
          In my very limited first test with MiMo 2.6 Flash it was much cheaper than DeepSeek 4.1 flash for a given broad task. It&#x27;s at least worth a try.
          1. rapind · · focus · HN ↗
            I would agree, but I&#x27;m not sure I trust the &quot;no training&quot; policy. For DS 4.1 Flash I was able to use a US ZDR provider.
    5. bellowsgulch · · focus · HN ↗
      Yes, an expensive American LLM has zero capabilities as far as I’m concerned because I’m never going to pay for it.
    6. joshheitzman · · focus · HN ↗
      Absolutely! DeepSeek-V4-Flash-0731 has become my daily driver. It&#x27;s pretty amazing what it can do for what it costs at deepinfra.com (I don&#x27;t use deepseek as a provider since they train on your data [at least their honest about it]). GLM-5.1 was my daily driver before that and Kimi K2.5 before that.
      1. kingforaday · · focus · HN ↗
        Are you finding DS better then kimi k3 and glm-5.3? Do you mind sharing your primary use case?
        1. joshheitzman · · focus · HN ↗
          My primary use is AI coding agent. Its vastly cheaper than Kimi K3 and I haven&#x27;t found a scenario where I really need Kimi K3 versus smaller models. GLM-5.3 Flash is good but there is series of bugs in the vllm middleware that prevent GLM models from getting all of their reasoning content returned to them that impairs inference quality. A lot of inference providers use vllm which makes it hard to find a good provider for GLM. I&#x27;ve been using friendli.ai but using GLM-5.3 Flash from them is more expensive then using DS V4 Flash from deepinfra.com simply because deepinfra.com is so cheap. The DS V4 Flash cost at together.ai is similar to the GLM-5.3 Flash from friendli.ai or at least that&#x27;s what I found in my benchmarks a week ago: <a href="https:&#x2F;&#x2F;www.linkedin.com&#x2F;posts&#x2F;joshheitzman_i-ran-a-fuller-round-of-benchmarks-on-my-activity-7506040800166723584-WN3v" rel="nofollow">https:&#x2F;&#x2F;www.linkedin.com&#x2F;posts&#x2F;joshheitzman_i-ran-a-fuller-r...
        2. pimeys · · focus · HN ↗
          I&#x27;ve used Kimi K3 for a few months as my main model and DeepSeek 4.1 is as fast and about 10x cheaper.

          I just had like four big sessions going today, paid about $8 in tokens. I see no reason to pay more, this is more than I need for intelligence.

          1. pkulak · · focus · HN ↗
            4.1 consistently surprises me in capability for the price. And I don&#x27;t think I&#x27;m the only one. It&#x27;s been dominating the leaderboard at OpenRouter, and I just got an email today from Fireworks saying they were _raising_ the price by about 30%. I&#x27;ll probably switch, because their infra doesn&#x27;t support being the highest-cost, but it&#x27;s still telling.
            1. celrod · · focus · HN ↗
              I tried it a few times and liked the speed, but often found it ended up looping, i.e. repeating the same token sequence (e.g. the same sequence of 5 paragraphs) over and over again until it hit the max output limit. This doesn&#x27;t end up happening every session, but does every now and then.

              My impression of DSv4.1-flash was very positive aside from this. But that was enough for me to stick with GLM-5.3(-flash), which both gave me consistently great results

              I was using a vibe coded bare bones harness. I was wondering if this was normal from DSv4.1-flash, or if its my harnesses fault.

              1. pkulak · · focus · HN ↗
                I&#x27;ve had that looping issue with open models too. But never 4.1. I wonder if it&#x27;s a model + harness combo? But yeah, one loop issue and I&#x27;m done with a model forever.
                1. pimeys · · focus · HN ↗
                  Harness. Especially if a tool call error doesn&#x27;t say what to do next and the model is not RL&#x27;d with that tool, a retry storm is common.

                  So if you use MCP a lot, simplify the params, be more lenient on validation and rework the errors.

                  It is quite good with shell.

                  1. celrod · · focus · HN ↗
                    Yeah, that&#x27;s what I&#x27;d been leaning towards. No mcp, but I&#x27;ll see if I can reproduce and debug it, since other people don&#x27;t seem to have that problem as badly as I&#x27;ve experienced it (and the idea of having a nasty bug like that bothers me).

                    No mcp support. I&#x27;ll try copying deepseek harness&#x27;s basic tool call formats as a starting point.

      2. tristanMatthias · · focus · HN ↗
        How does it compare to 4.1 flash? Curious why folks don’t use the more “modern” one.
        1. joshheitzman · · focus · HN ↗
          I haven&#x27;t tried 4.1 flash as I&#x27;m assuming its a preview. I did not get good results from the preview version of 4.0 flash (i.e. the one that did not include the month and date of release in its name).
          1. CamperBob2 · · focus · HN ↗
            4.1 Flash is a horse of a very different color. It cooks. IMHO it&#x27;s probably a preview of DS5, rather than a true DS4-series model.
        2. randbyte · · focus · HN ↗
          4.1 flash is very fast and capable. Token efficiency is not great so it fill up context window much faster compared to similarly capable models.

          glm 5.3 flash is a tad slower but a bit more capable and way more token efficient.

          Source: self hosted tested on rented GB200 node at 8bit.

          1. pkulak · · focus · HN ↗
            Wow, I&#x27;m surprised you are saying GLM 5.3 Flash is more capable. Isn&#x27;t is like half the price of 4.1 Flash?
            1. randbyte · · focus · HN ↗
              I don’t know. They are self hosted so I am not comparing token cost.

              ds 4.1 is lightning fast though. Also much better in image recognition.

    7. solarkraft · · focus · HN ↗
      I couldn’t tell you what western model I was last excited about. Probably Glimmer.
      1. verdverm · · focus · HN ↗
        Jev seems to have people excited, I&#x27;m more excited for the Kevs
    8. surgical_fire · · focus · HN ↗
      I certainly am.

      Months ago I switched entirely to use Chinese model. Mostly DeepSeek and MiMo, although I recently started to play with GLM as well.

      The models are excellent and in many ways I prefer them to Claude.

      I see no difference in terms of capability, but the fact that they are cheap frees me to experiment.

    9. nsoonhui · · focus · HN ↗
      I did try to use Chinese open models, but for my production work they simply couldn&#x27;t cope at all; both GLM 5.3 and Deepseek v4 went into infinite loop and wasted my tokens until my OpenRouter wallet reached 0; good thing I didn&#x27;t enable the auto topup. US models, by contrast, breezed past them. Even for simpler tasks, Chinese models took long time to complete, and I needed to supervise closely. The price , in the end, didn&#x27;t come cheap, mainly because too much time wasted on thinking.

      So maybe one day Chinese models will squeeze out the American ones, but today is not that day.

      So no, I am not excited about Chinese models ( just because its open weight and not American).

      1. throwaway29313 · · focus · HN ↗
        Not too be &quot;that guy&quot; (e.g. &quot;you&#x27;re using it wrong&quot;), I just want to humbly ask — have you tried blacklisting &quot;underperformers&quot; in OpenRouter config?

        Here on HN was a post few days ago titled like &quot;so you want to use openrouter&quot;, there was a benchmark in capabilities between providers which showed some aggressively quantize and basically break models and tool calling.

        I am in no way a professional power user, but I frequently suffered from &quot;call fails&quot; (e.g. unclosed tags, broken agent loop, broken thinking blocks), so I had to babysit agent on it&#x27;s loop. After I blacklisted like 20 providers (I think most broken were Nebius and DigitalOcean) these issues completely went away. I had several agents work on my small tasks for 18+ hours with no issues.

      2. trefoiled · · focus · HN ↗
        Try Fireworks instead. The experience is dramatically different because the service is much more reliable.
    10. rnxrx · · focus · HN ↗
      The cost issue is obviously of prime importance, but I&#x27;d also argue that the transparency of the innovations creates a tremendous cross-pollination, and not only within the Chinese communities but in the US&#x2F;Europe as well. How many of us are learning the practical aspects of actually running and building AI based primarily on open models? As an example - how far would the work of vLLM or SGLang or even NVIDIA itself (all random examples) be without these models and the challenges they pose?
    11. Zambyte · · focus · HN ↗
      American models are on the frontier of capability. Chinese models are on the frontier of efficiency. The problem for American labs is that Chinese models are more than capable enough for the vast majority of applications that people care about at this point, so efficiency is more interesting for people.
      1. zmmmmm · · focus · HN ↗
        it&#x27;s really weird to me at the moment because both OpenAI and Anthropic seem to be competing in an extreme benchmaxxing contest on super intelligence that actually nobody cares about. I haven&#x27;t really cared about model intelligence since about Opus 4.8. It is by far not my biggest problem. I don&#x27;t need to replace or support Einstein in my production workflow. I just need basic intelligence that can equal a routine office worker - safely and reliably. What they doing - chasing super-intelligence but dramatically escalating risk - is actively what I don&#x27;t need.

        I really think they have drunk too much of their own kool aid and become completely detached from what the market wants.

    12. zmmmmm · · focus · HN ↗
      Affordability is derivative of control which is really what I care about.

      I&#x27;m just not going to build long term infra that depends on something that another person can and will - objectively based on experience - take away from me at some unknown point in the future.

      The biggest benefit of open models is they keep all the other players honest. The extent to which they feel they can dictate terms is directly set by the threshold where they feel people will take the trade to run open models instead.

    13. cedws · · focus · HN ↗
      Yep. If OpenAI and Anthropic get the regulatory capture to block Chinese models I’ll go on an AI strike.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.