Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

> This matches my experience with DeepSeek V4 Pro at Max reasoning,

That was ages ago (in LLM release timelines). DeepSeek V4 Flash beats it now and a lot cheaper.

> On similar tasks, GLM 5.2 at Max reasoning screwed up maybe 20-30% of the time,

GLM 5.3 bridges this gap.

> I'd say as Chinese models get better, whatever moat Anthropic and OpenAI have dissipates.

Their moat, especially OpenAI is funding and hardware resources. They gain train models 10x as large and also serve at large scale. That's it.



> GLM 5.3 bridges this gap.

I’m sure the next models will only get better, when they’re released. Also super curious about what Moonshot will achieve and the full DeepSeek V4 Pro release!

> Their moat, especially OpenAI is funding and hardware resources. They gain train models 10x as large and also serve at large scale. That's it.

I’ve seen how much slower Kimi K3 can be and that part seems correct, their own GPU production still has ways to go and export restrictions definitely limit what they can do.

Not sure about the size part, if Kimi K3 achieves SOTA performance at 2.8T parameters, western models being >2x that size would be insanely bad in regards to efficiency. I bet they’re all within the same order of magnitude and below 10T and won’t really have a reason to go even that high for the foreseeable future.

As investors will start squeezing them for profitability, I suspect focusing more on efficiency will be commonplace.


> Not sure about the size part, if Kimi K3 achieves SOTA performance at 2.8T parameters, western models being >2x that size would be insanely bad in regards to efficiency. I bet they’re all within the same order of magnitude and below 10T and won’t really have a reason to go even that high for the foreseeable future.

They're a lot larger e.g. Fable. It is insanely bad. Do you know how much more resources "Western" companies have? Most in China don't have random GPUs to "play with" like every "frontier lab" employee does.

> As investors will start squeezing them for profitability, I suspect focusing more on efficiency will be commonplace.

They're born lucky though. Efficiency is "free". The next generation hardware e.g. Nvidia claims Blackwell -> Rubin is 10x efficiency (verified by Neoclouds apparently).




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: