3 ms·
It’s really cool that we have this proof that US companies are half year behind Chinese models in architecture.
by xiphias2 2mo ago
It’s really cool that we have this proof that US companies are half year behind Chinese models in architecture.
- geysersam 2mo agoWhat is the proof?
- cma 2mo agoNemotron was using hybrid with recurrence via mamba layers since around April 2025.
- firecall 2mo ago[dead]
- weird-eye-issue 2mo agoThen why are they (US frontier models) still so far ahead whenever I test them against the latest Chinese models? No bias here, I'd love them to be better for my own personal gain, but I haven't seen it
- regularfry 1mo agoBehind on architecture, ahead on training? It seemed pretty obvious to me that the opus 4.7 and 4.8 releases were more about trying to retain 4.6-level capabilities while being cheaper to run, which would fit. And they can burn so much money on training.
- weird-eye-issue 1mo agoI don't know I just care about the end result. And yeah what you're mentioning here is a pretty common conspiracy theory but you don't actually have any insight into that do you?
- regularfry 1mo agoApologies, you asked a question. I assumed that meant you were interested.
- weird-eye-issue 1mo agoYou didn't answer my question you literally just asked me another question and then parroted a common talking point about Opus models (which isn't even frontier - Fable is)
- regularfry 1mo agoMeh. You asked why, I suggested a possible reason, you them said you don't care about why. Congratulations, I now regret engaging.
- weird-eye-issue 1mo agoSorry I forgot to thank you for just adding more noise
- embedding-shape 1mo agoThere is so much misinformation in the ecosystem, parrots just hitting "Reply" without thinking one iota, you really cannot trust "human" opinions on the internet anymore, anywhere. Same with local LLMs, I'd love to use them for my day-to-day software engineering, and I'm not exactly GPU poor, then people with 12GB VRAM try to convince me their local setup is perfectly fine running latest Qwen and it does real engineering but whenever I try, they're a far cry from what Codex+GPT 5.x would do. Only way to be sure is creating your own private benchmarks and use those, and the difference in quality becomes very apparent, very quickly, for your specific use cases.