Register and share your invite link to earn from video plays and referrals.

Search results for chinesemodel
chinesemodel community
One keyword maps to one global community path.
Create community
People
Not Found
Tweets including chinesemodel
2026 Chinese model landscape in one list ▪️ Qwen3.8-27B, Qwen3.8-Max ▪️ GLM-5.3 ▪️ Kimi K3 ▪️ DeepSeek-V4 ▪️ ERNIE 5.1 ▪️ Baichuan-M3-235B, Baichuan-Omni-1.5 ▪️ Yi 1.5 ▪️ MiniCPM-o 4.5, MiniCPM-V 4.5, MiniCPM-V 4.6 They’re getting interesting in very different ways. Here we explained what makes each one stand out:
Show more
Another Chinese model just appeared on LMArena. Codename: “korrine” Likely the next Kimi (K3.1). K3 previously tested under “kivine”. This one is already live for testing.
Show more
The Chinese model also wore vintage Chanel couture and Givenchy by Sarah Burton to marry entrepreneur Mario Ho at Mont Saint-Michel—marking the first wedding to be held in its abbey in over 1,000 years—in Normandy, France.
Show more
A new open weight Chinese model has hit the timeline!! This time it’s from Tencent’s Hunyuan team with Hy4 preview. It’s a massive MoE with 770B total parameters / 49B active per token, a 1M context window, and the weights are released under Apache 2.0. Important caveat - Tencent says this is still an early Hy4 checkpoint with more pretraining + post training to come. I actually like how honest their benchmark sheet is. They show plenty of places where the model is still behind despite its size. Hy4 gets 85.4 on Terminal Bench 2.1, basically right in the frontier cluster, and jumps from Hy3’s 28.0 -> 64.3 on DeepSWE. But it’s still behind Kimi K3 at 74.0 and Claude Opus 5 at 74.7 there. On ProgramBench it gets 17.5 vs Claude’s 39.5, SWE Atlas Refactoring 53.3 vs 60.0, and Humanity’s Last Exam 43.4 vs 53.2. It’s a 49B active open weight model that is competitive for its size on a bunch of hard coding/agent benchmark. One thing Tencent also reported on GitHub is they ran a 163 person internal blind eval across 203 engineering tasks where Hy4 slightly beat GLM 5.3 and Kimi K3, which is pretty interesting.
Show more
Ox Alpha is a Chinese model Everyone is speculating it's GLM, but GLM is very capacity constrained It's more likely a Xiaomi model which seems very promising. Should be in ChatLLM shortly. Evals coming shortly
Show more
Noticed again that DeepSeek is the only Chinese model family that speaks fluent Russian even K3, 5.3 instantly slip into awkward Runglish the way Flash doesn't still probably the best general pretraining corpus over there (didn't check MiMo, Hy, Doubao, Qwen, were meh before)
Show more
in 7 days we are going to see the first Chinese model slaughtering all of the Western frontier lab models the flippening is happening (and they’re doing it with a fraction of the compute on huawei ascend clusters) 中国第一
Show more
Grok bot app used to make autonomous companies but instead powered by cheap subsidized Chinese model
My co-founder @ArmanHezarkhani broke down Kimi K3 (new 2.8T parameter Chinese model) on @FoxBusiness today with @cvpane. We had our engineers testing it all day before he went on air. Here's the gist: 1) K3 is very good. As good, if not better, than many of the American frontier models. 2) The cost story is being misread. Everyone expects a Chinese model to be the cheap one. Per token, K3 is competitive. Per run (the actual job you outsource to the model), it's as expensive as the frontier models. But it's supposedly open source, so developers will attack the cost curve. Give it weeks, not quarters. 3) The playbook should look familiar. It's the same one China ran on solar panels & EVs: flood the market with cheap supply, wait for the addiction, then move the price. 4) The internet & the space race were funded by the US government, and innovators competed on top of the platform. AI got funded by private markets, and now the same companies that spent those trillions are getting hamstrung right as China gives its models away. 5) On guardrails: the bad guys will have completely unconstrained tools no matter what we do. Foreign adversaries, and bad actors here at home. If the good guys don't have equally powerful tools, only one side is armed. 6) Commoditized models are actually good news for the hyperscalers. All that open source intelligence has to run somewhere, and Google & AWS will get paid a lot of money to run it. 7) The labs saw this coming too. It's why Anthropic & OpenAI are racing up the application layer instead of just selling tokens.
Show more
Ox Alpha seems pretty good to me. I don’t care if it’s a Chinese model, American model or a model I made up in my head. It’s cool.