LOL! The US government is now the new LLM benchmark authority 😂
They say Kimi K3 is much worse than US frontier models
LiveBench AI said the same thing a few days ago
Kimi is a very cheap sonnet / opus 4.6 class model for long running tasks - that’s not frontier intelligence