kimi
6 stories and discussions about kimi, aggregated from every source we track.
Ember-1 is a new specialized model from Fireworks Research that delivers Kimi K3’s quality with 40% fewer tokens.
How our Kimi K3 megakernel on TPU v7 reaches over 700 tokens/s with speculative decoding and nearly 2× GB200's batch-one decode throughput.
We pitted Kimi K3 against Claude Opus 5.5 on Pokémon Emerald. Both models found a novel link-cable zero-day and built a network worm that turns the game into Tetris, But their tool usage, safety refusals, and exploit…
Enterprise teams can now use open models like GLM-5.3 Flash and Kimi K3 natively in Codex and count spend against their OpenAI commit.
Linear RNNs based on the delta-rule enable efficient sequence modeling, but their linear updates with a low-rank correction constrain their expressivity. Prior work has shown that composing two delta-rule transitions…