13 stories tagged with #kimi-k, in publish-time order across the WeSearch catalog. Tag pages update as new stories ingest.
⌘ RSS feed for this tag → or search "Kimi K"
Alibaba prices Qwen3.8-Max at $2 per 1M input tokens and $6 per 1M output tokens via its API, below Kimi K3's $3/1M input tokens and $15/1M output tokens (Henry Siu/The Information)
Artificial Analysis: DeepSeek's V4-Flash costs $0.14/1M input and $0.28/1M output tokens, or $0.03 per test, far below Kimi K3's $0.86 and GPT-5.6 Sol's $1.86 (Eduardo Baptista/Reuters)
Alibaba says its 2.4T-parameter Qwen3.8-Max tops Kimi K3 on some benchmarks, and plans to release the open weights of Qwen3.8-Max and Qwen3.8-27B next week (Luz Ding/Bloomberg)
Luz Ding / Bloomberg : Alibaba says its 2.4T-parameter Qwen3.8-Max tops Kimi K3 on some benchmarks, and plans to release the open weights of Qwen3.8-Max and Qwen3.8-27B next week —…
China's free Kimi K3 AI model shakes up global tech market
Governments can now deploy top-tier AI locally, bypassing costly U.S. cloud rentals.…
Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
The fastest open source LLMs for enterprise.…
US lawmakers investigate DoorDash's use of Moonshot AI's Kimi K2.6 model
Food delivery giant latest targeted as rising adoption of China-built AI models pushes Washington to find strategies to combat trend.…
With Moonshot’s free Kimi K3, China changes the sovereign AI playbook
Governments can now deploy top-tier AI locally, bypassing costly U.S. cloud rentals.…
Kimi K3-256k
Kimi Code Documentation…
Self-hosting Kimi K3: 20% more hardware cost, 20% better task resolution
aistack is an open benchmarking initiative by imec, generating solid datapoints where real AI workloads meet real systems and silicon.…
China's Moonshot AI reportedly used Nvidia Blackwell chips for training Kimi K3 — company circumvented both U.S. export and Chinese import controls to acquire compute
The US gov't says Moonshot has purchased Blackwell systems and rented time on foreign clouds to help it train models.…
How Profitable Is LLM Inference? Doing the Math on Kimi K3
A look at LLM inference economics (batch size, GPU count, and the Pareto frontier that sets token prices) applied to Kimi K3 with back-of-the-envelope math.…
I gave Claude Opus 5 and Kimi K3 15 impossible prompts — the winner surprised me
The new Chinese model takes on Anthropic's latest release…
Running Kimi K3 on a M1 Mac
Run Kimi K3, a 2.8T-parameter Mixture-of-Experts LLM, on a single Apple Silicon Mac. Streams MXFP4 experts on demand over HTTP into a local disk cache — fused NEON kernels, Metal/M…