Navigate Select ESC Close

Kimi K3 Explained!

2026-07-17 Science & Technology
4.6k
131
27
Prompt Engineering
Prompt Engineering
245.0k subscribers

Unlock all features

FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.

Description

Kimi K3: The First Open-Weights Model Setting the Frontier (3T Params, MOE, Agentic Coding) I break down why Kimi K3 feels like a new “DeepSeek moment,” arguing it’s the first open-weights model that actually sets the frontier for specialized agentic coding. I cover its 3T-parameter scale, MOE architecture (896 experts, 16 routed per token), native multimodal vision, and 1M context window, plus how it ranks highly on the AI Index and excels in coding/web dev benchmarks like Terminal Bench 2.1 while lagging in general chat and some tests like Humanity’s Last Exam and hallucination rate. I explain why Kimi models are ecosystem-friendly bases for post-training (e.g., Cursor, Cognition/Windsor, Thinking Machines) and discuss frontier-level pricing, token efficiency, enterprise focus, and how this could pressure other labs to lower prices. I also preview upcoming comparisons vs Fable 5, GPT 5.6, Sol, and GLM 5.2. LINKS: https://www.kimi.com/blog/kimi-k3 https://artificialanalysis.ai/models/kimi-k3 My voice to text App: whryte.com Website: https://engineerprompt.ai/ RAG Beyond Basics Course: https://prompt-s-site.thinkific.com/courses/rag Signup for Newsletter, localgpt: https://tally.so/r/3y9bb0 Let's Connect: 🦾 Discord: https://discord.com/invite/t4eYQRUcXB ☕ Buy me a Coffee: https://ko-fi.com/promptengineering |🔴 Patreon: https://www.patreon.com/PromptEngineering 💼Consulting: https://calendly.com/engineerprompt/consulting-call 📧 Business Contact: [email protected] Become Member: http://tinyurl.com/y5h28s6h 💻 Pre-configured localGPT VM: https://bit.ly/localGPT (use Code: PromptEngineering for 50% off). Signup for Newsletter, localgpt: https://tally.so/r/3y9bb0 00:00 Kimi K3 Breakthrough 01:21 Frontier But Specialized 02:02 Agentic Coding Benchmarks 03:06 Ecosystem And Post Training 04:45 Architecture And Context 05:43 Pricing And Token Efficiency 07:36 What It Means For Users

Top Comments (10)

@Rationalific 2026-07-17

Thanks for the information! It's amazing how far open-weights models have come (even if this can't come close to running on personal hardware).

0
@RevtheDev1804 2026-07-18

Are you using it via terminal (Kimi Code) or App?

0
@dc33333 2026-07-17

the MAN

1
@ranjan561 2026-07-17

Great video as always! Can you tell me what tools you use to create the slides and materials that you present in your videos?

0 1 replies
@ValidatingUsername 2026-07-17

Trust me, the closed unreleased models are ASI self coding in air gapped black boxes

2
@cluelesssoldier 2026-07-18

What difference does it make if the cost to access it is close to the cost of frontier models? It’s basically just another frontier model, at this point lol.

0
@metrodyne 2026-07-17

Sonnet 4.6 was to best model to talk to, by far, for me. I really wish 5.6 and Fable 5 or Kimi 3 had the quality of output conversation Sonnet 4.6 had

1
@shApYT 2026-07-18

All western frontier labs: "What we propose to do is not to control content, but to create context."

0
@MeinDeutschkurs 2026-07-17

OK, grab now, run later (as soon we have more vram in our prosumer hardware. Reminds me to “OMG, 1 MB in an Amiga, so huge”. A CD with 700MB how can we ever fill this up? 😂😂😂

2 1 replies
@better_notes182 2026-07-17

Grea video as always ❤ no fluff, just good content. No hype. Helps get a good map of where things are. Thanks!

4 1 replies

Unlock the Data Inside
Turn Videos into Knowledge

  • Get FREE 10/day: transcripts, summaries, chats
  • Chat with videos, export text & PDF
  • $1 free API credit for RAG, chatbots & research

Free forever plan • All features unlocked

App screenshot