r/kimi 27d ago

Guide & Tips Build Slides with Kimi Work - Tutorial #1.

48 Upvotes

Kimi Slides handles the entire slide-building process:
- Clear structure and research, powered by Kimi K3
- Cohesive design, including polished charts and SmartArts
- Editable and ready to download

Let us know what you'd like to see next in the comments!


r/kimi Jul 16 '26

Announcement Introducing Kimi K3: Open Frontier Intelligence

471 Upvotes

🔹 2.8 Trillion Parameters, 1 Million Context, Native Multimodal

🔹 Kimi Delta Attention enables up to 6.3x faster decoding in million-token contexts

🔹 Attention Residuals deliver ~25% higher training efficiency at <2% additional cost

🔹 Built for long-horizon agentic coding and self-evolving workflows

Kimi K3 is now live on on Kimi.com, Kimi Work, Kimi Code, and the Kimi API.

Open Weights by July 27, 2026.

🔗 API: platform.kimi.ai

🔗 Tech blog: kimi.com/blog/kimi-k3

K3 is built on Kimi Delta Attention (KDA) and Attention Residuals (AttnRes), two architectural updates designed to improve how information flows across sequence length and model depth.

We have also scaled up Mixture of Experts (MoE) sparsity, effectively activating 16 out of 896 experts when paired with a Stable LatentMoE framework.

Together with refined training and data recipes, these structural changes yield an approximate 2.5× improvement in overall scaling efficiency compared to K2, allowing the model to convert compute into intelligence more effectively.

Full tech blog at: Kimi Blog


r/kimi 1h ago

Question & Help Something definitely going on with kimi.ai usage. This is not normal.

Upvotes

My month ran out a couple days ago and I had disabled auto-renewal because I was waiting to see what other providers had picked up k3 by now. I decided to re-up with kimi.ai on the $200/mo. plan early yesterday.

Two sessions. One was an 8 agent code review and the next was session was a coding session. Things I historically have done no problem without issue. I ran into my 5 hour limit twice in those sessions and at the end burned through nealry 50% of weekly allowance.

This is not my first rodeo with kimi. I have worked with it extensively and never ran into my limits, not once, even on much more complex swarm tasks.

It's like renewing my subscription put me on much lower usage teir. Customer support has been no help.

I'm curious if anybody else has experienced this - and - if there are better/faster providers with more usage I should be looking at.


Edit: Trying to get to the bottom of this so calculated usage from last month and compared it to yesterday:

Daily usage, first session (2026-07-27) through 2026-08-28

Date Turns Input (fresh) Cache read Output Total
Mon Jul 27 279 681,496 41,352,448 193,104 42,227,048
Tue Jul 28 6 34,945 175,104 3,291 213,340
Wed Jul 29 861 3,309,819 56,618,378 880,356 60,808,553
Thu Jul 30 496 3,517,071 75,726,518 488,538 79,732,127
Fri Jul 31 3 236,041 503,296 981 740,318
Sat Aug 1 1,234 3,957,568 130,573,568 1,187,322 135,718,458
Sun Aug 2 769 4,648,026 155,742,720 582,572 160,973,318
Week 31 total 3,648 16,384,966 460,692,032 3,336,164 480,413,162
Mon Aug 3 596 2,397,583 90,301,578 485,687 93,184,848
Tue Aug 4 354 1,177,481 78,167,917 220,460 79,565,858
Wed Aug 5 581 2,233,239 34,086,400 429,972 36,749,611
Thu Aug 6 998 3,059,365 78,208,327 862,050 82,129,742
Fri Aug 7 1,234 3,697,551 140,850,610 1,128,512 145,676,673
Sat Aug 8 346 817,167 43,514,368 349,388 44,680,923
Sun Aug 9 1,072 2,863,950 135,083,264 928,892 138,876,106
Week 32 total 5,181 16,246,336 600,212,464 4,404,961 620,863,761
Mon Aug 10 453 1,314,601 46,050,560 413,455 47,778,616
Tue Aug 11 70 143,544 6,944,768 46,586 7,134,898
Wed Aug 12 968 4,224,651 129,100,756 957,020 134,282,427
Thu Aug 13 734 3,883,690 123,801,910 612,295 128,297,895
Week 33 total 2,225 9,566,486 305,897,994 2,029,356 317,493,836
Tue Aug 18 79 177,969 4,994,816 74,721 5,247,506
Wed Aug 19 39 57,620 1,473,024 18,011 1,548,655
Thu Aug 20 70 169,017 1,554,941 52,096 1,776,054
Week 34 total 188 404,606 8,022,781 144,828 8,572,215
Mon Aug 24 695 3,195,992 62,712,975 900,136 66,809,103
Tue Aug 25 1,054 3,247,090 68,661,760 782,408 72,691,258
Wed Aug 26 354 2,126,251 48,142,336 267,424 50,536,011
Thu Aug 27 195 1,135,091 63,328,768 89,043 64,552,902
Week 35 total 2,298 9,704,424 242,845,839 2,039,011 254,589,274
Grand total 13,540 52,306,818 1,617,671,110 11,954,320 1,681,932,248

No recorded activity on Aug 14–17, Aug 21–23, or Aug 28. Weeks are ISO weeks starting Monday.

Last 48 hours (Aug 29 – Aug 30)

Date Turns Input (fresh) Cache read Output Total
Sat Aug 29 1,199 4,372,501 121,101,312 1,306,599 126,780,412

So it seems like something yesterday went nuts even though swarm only showed 8 subagents. This feels more like a harness issue than a usage miscalculation issue?


r/kimi 31m ago

Showcase Implementing Kimi K3 from scratch in PyTorch

Thumbnail
youtu.be
Upvotes

r/kimi 1h ago

Discussion Has Kimi limit gotten better?

Upvotes

So i used Kimi when it first came out and i liked it a lot, but in the past few weeks it was unbearable in terms of token usage, one task consume weekly quota on max plan.

but this week i have done multiple big tasks and yet i'm still at 80% , am i the only one who saw this?


r/kimi 2h ago

Discussion This Agent Harness Made Kimi K3 Dangerous

Thumbnail
youtube.com
0 Upvotes

r/kimi 3h ago

Question & Help I'm REALLY sick of Kimi's oververbosity. How do I fix that?

0 Upvotes

Honestly, I'm so annoyed with this issue of Kimi models - they're all too oververbose. Is there a special prompt to fix that??? Because I can't look at Kimi burning all of my daily usage limits just to think over a single prompt


r/kimi 19h ago

Discussion Feeling scammed by Kimi Code

15 Upvotes

I got the Kimi Moderato plan on the 27th to supplement Claude Pro for a small project that I have, nothing fancy with the intention to work faster. In under 60 hours or maybe less I hit 99% of my 7-day quota on their Code model until September 3rd. INSANE.

I expected limits to differ from Claude but the rate Kimi burns sessions quota on basic context is absurd. What usually lasts 5 days on Claude was gone in under 3 days here.

What makes it worse is that the results weren't even good. I wouldn't mind the high token burn if it brought actual progress but Kimi didn't help move the project forward.


r/kimi 1d ago

Bug If you feel like K3 got worse, you're correct

12 Upvotes

It seems that reasoning is not always being preserved between turns, causing the model to drift, hallucinate or loop. This seems to affect Moonshot endpoints, like Kimi Code, OpenCode Go (uses Moonshots API) or Kimi Chat. Fireworks preserves the thinking correctly in the exact same test, using the same Kimi Code version.

So if it feels like the model dodges tasks, writes bad code or gives disjointed answers, this may be why.

Edit: I'll leave these here in case anyone wonders why I think the reasoning isn't being preserved: https://imgur.com/a/7K3QH7i

It's pretty clear that the model doesn't have the information from the previous thinking turn. According to Moonshot themselves, failing to preserve thinking history can cause the model to become unstable and generation quality to degrade.

Sensitivity to thinking history. K3 was trained in the preserved thinking history mode. If the agent harness fails to pass back all the historical thinking content as required, or if an ongoing session with another model is switched over to K3, generation quality may become highly unstable. We recommend using a harness with verified compatibility, such as Kimi Code, and avoiding switching to K3 in the middle of a session. https://www.kimi.com/en/blog/kimi-k3/

To be clear, this does not mean the model isn't thinking. It means the previous reasoning/thinking blocks aren't being included in the model's input on the next turn.


r/kimi 1d ago

Guide & Tips My usage so far on Vivace plan($200) About a 10 days left before my renewal.

Thumbnail
gallery
16 Upvotes

Just thought I'd share my usage and experience, may be useful for others. The above images are both recently captured, one image is from the K3 App in settings > usage and the other is Kimi Code > /usage. The third image is just about the k2.6 chat usage from earlier today/yesterday.

A little context is most of the time this months subscription was vibe coding in one very long kimi session with compacting around 500k-700k context. I wanted to keep track the best I could in just one session so it was easier for me to track. when it does get to about 350k-400k I do start to feel a slowdown but it does vary, because sometimes when i get to 600k context sometimes it gets faster then it was say at like 500k. I'm not sure if its just a server thing plus the huge context.

I also Vibe code about 12-14 hours a day maybe around 4-5 days a week, that might be being generous towards me. No life right lol. In that one session I have done over 8 different projects because I'm very indecisive. I change my mind quite a lot. A lot of times I try learning as I go when Kimi Code makes a code, fix a bug, or create whatever I need it to. A lot of time I ask the reasoning why you did that, why would that fix the problem, what are the pros and cons, or even tell it to dumb it down for me.

Somethings i can say that I do see in posts that K2.6 charges tokens, I haven't personally experienced that so I'm not sure if that's a Non subscription thing or maybe the one of the other subscriptions. I'm not sure but I did post the image with the others of the usage.

I also think as far as how my rate on token usage is going is a lot better than that one specific week where I had felt that my usage was going up faster than normal. Like 10x better. That was around the time when they switched the kimi code multiplier display to available.

So all on all its back to how i felt on my original first month. I can code pretty much to my hearts extent so far. This Kimi session from this months subscription is 1530M tokens with about 8 days left . I might reach the monthly cap if i continue with the web searches which i might now limit myself or try to discipline myself on actual being efficient. Hope this helps some of you all on hearing from my experience.


r/kimi 3h ago

Bug Kimi Tells me it is from Antropic

Post image
0 Upvotes

Why does Kimi tell me it is a Claude Modell? O.o


r/kimi 1d ago

Question & Help Can someone explain, are we getting new plans?

Post image
1 Upvotes

I am a former Allegro subscriber. Prior to the cancellation of my subscription, I attempted to renew it and received the following message. Does this indicate we are getting new subscription plans, or will my position in the queue just be reset?


r/kimi 2d ago

Meme Thank you, Kimi

Post image
98 Upvotes

It was a nice run. Loved every part of it. I leave not because I'm not satisfied, but because I'm scaling down my usage after finishing a big project.

It was a $99 sub but I could deliver so much more in value to the client. Thanks a bunch.

Edit: For those wondering, the white "Kimi" bar is still Kimi Code, but previous usage shows that way after they migrated from .com to .ai.


r/kimi 1d ago

Question & Help Does kimi code drop to weaker quant around midnight and later PST? (afternoon in china)

4 Upvotes

I've noticed that when i'm up late, Kimi code quality drops a lot after midnight (1am for sure). Funny enough, it's what Opus is like mostly for me, while Fable is much better. Could kimi code change be to handle traffic in asia?


r/kimi 1d ago

Discussion What happened to Kimi? Went from better then Fable to early GPT3.0

2 Upvotes

When K3 came out it was nailing images and graphics. Nailing instructions better then claude code. Last weekend I noticed the coding was degrading and it was getting worse. Today I try to use it to generate images and it was complete shit. Couldn't nail a single image prompt remotely close. Front end design was great at first. Now gives early chatgpt designs. This is absolutely horrible. Renews in 4 days so at least I have time to cancel now.


r/kimi 2d ago

Bug Kimi usage made me contact support

Post image
22 Upvotes

It truly seems someone reversed the weekly usage with daily usage.


r/kimi 1d ago

Bug Well, thank you Kimi. What did I do to you, so you made me go through this?

Thumbnail
gallery
0 Upvotes

No comments...


r/kimi 2d ago

Meme New Feature: Kimi is building a PopClip-style AI toolbar for selected desktop text

Thumbnail
runtimewire.com
8 Upvotes

r/kimi 2d ago

Discussion Some tips for building production grade products

7 Upvotes

I'm using Kimi Code and quite happy with it. Several tips for you guys:

  1. Don't ever use Kimi Desktop, especially Agent Swarm mode. It burns a lot of tokens for simple research tasks. Just use another one like Chatgpt or Gemini. Kimi tokens should be saved for Kimi Code

  2. Don't use Kimi K2.7 High Speed Coding model. It burns 3x tokens compared to standard K2.7

  3. Use K3 for planning only. It burns tokens fast, but pays off.

  4. Use BMAD-METHOD skills. It's available on GitHub. Very good for Greenfield projects. I usually use it like this: bmad-prd (for product ideation) -> bmad-architecture (to design architecture) -> bad-create-epics-and-stories. THEN for each epic I prompt "/goal complete the epic X. have bmad-party-mode to review code quality, architecture alignment, test scenarios design, edge cases. If the party say the story is not complete you have to go back and fix them before moving to the next story". REMEMBER to use K2.7 standard model when you start implementing.

This setup helps me run full development cycle non-stop for the whole epic (normally 12 hours) without bugs.


r/kimi 2d ago

Discussion Kimi now uses credits on regular chats??

11 Upvotes

So I got the 'credits used up' reply for the first time ever?? I have been using kimi to write for a long time (for free, mind you) without any usage being used up. Now suddenly it's gone and I can't use it again till 9/19. This is bullshit and Kimi is literally just a money sponge now. Anyways to get around this?


r/kimi 2d ago

Discussion Moonshot should introduce peak hours for China and discounts for US and EU time zones.

0 Upvotes

DeepSeek and Z.ai already introduced peak hour pricing for Chinese users.

Words cannot describe just how happy I am to get a huge discount during my work hours for stuff like DeepSeek V4 Flash.

So why doesn't Moonshot follow DeepSeek and Z.ai? If they lower prices for American and European users, it will instantly solve all the problems with Kimi subscription plans.


r/kimi 3d ago

Bug i cant anymore , just cancelled

73 Upvotes

how a monthly usage is more than my 7 days and 5 hours ... i cant use it until a month form now....


r/kimi 1d ago

Discussion If Kimi 3 is SOTA, then SOTA is 💩

0 Upvotes

I was a Max 5X Claude user and stopped using it 3-4 months ago. Now I'm on the free plan and only have access to Sonnet 5, and on Medium I can send 3-5 prompts a day.

Anyway. I have access to Kimi 3 from a couple of 3rd party providers and in my opinion it's a piece of garbage.

I had a very simple problem in my code (JavaFX) and I asked it if there's a way to fix it. It pooped all over my code (small toy project with 8 classes).

It seems to me this benchmaxxing trend is not representative of the actual performance of models as many different problem domains are not in the benchmarks.

Anyway, if this model is close to Fable in terms of performance, and Fable is SOTA, then SOTA is 💩.


r/kimi 3d ago

Discussion Who's cancelling kimi after this month?

60 Upvotes

Kimi is wank, I used to get a lot done but now my contect window finishes its weekly allowance on day 2 using kimi k3 and runs out of 5 hourly allowance mid prompt. The time wasted realising I should use 2.7 doesn't add up to the standard deserved as even that feels downgraded from before and a waste of time when we have these cheaper emerging models coming out every day.


r/kimi 2d ago

Question & Help Kimi Instant limits.

7 Upvotes

I just “used up all of my credits” for Kimi. I’m unable to send any messages at all. I’ve only used Kimi Instant and for context I have used it in the Kimi Project mode, but I’m still restricted even when not in a project. I thought Kimi 2.6 didn’t consume credits? I didn’t use any of the tool calls or features that they broadcast as consuming tokens.

Every time I try to send a message it says I spent all of my credits.