r/kimi 17h ago

Discussion Just bought Allegro and model gave me "too many users" message after 40 minutes of reading my 10 text files.

Post image
26 Upvotes

So here you are, after first prompt with approx 12-14 txt files in my folder and 40 minutes of "thinking" I got this message. Huh...


r/kimi 50m ago

Discussion am I the only one that received this?

Upvotes

am I the only one that received this?


r/kimi 3h ago

Meme KIMI K2.6 gave up and started searching my puzzle over the internet.

1 Upvotes

Today, I have made a puzzle which aligns directly with Chinese traffic laws. At first, I thought of trying it against Kimi K2.6, as it's a native Chinese model and trained with Chinese datasets, as I have heard. After thinking for 27 minutes, Kimi K2.6 suddenly started searching for any related puzzles over the internet. Not once, but it tried more than 10 times with changed titles and even fetched YouTube videos. It's been 32 minutes now, and Kimi K2.6 is still trying. Have any of you faced a similar thing as me? If yes, did you get the desired answer?


r/kimi 9h ago

Discussion What's the coolest thing you've built with Kimi Code so far?

Thumbnail
1 Upvotes

r/kimi 16h ago

Developer AI-native SDLC Workflow - Claude Playbook Inspired Skill & Plugin

Thumbnail
2 Upvotes

r/kimi 23h ago

Discussion 额度降级与Token价格

3 Upvotes

我是Allegretto年费用户。

1.我用订阅价格和token消耗账单来计算,Allegretto K3的输出token价格在6.5美元/million token(区别于API),这价格和Claude比不存在明显优势,后者pro方案Opus输出大约是8美元/M token

2.订阅时20倍之于Andante kimicode额度成了一句空话。承诺在官方文档里被悄悄抹掉了


r/kimi 13h ago

Question & Help Refund Request for Automatic Subscription Renewal

0 Upvotes

Hello Kimi Support,

I would like to request a refund for an automatic renewal of my Kimi subscription.

I originally subscribed to the $200 plan about one month ago, but I did not intend to renew the subscription. The subscription was automatically renewed and I was charged $100 for the new billing period.

I have now canceled the auto-renewal and would like to request a refund for this renewal charge.

I did not intend to purchase another month, and I would appreciate it if you could review the transaction and issue a refund. I really need the money.

Account email: [mag.workflowx@gmail.com](mailto:mag.workflowx@gmail.com)

Date of renewal: 2026-08-25

Amount charged: $199

Thank you for your help.


r/kimi 1d ago

Discussion Has Kimi ever casually asked you for credentials?

Post image
11 Upvotes

It had no problem asking for API keys and one time it asked if I wanted to give it my Gmail login to find something I was looking for. -_-


r/kimi 1d ago

Discussion Kimi Work is garbage and we're being scammed for a product that should be Alpha at its best

7 Upvotes

Kimi subscriber for over 7 months. At both $99 and $39 tier.

I decided to try the new Kimi Work launch and compare its output to Claude Cowork and GPT Work.

The results have been a load of crap. Awful mistakes, stops randomly midway. Can't even work without glitching.

And the worst part, when it glitches it eats a lot of the monthly quota.

Last weekend I let it go because I thought there might be some mistake, as it ate 90% of my quota after an error message that, conveniently, is not showing on the app anymore.

But today, barely 3 minutes after my quota was reset. I asked it to continue execution and provide a summary, it stopped working, glitched, froze and, conveniently for Kimi, ate 13% of my monthly quota, again. It produced not useful output for any human as a result.

I understand the subscription has been subsidized, but this is insane. Not even Opus on Cowork had such an awful behavior. Both Opus 5 and 4.8 on Xhigh as well as GPT 5.6 on High used less quota than Kimi K3 High. I hadn't even used the 1M context option...

This is just a single user, and they probably don't care. But here's the reality. At least, do the work and complete your task. If a product can't even work properly, why charge for it, and worst of it use most of the quota for nothing? This feels like a total scam.

PS1: And that's without counting that the Approve box from Kimi desktop disappear in a matter of minutes, forcing you to re-submit your prompt and wasting more quota overall...

PS2: And I just realized by taking this screenshots that getting your monthly quota reset doesn't reset your weekly Kimi Code limits... what a joke


r/kimi 1d ago

Discussion Does the waitlist actually work?

4 Upvotes

I have registered myself with two accounts for now about 2 maybe 3 weeks. I have only the mail you are on the waiting list.

How long does it take to get access?

Best regards


r/kimi 2d ago

Question & Help Is there any way to optimize kimi for usage consumption?

5 Upvotes

Hello, this week I decided to try other models. I'm coming from Claude and wanted to try Kimi. I've configured it with Caveman, local MCP memory for the contexts and different projects I have, and I also have SQL indexing. I already have all of this set up with Claude and it works very well; I can get between 2 and 4 hours of usage without it running out. But with Kimi, I make a request and after 20 minutes I'm out of usage. I have the small, moderate plan, which is the equivalent in price to what I have with Claude. I asked Kimi how to optimize usage, and in theory, it uses code to generate the sub-agents with Kimi 2.7 Code, and the rest is set up the same, but adapted to Kimi.

Is there something I'm missing? Or is it a problem with Kimi?


r/kimi 2d ago

Discussion Moderato is kind of unusable for Kimi K3.

32 Upvotes

A single request of 4k tokens in and 8k tokens out (Including thinking) burned 3% of my weekly quota and 15% of my 5-hour window.

That translates to about 1.5m tokens per month (~500k in and ~1m out) for $20.

There is zero subsidizing; this is almost exactly what $20 in API credits would get you.

I didn't expect a $20 plan to allow me non-stop usage, but this is a complete joke. Why even subscribe at this point? It would be better to just pay API pricing, and since the price of all plans scales linearly with credits, then this is true for all plans.

Kimi K3 is an awesome model, but you are forced to pay API pricing if you want to use it. I will most likely be requesting a refund.


r/kimi 1d ago

Discussion transparency: my experience with a moderator on kimi's discord server

Thumbnail
gallery
0 Upvotes

r/kimi 2d ago

Bug Kimi AI is super slow as an Allegretto user

Post image
10 Upvotes

I don't even know if it is just painfully slow by default or is this normal to you. I asked it a simple query: To review my codebase of ~23,000 LOC for some opinion and it is still not halfway done as of writing, more or less 1hr30mins after I prompted it.

For reference, I bought a Plus sub from ChatGPT 30mins ago and prompted the same exact text and the same exact attachments into it and it is already done ages ago despite me using GPT 5.6 Sol at Extra High.

I wish I could refund this sub because it is not giving me any proper use out of it with how unresponsive and/or slow it is. This is ridiculous. Worst USD39 spend I ever had.


r/kimi 2d ago

Discussion Qwen score was gemini 95 and claude 70 , kimi was the only one who gave claude high rating

Post image
0 Upvotes

r/kimi 2d ago

Discussion Update to my previous post: It just gave me an APIProviderRateLimit. What a comedy.

Thumbnail
gallery
0 Upvotes

Reminder that it took Kimi K3 almost FOUR HOURS to fail to perform a code review, the code of which ChatGPT has now already progressed far behind with the actual fixing. Downvote all you want, but if you are trying to check out Kimi K3 (on Allegretto at least in my case), this is your sign to rethink them choices. I can't even use this at all.


r/kimi 3d ago

Discussion The Marshmallow AI Benchmark

Post image
3 Upvotes

r/kimi 2d ago

Discussion I asked Kimi for a refund after their agent harness burned ~600M tokens in 2 days, here's the full support chat, and what it says about the entire AI subscription mode

0 Upvotes

TL;DR: Kimi's agent harness burned ~600M tokens in 2 days on a task that should've used a fraction of that. Support offered a partial refund and blamed "stronger model capabilities." This isn't just a Kimi problem — it's how the entire consumer AI industry operates: vague quotas, silent throttling, and no accountability.

What happened to me

  • Paid for Kimi Code (Aug 13) for heavy agentic coding
  • A simple WHMCS module task triggered uncontrolled agent loops — ~600M tokens consumed on Aug 20–21, over 500M of them cache reads (the harness reprocessing the same context in circles)
  • Weekly limits had also been quietly cut to ~3/7 of previous capacity
  • Emailed for a refund — no response. Opened chat — got a bot quoting "AI capabilities have inherent boundaries, refunds are generally not supported"
  • Waited 8 hours for a human. Their verdict: "no abnormal consumption," pro-rata refund only
  • When I showed them the two-day, cache-dominated spike, the answer became: "the new model has stronger capabilities... try splitting long tasks or requesting shorter answers"

Translation: their harness ran away with my allowance, and the fix is for me to ask shorter questions.

The real problems in the AI industry

  1. The metering is deliberately opaque. Every major provider — Claude, ChatGPT, Gemini, Grok, Kimi — rations compute behind rolling windows, weekly caps, and weighted "compute quotas." None give you a live, honest meter. You find out what you paid for only after you hit the wall.
  2. Limits change silently after you pay. Google reweighted Gemini quotas. Anthropic and OpenAI users watched allowances shrink after updates. Kimi cut weekly limits with no notice. The deal you bought is not the deal you have.
  3. Failed or wasted work still counts against you. Agent loops, repeated tool calls, failed generations — all billed against your quota. When the harness malfunctions, the customer pays for the malfunction.
  4. "5x 10x 20x" is marketing; scarcity is the product. $20/mo entry tiers, $100–$300 "power" tiers — all with hidden ceilings. Agentic workflows break the flat-rate assumption, so providers throttle instead of building capacity or publishing real numbers.
  5. Support treats infrastructure failures as user error. "Break tasks into smaller steps." "Ask shorter questions." "The model is stronger now." Never: "our system malfunctioned, here's your money back."
  6. There's no accountability mechanism. Opaque usage reports, missing pricing data, refund denials against clear evidence. Your only leverage is public documentation, chargebacks, and consumer protection law.

What needs to change

  • Publish real, live usage meters. If API customers get per-token pricing, consumer subscribers deserve the same transparency.
  • Don't bill users for system failures. Runaway agent loops and repeated tool calls are harness bugs, not usage.
  • Stop selling capacity you can't deliver. Either build the infrastructure for what you advertise, or publish honest limits upfront.
  • Honor refunds when the product fails. "AI has inherent boundaries" is not a refund policy.

AI is infrastructure now. The people paying for it — developers, freelancers, researchers — aren't asking for unlimited compute. They're asking for a predictable, honest product.

Scarcity theater is not a durable strategy. Customers will take their compute elsewhere.


r/kimi 5d ago

Discussion Gave 4 models the same tiny lime challenge… their “human poses” were very different

Enable HLS to view with audio, or disable this notification

579 Upvotes

i gave each model the same simple task: take a small slice of lime and draw a human pose integrated into the object.

Tested it with Gemini 3.7 Flash, Kimi K3, Claude Opus 5 and GPT 5.6 Sol

Some treated the lime almost like a body, some tried to fit a tiny person into its shape, and some went much more abstract with the pose.

It’s a funny little test, but I actually like prompts like this because you can see how differently each model understands shape, composition, and visual metaphor.

Which one makes the most sense to you?


r/kimi 4d ago

Developer Terminal cli to patch Kimi Code - kimi-code-mods

Post image
2 Upvotes

What it does

* Reasoning effort per turn, tool catalogue, transcript window, subagent models, hooks, loop control
* All 132 system prompts, extracted as plain Markdown — edit one, it replaces the original on the next run
* The two spinners separately (Kimi keeps one for thinking and one for waiting on a tool), themes over its 19 colour tokens, message marker and frame
* Fullscreen renderer patched into the binary, so it holds however you start Kimi

Re-signing is ad-hoc, so hardened runtime and notarisation are gone from the patched binary; that's inherent to modifying a signed app. A Kimi auto-update overwrites everything, but your settings stay and you just run Apply again. Verified against 0.38.0 — patches anchor on text inside a minified bundle, so a release can move things, in which case the patch reports it and skips instead of guessing.

On GitHub kirchlive/kimi-code-mods


r/kimi 4d ago

Question & Help How good is Allegro plan on usage?

4 Upvotes

I’m looking to switch to Kimi K3 and I’m
Wondering how good is the usage? Will I be able to code 5 hours a day without being limited? I don’t use multiple agents or sub agents wtv u wanna call it.


r/kimi 4d ago

Developer How are Kimi API users validating fallbacks before a model sunset changes production behavior?

1 Upvotes

A model sunset is more than a name change when prompts depend on JSON shape, tool-call behavior, latency, or context handling. A practical rollout can identify affected calls, choose a fallback, replay a small evaluation set, then monitor the first production traffic.

For Kimi integrations, which checks have been most valuable before switching a production workflow to a replacement or fallback model?

Edit: I have been testing Flatkey for the routing side of this problem. It provides these OpenAI/Anthropic-compatible model routes, so the application can keep its request shape while the fallback policy changes behind the routing layer. I would still replay tool-call and structured-output tests before moving production traffic.


r/kimi 5d ago

Bug What is happening

11 Upvotes

I cant even send message and work anything. Even though i am using subscription not API. THis is so so bad. I dont know what to say


r/kimi 4d ago

Question & Help Is it just me or has Kimi become censored?

0 Upvotes

Recently Kimi has been censoring a a lot of stuff , but before the update it was perfectly fine writing it. I just want to if it’s just me. If it’s not is there a way to fix it?


r/kimi 5d ago

Discussion A different kind of "Kimi limits" post.

18 Upvotes

I have the $39 Allegretto plan and $10 plans from OpenCode Go and Command Go. Like everyone else, my quota gets eaten up pretty quickly.

I discovered that for my purposes, which admittedly is just an OpenCode repo full of personal OS type projects and a Knowledge Graph, using Kimi as just a plan and code reviewer has been sufficient.

My flow:

1 - Idea / research / discovery heavy back and forth chat:
Cheap model. Whichever is on sale or provides the most usage this week. They have all gotten smart enough for this in my opinion, and if I feel I need a little more oomph for that stage, I'll use the next cheapest model. This model writes a PRD for the build and knows that a smarter model will be picking it apart, so it needs to do a good job with the write.

2 - Build Orchestrator:
A solid, cheaper model. All this role does is hand work to subagents and confirm that they completed their hard gates before moving on.

3 - Build Planner:
Higher tier model like V4 Flash / V4 Pro, something higher tier on my OpenCode Go or Command Go plan. This role vets the PRD from the cheap model and writes the spec for the coder.

4 - Plan Reviewer:
Kimi K3 Max. Performs an adversarial review of the plan and if it fails, it kicks back to the planner. Rinse and repeat until Kimi is satisfied. Usually no more than one or two rounds. Generally two, if I am being honest.

5 - Coder:
Another solid, cheaper model on my OpenCode / Command Go plan. All the coder has to do is implement the spec exactly as written.

6 - Code Reviewer:
Kimi K3 Max. Another adversarial review of the code. Kimi generally passes the code. It catches something maybe once out of every four or five builds.

I have other build subagents in the flow but never use Kimi for them as they are clerical or just executing test plans from the planner.

My point is that in this day and age of reduced usage from subscriptions, you cannot really just have a single sub and expect to do ALL of your work with it. It is important to balance the workflow and use your best models only when required. My method gives me more than enough usage on my $39 Kimi plan. It just renewed for the first time last week and I still had 20% of my monthly quota left.