Posted by serialx 1 day ago
Will have to watch cost carefully though. If only I could get it at DeepSeek adjacent prices.
I couldn't find any info on how they (siliconflow) quantized the model. I think some of the other ineference service aggregators might offer it as well, already.
It’s starting to sound like Claude.
The token limits do feel a bit less than I get with Anthropic Max 5x, but maybe that's because I've mostly been running it on Max reasoning (oh and Anthropic is also temporarily boosting the limits, who knows, it's hard to keep track of all of this stuff exactly) and there's plenty of tasks where High is still close enough in performance. The token limits still feel a bit more generous proportionally to the price compared to what I got when trying out the 65 USD tier of GLM Coding Subscription with GLM 5.2, and that was with the ZCode usage discount as well, though I did enjoy that harness.
Plus, if I decide to go with Kimi's annual pricing, then it'd come out to only around 159 USD per month or 139 EUR per month, which is really good and pretty close to what I pay Anthropic anyways: https://www.kimi.com/help/membership/membership-pricing
I hope they wire up new hardware quickly to handle the demand.
Total usage
12.85%
Resets in 2026-08-17
5-hour usage
71.97%
Resets in 07-20 03:14
7-day usage
31.29%
Resets in 07-24 10:14
Still, I can respect the commitment to not over provisioning.