Top
Best
New

Posted by altertable 12 hours ago

Qwen 3.8 27B available on Cerebras at 1500 tokens/s(inference-docs.cerebras.ai)
509 points | 158 commentspage 3
grav 11 hours ago|
Should be available in OpenCode once this lands: https://github.com/anomalyco/models.dev/pull/6199/changes
irthomasthomas 10 hours ago|
It's going to cost a fortune in opencode without prompt caching.
forlorn 3 hours ago||
Is Kimi 3 available anywhere like that?
darkbatman 11 hours ago||
I have been their user for more than year even used coding plans, though for normal coding the quota will definitely be a blocker if you are using opencode because rpm are bit less. Good for products/api though.
polygot 12 hours ago||
Ut oh, might be down: "Unable to connect to the server. Please check your connection and try again." when sending a message to Qwen 3.8 27B.
vb-8448 11 hours ago||
At that speed it's too pricey for agentinc tasks.
yipinwong 11 hours ago|
The target audience is who needs raw speed.

Having the choice is good as you can make a trade-off between speed, perf, and quality.

Until last year, people had a single AI god they believed in (mostly Anthropic stuff). Now we have power to make choices (open-weights, SOTA, speed-optimized, etc) the same way you do for system designs.

vb-8448 11 hours ago||
It's not a criticism, I was really looking forward to trying out such a powerful model at this speed.

But I burn my 5$ allowance in 10 minutes ... and only because I was hitting rate limits, without it would probably be less than a minute.

yipinwong 10 hours ago||
I hear ya... the best option is to use company budget as normies will rack up ridciulous amount soon with that raw speed.
karim79 6 hours ago||
Tokens are the new latest and greatest nonsensical shit on the planet. It's amusing. I can't wait to see the world in 1-2 years and the hilarity of looking back on this day.
fulafel 11 hours ago||
What are the best benchmarks/leaderboards that compare task completion time between provider+model combos?
srcreigh 11 hours ago||
How many years until chips like this are available to consumers?
nicce 11 hours ago|
Many. Too lucrative for certain companies and even governments to allow that to happen
drchaim 11 hours ago||
The idea of custom software on the fly is coming
Marciplan 12 hours ago|
used their Code product with GLM4.7. its fun but if the model is bad it just doesn’t do much useful.

Hope they add such models to Code too :)

altertable 12 hours ago|
Yeah GLM 4.7 is from another decade at the speed we're going
More comments...