The CEO was on Gradient Dissent a couple years ago: https://www.youtube.com/watch?v=qNXebAQ6igs
[1] https://www.cerebras.ai/blog/accelerating-gpt-5-6-sol-ultraf...
Hope they add such models to Code too :)
update: I got my answer. support@ replied and said my email domain is on their blacklist.
Psychopaths: tok/SEC
For those who haven't noticed though, the context size they allow for Qwen is just 128k. Still interesting as a specialized sub-agent but not really well suited for long tasks.
They do appear to host other models on OpenRouter so maybe Qwen3.8 will be there soon: https://openrouter.ai/provider/cerebras
https://mixlayer.com, LAUNCH-Q38-27B gets you $5 in credits if you want to kick the tires.
It would look bad for cerebras if other people are hosting the 27b version and show a higher TPS than cerebras.
gardnr•35m ago
Edit: it looks like this is only available on a API token pricing. Does anyone know if they have rolled out prompt caching yet? It used to get pretty expensive for agentic coding tasks with no prompt caching.
altertable•33m ago
jasongill•27m ago
cute_boi•26m ago
eli•9m ago
singpolyma3•9m ago