Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Assuming all 64 subagents were running for a full hour (the tweet states just under an hour):

  Throughput                    Output tokens   Output cost
  ----------------------------  -------------   -----------
  40 tok/s  (5.5 low)                   ~9.2M         ~$275
  55 tok/s  (5.5 base)                 ~12.7M         ~$380
  70 tok/s  (5.5 high)                 ~16.1M         ~$485
  750 tok/s (Sol Fast, $75/M)         ~172.8M       ~$13,000
Claude estimates that tool use / input tokens might add 10-15% on top of that depending on exactly how the model went about the task.

Edit: better tok/s estimate buckets based on GPT 5.5 actual speeds since I couldn't find real benchmarks on 5.6 published anywhere. Also account for Sol Fast pricing.



Sol fast isn't the Cerebras 750 tok/s version, it's just 1.5x speed at 2.5x price

I assume they didn't use the Cerebras version for this since it's probably very supply-constrained right now


But Sol is running on Cerebras. That’s the whole point of this. That’s how they get 750 tokens per second. There is no other way.


Regular Sol does not run on Cerebra’s. I don’t think anyone public has access to that.

https://x.com/thsottiaux/status/2075596669958472146?s=46&t=Z...


Yeah they posted an update below it

> apparently that is just Sol being Sol on fast mode, not 750 tps. :x O.o

> This is real and it is not 750 TPS. Anyone with 5.6 Sol Ultra on fast mode can reproduce this! It’s all GUI interactions with CUA. No MCP or bpy needed.

https://x.com/kimmonismus/status/2075493505011482922?s=20


Who’s they, who’s chubby, why would I care.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: