Comment by riknos314
2 months ago
Assuming all 64 subagents were running for a full hour (the tweet states just under an hour):
Throughput Output tokens Output cost
---------------------------- ------------- -----------
40 tok/s (5.5 low) ~9.2M ~$275
55 tok/s (5.5 base) ~12.7M ~$380
70 tok/s (5.5 high) ~16.1M ~$485
750 tok/s (Sol Fast, $75/M) ~172.8M ~$13,000
Claude estimates that tool use / input tokens might add 10-15% on top of that depending on exactly how the model went about the task.
Edit: better tok/s estimate buckets based on GPT 5.5 actual speeds since I couldn't find real benchmarks on 5.6 published anywhere. Also account for Sol Fast pricing.
Sol fast isn't the Cerebras 750 tok/s version, it's just 1.5x speed at 2.5x price
I assume they didn't use the Cerebras version for this since it's probably very supply-constrained right now
But Sol is running on Cerebras. That’s the whole point of this. That’s how they get 750 tokens per second. There is no other way.
Regular Sol does not run on Cerebra’s. I don’t think anyone public has access to that.
https://x.com/thsottiaux/status/2075596669958472146?s=46&t=Z...
2 replies →