Comment by embedding-shape
4 days ago
> Quasar is not only intelligent, it is also fast. It returns 500 tokens, thinking time included, in 15.3 seconds.
Seems to be worded a bit strange, is "thinking time" referring to prompt processing or something? Otherwise "reasoning/thinking" is typically part of the returned tokens, at least for most non-OpenAI/non-Anthropic platforms, so you can see the actual reasoning. But here it seems either they word this weirdly, or "thinking" is somehow separate from the actual chat completion request?
No comments yet
Contribute on Hacker News ↗