Comment by retrobox
7 hours ago
Anecdotally, +1. I’d also say this benchmark matches my experiences and how much I trust the model output
7 hours ago
Anecdotally, +1. I’d also say this benchmark matches my experiences and how much I trust the model output
No comments yet
Contribute on Hacker News ↗