Comment by retrobox
6 hours ago
Anecdotally, +1. I’d also say this benchmark matches my experiences and how much I trust the model output
6 hours ago
Anecdotally, +1. I’d also say this benchmark matches my experiences and how much I trust the model output
No comments yet
Contribute on Hacker News ↗