Comment by CROON_tv
13 hours ago
What I'd want to see next to accuracy is tail latency. In a real-time use, deciding when a spoken sentence is finished, a general LLM with the same prompt was slower and more hesitant for us than Jev, even though both cost about the same.
No comments yet
Contribute on Hacker News ↗