Comment by roseway4
3 hours ago
If you don’t see good performance with LRs, you may want to try RBF SVMs. We’ve found they work super well for our use cases with the embeddinggemma model as they can better separate classes in the non-linear embedding space.
Our resulting RBF models are tiny and fit in L1 cache, with microsecond inference latency.
No comments yet
Contribute on Hacker News ↗