Comment by sigmar
20 hours ago
for anyone wanting a glossary to explain the naming scheme here:
E4B = 4B effective parameters (using per-layer embeddings)
E2B = 2B (like above)
it = instruction tuned (rlhf and all that jazz)
assistant = Multi-token drafters (the new 2x speed up)
No comments yet
Contribute on Hacker News ↗