these are the updated models: google/gemma-4-31B-it-assistant google/gemma-4-26B-A4B-i...

AbuAssar • yesterday at 6:09 PM • 1 reply • view on HN

these are the updated models:

google/gemma-4-31B-it-assistant

google/gemma-4-26B-A4B-it-assistant

google/gemma-4-E4B-it-assistant

google/gemma-4-E2B-it-assistant

sigmar • yesterday at 7:07 PM

for anyone wanting a glossary to explain the naming scheme here:

E4B = 4B effective parameters (using per-layer embeddings)

E2B = 2B (like above)

it = instruction tuned (rlhf and all that jazz)

assistant = Multi-token drafters (the new 2x speed up)

alt Hacker News