gemma3 is absolutely trash.

by Maani - opened 2 days ago

2 days ago

while this model isn't that, it still uses the same implementation of attention. it's trash for quantization, it's trash for inference, it's trash for any sorta computation beyond take it and inference it on TPUs.
I might be alone in asking for this, but I'd very much appreciate it if you left a dumpster fire like google behind and used a normal implementation that can work with the rest of the libraries instead of fighting with the developer because google's fucking shares might drop. what has google done for us lately?! except sucking our bones dry like the parasite it is?

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment