Google DeepMind Releases Gemma 4 12B: An Encoder-Free Multimodal Model with Native audio that runs on a 16 GB laptop
Google DeepMind released Gemma 4 12B, an encoder-free multimodal model that processes text, images, audio, and video natively without separate vision or audio encoders. The 12-billion-parameter model …