AI huggingface tools
2 picks we've reviewed.
open-source · Apache 2.0 · 128K tokens (E2B/E4B)2mo ago
model multimodal
Gemma 4
Google DeepMind's fourth-generation open-weight model family — five sizes from 2B to 31B, Apache 2.0 licensed, with the 12B Unified variant accepting text, image, audio, and video in a single encoder-free architecture.
Google DeepMind
paid · 1M tokens2mo ago
model multimodal
MiniMax M3
MiniMax's third-generation flagship — M3 is an open-weight 428B-parameter MoE (~23B active per token) using MiniMax Sparse Attention (MSA), with a 1M-token context window and native multimodality (text, image, and video input). It's positioned as the first open-weight model to combine frontier coding, a 1M context, and native multimodality — and can even operate a desktop computer.
MiniMax