Vidu S1: A Real-Time Interactive Video Generation Model Paper • 2607.03118 • Published Jul 3 • 147
Anthropogenic Regional Adaptation in Multimodal Vision-Language Model Paper • 2604.11490 • Published Apr 13 • 16
view article Article KV Caching Explained: Optimizing Transformer Inference Efficiency not-lain • Jan 30, 2025 • 402
view article Article 🪆 Introduction to Matryoshka Embedding Models +1 tomaarsen, Xenova, osanseviero • Feb 23, 2024 • 219
Gemma 4 Collection Our most intelligent open models to date • 16 items • Updated 14 days ago • 1.11k
view article Article Welcome Gemma 4: Frontier multimodal intelligence on device +5 merve, pcuenq, sergiopaniego, burtenshaw, Steveeeeeeen, alvarobartt, SaylorTwift • Apr 2 • 927
CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data Paper • 2601.18026 • Published Jan 25