AI & ML interests
Google ❤️ Open Source AI
Recent Activity
Papers
Enabling Creative Exploration for Vibe Design Agents
Dream-RSI: Recursive Self-Improvement through Evolving Worlds
-
google/gemma-4-E2B-it-qat-mobile-transformers
Any-to-Any • 2B • Updated • 4.81k • 151 -
google/gemma-4-E4B-it-qat-mobile-transformers
Any-to-Any • 3B • Updated • 1.38k • 28 -
google/gemma-4-E2B-it-qat-mobile-ct
Any-to-Any • 6B • Updated • 4.34k • 31 -
google/gemma-4-E4B-it-qat-mobile-ct
Any-to-Any • 9B • Updated • 7.93k • 27
-
google/tipsv2-b14
Zero-Shot Image Classification • 0.2B • Updated • 256k • 125 -
google/tipsv2-l14
Zero-Shot Image Classification • 0.5B • Updated • 4.89k • 23 -
google/tipsv2-so400m14
Zero-Shot Image Classification • 0.9B • Updated • 276k • 19 -
google/tipsv2-g14
Zero-Shot Image Classification • 2B • Updated • 1.71k • 30
-
Compare Siglip1 Siglip2
🚀53Compare SigLIP1 and SigLIP2 on zero shot classification
-
SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Paper • 2502.14786 • Published • 169 -
google/siglip2-base-patch16-224
Zero-Shot Image Classification • 0.4B • Updated • 1.59M • 139 -
google/siglip2-base-patch16-256
Zero-Shot Image Classification • 0.4B • Updated • 3.89M • 16
-
PaliGemma 2: A Family of Versatile VLMs for Transfer
Paper • 2412.03555 • Published • 136 -
google/paligemma2-3b-pt-224
Image-Text-to-Text • 3B • Updated • 12.6k • 183 -
google/paligemma2-3b-pt-448
Image-Text-to-Text • 3B • Updated • 4.95k • 53 -
google/paligemma2-3b-pt-896
Image-Text-to-Text • 3B • Updated • 707 • 28
-
google/timesfm-1.0-200m
Time Series Forecasting • Updated • 1.01k • 840 -
google/timesfm-1.0-200m-pytorch
Time Series Forecasting • Updated • 2.08k • 35 -
google/timesfm-2.0-500m-jax
Time Series Forecasting • Updated • 22 • 21 -
google/timesfm-2.0-500m-pytorch
Time Series Forecasting • 0.5B • Updated • 86.3k • 264
-
google/gemma-4-E2B-it-qat-q4_0-unquantized
Any-to-Any • 5B • Updated • 26.5k • 34 -
google/gemma-4-E4B-it-qat-q4_0-unquantized
Any-to-Any • 8B • Updated • 4.46k • 29 -
google/gemma-4-12B-it-qat-q4_0-unquantized
Any-to-Any • 12B • Updated • 10.3k • 79 -
google/gemma-4-26B-A4B-it-qat-q4_0-unquantized
Image-Text-to-Text • 27B • Updated • 530k • • 52
-
MedGemma - Radiology Explainer Demo
🩺248Radiology Image & Report Explainer Demo. Built with MedGemma
-
Appoint Ready - MedGemma Demo
📋208Simulated Pre-visit Intake Demo built using MedGemma
-
Radiology Learning Companion
🏃28A demo showcasing a medical learning experience of CXR image
-
EHR Navigator Agent With MedGemma
🩺66Explore patient records with an interactive EHR assistant
-
VideoPrism: A Foundational Visual Encoder for Video Understanding
Paper • 2402.13217 • Published • 41 -
google/videoprism-base-f16r288
Video Classification • Updated • 2.39k • 109 -
google/videoprism-large-f8r288
Video Classification • Updated • 206 • 21 -
google/videoprism-lvt-base-f16r288
Video Classification • Updated • 2.23k • 16
-
Path Foundation Demo
🔬47Explore pathology image library
-
CXR Foundation Demo
🩻22Demo usage of the CXR Foundation model embeddings
-
MedGemma - Radiology Explainer Demo
🩺248Radiology Image & Report Explainer Demo. Built with MedGemma
-
Appoint Ready - MedGemma Demo
📋208Simulated Pre-visit Intake Demo built using MedGemma
-
google/gemma-3-4b-it-qat-q4_0-gguf
Image-Text-to-Text • 4B • Updated • 2.49k • 287 -
google/gemma-3-4b-pt-qat-q4_0-gguf
Image-Text-to-Text • 4B • Updated • 25 • 28 -
google/gemma-3-1b-it-qat-q4_0-gguf
Text Generation • 1.0B • Updated • 874 • 153 -
google/gemma-3-1b-pt-qat-q4_0-gguf
Text Generation • 1.0B • Updated • 105 • 16
-
Paligemma2 Mix
🌖98Generate text answers or segment objects from images
-
google/paligemma2-3b-mix-224
Image-Text-to-Text • 3B • Updated • 22.3k • 58 -
google/paligemma2-3b-mix-448
Image-Text-to-Text • 3B • Updated • 3.48k • 66 -
google/paligemma2-10b-mix-224
Image-Text-to-Text • 10B • Updated • 100 • 10
-
google-t5/t5-base
Translation • 0.2B • Updated • 2.79M • • 793 -
google-t5/t5-small
Translation • 60.5M • Updated • 25M • • 632 -
google-t5/t5-large
Translation • 0.7B • Updated • 374k • • 264 -
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Paper • 1910.10683 • Published • 20
-
google/siglip-so400m-patch14-384
Zero-Shot Image Classification • 0.9B • Updated • 1.32M • 691 -
google/siglip-so400m-patch14-224
Zero-Shot Image Classification • 0.9B • Updated • 15.7k • 59 -
google/siglip-so400m-patch16-256-i18n
Zero-Shot Image Classification • 1B • Updated • 679 • 31 -
google/siglip-base-patch16-256-multilingual
Zero-Shot Image Classification • 0.4B • Updated • 20.9k • 55
-
google/gemma-4-E2B-it-qat-mobile-transformers
Any-to-Any • 2B • Updated • 4.81k • 151 -
google/gemma-4-E4B-it-qat-mobile-transformers
Any-to-Any • 3B • Updated • 1.38k • 28 -
google/gemma-4-E2B-it-qat-mobile-ct
Any-to-Any • 6B • Updated • 4.34k • 31 -
google/gemma-4-E4B-it-qat-mobile-ct
Any-to-Any • 9B • Updated • 7.93k • 27
-
google/gemma-4-E2B-it-qat-q4_0-unquantized
Any-to-Any • 5B • Updated • 26.5k • 34 -
google/gemma-4-E4B-it-qat-q4_0-unquantized
Any-to-Any • 8B • Updated • 4.46k • 29 -
google/gemma-4-12B-it-qat-q4_0-unquantized
Any-to-Any • 12B • Updated • 10.3k • 79 -
google/gemma-4-26B-A4B-it-qat-q4_0-unquantized
Image-Text-to-Text • 27B • Updated • 530k • • 52
-
google/tipsv2-b14
Zero-Shot Image Classification • 0.2B • Updated • 256k • 125 -
google/tipsv2-l14
Zero-Shot Image Classification • 0.5B • Updated • 4.89k • 23 -
google/tipsv2-so400m14
Zero-Shot Image Classification • 0.9B • Updated • 276k • 19 -
google/tipsv2-g14
Zero-Shot Image Classification • 2B • Updated • 1.71k • 30
-
MedGemma - Radiology Explainer Demo
🩺248Radiology Image & Report Explainer Demo. Built with MedGemma
-
Appoint Ready - MedGemma Demo
📋208Simulated Pre-visit Intake Demo built using MedGemma
-
Radiology Learning Companion
🏃28A demo showcasing a medical learning experience of CXR image
-
EHR Navigator Agent With MedGemma
🩺66Explore patient records with an interactive EHR assistant
-
VideoPrism: A Foundational Visual Encoder for Video Understanding
Paper • 2402.13217 • Published • 41 -
google/videoprism-base-f16r288
Video Classification • Updated • 2.39k • 109 -
google/videoprism-large-f8r288
Video Classification • Updated • 206 • 21 -
google/videoprism-lvt-base-f16r288
Video Classification • Updated • 2.23k • 16
-
Path Foundation Demo
🔬47Explore pathology image library
-
CXR Foundation Demo
🩻22Demo usage of the CXR Foundation model embeddings
-
MedGemma - Radiology Explainer Demo
🩺248Radiology Image & Report Explainer Demo. Built with MedGemma
-
Appoint Ready - MedGemma Demo
📋208Simulated Pre-visit Intake Demo built using MedGemma
-
google/gemma-3-4b-it-qat-q4_0-gguf
Image-Text-to-Text • 4B • Updated • 2.49k • 287 -
google/gemma-3-4b-pt-qat-q4_0-gguf
Image-Text-to-Text • 4B • Updated • 25 • 28 -
google/gemma-3-1b-it-qat-q4_0-gguf
Text Generation • 1.0B • Updated • 874 • 153 -
google/gemma-3-1b-pt-qat-q4_0-gguf
Text Generation • 1.0B • Updated • 105 • 16
-
Compare Siglip1 Siglip2
🚀53Compare SigLIP1 and SigLIP2 on zero shot classification
-
SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Paper • 2502.14786 • Published • 169 -
google/siglip2-base-patch16-224
Zero-Shot Image Classification • 0.4B • Updated • 1.59M • 139 -
google/siglip2-base-patch16-256
Zero-Shot Image Classification • 0.4B • Updated • 3.89M • 16
-
Paligemma2 Mix
🌖98Generate text answers or segment objects from images
-
google/paligemma2-3b-mix-224
Image-Text-to-Text • 3B • Updated • 22.3k • 58 -
google/paligemma2-3b-mix-448
Image-Text-to-Text • 3B • Updated • 3.48k • 66 -
google/paligemma2-10b-mix-224
Image-Text-to-Text • 10B • Updated • 100 • 10
-
PaliGemma 2: A Family of Versatile VLMs for Transfer
Paper • 2412.03555 • Published • 136 -
google/paligemma2-3b-pt-224
Image-Text-to-Text • 3B • Updated • 12.6k • 183 -
google/paligemma2-3b-pt-448
Image-Text-to-Text • 3B • Updated • 4.95k • 53 -
google/paligemma2-3b-pt-896
Image-Text-to-Text • 3B • Updated • 707 • 28
-
google-t5/t5-base
Translation • 0.2B • Updated • 2.79M • • 793 -
google-t5/t5-small
Translation • 60.5M • Updated • 25M • • 632 -
google-t5/t5-large
Translation • 0.7B • Updated • 374k • • 264 -
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Paper • 1910.10683 • Published • 20
-
google/siglip-so400m-patch14-384
Zero-Shot Image Classification • 0.9B • Updated • 1.32M • 691 -
google/siglip-so400m-patch14-224
Zero-Shot Image Classification • 0.9B • Updated • 15.7k • 59 -
google/siglip-so400m-patch16-256-i18n
Zero-Shot Image Classification • 1B • Updated • 679 • 31 -
google/siglip-base-patch16-256-multilingual
Zero-Shot Image Classification • 0.4B • Updated • 20.9k • 55
-
google/timesfm-1.0-200m
Time Series Forecasting • Updated • 1.01k • 840 -
google/timesfm-1.0-200m-pytorch
Time Series Forecasting • Updated • 2.08k • 35 -
google/timesfm-2.0-500m-jax
Time Series Forecasting • Updated • 22 • 21 -
google/timesfm-2.0-500m-pytorch
Time Series Forecasting • 0.5B • Updated • 86.3k • 264