Indian Flag
Government Of India
A-
A
A+
ORGANISATION

MedGemma-4b-it

google/medgemma-4b-it

About Model

MedGemma is a collection of Gemma 3 variants that are trained for performance on medical text and image comprehension. Developers can use MedGemma to accelerate building healthcare-based AI applications. MedGemma currently comes in three variants: a 4B multimodal version and 27B text-only and multimodal versions.

Both MedGemma multimodal versions utilize a SigLIP image encoder that has been specifically pre-trained on a variety of de-identified medical data, including chest X-rays, dermatology images, ophthalmology images, and histopathology slides. Their LLM components are trained on a diverse set of medical data, including medical text, medical question-answer pairs, FHIR-based electronic health record data (27B multimodal only), radiology images, histopathology patches, ophthalmology images, and dermatology images.

MedGemma 4B is available in both pre-trained (suffix: -pt) and instruction-tuned (suffix -it) versions. The instruction-tuned version is a better starting point for most applications. The pre-trained version is available for those who want to experiment more deeply with the models.

MedGemma 27B multimodal has pre-training on medical image, medical record and medical record comprehension tasks. MedGemma 27B text-only has been trained exclusively on medical text. Both models have been optimized for inference-time computation on medical reasoning. This means it has slightly higher performance on some text benchmarks than MedGemma 27B multimodal. Users who want to work with a single model for both medical text, medical record and medical image tasks are better suited for MedGemma 27B multimodal. Those that only need text use-cases may be better served with the text-only variant. Both MedGemma 27B variants are only available in instruction-tuned versions.

MedGemma variants have been evaluated on a range of clinically relevant benchmarks to illustrate their baseline performance. These evaluations are based on both open benchmark datasets and curated datasets. Developers can fine-tune MedGemma variants for improved performance. Consult the Intended Use section below for more details.

MedGemma is optimized for medical applications that involve a text generation component. For medical image-based applications that do not involve text generation, such as data-efficient classification, zero-shot classification, or content-based or semantic image retrieval, the MedSigLIP image encoder is recommended. MedSigLIP is based on the same image encoder that powers MedGemma.

Please consult the MedGemma Technical Report for more details.

MedGemma-4b-it

Metadata Metadata

Other

Google

Large Language Models

Transformers

Open

Google LLC

Healthcare, Wellness and Family Welfare

18/09/25 10:28:12

Amrita Kamat

0

Activity Overview Activity Overview

  • Downloads0
  • Redirect 26
  • File Size 0
  • Views 1,567

Tags Tags

  • Transformers
  • safetensors
  • endpoints_compatible
  • image-text-to-text
  • text-generation-inference
  • conversational
  • radiology
  • medical
  • gemma3
  • region:us
  • license:other
  • arxiv:2106.14463
  • clinical-reasoning
  • base_model:google/medgemma-4b-pt
  • base_model:finetune:google/medgemma-4b-pt
  • arxiv:2404.05590
  • dermatology
  • arxiv:2405.03162
  • arxiv:2501.19393
  • chest-x-ray
  • arxiv:2009.13081
  • arxiv:2412.03555
  • arxiv:2303.15343
  • ophthalmology
  • arxiv:2507.05201
  • arxiv:2501.18362
  • arxiv:2102.09542
  • arxiv:2411.15640
  • pathology

License Control License Control

Other

More Models from Google LLC More Models from Google LLC

TimesFM 3 PyTorch
TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting.
forecasting
PyTorch
time-series
google
pretrained
  • See Upvoters0
  • Downloads0
  • File Size0
  • Views3
Updated 1 day(s) ago

GOOGLE LLC

GNM Generative aNthropometric Model - v3
GNM (pronounced genome) is a state-of-the-art parametric 3D statistical model of the human head developed by Google. Learned from a large dataset of high-resolution 3D scans, GNM provides fine-grained, disentangled control over facial identity, expressions, and head pose, complete with controllable internal anatomy (eyeballs, teeth, and tongue). The model is released under the Apache 2.0 permissive license, suitable for both academic research and commercial applications.
mesh
3d
3dmm
computer vision
head-tracking
digital-humans
parametric-model
human-head
  • See Upvoters0
  • Downloads0
  • File Size0
  • Views3
Updated 1 day(s) ago

GOOGLE LLC

Gemma-4-12B-it-assistant
This model card is for the Multi-Token Prediction (MTP) drafters for the Gemma 4 models. MTP is implemented by extending the base model with a smaller, faster draft model. When used in a Speculative Decoding pipeline, the draft model predicts several tokens ahead, which the target model then verifies in parallel. This results in significant decoding speedups (up to 3x) while guaranteeing the exact same quality as standard generation, making these checkpoints perfect for low-latency and ODML.
Transformers
license:apache-2.0
arxiv:2607.02770
gemma4_unified_assistant
region:us
any-to-any
Text Generation
endpoints_compatible
safetensors
  • See Upvoters0
  • Downloads0
  • File Size0
  • Views3
Updated 1 day(s) ago

GOOGLE LLC

TIPS Text-Image Pre-training with Spatial awareness ICLR 2025
TIPS (Text-Image Pre-training with Spatial awareness, ICLR 2025) is a family of contrastive vision-language models that produce spatially rich image features aligned with text embeddings. This is the original (v1) g/14 release with 1.1B vision params and 389M text params, converted from the official checkpoints.
Feature Extraction
arxiv:2410.16512
zero-shot
contrastive-learning
image-text
vision
  • See Upvoters0
  • Downloads0
  • File Size0
  • Views3
Updated 1 day(s) ago

GOOGLE LLC

DiffusionGemma
DiffusionGemma is a generative model built by Google DeepMind. Based on the 26B A4B Mixture-of-Experts (MoE) Gemma 4 architecture, DiffusionGemma generates tokens using discrete diffusion. This open-weights model is multimodal, handling text, image, and video inputs to generate text output.
region:us
conversational
image-text-to-text
endpoints_compatible
safetensors
Transformer
license:apache-2.0
deploy:sagemaker
diffusion_gemma
Public Space AI
google
  • See Upvoters1
  • Downloads3
  • File Size0
  • Views34
Updated 11 day(s) ago

GOOGLE LLC

google/embeddinggemma-300m-qat-q8_0-unquantized
google/embeddinggemma-300m-qat-q8_0-unquantized
gemma3_text
autotrain_compatible
Sentence Similarity
endpoints_compatible
safetensors
Feature Extraction
region:us
sentence-transformers
license:gemma
  • See Upvoters0
  • Downloads2
  • File Size0
  • Views315
Updated 1 year(s) ago

GOOGLE LLC

txgemma-27b-chat
google/txgemma-27b-chat
conversational
Transformers
safetensors
endpoints_compatible
Text Generation
text-generation-inference
en
autotrain_compatible
gemma2
region:us
license:other
therapeutics
drug-development
arxiv:2406.06316
arxiv:2504.06196
  • See Upvoters0
  • Downloads8
  • File Size0
  • Views242
Updated 1 year(s) ago

GOOGLE LLC

txgemma-9b-chat
google/txgemma-9b-chat
arxiv:2504.06196
arxiv:2406.06316
Transformers
safetensors
endpoints_compatible
Text Generation
text-generation-inference
conversational
en
autotrain_compatible
gemma2
region:us
license:other
therapeutics
drug-development
  • See Upvoters0
  • Downloads6
  • File Size0
  • Views187
Updated 1 year(s) ago

GOOGLE LLC

txgemma-9b-predict
google/txgemma-9b-predict
therapeutics
arxiv:2406.06316
arxiv:2504.06196
drug-development
license:other
region:us
gemma2
autotrain_compatible
en
text-generation-inference
Text Generation
endpoints_compatible
safetensors
Transformers
  • See Upvoters0
  • Downloads3
  • File Size0
  • Views219
Updated 1 year(s) ago

GOOGLE LLC

medgemma-27b-text-it
google/medgemma-27b-text-it
arxiv:2303.15343
arxiv:2102.09542
arxiv:2009.13081
arxiv:2404.05590
base_model:google/gemma-3-27b-pt
arxiv:2501.18362
arxiv:2411.15640
arxiv:2501.19393
endpoints_compatible
Text Generation
text-generation-inference
conversational
safetensors
autotrain_compatible
medical
region:us
license:other
Transformers
gemma3_text
clinical-reasoning
thinking
base_model:finetune:google/gemma-3-27b-pt
  • See Upvoters0
  • Downloads5
  • File Size0
  • Views864
Updated 1 year(s) ago

GOOGLE LLC