SandLogicTechnologies/MedGemma-4B-IT-GGUF

https://huggingface.co/SandLogicTechnologies/MedGemma-4B-IT-GGUF
Idleby SandLogicTechnologies1065updated 1 year ago

This repository provides quantized GGUF versions of the google/medgemma-4b-it model. These 4-bit and 5-bit quantized variants retain the original model’s strengths in multimodal medical reasoning, while reducing memory and compute requirements—ideal for efficient inference on resource-constrained…

Sourced from

  • HuggingFace — SandLogicTechnologies/MedGemma-4B-IT-GGUF

Related resources

Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants.

Idle17.7K8 months ago
Python

Madrimed1.2 2B GGUF Optimized GGUF Quantization Tiers for MadriMed 1.2 (2B Medical Mulitmodel)

Active7786 days ago

# fernandoruiz/medgemma-4b-it-Q4_0-GGUF This model was converted to GGUF format from google/medgemma-4b-it using llama.cpp via the ggml.ai's GGUF-my-repo space. Refer to the original model card for more details on the model.

Idle371 year ago
Python

Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants.

Idle17.7K1 year ago
Python

Using llama.cpp release b2440 for quantization.

Stale6152 years ago
Python

A 2B-parameter multimodal medical vision-language model, trained for medical image understanding, radiology report assistance, clinical visual question answering, and medical text reasoning.

Active1751 week ago