SandLogicTechnologies/MedGemma-4B-IT-GGUF

https://huggingface.co/SandLogicTechnologies/MedGemma-4B-IT-GGUF
Idleby SandLogicTechnologies755updated 1 year ago

This repository provides quantized GGUF versions of the google/medgemma-4b-it model. These 4-bit and 5-bit quantized variants retain the original model’s strengths in multimodal medical reasoning, while reducing memory and compute requirements—ideal for efficient inference on resource-constrained…

Sourced from

  • HuggingFaceSandLogicTechnologies/MedGemma-4B-IT-GGUF

Related resources

Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants.

Idle8.8K7 months ago
Python

Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants.

Idle17.7K1 year ago
Python

# fernandoruiz/medgemma-4b-it-Q4_0-GGUF This model was converted to GGUF format from google/medgemma-4b-it using llama.cpp via the ggml.ai's GGUF-my-repo space. Refer to the original model card for more details on the model.

Idle371 year ago
Python
Active240.8K4 months ago
Python

Using llama.cpp release b2440 for quantization.

Stale6152 years ago
Python
Active6.2K3 months ago