HarshBhanushali7705/medgemma-27b-text-it-GPTQ-4bit

https://huggingface.co/HarshBhanushali7705/medgemma-27b-text-it-GPTQ-4bit
Activeby HarshBhanushali77053463updated 1 month ago

This repository contains a 4-bit GPTQ quantized version of google/medgemma-27b-text-it, optimized for high-throughput inference using vLLM and the Marlin kernel.

Sourced from

  • HuggingFaceHarshBhanushali7705/medgemma-27b-text-it-GPTQ-4bit

Related resources

Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants.

Idle1.2K1 year ago
Python

This model mlx-community/medgemma-27b-text-it-bf16 was converted to MLX format from google/medgemma-27b-text-it using mlx-lm version 0.25.1.

Idle3611 year ago
Python

> [!NOTE] > Inspired by the thought of: what if you could speak to an offline medical assistant that doesn't decline to answer some of your questions?

Idle268 months ago
Python

GRamma-12B is a 12-billion-parameter instruction-tuned language model specialized for the Greek medical domain. It is built on top of Gemma 3 12B Instruct and adapted through parameter-efficient fine-tuning on a collection of Greek and bilingual medical question-answering data.

Active192 months ago
Python

Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants.

Idle5.7K1 year ago
Python