lastmass/Qwen3_Medical_GRPO

https://huggingface.co/lastmass/Qwen3_Medical_GRPO
Idleby lastmass14722updated 1 year ago
Python

中文版说明

Sourced from

  • HuggingFace — lastmass/Qwen3_Medical_GRPO

Related resources

HuatuoGPT-3-Grader-8B GitHub | Paper

Active6552 weeks ago
Python

Diabetica: Adapting Large Language Model to Enhance Multiple Medical Tasks in Diabetes Care and Management

Idle711 year ago
Python

MedVAL-4B (medical text validator) is a language model fine-tuned to assess AI-generated medical text outputs at near physician-level reliability.

Idle2031 year ago
Python

This is https://huggingface.co/kingabzpro/Qwen-3-32B-Medical-Reasoning applied to https://huggingface.co/Qwen/Qwen3-32B Original model card created by @kingabzpro

Idle3421 year ago
Python

Model Card for "medllama" ---------------------------

Stale152 years ago
Python

This is an official model checkpoint for Asclepius-Mistral-7B-v0.3 (arxiv). This model is an enhanced version of Asclepius-7B, by replacing the base model with Mistral-7B-v0.3 and increasing the max sequence length to 8192.

Stale2462 years ago
Python