lastmass/Qwen3_Medical_GRPO
https://huggingface.co/lastmass/Qwen3_Medical_GRPO中文版说明
Sourced from
- HuggingFace — lastmass/Qwen3_Medical_GRPO
Related resources
FreedomIntelligence/HuatuoGPT-3-Grader-8B
by FreedomIntelligenceHuatuoGPT-3-Grader-8B GitHub | Paper
WaltonFuture/Diabetica-7B
by WaltonFutureDiabetica: Adapting Large Language Model to Enhance Multiple Medical Tasks in Diabetes Care and Management
stanfordmimi/MedVAL-4B
by stanfordmimiMedVAL-4B (medical text validator) is a language model fine-tuned to assess AI-generated medical text outputs at near physician-level reliability.
This is https://huggingface.co/kingabzpro/Qwen-3-32B-Medical-Reasoning applied to https://huggingface.co/Qwen/Qwen3-32B Original model card created by @kingabzpro
tmberooney/medllama-merged
by tmberooneyModel Card for "medllama" ---------------------------
This is an official model checkpoint for Asclepius-Mistral-7B-v0.3 (arxiv). This model is an enhanced version of Asclepius-7B, by replacing the base model with Mistral-7B-v0.3 and increasing the max sequence length to 8192.