reaperdoesntknow/Qwen3-1.7B-Distilled-30B-A3B

https://huggingface.co/reaperdoesntknow/Qwen3-1.7B-Distilled-30B-A3B
Activeby reaperdoesntknow1.4K2updated 1 month ago
Python

A 1.7B-parameter causal language model distilled from Qwen3-30B-A3B on 6,122 STEM chain-of-thought samples using discrepancy-informed knowledge distillation. The training objective emphasizes proof structure, detects reasoning pivot tokens through token-level divergence dynamics, smooths…

Sourced from

  • HuggingFacereaperdoesntknow/Qwen3-1.7B-Distilled-30B-A3B

Related resources

A compact protein language model distilled from ProtGPT2 using complementary-regularizer distillation---a method that combines uncertainty-aware position weighting with calibration-aware label smoothing to achieve 54% better perplexity than standard knowledge distillation at 9.4x compression.

Idle56 months ago
Python

A compact protein language model distilled from ProtGPT2 using complementary-regularizer distillation---a method that combines uncertainty-aware position weighting with calibration-aware label smoothing to achieve 31% better perplexity than standard knowledge distillation at 3.8x compression.

Idle706 months ago
Python

This is https://huggingface.co/kingabzpro/Qwen-3-32B-Medical-Reasoning applied to https://huggingface.co/Qwen/Qwen3-32B Original model card created by @kingabzpro

Idle3421 year ago
Python

MedVAL-4B (medical text validator) is a language model fine-tuned to assess AI-generated medical text outputs at near physician-level reliability.

Idle20310 months ago
Python

中文版说明

Idle14711 months ago
Python