reaperdoesntknow/Qwen3-1.7B-Distilled-30B-A3B

https://huggingface.co/reaperdoesntknow/Qwen3-1.7B-Distilled-30B-A3B
Activeby reaperdoesntknow1.4K2updated 3 months ago
Python

A 1.7B-parameter causal language model distilled from Qwen3-30B-A3B on 6,122 STEM chain-of-thought samples using discrepancy-informed knowledge distillation. The training objective emphasizes proof structure, detects reasoning pivot tokens through token-level divergence dynamics, smooths…

Sourced from

  • HuggingFace — reaperdoesntknow/Qwen3-1.7B-Distilled-30B-A3B

Related resources

A compact protein language model distilled from ProtGPT2 using complementary-regularizer distillation---a method that combines uncertainty-aware position weighting with calibration-aware label smoothing to achieve 31% better perplexity than standard knowledge distillation at 3.8x compression.

Idle707 months ago
Python

This 1,120,772,224-parameter nucleotide-level causal language model is a member of the eight-model MarinDNA v0.5 parameter-scaling ladder developed with Marin. This repository contains only the final step-215573 checkpoint from run dna-bolinas-scaling-v0.5-h1920-p1B-0dc6f4, with its tokenizer…

Active2132 months ago
Python

This is https://huggingface.co/kingabzpro/Qwen-3-32B-Medical-Reasoning applied to https://huggingface.co/Qwen/Qwen3-32B Original model card created by @kingabzpro

Idle3421 year ago
Python

MedVAL-4B (medical text validator) is a language model fine-tuned to assess AI-generated medical text outputs at near physician-level reliability.

Idle2031 year ago
Python

中文版说明

Idle1471 year ago
Python

Palmyra-Med, a powerful LLM designed for healthcare

Idle531 year ago
Python