Skip to content
#

swanlab

Here are 4 public repositories matching this topic...

Language: All
Filter by language

Supervised fine-tuning (SFT) of Qwen3 for structured medical reasoning QA — teaching models to "think before answering" (<think>...</think>). Supports full fine-tuning & LoRA, with an end-to-end data→train→eval→compare pipeline and quantitative benchmarks (PPL, format compliance, semantic similarity, latency/throughput).

  • Updated Jul 16, 2026
  • Python

Improve this page

Add a description, image, and links to the swanlab topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the swanlab topic, visit your repo's landing page and select "manage topics."

Learn more