Skip to content
#

unsloth

Here are 559 public repositories matching this topic...

本项目利用医学领域的 CoT 数据对 Deepseek-R1-Distill-Qwen-7B 进行微调,通过 QLoRA 量化和 Unsloth 加速训练,显著提升模型在复杂医学推理任务中的慢思考能力。知识蒸馏技术使轻量级模型获得大模型的推理优势,实现高效、准确且具有解释性的医学问答系统。

  • Updated Mar 10, 2025
  • Python
unsloth-llama3-alpaca-lora

Custom model training using modern architectures. 4-bit QLoRA fine-tuning pipeline for LLaMA 3 8B with production-grade optimization. Memory-efficient training on consumer GPUs. Published adapter on HuggingFace. From training pipeline to deployed model.

  • Updated Apr 2, 2026
  • Jupyter Notebook

Add this topic to your repo

To associate your repository with the unsloth topic, visit your repo's landing page and select "manage topics."

Learn more