Week 3 Plan
This issue will track goals and tasks for Week 3.
Mục tiêu tuần 3: thực hành semi-hands-on với LoRA/QLoRA, thiết lập repo evaluation, agent orchestration, và mở rộng fine-tuning configs.
🎯 Goals
- Semi-hands-on LoRA/QLoRA experiments on open-weight LLMs.
- Prepare evaluation scripts & probes for base vs LoRA models.
- Minimal agent orchestration loop with mock tools & planner.
- Extend fine-tuning experiments with QLoRA (8-bit) configs.
- Log example outputs (training logs, eval results) for reproducibility.
✅ Tasks
📝 Notes
- Focus on semi-hands-on skeletons; all code/configs should be runnable, even if on tiny datasets.
- Keep GPU memory modest (8–24GB) for LoRA/QloRA.
- Placeholder datasets and logs are fine; the goal is reproducible skeletons.
- Progress to be logged in each repo’s Progress Log in README.md.
Feel free to add updates below as comments.
Week 3 Plan
This issue will track goals and tasks for Week 3.
Mục tiêu tuần 3: thực hành semi-hands-on với LoRA/QLoRA, thiết lập repo evaluation, agent orchestration, và mở rộng fine-tuning configs.
🎯 Goals
✅ Tasks
📝 Notes
Feel free to add updates below as comments.