Chọn base, fine-tune, harness, và serving — GPU riêng khi corpus không rời đi.
Ghi chú model từ chương trình research-to-prod.