Fine-tuning hay bị kẹt ở mấy phần nhàm chán: dòng JSONL lỗi, dữ liệu trùng lặp, rồi ngồi canh job chạy. Pipeline này xử lý toàn bộ quy trình từ một thư mục Drive.
Fine-tuning hay bị kẹt ở mấy phần nhàm chán: dòng JSONL lỗi, dữ liệu trùng lặp, rồi ngồi canh job chạy. Pipeline này xử lý toàn bộ quy trình từ một thư mục Drive.
Nó kéo dataset về, kiểm tra và làm sạch định dạng, chia tập train và validation, rồi khởi chạy fine-tune và theo dõi đến khi hoàn tất. Bạn nhận lại model ID và validation loss, việc duy nhất cần làm là trỏ đến thư mục và chọn base model.
Cách vận hành
Used at step 01 to kick off the pipeline.
Write
Used at step 01 to kick off the pipeline.
WebFetch
Used at step 01 to kick off the pipeline.
WebSearch
Used at step 01 to kick off the pipeline.
Làm sạch các dòng lỗi, loại bỏ bản sao trùng lặp và chia tập dữ liệu thành phần huấn luyện và kiểm định.
pending
Tải file đã chuẩn bị lên và khởi chạy tác vụ fine-tuning trên OpenAI với mô hình nền đã chọn.
pending
Theo dõi tiến trình đến khi hoàn tất và báo cáo ID mô hình cuối cùng cùng độ mất mát (loss) ở tập kiểm định.
pending
Ví dụ đầu ra
json
// Sample output
// (generated when the pipeline finishes)
Take a Drive dataset folder, validate and prepare it, launch an OpenAI fine-tune, and report the model ID and loss.
Unlock the rest
The full agent definition, install snippet, and starter task are gated for community members.
Members get the full `.md` agent file, the npm / pnpm install one-liners, a starter prompt that we've tuned against real runs, and the open-source repo when this automation ships there. One email, magic link, done.