Hi,
Could you share some reference information for reproduction:
- How many GPUs are recommended for
ptp_pregenerate (pre‑generate base‑model completions) and ptp_train respectively?
- What is the approximate runtime for each step with TinyLlama‑1.1B (under your experimental setup)?
Hi,
Could you share some reference information for reproduction:
ptp_pregenerate(pre‑generate base‑model completions) andptp_trainrespectively?