Empirical notes on fine-tuning π0.5 for a real manufacturing task: training duration, LoRA vs full fine-tuning, batch size, data quantity and quality, and how much your evaluation can actually resolve.

1 points•dopaul•7 days ago•0 comments•

0 comments

No comments yet.

Read the full thread on Hacker News →

Related stories