Decide, Then Baseline
Establish whether this problem needs new weights at all, and what number would prove it.
$87
this phase
- Open1.1
The Cheaper Answers First
Free previewA better prompt, few-shot examples, or retrieval beats a tune more often than anyone admits.
- 1.2
What Tuning Actually Buys
Format compliance, tone, latency from a smaller model, and private behavior no prompt encodes.
- 1.3
Writing the Task Down
Turn 'make it better' into a scored task with a holdout set a stranger could grade the same way.
- 1.4
Baseline Scorecard
Measure the base model on your task, on IFEval, and on MMLU before touching a single weight.
- 1.5
Hardware and Cost Math
Bytes per parameter, H100 near 2 to 3 dollars an hour, and what an 8B LoRA run really costs.