Discussion about this post

User's avatar
Mohamed F. Ahmed's avatar

I'd push back slightly on treating this as pure convergence. I built an agent harness for a narrow legal-adjacent workflow last year, and fine-tuning a small model on 50k labeled examples beat GPT-4 orchestration on both cost and latency. The Workload-Harness Fit framework is useful, but in my experience the deciding factor isn't workload type so much as data density: if you've got a tight, repetitive task with clean historical examples, training wins fast.

No posts

Ready for more?