castform founder here. the roi optimization makes sense. i think there are lots of usecases for which even a 2% gain in accuracy can be quite useful. off the top of my head
- high volume customer support. higher accuracy means fewer escalation, reducing labor costs
- fraud detection. catching even one extra fraud attempt could mean a lot in savings
- and ofc the classic ads use-case where at scale bps in improvement could mean millions in revenue :)
the bigger model would still cost more :)
at the same time, i see prompting as being orthogonal to post-training. i'd imagine post-training a smaller model with a better prompt would make it perform even better
- high volume customer support. higher accuracy means fewer escalation, reducing labor costs - fraud detection. catching even one extra fraud attempt could mean a lot in savings - and ofc the classic ads use-case where at scale bps in improvement could mean millions in revenue :)