Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I found that keeping current context utilization at 18% of total context length was best for minimizing spend, across all models with 400k context length or more


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: