Everyone’s Optimizing Prompts. Nobody’s Optimizing the Data Going Into Them
The discourse around trimming AI costs has focused on prompts. Lesser attention has been paid to the data the tools send back, even though that's where most tokens are spent. While developers focus on creating prompt libraries, counting tokens, and aiming for shorter messages, the payloads agents pull from APIs barely register a mention. It's a real blind spot.
What's Going On?
Prompt tuning can feel like real work. You can cut down your prompt length. You can shave away words. The response you get back seems like it's beyond your control and so you don't think you need to clean it up, even though you'll wind up paying for every token it entails.
Probably because prompts seem to be the most under our control, the internet is chock-full of prompt-engineering tips. The shape of the incoming data is less of a concern. But the bill doesn't really care whether...
Copyright of this story solely belongs to hackernoon.com. To see the full text click HERE