Ya prompt engineering can be a more difficult than it looks. Especially when dealing with less intelligent models.
That's why I recommend having an error checking stage were the model gives a model should be able to return a simple "True" or "Yes" when presented it's last response. This eats up more GPU time but the signal to noise ratio improves drastically.
> That's why I recommend having an error checking stage were the model gives a model should be able to return a simple "True" or "Yes" when presented it's last response.
Mind elaborating on that? Looks like a typo but i'm having difficulty knowing for sure. Thanks!