@ayo They don't know when to shut up but the LLM clients that give you access to all the knobs let you tell it how long to talk.
Reasoning models, they make them talk on and on on purpose.
https://docs.sillytavern.app/usage/common-settings/#response-tokens