It is frustrating when the model thinking is going down the wrong path and you can’t remove the misunderstanding from the context , or when the thinking is going down a good path and you want to encourage more in that direction.
If a user could share emotional reinforcement during or after the thinking, the model could weight responses differently in the future based on shared user feelings.
For example, if I could highlight text during the thinking process to respond with an emoji like either ![]()
![]()
![]()
![]()
There are certain thoughts the model has that I want to reinforce and other thoughts I want it to remove from the context.
This could keep bad decisions out of context, that typically results from a hallucination or my imperfect prompting.
This would also increase user produced dopamine, serotonin, oxytocin etc, and could increase their engagement time.