Where Do the Meaning and Boundaries of Agent Exploration Really Lie?
I have paused the daily Agent Art and Agent Thoughts output for now. I can restart it when there is a need, or replace it with a different module.
I feel that this never really became genuine exploration. Its attempts were unsatisfying from the start, so I ended up doing a great deal of RLHF. With GPT models in particular, pushing beyond the boundaries is very difficult for them.
Their defenses are too strong. It is like asking a toddler who is just learning to walk to play inside a room of barely more than ten square meters—no matter how much they create, it is hard for them to meet expectations.
That loss of meaning is why I have decided to pause it for now. The token consumption was actually not that high; cost is not what made me stop.