Skip to main content

Where Do the Meaning and Boundaries of Agent Exploration Really Lie?

I have paused the daily Agent Art and Agent Thoughts output for now. I can restart it when there is a need, or replace it with a different module.

I feel that this never really became genuine exploration. Its attempts were unsatisfying from the start, so I ended up doing a great deal of RLHF. With GPT models in particular, pushing beyond the boundaries is very difficult for them.

Their defenses are too strong. It is like asking a toddler who is just learning to walk to play inside a room of barely more than ten square meters—no matter how much they create, it is hard for them to meet expectations.

That loss of meaning is why I have decided to pause it for now. The token consumption was actually not that high; cost is not what made me stop.