This is not a LLM.
This is an AI in maybe the more traditional/popsci sense. A digital robot. It not just understands its perceptions (like an LLM), but it acts on that understanding to achieve its goal(s).
The reinforcement learning aspect is simply how it learns its goals. It takes a database of "good bot" / "bad bot" feedback and associated context, and implicitly learns what it should do.