Autonomous 3D agent learning to navigate a spatial environment.
The agent perceives the world from a first-person perspective, avoids collisions with walls and objects, and independently explores the arena to find randomly placed rewards (apples). Learning emerges through continuous interaction with the environment and simple reinforcement, without predefined paths or scripted intelligence.