Multi-Agent Orchestration with the Common-Sense Reasoning Capabilities of LLMs for Autonomous Driving
August 21, 2026
A hybrid framework coordinates PPO-trained reinforcement learning and PID control through an LLM-based orchestrator to mitigate latency and hallucination risks. The system applies LLM common-sense reasoning to iteratively refine RL reward functions during dynamic driving tasks. Evaluation was conducted using highly randomized scenarios within the CARLA simulator.