Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

TreeIRL: Safe Urban Driving with Tree Search and Inverse Reinforcement Learning

Domain:

mobility

Record type:

papersoftware
Creator:
TomLeeHenHuh
Host:avatar
We present TreeIRL, a novel planner for autonomous driving that combines Monte Carlo tree search (MCTS) and inverse reinforcement learning (IRL) to achieve state-of-the-art performance in simulation and in real-world driving. The core idea is to use MCTS to find a promising set of safe candidate trajectories and a deep IRL scoring function to select the most human-like among them. We evaluate TreeIRL against both classical and state-of-the-art planners in large-scale simulations and on 500+ miles of real-world autonomous driving in the Las Vegas metropolitan area. Test scenarios include dense urban traffic, adaptive cruise control, cut-ins, and traffic lights. TreeIRL achieves the best overall performance, striking a balance between safety, progress, comfort, and human-likeness. To our knowledge, our work is the first demonstration of MCTS-based planning on public roads and underscores the importance of evaluating planners across a diverse set of metrics and in real-world environments. TreeIRL is highly extensible and could be further improved with reinforcement learning and imitation learning, providing a framework for exploring different combinations of classical and learning-based approaches to solve the planning bottleneck in autonomous driving.

Visit

arxiv.org

Tags

RoboticsArtificial IntelligenceMachine Learning

Similar

Safe Trajectory Sampling in Model-based Reinforcement LearningMulti-agents Ultimatum Game with Reinforcement LearningM-Walk: Learning to Walk over Graphs using Monte Carlo Tree SearchReinforcement Learning in an Environment Synthetically Augmented with Digital PheromonesDEEP REINFORCEMENT LEARNING WITH HIDDEN MARKOV MODEL FOR SPEECH RECOGNITIONgilbert215/microgrid-reinforcement-Learning-

Safe Trajectory Sampling in Model-based Reinforcement Learning

Safe Trajectory Sampling in Model-based Reinforcement Learning

Poster presented at the Deep Learning Indaba 2023 by Sicelukwanda Zwane

Multi-agents Ultimatum Game with Reinforcement Learning

International Workshops of PAAMS 2020, L'Aquila, Italy, October 7–9, 2020, Proceedings International

M-Walk: Learning to Walk over Graphs using Monte Carlo Tree Search

Learning to walk over a graph towards a target node for a given query and a source node is an import

Reinforcement Learning in an Environment Synthetically Augmented with Digital Pheromones

Reinforcement learning requires information about states, actions, and outcomes as the basis for lea

DEEP REINFORCEMENT LEARNING WITH HIDDEN MARKOV MODEL FOR SPEECH RECOGNITION

Nowadays, many applications uses speech recognition especially the field of computer science and ele

gilbert215/microgrid-reinforcement-Learning-

Adaptive energy dispatch for solar-powered microgrids using Proximal Policy Optimization (PPO) on re