SwapTransformer: Highway Overtaking Tactical Planner Model via Imitation Learning on OSHA Dataset
Alireza Shamsoshoara, safin.salih@vw.com, pedram.aghazadeh@vw.com
OpenReview ground truth
Abstract
This paper investigates the high-level decision-making problem in highway scenarios regarding lane changing and over-taking other slower vehicles. In particular, this paper aims to improve the Travel Assist feature for automatic overtaking and lane changes on highways. About 9 million samples including lane images and other dynamic objects are collected in simulation. This data; Overtaking on Simulated HighwAys (OSHA) dataset is released to tackle this challenge. To solve this problem, an architecture called SwapTransformer is designed and implemented as an imitation learning approach on the OSHA dataset. Moreover, auxiliary tasks such as future points and car distance network predictions are proposed to aid the model in better understanding the surrounding environment. The performance of the proposed solution is compared with a multi-layer perceptron (MLP) and multi-head self-attention networks as baselines in a simulation environment. We also demonstrate the performance of the model with and without auxiliary tasks. All models are evaluated based on different metrics such as time to finish each lap, number of overtakes, and speed difference with speed limit. The evaluation shows that the SwapTransformer model outperforms other models in different traffic densities in the inference phase.
Author context
Most prolific author: 1 submissions (credibility 1.00).
No mass-submission penalty for this paper (authors within normal submission volume).
Aggregate statistics only — no individual author rankings.
Ranking trajectory
Percentile by tournament round — convergence indicates rating stability.
Battle history — 36 comparisons
Ranked above opponent in 34% of matchups.
- ▼ lost to Deep Learning-based Discrimination of Paus… ×12
- ▲ beat TABLEYE: SEEING SMALL TABLES THROUGH THE L… ×8
- ▼ lost to Exploring High-Order Message-Passing in Gr… ×6
- ▼ lost to Modulate Your Spectrum in Self-Supervised … ×4
- ▼ lost to Fixed Non-negative Orthogonal Classifier: … ×4
Judge assessments
Mean overall score 0.0 ± 0.0 (n = 36)