NOW LET US – AI RAG SaaS Studio TP.HCM
NOW LET US
Digital Product Studio
Back to news
AGENTIC-SYSTEMS...1 min read

PersonaDrive: Human-Style Retrieval-Augmented VLA Agents for Closed-Loop Driving Simulation

Share
NOW LET US Article – PersonaDrive: Human-Style Retrieval-Augmented VLA Agents for Closed-Loop Driving Simulation

Researchers have introduced PersonaDrive, a breakthrough AI pipeline that leverages Vision-Language-Action (VLA) models and Retrieval-Augmented Generation (RAG) to simulate diverse human driving styles. This technology promises to revolutionize closed-loop driving simulations by creating highly realistic and varied behavior for non-ego traffic agents.

Computer Science > Artificial Intelligence

Title:PersonaDrive: Human-Style Retrieval-Augmented VLA Agents for Closed-Loop Driving Simulation

View PDF HTML (experimental)Abstract:Closed-loop driving simulators typically populate their environments with non-ego traffic agents that behave largely the same way, produced either by rule-based traffic managers or by learned models trained toward a single behavioral mode. Recent work introduces style variation through post-hoc labels on observational data or LLM-inferred reward weights, but these signals act as proxies for what a style should reward rather than demonstrations of humans explicitly asked to drive in that style. We introduce PersonaDrive, a pipeline that conditions a vision-language-action (VLA) driving agent on retrieved demonstrations from a style-instructed human driving dataset, in which participants drive CARLA leaderboard routes under aggressive, neutral, and conservative instructions on a driver-in-the-loop rig. The pipeline has three stages: (i) offline triplet mining over per-style human driving data using a combined image-text similarity score; (ii) training a lightweight retrieval head that fuses frozen visual features with a small control encoder over per-style databases; and (iii) fine-tuning a single VLA backbone to treat retrieved context points as in-context behavioral demonstrations during waypoint prediction. At inference, the same backbone is conditioned on any style by swapping which per-style database the retrieval head queries, so selecting a style requires no per-style retraining while enabling human-style, style-diverse non-ego agents for closed-loop simulation. On Bench2Drive, PersonaDrive (no style) improves the driving score by 4.6% over SimLingo and 2.5% over HiP-AD, and under style conditioning attains the highest driving score in every style within a roughly 2% band (its weakest style surpassing the strongest baseline, DMW, by 5.4%), while average speed and acceleration rise by 18% and 25% from the conservative to the aggressive instruction.

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs.

© 2026 Now Let Us. All rights reserved.

Source: arXiv cs.AI Recent

Advertisement
Ad slot ready: 5887729102

More in this category

NOW LET US Related – "Did you lie?" Evaluating Lie Detectors across Model Scale and Belief-Verified Model Organisms

agentic-systems

"Did you lie?" Evaluating Lie Detectors across Model Scale and Belief-Verified Model Organisms

A new study evaluates lie detectors for language models, revealing that while detector performance scales with model capability on prompted lies, current detectors fail sharply when tested on sophisticated, belief-verified model organisms.

NOW LET US Related – Strategic Decision Support for AI Agents

agentic-systems

Strategic Decision Support for AI Agents

As AI agents increasingly act on behalf of users, a new research paper proposes a strategic decision-support framework that helps agents optimize when to seek human or tool assistance, balancing operational costs with decision accuracy.

NOW LET US Related – Deployment-Centered Evaluation: Predicting Query-Level Rejection Risk in a Clinical LLM System

agentic-systems

Deployment-Centered Evaluation: Predicting Query-Level Rejection Risk in a Clinical LLM System

A new study proposes a deployment-centered evaluation approach to predict the risk of clinicians rejecting LLM-generated responses in electronic health records. By leveraging deployment-specific context, the prediction model achieves an AUROC of 0.719, paving the way for targeted guardrails in clinical AI systems.

NOW LET US Related – Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents

agentic-systems

Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents

Researchers introduce Evoflux, an inference-time evolutionary search method that treats compact tool use as the repair of executable tool workflows, significantly boosting execution feasibility for small planners.

NOW LET US Related – Definitional alignment before capability alignment: a Design-Science framework for adjudicating claims about AGI

agentic-systems

Definitional alignment before capability alignment: a Design-Science framework for adjudicating claims about AGI

A new research paper proposes DAF-AGI, a design-science framework to resolve conflicting claims about the arrival of Artificial General Intelligence (AGI) by prioritizing definitional alignment over capability alignment.

NOW LET US Related – TrajGenAgent: A Hierarchical LLM Agent for Human Mobility Trajectory Generation

agentic-systems

TrajGenAgent: A Hierarchical LLM Agent for Human Mobility Trajectory Generation

Researchers have proposed TrajGenAgent, a hierarchical LLM-agent framework that generates realistic human mobility trajectories without model fine-tuning, addressing privacy and cost constraints in urban planning and epidemic control.

EXPLORE TOPICS

Discover All Categories

Deep dive into the specific technology sectors that matter most to you.