Research Papers research paper arxiv ai artificial-intelligence

Drive My Way: Preference Alignment of Vision-Language-Action Model for Personalized Driving

arXivMarch 26, 202610 min read0 views

Human driving behavior is inherently personal, which is shaped by long-term habits and influenced by short-term intentions. Individuals differ in how they accelerate, brake, merge, yield, and overtake across diverse situations. However, existing end-to-end autonomous driving systems either optimize for generic objectives or rely on fixed driving modes, lacking the ability to adapt to individual preferences or interpret natural language intent. To address this gap, we propose Drive My Way (DMW), a personalized Vision-Language-Action (VLA) driving framework that aligns with users' long-term driv — Zehao Wang, Huaide Jiang, Shuaiwu Dong

View PDF HTML (experimental)

Abstract:Human driving behavior is inherently personal, which is shaped by long-term habits and influenced by short-term intentions. Individuals differ in how they accelerate, brake, merge, yield, and overtake across diverse situations. However, existing end-to-end autonomous driving systems either optimize for generic objectives or rely on fixed driving modes, lacking the ability to adapt to individual preferences or interpret natural language intent. To address this gap, we propose Drive My Way (DMW), a personalized Vision-Language-Action (VLA) driving framework that aligns with users' long-term driving habits and adapts to real-time user instructions. DMW learns a user embedding from our personalized driving dataset collected across multiple real drivers and conditions the policy on this embedding during planning, while natural language instructions provide additional short-term guidance. Closed-loop evaluation on the Bench2Drive benchmark demonstrates that DMW improves style instruction adaptation, and user studies show that its generated behaviors are recognizable as each driver's own style, highlighting personalization as a key capability for human-centered autonomous driving. Our data and code are available at this https URL.

Comments: IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2026); Project website: this https URL

Subjects:

Robotics (cs.RO); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Multiagent Systems (cs.MA)

Cite as: arXiv:2603.25740 [cs.RO]

(or arXiv:2603.25740v1 [cs.RO] for this version)

https://doi.org/10.48550/arXiv.2603.25740

arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Jiachen Li [view email] [v1] Thu, 26 Mar 2026 17:59:54 UTC (3,392 KB)

Original source

arXiv

https://arxiv.org/abs/2603.25740v1

Was this article helpful?

Ask AI about this article

Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

More about

researchpaperarxiv

Research Papers

Humboldt Fellow from the US conducts research in robotics to one day harvest energy from ocean waves

is.mpg.de

1m8 months ago

Models

Howard University and Google Research Enhance A.I. Speech Recognition of African American English - The Dig at Howard University

<a href="https://news.google.com/rss/articles/CBMiygFBVV95cUxQRTh4T2h6cVRsdEF2cjlkWGQyT2tWZnVTTmh4czBJV3ZpSmd1T1Z2eG5Ld1dvQWhNckpjRDItVEtiZ2hMdjBVLWJ0b0xTY0pieG82U0VibXFBLWVUN0tlQ3J1dzBFa2ZBekF1YXJPZlpHNGtkOWZjdWFCSlVTQTctcTNvcURtOER4MnhnYk1BQUt4WllmekE4WkVERTA4Wi1VcnFCY2xYSml6ak9GM1o1NmI0VWtXb2xERlVZVFNBTTQyQ1FBWThESk53?oc=5" target="_blank">Howard University and Google Research Enhance A.I. Speech Recognition of African American English</a> The Dig at Howard University

GNews AI voice

1m9 months ago

Products

Speech-to-Retrieval (S2R): A new approach to voice search - research.google

<a href="https://news.google.com/rss/articles/CBMijAFBVV95cUxQekN0T0VkREpJVGk0U25zMVcyX0VYV0V4eVRJY2ozVW02ampCVXFMRDJybk56blpMdWVhdkRsWWI2S19JemlYM3dHd2dBSkx0SWxtNnNfN18zcjBKLWVXN3JZUnVFdndndTBnSVlVSGhVdWwyS1V3TkRCSUJ5SnRkYXJBV1NfZWUwa3ByWA?oc=5" target="_blank">Speech-to-Retrieval (S2R): A new approach to voice search</a> research.google

GNews AI voice

1m6 months ago

Knowledge Map

TopicsEntitiesSource

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Knowledge Graph100 articles · 97 connections

Scroll to zoom · drag to pan · click to open

Discussion

No comments yet — be the first to share your thoughts!

More in Research Papers

Research Papers

Humboldt Fellow from the US conducts research in robotics to one day harvest energy from ocean waves

is.mpg.de

1m8 months ago

Research Papers

AI-driven digital manipulation ‘tested’ Dutch election integrity, researchers warn - EUobserver

<a href="https://news.google.com/rss/articles/CBMirwFBVV95cUxQcERTcUc5ZndxZ054endXTXNwTlhtYjRyLXBHWVJmRXloNV9JUUpFZnBrLUdDeUpSNklZRFJuUXl0bThIT2ZzbFd6ZU02TW9yaXBPbHducUlHaXVUbWprS0pla0JENkxpSkZfWW9vdTRvcjIzc2ZzWGF6ZmJPMXRVRkFnNmp5NWpLZTBIRk9LamF2RUtkdnQ2bFJXRVZMdVkxZWNHVUl1SzZZeE1JT3R3?oc=5" target="_blank">AI-driven digital manipulation ‘tested’ Dutch election integrity, researchers warn</a> EUobserver

GNews AI Netherlands

1m2 months ago

Research PapersLive

Why Drug Toxicity Can’t Be Predicted in Isolation — Building EIRION with Graph Neural Networks

How we built a graph neural network that finally sees the whole play — not just the audition Every year, drugs that passed early safety tests go on to harm people in ways nobody predicted. Not because the chemistry was wrong. Not because the researchers were careless. But because we kept evaluating drugs the way a talent agent judges an actor from a solo audition tape. Isolated. Out of context. No script. No co-stars. No stage. In real theatre, a performance is never just about one actor. It depends on who they share the stage with, which scene they appear in, what the story demands at that moment. A brilliant performer in the wrong play, surrounded by the wrong cast, in the wrong context — can still wreck the whole production. That is exactly how drug toxicity works. And that is exactly t

Towards AI

17mabout 1 hour ago

Research PapersLive

It's Not Smarter Models — It's Cheaper Memory: TurboQuant's Real Impact, Wall Street Panic & Academic Storm

<blockquote> One-line summary: TurboQuant is a genuinely important engineering breakthrough — but Google's marketing, academic ethics controversy, and Wall Street's overreaction made the story far more dramatic than the technology itself. </blockquote> <h2> 0. What This Article Answers </h2> Google Research published TurboQuant at ICLR 2026 (<a href="https://arxiv.org/abs/2504.19874" rel="noopener noreferrer">arXiv 2504.19874</a>), claiming 6x memory compression, 8x speedup, and zero accuracy loss for LLM KV caches. Then, in the same week: <ol> <li>Global memory stocks lost over $90 billion in market cap</li> <li>An ETH Zürich researcher publicly accused the paper of academic plagiarism and experimental fraud </li> <li

DEV Community

10mabout 1 hour ago

Drive My Way: Preference Alignment of Vision-Language-Action Model for Personalized Driving

Submission history

Daily AI Digest

More about

Humboldt Fellow from the US conducts research in robotics to one day harvest energy from ocean waves

Howard University and Google Research Enhance A.I. Speech Recognition of African American English - The Dig at Howard University

​​Speech-to-Retrieval (S2R): A new approach to voice search - research.google

Knowledge Map

Connected Articles — Knowledge Graph

Discussion

More in Research Papers

Humboldt Fellow from the US conducts research in robotics to one day harvest energy from ocean waves

AI-driven digital manipulation ‘tested’ Dutch election integrity, researchers warn - EUobserver

Why Drug Toxicity Can’t Be Predicted in Isolation — Building EIRION with Graph Neural Networks

It's Not Smarter Models — It's Cheaper Memory: TurboQuant's Real Impact, Wall Street Panic & Academic Storm

Speech-to-Retrieval (S2R): A new approach to voice search - research.google