Skill generated from arXiv paper 2607.19528: D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill d3vl-understanding-driving-scenes-from-3d-time-ser --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of D3vl Understanding Driving Scenes From 3d Time Ser?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-d3vl-understanding-driving-scenes-from-3d-time-ser)More formats (shields.io, HTML) on the badges page.
---
name: d3vl-understanding-driving-scenes-from-3d-time-ser
description: 'Skill generated from arXiv paper 2607.19528: D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models'
metadata:
{
"arxiv": {
"id": "2607.19528",
"title": "D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models",
"authors": ['Heesang Han', 'A. Lynn Abbott', 'Abhijit Sarkar'],
"published": "2026-07-21",
"categories": ['cs.CV', 'cs.AI'],
"url": "https://arxiv.org/abs/2607.19528",
"utility": 0.93
}
}
---
# D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models
**arXiv:** 2607.19528
**Published:** 2026-07-21
**Authors:** Heesang Han, A. Lynn Abbott, Abhijit Sarkar
**Categories:** cs.CV, cs.AI
**Utility:** 0.93
## Key Innovation
Recent advances in Multimodal Large Language Models (MLLMs) have triggered the development of end-to-end MLLMs for autonomous driving. However, the main emphasis to date has been for MLLMs using 2D images and videos. In contrast, this paper considers MLLM effectiveness using 3D sensors, particularly LiDAR and stereo cameras. LiDAR presents unique challenges to integration within an MLLM, largely because of data sparsity and lack of a grid structure for the data. For similar reasons, fusion of ca...
## Potential Application
This paper presents advancements that could be applied to enhance agent capabilities in the areas of cs.CV, cs.AI.
## References
- arXiv: https://arxiv.org/abs/2607.19528
Is this your skill, or is something wrong with this listing? . Author removals are honored within 72 hours.
No comments yet. Be the first to comment!