Advancing Multimodal LLMs by Large-Scale 3D Visual Instruction Dataset Generation
Fuente:
arXiv
Saved in:
| Main Authors: | He, Liu, Zeng, Xiao, Song, Yizhi, Chen, Albert Y. C., Xia, Lu, Verma, Shashwat, Dayal, Sankalp, Sun, Min, Kuo, Cheng-Hao, Aliaga, Daniel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation
by: He, Liu, et al.
Published: (2024)
by: He, Liu, et al.
Published: (2024)
COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generation
by: He, Liu, et al.
Published: (2024)
by: He, Liu, et al.
Published: (2024)
Adaptive Frameless Rendering
by: Dayal, Abhinav, et al.
Published: (2025)
by: Dayal, Abhinav, et al.
Published: (2025)
Adaptive Multi-Resolution Encoding for Interactive Large-Scale Volume Visualization through Functional Approximation
by: Sun, Jianxin, et al.
Published: (2024)
by: Sun, Jianxin, et al.
Published: (2024)
ACT-R: Adaptive Camera Trajectories for Single View 3D Reconstruction
by: Wang, Yizhi, et al.
Published: (2025)
by: Wang, Yizhi, et al.
Published: (2025)
Counterpoint: Orchestrating Large-Scale Custom Animated Visualizations
by: Sivaraman, Venkatesh, et al.
Published: (2024)
by: Sivaraman, Venkatesh, et al.
Published: (2024)
Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation
by: Wang, Feifan, et al.
Published: (2025)
by: Wang, Feifan, et al.
Published: (2025)
MAIDR: Making Statistical Visualizations Accessible with Multimodal Data Representation
by: Seo, JooYoung, et al.
Published: (2024)
by: Seo, JooYoung, et al.
Published: (2024)
ProHap Explorer: Visualizing Haplotypes in Proteogenomic Datasets
by: Vašíček, Jakub, et al.
Published: (2025)
by: Vašíček, Jakub, et al.
Published: (2025)
VizTA: Enhancing Comprehension of Distributional Visualization with Visual‐Lexical Fused Conversational Interface
by: Liangwei Wang, et al.
Published: (2025)
by: Liangwei Wang, et al.
Published: (2025)
3D Gaussian Particle Approximation of VDB Datasets: A Study for Scientific Visualization
by: Sharma, Isha, et al.
Published: (2025)
by: Sharma, Isha, et al.
Published: (2025)
Reliving the Dataset: Combining the Visualization of Road Users' Interactions with Scenario Reconstruction in Virtual Reality
by: Töttel, Lars, et al.
Published: (2021)
by: Töttel, Lars, et al.
Published: (2021)
Holo360D: A Large-Scale Real-World Dataset with Continuous Trajectories for Advancing Panoramic 3D Reconstruction and Beyond
by: Ou, Jing, et al.
Published: (2026)
by: Ou, Jing, et al.
Published: (2026)
Rendering Large Volume Datasets in Unreal Engine 5: A Survey
by: Schlüter, Markus, et al.
Published: (2025)
by: Schlüter, Markus, et al.
Published: (2025)
Concurrent Binary Trees for Large-Scale Game Components
by: Benyoub, Anis, et al.
Published: (2024)
by: Benyoub, Anis, et al.
Published: (2024)
Detail Enhanced Gaussian Splatting for Large-Scale Volumetric Capture
by: Philip, Julien, et al.
Published: (2025)
by: Philip, Julien, et al.
Published: (2025)
FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On
by: Karras, Johanna, et al.
Published: (2026)
by: Karras, Johanna, et al.
Published: (2026)
Interactive Hypergraph Visual Analytics for Exploring Large and Complex Image Collections
by: Gisolf, Floris, et al.
Published: (2025)
by: Gisolf, Floris, et al.
Published: (2025)
VibraVerse: A Large-Scale Geometry-Acoustics Alignment Dataset for Physically-Consistent Multimodal Learning
by: Pang, Bo, et al.
Published: (2025)
by: Pang, Bo, et al.
Published: (2025)
Animator-Centric Skeleton Generation on Objects with Fine-Grained Details
by: Sun, Mingze, et al.
Published: (2026)
by: Sun, Mingze, et al.
Published: (2026)
Palace: A Library for Interactive GPU-Accelerated Large Tensor Processing and Visualization
by: Drees, Dominik, et al.
Published: (2025)
by: Drees, Dominik, et al.
Published: (2025)
From Cluster to Desktop: A Cache-Accelerated INR framework for Interactive Visualization of Tera-Scale Data
by: Zavorotny, Daniel, et al.
Published: (2025)
by: Zavorotny, Daniel, et al.
Published: (2025)
CelloCut: Constructive Watertight Remeshing via Tetrahedral Cell Cuts
by: Yang, Xuan, et al.
Published: (2026)
by: Yang, Xuan, et al.
Published: (2026)
Enhance Comprehension of Over-the-Counter Drug Instructions for the General Public and Medical Professionals through Visualization Design
by: Fan, Mengjie, et al.
Published: (2026)
by: Fan, Mengjie, et al.
Published: (2026)
CCWSIM: An Efficient and Fast Wavelet-Based CCSIM for Categorical Characterization of Large-Scale
by: Bavandsavadkoohi, Mojtaba, et al.
Published: (2024)
by: Bavandsavadkoohi, Mojtaba, et al.
Published: (2024)
M2fNet: Multi-modal Forest Monitoring Network on Large-scale Virtual Dataset
by: Lu, Yawen, et al.
Published: (2024)
by: Lu, Yawen, et al.
Published: (2024)
Towards Scaling‐Invariant Projections for Data Visualization
by: Joel Dierkes, et al.
Published: (2025)
by: Joel Dierkes, et al.
Published: (2025)
One Model to Rig Them All: Diverse Skeleton Rigging with UniRig
by: Zhang, Jia-Peng, et al.
Published: (2025)
by: Zhang, Jia-Peng, et al.
Published: (2025)
SDGraph: Multi-Level Sketch Representation Learning by Sparse-Dense Graph Architecture
by: Cheng, Xi, et al.
Published: (2025)
by: Cheng, Xi, et al.
Published: (2025)
InterChat: Enhancing Generative Visual Analytics using Multimodal Interactions
by: Juntong Chen, et al.
Published: (2025)
by: Juntong Chen, et al.
Published: (2025)
GHAR: GeoPose-based Handheld Augmented Reality for Architectural Positioning, Manipulation and Visual Exploration
by: Israr, Sabahat, et al.
Published: (2025)
by: Israr, Sabahat, et al.
Published: (2025)
Large-Scale Photogrammetric Documentation of St. John's Co-Cathedral: A Workflow for Cultural Heritage Preservation
by: Kenely, Matthew, et al.
Published: (2026)
by: Kenely, Matthew, et al.
Published: (2026)
From Visual Synthesis to Interactive Worlds: Toward Production-Ready 3D Asset Generation
by: Wu, Jiafeng, et al.
Published: (2026)
by: Wu, Jiafeng, et al.
Published: (2026)
QuadricsReg: Large-Scale Point Cloud Registration using Quadric Primitives
by: Wu, Ji, et al.
Published: (2024)
by: Wu, Ji, et al.
Published: (2024)
Accelerating local laplacian filters on FPGAs
by: Khandelwal, Shashwat, et al.
Published: (2024)
by: Khandelwal, Shashwat, et al.
Published: (2024)
SpringTime: Learning Simulatable Models of Cloth with Spatially-varying Constitutive Properties
by: Chen, Guanxiong, et al.
Published: (2025)
by: Chen, Guanxiong, et al.
Published: (2025)
Comparative Study of Four Visualization Techniques and Positional Variations for Displaying Exercise Data on Smartwatches
by: Yu Liu, et al.
Published: (2025)
by: Yu Liu, et al.
Published: (2025)
SVGEditBench V2: A Benchmark for Instruction-based SVG Editing
by: Nishina, Kunato, et al.
Published: (2025)
by: Nishina, Kunato, et al.
Published: (2025)
OLATverse: A Large-scale Real-world Object Dataset with Precise Lighting Control
by: Zhou, Xilong, et al.
Published: (2025)
by: Zhou, Xilong, et al.
Published: (2025)
A Study on Activity Visualization for Smart Watches
by: Xia, Zhouxuan, et al.
Published: (2024)
by: Xia, Zhouxuan, et al.
Published: (2024)
Similar Items
-
Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation
by: He, Liu, et al.
Published: (2024) -
COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generation
by: He, Liu, et al.
Published: (2024) -
Adaptive Frameless Rendering
by: Dayal, Abhinav, et al.
Published: (2025) -
Adaptive Multi-Resolution Encoding for Interactive Large-Scale Volume Visualization through Functional Approximation
by: Sun, Jianxin, et al.
Published: (2024) -
ACT-R: Adaptive Camera Trajectories for Single View 3D Reconstruction
by: Wang, Yizhi, et al.
Published: (2025)