AirLetters: An Open Video Dataset of Characters Drawn in the Air
Fuente:
arXiv
Guardado en:
| Autores principales: | Dagli, Rishit, Berger, Guillaume, Materzynska, Joanna, Bax, Ingo, Memisevic, Roland |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Can Vision-Language Models Answer Face to Face Questions in the Real-World?
por: Pourreza, Reza, et al.
Publicado: (2025)
por: Pourreza, Reza, et al.
Publicado: (2025)
DiffuseRAW: End-to-End Generative RAW Image Processing for Low-Light Images
por: Dagli, Rishit
Publicado: (2023)
por: Dagli, Rishit
Publicado: (2023)
What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction
por: Panchal, Sunny, et al.
Publicado: (2024)
por: Panchal, Sunny, et al.
Publicado: (2024)
FairyGen: Storied Cartoon Video from a Single Child-Drawn Character
por: Zheng, Jiayi, et al.
Publicado: (2025)
por: Zheng, Jiayi, et al.
Publicado: (2025)
NeRF-US: Removing Ultrasound Imaging Artifacts from Neural Radiance Fields in the Wild
por: Dagli, Rishit, et al.
Publicado: (2024)
por: Dagli, Rishit, et al.
Publicado: (2024)
Opt-In Art: Learning Art Styles Only from Few Examples
por: Ren, Hui, et al.
Publicado: (2024)
por: Ren, Hui, et al.
Publicado: (2024)
Squeeze3D: Your 3D Generation Model is Secretly an Extreme Neural Compressor
por: Dagli, Rishit, et al.
Publicado: (2025)
por: Dagli, Rishit, et al.
Publicado: (2025)
SEE-2-SOUND: Zero-Shot Spatial Environment-to-Spatial Sound
por: Dagli, Rishit, et al.
Publicado: (2024)
por: Dagli, Rishit, et al.
Publicado: (2024)
NewMove: Customizing text-to-video models with novel motions
por: Materzynska, Joanna, et al.
Publicado: (2023)
por: Materzynska, Joanna, et al.
Publicado: (2023)
OpenCOOD-Air: Prompting Heterogeneous Ground-Air Collaborative Perception with Spatial Conversion and Offset Prediction
por: Wu, Xianke, et al.
Publicado: (2026)
por: Wu, Xianke, et al.
Publicado: (2026)
Unified Concept Editing in Diffusion Models
por: Gandikota, Rohit, et al.
Publicado: (2023)
por: Gandikota, Rohit, et al.
Publicado: (2023)
SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary
por: He, Jingyi, et al.
Publicado: (2026)
por: He, Jingyi, et al.
Publicado: (2026)
Common Corruptions for Enhancing and Evaluating Robustness in Air-to-Air Visual Object Detection
por: Arsenos, Anastasios, et al.
Publicado: (2024)
por: Arsenos, Anastasios, et al.
Publicado: (2024)
Model-Agnostic Open-Set Air-to-Air Visual Object Detection for Reliable UAV Perception
por: Loukovitis, Spyridon, et al.
Publicado: (2025)
por: Loukovitis, Spyridon, et al.
Publicado: (2025)
Character Mixing for Video Generation
por: Liao, Tingting, et al.
Publicado: (2025)
por: Liao, Tingting, et al.
Publicado: (2025)
FreeForm: Reduced-Order Deformable Simulation from Particle-Based Skinning Eigenmodes
por: Xiang, Donglai, et al.
Publicado: (2026)
por: Xiang, Donglai, et al.
Publicado: (2026)
PM25Vision: A Large-Scale Benchmark Dataset for Visual Estimation of Air Quality
por: Han, Yang
Publicado: (2025)
por: Han, Yang
Publicado: (2025)
SCT-MOT: Enhancing Air-to-Air Multiple UAVs Tracking with Swarm-Coupled Motion and Trajectory Guidance
por: Chu, Zhaochen, et al.
Publicado: (2026)
por: Chu, Zhaochen, et al.
Publicado: (2026)
AirZoo: A Unified Large-Scale Dataset for Grounding Aerial Geometric 3D Vision
por: Cheng, Xiaoya, et al.
Publicado: (2026)
por: Cheng, Xiaoya, et al.
Publicado: (2026)
MovieCharacter: A Tuning-Free Framework for Controllable Character Video Synthesis
por: Qiu, Di, et al.
Publicado: (2024)
por: Qiu, Di, et al.
Publicado: (2024)
Detecting Pen-In-Air States from Video: A Proof-of-Concept Toward Complementary Handwriting Analysis
por: Sismeiro, Lauren, et al.
Publicado: (2026)
por: Sismeiro, Lauren, et al.
Publicado: (2026)
Look, Remember and Reason: Grounded reasoning in videos with language models
por: Bhattacharyya, Apratim, et al.
Publicado: (2023)
por: Bhattacharyya, Apratim, et al.
Publicado: (2023)
AirRoom: Objects Matter in Room Reidentification
por: Yao, Runmao, et al.
Publicado: (2025)
por: Yao, Runmao, et al.
Publicado: (2025)
Rotate Your Character: Revisiting Video Diffusion Models for High-Quality 3D Character Generation
por: Wang, Jin, et al.
Publicado: (2026)
por: Wang, Jin, et al.
Publicado: (2026)
AirCast: Improving Air Pollution Forecasting Through Multi-Variable Data Alignment
por: Nedungadi, Vishal, et al.
Publicado: (2025)
por: Nedungadi, Vishal, et al.
Publicado: (2025)
Enhancing 3D-Air Signature by Pen Tip Tail Trajectory Awareness: Dataset and Featuring by Novel Spatio-temporal CNN
por: Atreya, Saurabh, et al.
Publicado: (2024)
por: Atreya, Saurabh, et al.
Publicado: (2024)
Exploring the Efficacy of Modified Transfer Learning in Identifying Parkinson's Disease Through Drawn Image Patterns
por: Daiyan, Nabil, et al.
Publicado: (2025)
por: Daiyan, Nabil, et al.
Publicado: (2025)
Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?
por: Bhattacharyya, Apratim, et al.
Publicado: (2025)
por: Bhattacharyya, Apratim, et al.
Publicado: (2025)
AirShot: Efficient Few-Shot Detection for Autonomous Exploration
por: Wang, Zihan, et al.
Publicado: (2024)
por: Wang, Zihan, et al.
Publicado: (2024)
Do Generalised Classifiers really work on Human Drawn Sketches?
por: Bandyopadhyay, Hmrishav, et al.
Publicado: (2024)
por: Bandyopadhyay, Hmrishav, et al.
Publicado: (2024)
MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling
por: Men, Yifang, et al.
Publicado: (2024)
por: Men, Yifang, et al.
Publicado: (2024)
AnimationBench: Are Video Models Good at Character-Centric Animation?
por: Wu, Leyi, et al.
Publicado: (2026)
por: Wu, Leyi, et al.
Publicado: (2026)
Gloria: Consistent Character Video Generation via Content Anchors
por: Yang, Yuhang, et al.
Publicado: (2026)
por: Yang, Yuhang, et al.
Publicado: (2026)
AirV2X: Unified Air-Ground Vehicle-to-Everything Collaboration
por: Gao, Xiangbo, et al.
Publicado: (2025)
por: Gao, Xiangbo, et al.
Publicado: (2025)
OpenMarcie: Dataset for Multimodal Action Recognition in Industrial Environments
por: Bello, Hymalai, et al.
Publicado: (2026)
por: Bello, Hymalai, et al.
Publicado: (2026)
Oracle-MNIST: a Dataset of Oracle Characters for Benchmarking Machine Learning Algorithms
por: Wang, Mei, et al.
Publicado: (2022)
por: Wang, Mei, et al.
Publicado: (2022)
Transport Network, Graph, and Air Pollution
por: Xu, Nan
Publicado: (2025)
por: Xu, Nan
Publicado: (2025)
MotionCharacter: Fine-Grained Motion Controllable Human Video Generation
por: Fang, Haopeng, et al.
Publicado: (2024)
por: Fang, Haopeng, et al.
Publicado: (2024)
BachVid: Training-Free Video Generation with Consistent Background and Character
por: Yan, Han, et al.
Publicado: (2025)
por: Yan, Han, et al.
Publicado: (2025)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
por: Hu, Li, et al.
Publicado: (2023)
por: Hu, Li, et al.
Publicado: (2023)
Ejemplares similares
-
Can Vision-Language Models Answer Face to Face Questions in the Real-World?
por: Pourreza, Reza, et al.
Publicado: (2025) -
DiffuseRAW: End-to-End Generative RAW Image Processing for Low-Light Images
por: Dagli, Rishit
Publicado: (2023) -
What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction
por: Panchal, Sunny, et al.
Publicado: (2024) -
FairyGen: Storied Cartoon Video from a Single Child-Drawn Character
por: Zheng, Jiayi, et al.
Publicado: (2025) -
NeRF-US: Removing Ultrasound Imaging Artifacts from Neural Radiance Fields in the Wild
por: Dagli, Rishit, et al.
Publicado: (2024)