How Video Meetings Change Your Expression
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sarin, Sumit, Mall, Utkarsh, Tendulkar, Purva, Vondrick, Carl |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evolving Interpretable Visual Classifiers with Large Language Models
von: Chiquier, Mia, et al.
Veröffentlicht: (2024)
von: Chiquier, Mia, et al.
Veröffentlicht: (2024)
DiSciPLE: Learning Interpretable Programs for Scientific Visual Discovery
von: Mall, Utkarsh, et al.
Veröffentlicht: (2025)
von: Mall, Utkarsh, et al.
Veröffentlicht: (2025)
4D Gaussian Splatting as a Learned Dynamical System
von: Asiimwe, Arnold Caleb, et al.
Veröffentlicht: (2025)
von: Asiimwe, Arnold Caleb, et al.
Veröffentlicht: (2025)
Scale-Aware Recognition in Satellite Images under Resource Constraints
von: Revankar, Shreelekha, et al.
Veröffentlicht: (2024)
von: Revankar, Shreelekha, et al.
Veröffentlicht: (2024)
MONITRS: Multimodal Observations of Natural Incidents Through Remote Sensing
von: Revankar, Shreelekha, et al.
Veröffentlicht: (2025)
von: Revankar, Shreelekha, et al.
Veröffentlicht: (2025)
Whiteboard-of-Thought: Thinking Step-by-Step Across Modalities
von: Menon, Sachit, et al.
Veröffentlicht: (2024)
von: Menon, Sachit, et al.
Veröffentlicht: (2024)
MedAutoCorrect: Image-Conditioned Autocorrection in Medical Reporting
von: Asiimwe, Arnold Caleb, et al.
Veröffentlicht: (2024)
von: Asiimwe, Arnold Caleb, et al.
Veröffentlicht: (2024)
CAViAR: Critic-Augmented Video Agentic Reasoning
von: Menon, Sachit, et al.
Veröffentlicht: (2025)
von: Menon, Sachit, et al.
Veröffentlicht: (2025)
EraseDraw: Learning to Draw Step-by-Step via Erasing Objects from Images
von: Canberk, Alper, et al.
Veröffentlicht: (2024)
von: Canberk, Alper, et al.
Veröffentlicht: (2024)
Video, How Do Your Tokens Merge?
von: Pollard, Sam, et al.
Veröffentlicht: (2025)
von: Pollard, Sam, et al.
Veröffentlicht: (2025)
Differentiable Robot Rendering
von: Liu, Ruoshi, et al.
Veröffentlicht: (2024)
von: Liu, Ruoshi, et al.
Veröffentlicht: (2024)
Controlling the World by Sleight of Hand
von: Sudhakar, Sruthi, et al.
Veröffentlicht: (2024)
von: Sudhakar, Sruthi, et al.
Veröffentlicht: (2024)
CRAM: Large-scale Video Continual Learning with Bootstrapped Compression
von: Mall, Shivani, et al.
Veröffentlicht: (2025)
von: Mall, Shivani, et al.
Veröffentlicht: (2025)
Dreamitate: Real-World Visuomotor Policy Learning via Video Generation
von: Liang, Junbang, et al.
Veröffentlicht: (2024)
von: Liang, Junbang, et al.
Veröffentlicht: (2024)
Teaching Humans Subtle Differences with DIFFusion
von: Chiquier, Mia, et al.
Veröffentlicht: (2025)
von: Chiquier, Mia, et al.
Veröffentlicht: (2025)
Sin3DM: Learning a Diffusion Model from a Single 3D Textured Shape
von: Wu, Rundi, et al.
Veröffentlicht: (2023)
von: Wu, Rundi, et al.
Veröffentlicht: (2023)
How to Design and Train Your Implicit Neural Representation for Video Compression
von: Gwilliam, Matthew, et al.
Veröffentlicht: (2025)
von: Gwilliam, Matthew, et al.
Veröffentlicht: (2025)
AllClear: A Comprehensive Dataset and Benchmark for Cloud Removal in Satellite Imagery
von: Zhou, Hangyu, et al.
Veröffentlicht: (2024)
von: Zhou, Hangyu, et al.
Veröffentlicht: (2024)
Let Your Video Listen to Your Music!
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
Your Interest, Your Summaries: Query-Focused Long Video Summarization
von: Patel, Nirav, et al.
Veröffentlicht: (2024)
von: Patel, Nirav, et al.
Veröffentlicht: (2024)
MINERVA: Evaluating Complex Video Reasoning
von: Nagrani, Arsha, et al.
Veröffentlicht: (2025)
von: Nagrani, Arsha, et al.
Veröffentlicht: (2025)
VideoMAC: Video Masked Autoencoders Meet ConvNets
von: Pei, Gensheng, et al.
Veröffentlicht: (2024)
von: Pei, Gensheng, et al.
Veröffentlicht: (2024)
TIIF-Bench: How Does Your T2I Model Follow Your Instructions?
von: Wei, Xinyu, et al.
Veröffentlicht: (2025)
von: Wei, Xinyu, et al.
Veröffentlicht: (2025)
pix2gestalt: Amodal Segmentation by Synthesizing Wholes
von: Ozguroglu, Ege, et al.
Veröffentlicht: (2024)
von: Ozguroglu, Ege, et al.
Veröffentlicht: (2024)
OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?
von: Li, Yifei, et al.
Veröffentlicht: (2025)
von: Li, Yifei, et al.
Veröffentlicht: (2025)
Do multimodal models imagine electric sheep?
von: Ramakrishnan, Santhosh Kumar, et al.
Veröffentlicht: (2026)
von: Ramakrishnan, Santhosh Kumar, et al.
Veröffentlicht: (2026)
Open-Set Video-based Facial Expression Recognition with Human Expression-sensitive Prompting
von: Liu, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Liu, Yuanyuan, et al.
Veröffentlicht: (2024)
Waver: Wave Your Way to Lifelike Video Generation
von: Zhang, Yifu, et al.
Veröffentlicht: (2025)
von: Zhang, Yifu, et al.
Veröffentlicht: (2025)
Generate Your Talking Avatar from Video Reference
von: Guo, Zujin, et al.
Veröffentlicht: (2026)
von: Guo, Zujin, et al.
Veröffentlicht: (2026)
Yo'LLaVA: Your Personalized Language and Vision Assistant
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
VideoMolmo: Spatio-Temporal Grounding Meets Pointing
von: Ahmad, Ghazi Shazan, et al.
Veröffentlicht: (2025)
von: Ahmad, Ghazi Shazan, et al.
Veröffentlicht: (2025)
ViFeEdit: A Video-Free Tuner of Your Video Diffusion Transformer
von: Yu, Ruonan, et al.
Veröffentlicht: (2026)
von: Yu, Ruonan, et al.
Veröffentlicht: (2026)
End-to-End Facial Expression Detection in Long Videos
von: Fang, Yini, et al.
Veröffentlicht: (2025)
von: Fang, Yini, et al.
Veröffentlicht: (2025)
SwInception -- Local Attention Meets Convolutions
von: Hagerman, David, et al.
Veröffentlicht: (2026)
von: Hagerman, David, et al.
Veröffentlicht: (2026)
Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery
von: Menn, Dennis, et al.
Veröffentlicht: (2026)
von: Menn, Dennis, et al.
Veröffentlicht: (2026)
Animate Your Motion: Turning Still Images into Dynamic Videos
von: Li, Mingxiao, et al.
Veröffentlicht: (2024)
von: Li, Mingxiao, et al.
Veröffentlicht: (2024)
StyleMaster: Stylize Your Video with Artistic Generation and Translation
von: Ye, Zixuan, et al.
Veröffentlicht: (2024)
von: Ye, Zixuan, et al.
Veröffentlicht: (2024)
Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly
von: Chetan, Aditya, et al.
Veröffentlicht: (2026)
von: Chetan, Aditya, et al.
Veröffentlicht: (2026)
Bring Your Dreams to Life: Continual Text-to-Video Customization
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
ALIVE: Animate Your World with Lifelike Audio-Video Generation
von: Guo, Ying, et al.
Veröffentlicht: (2026)
von: Guo, Ying, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Evolving Interpretable Visual Classifiers with Large Language Models
von: Chiquier, Mia, et al.
Veröffentlicht: (2024) -
DiSciPLE: Learning Interpretable Programs for Scientific Visual Discovery
von: Mall, Utkarsh, et al.
Veröffentlicht: (2025) -
4D Gaussian Splatting as a Learned Dynamical System
von: Asiimwe, Arnold Caleb, et al.
Veröffentlicht: (2025) -
Scale-Aware Recognition in Satellite Images under Resource Constraints
von: Revankar, Shreelekha, et al.
Veröffentlicht: (2024) -
MONITRS: Multimodal Observations of Natural Incidents Through Remote Sensing
von: Revankar, Shreelekha, et al.
Veröffentlicht: (2025)