Several questions of visual generation in 2024
Fuente:
arXiv
Saved in:
| Main Author: | Gu, Shuyang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can ChatGPT assist visually impaired people with micro-navigation?
by: He, Junxian, et al.
Published: (2024)
by: He, Junxian, et al.
Published: (2024)
The 3D-PC: a benchmark for visual perspective taking in humans and machines
by: Linsley, Drew, et al.
Published: (2024)
by: Linsley, Drew, et al.
Published: (2024)
Influence of field of view in visual prostheses design: Analysis with a VR system
by: Sanchez-Garcia, Melani, et al.
Published: (2025)
by: Sanchez-Garcia, Melani, et al.
Published: (2025)
CoCoG-2: Controllable generation of visual stimuli for understanding human concept representation
by: Wei, Chen, et al.
Published: (2024)
by: Wei, Chen, et al.
Published: (2024)
More than One Step at a Time: Designing Procedural Feedback for Non-visual Makeup Routines
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
SPGen: Stochastic scanpath generation for paintings using unsupervised domain adaptation
by: Kerkouri, Mohamed Amine, et al.
Published: (2026)
by: Kerkouri, Mohamed Amine, et al.
Published: (2026)
Category-aware EEG image generation based on wavelet transform and contrast semantic loss
by: Zhang, Enshang, et al.
Published: (2025)
by: Zhang, Enshang, et al.
Published: (2025)
E-ANT: A Large-Scale Dataset for Efficient Automatic GUI NavigaTion
by: Wang, Ke, et al.
Published: (2024)
by: Wang, Ke, et al.
Published: (2024)
Unsupervised visualization of image datasets using contrastive learning
by: Böhm, Jan Niklas, et al.
Published: (2022)
by: Böhm, Jan Niklas, et al.
Published: (2022)
Multi-scale structural complexity as a quantitative measure of visual complexity
by: Kravchenko, Anna, et al.
Published: (2024)
by: Kravchenko, Anna, et al.
Published: (2024)
IVISIT: An Interactive Visual Simulation Tool for system simulation, visualization, optimization, and parameter management
by: Knoblauch, Andreas
Published: (2024)
by: Knoblauch, Andreas
Published: (2024)
Generalizability analysis of deep learning predictions of human brain responses to augmented and semantically novel visual stimuli
by: Piskovskyi, Valentyn, et al.
Published: (2024)
by: Piskovskyi, Valentyn, et al.
Published: (2024)
ASAP: Interpretable Analysis and Summarization of AI-generated Image Patterns at Scale
by: Huang, Jinbin, et al.
Published: (2024)
by: Huang, Jinbin, et al.
Published: (2024)
Evaluating how interactive visualizations can assist in finding samples where and how computer vision models make mistakes
by: Song, Hayeong, et al.
Published: (2023)
by: Song, Hayeong, et al.
Published: (2023)
How good are humans at detecting AI-generated images? Learnings from an experiment
by: Roca, Thomas, et al.
Published: (2025)
by: Roca, Thomas, et al.
Published: (2025)
MILE: A Mechanically Isomorphic Exoskeleton Data Collection System with Fingertip Visuotactile Sensing for Dexterous Manipulation
by: Du, Jinda, et al.
Published: (2025)
by: Du, Jinda, et al.
Published: (2025)
ExeChecker: Where Did I Go Wrong?
by: Gu, Yiwen, et al.
Published: (2024)
by: Gu, Yiwen, et al.
Published: (2024)
A Dataset for Crucial Object Recognition in Blind and Low-Vision Individuals' Navigation
by: Islam, Md Touhidul, et al.
Published: (2024)
by: Islam, Md Touhidul, et al.
Published: (2024)
Resource-Efficient Gesture Recognition using Low-Resolution Thermal Camera via Spiking Neural Networks and Sparse Segmentation
by: Safa, Ali, et al.
Published: (2024)
by: Safa, Ali, et al.
Published: (2024)
Towards Context-aware Support for Color Vision Deficiency: An Approach Integrating LLM and AR
by: Morita, Shogo, et al.
Published: (2024)
by: Morita, Shogo, et al.
Published: (2024)
Stratified Avatar Generation from Sparse Observations
by: Feng, Han, et al.
Published: (2024)
by: Feng, Han, et al.
Published: (2024)
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
by: Jiao, Chuhan, et al.
Published: (2024)
by: Jiao, Chuhan, et al.
Published: (2024)
Generalized Pose Space Embeddings for Training In-the-Wild using Anaylis-by-Synthesis
by: Borer, Dominik, et al.
Published: (2024)
by: Borer, Dominik, et al.
Published: (2024)
Don't Look at the Camera: Achieving Perceived Eye Contact
by: Gao, Alice, et al.
Published: (2024)
by: Gao, Alice, et al.
Published: (2024)
Real-Time Hand Gesture Recognition: Integrating Skeleton-Based Data Fusion and Multi-Stream CNN
by: Yusuf, Oluwaleke, et al.
Published: (2024)
by: Yusuf, Oluwaleke, et al.
Published: (2024)
Enhanced Automated Quality Assessment Network for Interactive Building Segmentation in High-Resolution Remote Sensing Imagery
by: Zhang, Zhili, et al.
Published: (2024)
by: Zhang, Zhili, et al.
Published: (2024)
AccessLens: Auto-detecting Inaccessibility of Everyday Objects
by: Kwon, Nahyun, et al.
Published: (2024)
by: Kwon, Nahyun, et al.
Published: (2024)
Leveraging Digital Perceptual Technologies for Remote Perception and Analysis of Human Biomechanical Processes: A Contactless Approach for Workload and Joint Force Assessment
by: Omidokun, Jesudara, et al.
Published: (2024)
by: Omidokun, Jesudara, et al.
Published: (2024)
Enhancing Saliency Prediction in Monitoring Tasks: The Role of Visual Highlights
by: Wu, Zekun, et al.
Published: (2024)
by: Wu, Zekun, et al.
Published: (2024)
NarrativeBridge: Enhancing Video Captioning with Causal-Temporal Narrative
by: Nadeem, Asmar, et al.
Published: (2024)
by: Nadeem, Asmar, et al.
Published: (2024)
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
by: Xuan, Xiwei, et al.
Published: (2024)
by: Xuan, Xiwei, et al.
Published: (2024)
Navigate Biopsy with Ultrasound under Augmented Reality Device: Towards Higher System Performance
by: Li, Haowei, et al.
Published: (2024)
by: Li, Haowei, et al.
Published: (2024)
Finger in Camera Speaks Everything: Unconstrained Air-Writing for Real-World
by: Wu, Meiqi, et al.
Published: (2024)
by: Wu, Meiqi, et al.
Published: (2024)
Detecting Activities of Daily Living in Egocentric Video to Contextualize Hand Use at Home in Outpatient Neurorehabilitation Settings
by: Kadambi, Adesh, et al.
Published: (2024)
by: Kadambi, Adesh, et al.
Published: (2024)
FathomVerse: A community science dataset for ocean animal discovery
by: Patterson, Genevieve, et al.
Published: (2024)
by: Patterson, Genevieve, et al.
Published: (2024)
VFA: Vision Frequency Analysis of Foundation Models and Human
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
GAZEploit: Remote Keystroke Inference Attack by Gaze Estimation from Avatar Views in VR/MR Devices
by: Wang, Hanqiu, et al.
Published: (2024)
by: Wang, Hanqiu, et al.
Published: (2024)
FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs
by: Chen, Haodong, et al.
Published: (2024)
by: Chen, Haodong, et al.
Published: (2024)
Allowing humans to interactively guide machines where to look does not always improve human-AI team's classification accuracy
by: Nguyen, Giang, et al.
Published: (2024)
by: Nguyen, Giang, et al.
Published: (2024)
Do Object Detection Localization Errors Affect Human Performance and Trust?
by: de Witte, Sven, et al.
Published: (2024)
by: de Witte, Sven, et al.
Published: (2024)
Similar Items
-
Can ChatGPT assist visually impaired people with micro-navigation?
by: He, Junxian, et al.
Published: (2024) -
The 3D-PC: a benchmark for visual perspective taking in humans and machines
by: Linsley, Drew, et al.
Published: (2024) -
Influence of field of view in visual prostheses design: Analysis with a VR system
by: Sanchez-Garcia, Melani, et al.
Published: (2025) -
CoCoG-2: Controllable generation of visual stimuli for understanding human concept representation
by: Wei, Chen, et al.
Published: (2024) -
More than One Step at a Time: Designing Procedural Feedback for Non-visual Makeup Routines
by: Li, Franklin Mingzhe, et al.
Published: (2025)