Negative Shanshui: Real-time Interactive Ink Painting Synthesis
Fuente:
arXiv
Saved in:
| Main Author: | Zhou, Aven-Le |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RITA: A Real-time Interactive Talking Avatars Framework
by: Cheng, Wuxinlin, et al.
Published: (2024)
by: Cheng, Wuxinlin, et al.
Published: (2024)
StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars
by: Sun, Zhiyao, et al.
Published: (2025)
by: Sun, Zhiyao, et al.
Published: (2025)
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
Using Salient Object Detection to Identify Manipulative Cookie Banners that Circumvent GDPR
by: Grossman, Riley, et al.
Published: (2025)
by: Grossman, Riley, et al.
Published: (2025)
From Image Generation to Infrastructure Design: a Multi-agent Pipeline for Street Design Generation
by: Wang, Chenguang, et al.
Published: (2025)
by: Wang, Chenguang, et al.
Published: (2025)
Egocentric Co-Pilot: Web-Native Smart-Glasses Agents for Assistive Egocentric AI
by: Yang, Sicheng, et al.
Published: (2026)
by: Yang, Sicheng, et al.
Published: (2026)
Joining Forces for Pathology Diagnostics with AI Assistance: The EMPAIA Initiative
by: Zerbe, Norman, et al.
Published: (2023)
by: Zerbe, Norman, et al.
Published: (2023)
Gaze patterns predict preference and confidence in pairwise AI image evaluation
by: Papadopoulos, Nikolas, et al.
Published: (2026)
by: Papadopoulos, Nikolas, et al.
Published: (2026)
AI-Based Facial Emotion Recognition Solutions for Education: A Study of Teacher-User and Other Categories
by: Ravenor, R. Yamamoto
Published: (2023)
by: Ravenor, R. Yamamoto
Published: (2023)
Refusal as Silence: Gendered Disparities in Vision-Language Model Responses
by: Luo, Sha, et al.
Published: (2024)
by: Luo, Sha, et al.
Published: (2024)
Mask-up: Investigating Biases in Face Re-identification for Masked Faces
by: Jaiswal, Siddharth D, et al.
Published: (2024)
by: Jaiswal, Siddharth D, et al.
Published: (2024)
Handwritten Code Recognition for Pen-and-Paper CS Education
by: Islam, Md Sazzad, et al.
Published: (2024)
by: Islam, Md Sazzad, et al.
Published: (2024)
BLK-Assist: A Methodological Framework for Artist-Led Co-Creation with Generative AI Models
by: Grimes, Daniel, et al.
Published: (2026)
by: Grimes, Daniel, et al.
Published: (2026)
Transfer Learning-based Real-time Handgun Detection
by: Elmir, Youssef
Published: (2023)
by: Elmir, Youssef
Published: (2023)
Generating Robot Constitutions & Benchmarks for Semantic Safety
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
Mapping User Trust in Vision Language Models: Research Landscape, Challenges, and Prospects
by: Chiatti, Agnese, et al.
Published: (2025)
by: Chiatti, Agnese, et al.
Published: (2025)
Determining the Difficulties of Students With Dyslexia via Virtual Reality and Artificial Intelligence: An Exploratory Analysis
by: Yeguas-Bolívar, Enrique, et al.
Published: (2024)
by: Yeguas-Bolívar, Enrique, et al.
Published: (2024)
Yume: An Interactive World Generation Model
by: Mao, Xiaofeng, et al.
Published: (2025)
by: Mao, Xiaofeng, et al.
Published: (2025)
Vision-Language Models Suppress Female Representations Under Ambiguous Input
by: Marin-Llobet, Arnau, et al.
Published: (2026)
by: Marin-Llobet, Arnau, et al.
Published: (2026)
SSSUMO: Real-Time Semi-Supervised Submovement Decomposition
by: Rudakov, Evgenii, et al.
Published: (2025)
by: Rudakov, Evgenii, et al.
Published: (2025)
Explorer: Robust Collection of Interactable GUI Elements
by: Chaimalas, Iason, et al.
Published: (2025)
by: Chaimalas, Iason, et al.
Published: (2025)
Real-Time Feedback and Benchmark Dataset for Isometric Pose Evaluation
by: Jaiswal, Abhishek, et al.
Published: (2025)
by: Jaiswal, Abhishek, et al.
Published: (2025)
Let's Go Real Talk: Spoken Dialogue Model for Face-to-Face Conversation
by: Park, Se Jin, et al.
Published: (2024)
by: Park, Se Jin, et al.
Published: (2024)
OmniResponse: Online Multimodal Conversational Response Generation in Dyadic Interactions
by: Luo, Cheng, et al.
Published: (2025)
by: Luo, Cheng, et al.
Published: (2025)
"I Can See Forever!": Evaluating Real-time VideoLLMs for Assisting Individuals with Visual Impairments
by: Zhang, Ziyi, et al.
Published: (2025)
by: Zhang, Ziyi, et al.
Published: (2025)
SelfReDepth: Self-Supervised Real-Time Depth Restoration for Consumer-Grade Sensors
by: Duarte, Alexandre, et al.
Published: (2024)
by: Duarte, Alexandre, et al.
Published: (2024)
See-Control: A Multimodal Agent Framework for Smartphone Interaction with a Robotic Arm
by: Zhao, Haoyu, et al.
Published: (2025)
by: Zhao, Haoyu, et al.
Published: (2025)
Achieving Effective Virtual Reality Interactions via Acoustic Gesture Recognition based on Large Language Models
by: Zhang, Xijie, et al.
Published: (2025)
by: Zhang, Xijie, et al.
Published: (2025)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
by: Song, Jookyung, et al.
Published: (2025)
by: Song, Jookyung, et al.
Published: (2025)
PupilSense: A Novel Application for Webcam-Based Pupil Diameter Estimation
by: Shah, Vijul, et al.
Published: (2024)
by: Shah, Vijul, et al.
Published: (2024)
A Survey on Trustworthiness in Foundation Models for Medical Image Analysis
by: Shi, Congzhen, et al.
Published: (2024)
by: Shi, Congzhen, et al.
Published: (2024)
Visual Evaluative AI: A Hypothesis-Driven Tool with Concept-Based Explanations and Weight of Evidence
by: Le, Thao, et al.
Published: (2024)
by: Le, Thao, et al.
Published: (2024)
FluentLip: A Phonemes-Based Two-stage Approach for Audio-Driven Lip Synthesis with Optical Flow Consistency
by: Liu, Shiyan, et al.
Published: (2025)
by: Liu, Shiyan, et al.
Published: (2025)
T2I-Copilot: A Training-Free Multi-Agent Text-to-Image System for Enhanced Prompt Interpretation and Interactive Generation
by: Chen, Chieh-Yun, et al.
Published: (2025)
by: Chen, Chieh-Yun, et al.
Published: (2025)
SkinGEN: an Explainable Dermatology Diagnosis-to-Generation Framework with Interactive Vision-Language Models
by: Lin, Bo, et al.
Published: (2024)
by: Lin, Bo, et al.
Published: (2024)
DiMo-GUI: Advancing Test-time Scaling in GUI Grounding via Modality-Aware Visual Reasoning
by: Wu, Hang, et al.
Published: (2025)
by: Wu, Hang, et al.
Published: (2025)
Summary of the Unusual Activity Recognition Challenge for Developmental Disability Support
by: Garcia, Christina, et al.
Published: (2026)
by: Garcia, Christina, et al.
Published: (2026)
UI-TARS: Pioneering Automated GUI Interaction with Native Agents
by: Qin, Yujia, et al.
Published: (2025)
by: Qin, Yujia, et al.
Published: (2025)
PaintCopilot: Modeling Painting as Autonomous Artistic Continuation
by: Wen, Yunge, et al.
Published: (2026)
by: Wen, Yunge, et al.
Published: (2026)
Context-Awareness and Interpretability of Rare Occurrences for Discovery and Formalization of Critical Failure Modes
by: Polavaram, Sridevi, et al.
Published: (2025)
by: Polavaram, Sridevi, et al.
Published: (2025)
Similar Items
-
RITA: A Real-time Interactive Talking Avatars Framework
by: Cheng, Wuxinlin, et al.
Published: (2024) -
StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars
by: Sun, Zhiyao, et al.
Published: (2025) -
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
by: Becerra, Alvaro, et al.
Published: (2025) -
Using Salient Object Detection to Identify Manipulative Cookie Banners that Circumvent GDPR
by: Grossman, Riley, et al.
Published: (2025) -
From Image Generation to Infrastructure Design: a Multi-agent Pipeline for Street Design Generation
by: Wang, Chenguang, et al.
Published: (2025)