OregairuChar: A Benchmark Dataset for Character Appearance Frequency Analysis in My Teen Romantic Comedy SNAFU
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Qi, Zhou, Dingju, Zhang, Lina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
Learning to Align: Addressing Character Frequency Distribution Shifts in Handwritten Text Recognition
von: Kaliosis, Panagiotis, et al.
Veröffentlicht: (2025)
von: Kaliosis, Panagiotis, et al.
Veröffentlicht: (2025)
EGGS: Exchangeable 2D/3D Gaussian Splatting for Geometry-Appearance Balanced Novel View Synthesis
von: Zhang, Yancheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yancheng, et al.
Veröffentlicht: (2025)
Nepali Sign Language Characters Recognition: Dataset Development and Deep Learning Approaches
von: Poudel, Birat, et al.
Veröffentlicht: (2025)
von: Poudel, Birat, et al.
Veröffentlicht: (2025)
PsOCR: Benchmarking Large Multimodal Models for Optical Character Recognition in Low-resource Pashto Language
von: Haq, Ijazul, et al.
Veröffentlicht: (2025)
von: Haq, Ijazul, et al.
Veröffentlicht: (2025)
Evaluating Dataset Watermarking for Fine-tuning Traceability of Customized Diffusion Models: A Comprehensive Benchmark and Removal Approach
von: Wang, Xincheng, et al.
Veröffentlicht: (2025)
von: Wang, Xincheng, et al.
Veröffentlicht: (2025)
A Computer Vision Pipeline for Individual-Level Behavior Analysis: Benchmarking on the Edinburgh Pig Dataset
von: Yang, Haiyu, et al.
Veröffentlicht: (2025)
von: Yang, Haiyu, et al.
Veröffentlicht: (2025)
Character-Adapter: Prompt-Guided Region Control for High-Fidelity Character Customization
von: Ma, Yuhang, et al.
Veröffentlicht: (2024)
von: Ma, Yuhang, et al.
Veröffentlicht: (2024)
January Food Benchmark (JFB): A Public Benchmark Dataset and Evaluation Suite for Multimodal Food Analysis
von: Hosseinian, Amir, et al.
Veröffentlicht: (2025)
von: Hosseinian, Amir, et al.
Veröffentlicht: (2025)
ClimateIQA: A New Dataset and Benchmark to Advance Vision-Language Models in Meteorology Anomalies Analysis
von: Chen, Jian, et al.
Veröffentlicht: (2024)
von: Chen, Jian, et al.
Veröffentlicht: (2024)
PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos
von: Zhou, Zhiyu, et al.
Veröffentlicht: (2026)
von: Zhou, Zhiyu, et al.
Veröffentlicht: (2026)
My Body My Choice: Human-Centric Full-Body Anonymization
von: Ciftci, Umur Aybars, et al.
Veröffentlicht: (2024)
von: Ciftci, Umur Aybars, et al.
Veröffentlicht: (2024)
Beyond Visual Appearances: Privacy-sensitive Objects Identification via Hybrid Graph Reasoning
von: Jiang, Zhuohang, et al.
Veröffentlicht: (2024)
von: Jiang, Zhuohang, et al.
Veröffentlicht: (2024)
Zero-shot High-fidelity and Pose-controllable Character Animation
von: Zhu, Bingwen, et al.
Veröffentlicht: (2024)
von: Zhu, Bingwen, et al.
Veröffentlicht: (2024)
TowerDataset: A Heterogeneous Benchmark for Transmission Corridor Segmentation with a Global-Local Fusion Framework
von: Cui, Xu, et al.
Veröffentlicht: (2026)
von: Cui, Xu, et al.
Veröffentlicht: (2026)
2K-Characters-10K-Stories: A Quality-Gated Stylized Narrative Dataset with Disentangled Control and Sequence Consistency
von: Yin, Xingxi, et al.
Veröffentlicht: (2025)
von: Yin, Xingxi, et al.
Veröffentlicht: (2025)
Animate Any Character in Any World
von: Wang, Yitong, et al.
Veröffentlicht: (2025)
von: Wang, Yitong, et al.
Veröffentlicht: (2025)
OneActor: Consistent Character Generation via Cluster-Conditioned Guidance
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
Gastric-X: A Multimodal Multi-Phase Benchmark Dataset for Advancing Vision-Language Models in Gastric Cancer Analysis
von: Lu, Sheng, et al.
Veröffentlicht: (2026)
von: Lu, Sheng, et al.
Veröffentlicht: (2026)
UniLat3D: Geometry-Appearance Unified Latents for Single-Stage 3D Generation
von: Wu, Guanjun, et al.
Veröffentlicht: (2025)
von: Wu, Guanjun, et al.
Veröffentlicht: (2025)
STAR: A First-Ever Dataset and A Large-Scale Benchmark for Scene Graph Generation in Large-Size Satellite Imagery
von: Li, Yansheng, et al.
Veröffentlicht: (2024)
von: Li, Yansheng, et al.
Veröffentlicht: (2024)
AllClear: A Comprehensive Dataset and Benchmark for Cloud Removal in Satellite Imagery
von: Zhou, Hangyu, et al.
Veröffentlicht: (2024)
von: Zhou, Hangyu, et al.
Veröffentlicht: (2024)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
Eye-for-an-eye: Appearance Transfer with Semantic Correspondence in Diffusion Models
von: Go, Sooyeon, et al.
Veröffentlicht: (2024)
von: Go, Sooyeon, et al.
Veröffentlicht: (2024)
Multi-Modal Character Localization and Extraction for Chinese Text Recognition
von: Li, Qilong, et al.
Veröffentlicht: (2026)
von: Li, Qilong, et al.
Veröffentlicht: (2026)
Unified and Dynamic Graph for Temporal Character Grouping in Long Videos
von: Shu, Xiujun, et al.
Veröffentlicht: (2023)
von: Shu, Xiujun, et al.
Veröffentlicht: (2023)
LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation
von: Deng, Fan, et al.
Veröffentlicht: (2024)
von: Deng, Fan, et al.
Veröffentlicht: (2024)
MOT FCG++: Enhanced Representation of Spatio-temporal Motion and Appearance Features
von: Fang, Yanzhao
Veröffentlicht: (2024)
von: Fang, Yanzhao
Veröffentlicht: (2024)
LOCR: Location-Guided Transformer for Optical Character Recognition
von: Sun, Yu, et al.
Veröffentlicht: (2024)
von: Sun, Yu, et al.
Veröffentlicht: (2024)
SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis
von: Zhao, Shuxian, et al.
Veröffentlicht: (2026)
von: Zhao, Shuxian, et al.
Veröffentlicht: (2026)
PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
OODFace: Benchmarking Robustness of Face Recognition under Common Corruptions and Appearance Variations
von: Kang, Caixin, et al.
Veröffentlicht: (2024)
von: Kang, Caixin, et al.
Veröffentlicht: (2024)
Deep Learning in Dental Image Analysis: A Systematic Review of Datasets, Methodologies, and Emerging Challenges
von: Zhou, Zhenhuan, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenhuan, et al.
Veröffentlicht: (2025)
MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance
von: Zhao, Kaikai, et al.
Veröffentlicht: (2025)
von: Zhao, Kaikai, et al.
Veröffentlicht: (2025)
AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
Exploring Cross-Domain Few-Shot Classification via Frequency-Aware Prompting
von: Zhang, Tiange, et al.
Veröffentlicht: (2024)
von: Zhang, Tiange, et al.
Veröffentlicht: (2024)
Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset
von: Chen, Qian, et al.
Veröffentlicht: (2026)
von: Chen, Qian, et al.
Veröffentlicht: (2026)
PointCloud-Text Matching: Benchmark Datasets and a Baseline
von: Feng, Yanglin, et al.
Veröffentlicht: (2024)
von: Feng, Yanglin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration
von: Na, Kihyun, et al.
Veröffentlicht: (2025) -
Learning to Align: Addressing Character Frequency Distribution Shifts in Handwritten Text Recognition
von: Kaliosis, Panagiotis, et al.
Veröffentlicht: (2025) -
EGGS: Exchangeable 2D/3D Gaussian Splatting for Geometry-Appearance Balanced Novel View Synthesis
von: Zhang, Yancheng, et al.
Veröffentlicht: (2025) -
Nepali Sign Language Characters Recognition: Dataset Development and Deep Learning Approaches
von: Poudel, Birat, et al.
Veröffentlicht: (2025) -
PsOCR: Benchmarking Large Multimodal Models for Optical Character Recognition in Low-resource Pashto Language
von: Haq, Ijazul, et al.
Veröffentlicht: (2025)