PairHuman: A High-Fidelity Photographic Dataset for Customized Dual-Person Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Pan, Ting, Wang, Ye, Jing, Peiguang, Ma, Rui, Yi, Zili, Liu, Yu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Audio Generators Become Good Listeners: Generative Features for Understanding Tasks
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
SemanticVocoder: Bridging Audio Generation and Audio Understanding via Semantic Latents
by: Xie, Zeyu, et al.
Published: (2026)
by: Xie, Zeyu, et al.
Published: (2026)
Toward Storage-Aware Learning with Compressed Data An Empirical Exploratory Study on JPEG
by: Lee, Kichang, et al.
Published: (2025)
by: Lee, Kichang, et al.
Published: (2025)
Verifiable Dropout: Turning Randomness into a Verifiable Claim
by: Lee, Kichang, et al.
Published: (2025)
by: Lee, Kichang, et al.
Published: (2025)
VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall
by: Ruiz, Guillermo, et al.
Published: (2025)
by: Ruiz, Guillermo, et al.
Published: (2025)
Distributional Drift Adaptation with Temporal Conditional Variational Autoencoder for Multivariate Time Series Forecasting
by: He, Hui, et al.
Published: (2022)
by: He, Hui, et al.
Published: (2022)
AudioTime: A Temporally-aligned Audio-text Benchmark Dataset
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
FakeSound: Deepfake General Audio Detection
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
STAR: Speech-to-Audio Generation via Representation Learning
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
PicoAudio2: Temporal Controllable Text-to-Audio Generation with Natural Language Description
by: Zheng, Zihao, et al.
Published: (2025)
by: Zheng, Zihao, et al.
Published: (2025)
PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
FakeSound2: A Benchmark for Explainable and Generalizable Deepfake Sound Detection
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
Contemporary Agent Technology: LLM-Driven Advancements vs Classic Multi-Agent Systems
by: Bădică, Costin, et al.
Published: (2025)
by: Bădică, Costin, et al.
Published: (2025)
CAST-TTS: A Simple Cross-Attention Framework for Unified Timbre Control in TTS
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
Robust Multivariate Time Series Forecasting against Intra- and Inter-Series Transitional Shift
by: He, Hui, et al.
Published: (2024)
by: He, Hui, et al.
Published: (2024)
RoNFA: Robust Neural Field-based Approach for Few-Shot Image Classification with Noisy Labels
by: Xiang, Nan, et al.
Published: (2025)
by: Xiang, Nan, et al.
Published: (2025)
AI-Driven Innovations in Modern Cloud Computing
by: Kumar, Animesh
Published: (2024)
by: Kumar, Animesh
Published: (2024)
Enabling Trustworthy Federated Learning in Industrial IoT: Bridging the Gap Between Interpretability and Robustness
by: Jagatheesaperumal, Senthil Kumar, et al.
Published: (2024)
by: Jagatheesaperumal, Senthil Kumar, et al.
Published: (2024)
Redefining Finance: The Influence of Artificial Intelligence (AI) and Machine Learning (ML)
by: Kumar, Animesh
Published: (2024)
by: Kumar, Animesh
Published: (2024)
Superior Scoring Rules for Probabilistic Evaluation of Single-Label Multi-Class Classification Tasks
by: Ahmadian, Rouhollah, et al.
Published: (2024)
by: Ahmadian, Rouhollah, et al.
Published: (2024)
VALSE: A Task-Independent Benchmark for Vision and Language Models Centered on Linguistic Phenomena
by: Parcalabescu, Letitia, et al.
Published: (2021)
by: Parcalabescu, Letitia, et al.
Published: (2021)
On Measuring Faithfulness or Self-consistency of Natural Language Explanations
by: Parcalabescu, Letitia, et al.
Published: (2023)
by: Parcalabescu, Letitia, et al.
Published: (2023)
MM-SHAP: A Performance-agnostic Metric for Measuring Multimodal Contributions in Vision and Language Models & Tasks
by: Parcalabescu, Letitia, et al.
Published: (2022)
by: Parcalabescu, Letitia, et al.
Published: (2022)
Conversion rate prediction in online advertising: modeling techniques, performance evaluation and future directions
by: Xue, Tao, et al.
Published: (2025)
by: Xue, Tao, et al.
Published: (2025)
Toward a benchmark for CTR prediction in online advertising: datasets, evaluation protocols and perspectives
by: Gao, Shan, et al.
Published: (2025)
by: Gao, Shan, et al.
Published: (2025)
Temperature Scaling Attack Disrupting Model Confidence in Federated Learning
by: Lee, Kichang, et al.
Published: (2026)
by: Lee, Kichang, et al.
Published: (2026)
Dance of the ADS: Orchestrating Failures through Historically-Informed Scenario Fuzzing
by: Wang, Tong, et al.
Published: (2024)
by: Wang, Tong, et al.
Published: (2024)
Multimodal sensor fusion in the latent representation space
by: Piechocki, Robert J., et al.
Published: (2022)
by: Piechocki, Robert J., et al.
Published: (2022)
Emergence of heavy tails in homogenized stochastic gradient descent
by: Jiao, Zhe, et al.
Published: (2024)
by: Jiao, Zhe, et al.
Published: (2024)
Model-Based Soft Maximization of Suitable Metrics of Long-Term Human Power
by: Heitzig, Jobst, et al.
Published: (2025)
by: Heitzig, Jobst, et al.
Published: (2025)
Dynamic Cooperative Strategies in Search Engine Advertising Market: With and Without Retail Competition
by: Li, Huiran, et al.
Published: (2025)
by: Li, Huiran, et al.
Published: (2025)
Koopman operator learning using invertible neural networks
by: Meng, Yuhuang, et al.
Published: (2023)
by: Meng, Yuhuang, et al.
Published: (2023)
Expressivity of Representation Learning on Continuous-Time Dynamic Graphs: An Information-Flow Centric Review
by: Ennadir, Sofiane, et al.
Published: (2024)
by: Ennadir, Sofiane, et al.
Published: (2024)
Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Explanations?
by: Parcalabescu, Letitia, et al.
Published: (2024)
by: Parcalabescu, Letitia, et al.
Published: (2024)
Automated Quality Control System for Canned Tuna Production using Artificial Vision
by: Vera, Sendey, et al.
Published: (2024)
by: Vera, Sendey, et al.
Published: (2024)
From Benchmarks to Business Impact: Deploying IBM Generalist Agent in Enterprise Production
by: Shlomov, Segev, et al.
Published: (2025)
by: Shlomov, Segev, et al.
Published: (2025)
Any four real numbers are on all fours with analogy
by: Lepage, Yves, et al.
Published: (2024)
by: Lepage, Yves, et al.
Published: (2024)
Adaptive Dataset Quantization: A New Direction for Dataset Pruning
by: Yu, Chenyue, et al.
Published: (2025)
by: Yu, Chenyue, et al.
Published: (2025)
Recurrent Transformer-Based Near- and Far-Field THz Wideband Channel Estimation for UM-MIMO
by: Artemasov, Dmitry, et al.
Published: (2026)
by: Artemasov, Dmitry, et al.
Published: (2026)
Efficient Sliced Wasserstein Distance Computation via Adaptive Bayesian Optimization
by: Acharya, Manish, et al.
Published: (2025)
by: Acharya, Manish, et al.
Published: (2025)
Similar Items
-
When Audio Generators Become Good Listeners: Generative Features for Understanding Tasks
by: Xie, Zeyu, et al.
Published: (2025) -
SemanticVocoder: Bridging Audio Generation and Audio Understanding via Semantic Latents
by: Xie, Zeyu, et al.
Published: (2026) -
Toward Storage-Aware Learning with Compressed Data An Empirical Exploratory Study on JPEG
by: Lee, Kichang, et al.
Published: (2025) -
Verifiable Dropout: Turning Randomness into a Verifiable Claim
by: Lee, Kichang, et al.
Published: (2025) -
VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall
by: Ruiz, Guillermo, et al.
Published: (2025)