Saved in:
| Main Authors: | Chen, Yi-Hsin, Yao, Yi-Chen, Ho, Kuan-Wei, Wu, Chun-Hung, Phung, Huu-Tai, Benjak, Martin, Ostermann, Jörn, Peng, Wen-Hsiao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.02072 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Conditional Residual Coding with Explicit-Implicit Temporal Buffering for Learned Video Compression
by: Chen, Yi-Hsin, et al.
Published: (2025)
by: Chen, Yi-Hsin, et al.
Published: (2025)
On the Rate-Distortion-Complexity Trade-offs of Neural Video Coding
by: Chen, Yi-Hsin, et al.
Published: (2024)
by: Chen, Yi-Hsin, et al.
Published: (2024)
A Cross-Framework Study of Temporal Information Buffering Strategies for Learned Video Compression
by: Ho, Kuan-Wei, et al.
Published: (2025)
by: Ho, Kuan-Wei, et al.
Published: (2025)
MaskCRT: Masked Conditional Residual Transformer for Learned Video Compression
by: Chen, Yi-Hsin, et al.
Published: (2023)
by: Chen, Yi-Hsin, et al.
Published: (2023)
MH-LVC: Multi-Hypothesis Temporal Prediction for Learned Conditional Residual Video Coding
by: Phung, Huu-Tai, et al.
Published: (2025)
by: Phung, Huu-Tai, et al.
Published: (2025)
LANCE: Locally Adaptive Neural Context Estimation for Overfitted Image Compression
by: Benjak, Martin, et al.
Published: (2026)
by: Benjak, Martin, et al.
Published: (2026)
Exploring Autoregressive Vision Foundation Models for Image Compression
by: Phung, Huu-Tai, et al.
Published: (2025)
by: Phung, Huu-Tai, et al.
Published: (2025)
Investigating Zero-Shot Generalizability on Mandarin-English Code-Switched ASR and Speech-to-text Translation of Recent Foundation Models with Self-Supervision and Weak Supervision
by: Yang, Chih-Kai, et al.
Published: (2023)
by: Yang, Chih-Kai, et al.
Published: (2023)
A Linear Two‐Coordinate Cr(II) Complex: Synthesis, Characterization, and Reactivity
by: Kai‐Chin Hsiao, et al.
Published: (2024)
by: Kai‐Chin Hsiao, et al.
Published: (2024)
Can Large Audio-Language Models Truly Hear? Tackling Hallucinations with Multi-Task Assessment and Stepwise Audio Reasoning
by: Kuan, Chun-Yi, et al.
Published: (2024)
by: Kuan, Chun-Yi, et al.
Published: (2024)
AQUA-Bench: Beyond Finding Answers to Knowing When There Are None in Audio Question Answering
by: Kuan, Chun-Yi, et al.
Published: (2026)
by: Kuan, Chun-Yi, et al.
Published: (2026)
Gender Bias in Instruction-Guided Speech Synthesis Models
by: Kuan, Chun-Yi, et al.
Published: (2025)
by: Kuan, Chun-Yi, et al.
Published: (2025)
Teaching Audio-Aware Large Language Models What Does Not Hear: Mitigating Hallucinations through Synthesized Negative Samples
by: Kuan, Chun-Yi, et al.
Published: (2025)
by: Kuan, Chun-Yi, et al.
Published: (2025)
From Alignment to Advancement: Bootstrapping Audio-Language Alignment with Synthetic Data
by: Kuan, Chun-Yi, et al.
Published: (2025)
by: Kuan, Chun-Yi, et al.
Published: (2025)
The status of marine turtle conservation in Vietnam
by: Nguyen, Huu Phung
Published: (2000)
by: Nguyen, Huu Phung
Published: (2000)
Fish eggs and larvae from the sea of Vietnam
by: Nguyen, Huu Phung
Published: (1991)
by: Nguyen, Huu Phung
Published: (1991)
The classification of lizard-fishes larvae from Bac Bo gulf
by: Nguyen, Huu Phung
Published: (1980)
by: Nguyen, Huu Phung
Published: (1980)
Some newly revealed species of butterflyfishes (Chaetodontidae) from the sea of Vietnam
by: Nguyen, Huu Phung
Published: (1996)
by: Nguyen, Huu Phung
Published: (1996)
ConSep: a Noise- and Reverberation-Robust Speech Separation Framework by Magnitude Conditioning
by: Ho, Kuan-Hsun, et al.
Published: (2024)
by: Ho, Kuan-Hsun, et al.
Published: (2024)
A Sleep Monitoring System Based on Audio, Video and Depth Information
by: Chen, Lyn Chao-ling, et al.
Published: (2025)
by: Chen, Lyn Chao-ling, et al.
Published: (2025)
Nested Feature Spectrum Topology: Tripartite Topological Equivalence of Feature, Entanglement, and Wilson Loop Spectrum
by: Hung, Yi-Chun, et al.
Published: (2026)
by: Hung, Yi-Chun, et al.
Published: (2026)
Transformer-based Learned Image Compression for Joint Decoding and Denoising
by: Chen, Yi-Hsin, et al.
Published: (2024)
by: Chen, Yi-Hsin, et al.
Published: (2024)
Win‐Win Strategies Enable Efficient Anode‐Less Zinc‐Ion Hybrid Supercapacitors
by: Tai‐Feng Hung, et al.
Published: (2024)
by: Tai‐Feng Hung, et al.
Published: (2024)
Collaborative Hybrid Propagator for Temporal Misalignment in Audio-Visual Segmentation
by: Li, Kexin, et al.
Published: (2024)
by: Li, Kexin, et al.
Published: (2024)
TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation
by: Wang, Wenhao, et al.
Published: (2024)
by: Wang, Wenhao, et al.
Published: (2024)
Case report: Rapid diagnosis followed by rapid remission of neurosarcoidosis
by: En‐Ying Wang, et al.
Published: (2024)
by: En‐Ying Wang, et al.
Published: (2024)
All‐solid‐state Li‐ion battery: A study on the charge/discharge mechanism of an LMO‐BCD‐MgC system
by: Po‐Ting Wu, et al.
Published: (2024)
by: Po‐Ting Wu, et al.
Published: (2024)
Simple Self-Conditioning Adaptation for Masked Diffusion Models
by: Cardei, Michael, et al.
Published: (2026)
by: Cardei, Michael, et al.
Published: (2026)
HyCodePolicy: Hybrid Language Controllers for Multimodal Monitoring and Decision in Embodied Agents
by: Liu, Yibin, et al.
Published: (2025)
by: Liu, Yibin, et al.
Published: (2025)
Speech-IFEval: Evaluating Instruction-Following and Quantifying Catastrophic Forgetting in Speech-Aware Language Models
by: Lu, Ke-Han, et al.
Published: (2025)
by: Lu, Ke-Han, et al.
Published: (2025)
Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models
by: Kuan, Chun-Yi, et al.
Published: (2026)
by: Kuan, Chun-Yi, et al.
Published: (2026)
AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering
by: Kuan, Chun-Yi, et al.
Published: (2026)
by: Kuan, Chun-Yi, et al.
Published: (2026)
Understanding Sounds, Missing the Questions: The Challenge of Object Hallucination in Large Audio-Language Models
by: Kuan, Chun-Yi, et al.
Published: (2024)
by: Kuan, Chun-Yi, et al.
Published: (2024)
Time-Reversal Soliton Pairs In Even Spin-Chern-Number Higher-Order Topological Insulators
by: Hung, Yi-Chun, et al.
Published: (2023)
by: Hung, Yi-Chun, et al.
Published: (2023)
Area between trajectories: Insights into optimal group selection and trajectory heterogeneity in group-based trajectory modeling
by: Hsiao, Yi-Chen, et al.
Published: (2025)
by: Hsiao, Yi-Chen, et al.
Published: (2025)
Conditional Neural Video Coding with Spatial-Temporal Super-Resolution
by: Wang, Henan, et al.
Published: (2024)
by: Wang, Henan, et al.
Published: (2024)
Improved Belief Propagation Decoding Algorithms for Surface Codes
by: Chen, Jiahan, et al.
Published: (2024)
by: Chen, Jiahan, et al.
Published: (2024)
SegForestNet: Spatial-Partitioning-Based Aerial Image Segmentation
by: Gritzner, Daniel, et al.
Published: (2023)
by: Gritzner, Daniel, et al.
Published: (2023)
Pruning-aware Loss Functions for STOI-Optimized Pruned Recurrent Autoencoders for the Compression of the Stimulation Patterns of Cochlear Implants at Zero Delay
by: Hinrichs, Reemt, et al.
Published: (2025)
by: Hinrichs, Reemt, et al.
Published: (2025)
Some main resources of Bivalve (Bivalve-Mollusca) in marine waters of Vietnam
by: Nguyen, Huu Phung, et al.
Published: (1996)
by: Nguyen, Huu Phung, et al.
Published: (1996)
Similar Items
-
Conditional Residual Coding with Explicit-Implicit Temporal Buffering for Learned Video Compression
by: Chen, Yi-Hsin, et al.
Published: (2025) -
On the Rate-Distortion-Complexity Trade-offs of Neural Video Coding
by: Chen, Yi-Hsin, et al.
Published: (2024) -
A Cross-Framework Study of Temporal Information Buffering Strategies for Learned Video Compression
by: Ho, Kuan-Wei, et al.
Published: (2025) -
MaskCRT: Masked Conditional Residual Transformer for Learned Video Compression
by: Chen, Yi-Hsin, et al.
Published: (2023) -
MH-LVC: Multi-Hypothesis Temporal Prediction for Learned Conditional Residual Video Coding
by: Phung, Huu-Tai, et al.
Published: (2025)