An LLM-Enabled Frequency-Aware Flow Diffusion Model for Natural-Language-Guided Power System Scenario Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Zhenghao, Li, Yiyan, Xie, Fei, Wang, Lu, Wang, Bo, Wang, Jiansheng, Yan, Zheng, Chow, Mo-Yuen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unsupervised and Interpretable Synthesizing for Electrical Time Series Based on Information Maximizing Generative Adversarial Nets
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2024)
A Glass-Box Deep-Learning Method for Electrical Energy System Modeling Based on Kolmogorov-Arnold Network
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2024)
A Causal-Guided Multimodal Large Language Model for Generalized Power System Time-Series Data Analytics
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2025)
A Neural-Network-Embedded Equivalent Circuit Model for Lithium-ion Battery State Estimation
von: Guo, Zelin, et al.
Veröffentlicht: (2024)
von: Guo, Zelin, et al.
Veröffentlicht: (2024)
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
DECAF: Dynamic Envelope Context-Aware Fusion for Speech-Envelope Reconstruction from EEG
von: Thakkar, Karan, et al.
Veröffentlicht: (2026)
von: Thakkar, Karan, et al.
Veröffentlicht: (2026)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
BR-ASR: Efficient and Scalable Bias Retrieval Framework for Contextual Biasing ASR in Speech LLM
von: Gong, Xun, et al.
Veröffentlicht: (2025)
von: Gong, Xun, et al.
Veröffentlicht: (2025)
Modulation Feature Enhancement with a Multi-Stage Attention Network for Underwater Acoustic Target Recognition
von: Yu, Jiaping, et al.
Veröffentlicht: (2026)
von: Yu, Jiaping, et al.
Veröffentlicht: (2026)
ASVspoof 5: Evaluation of Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
von: Wang, Xin, et al.
Veröffentlicht: (2026)
von: Wang, Xin, et al.
Veröffentlicht: (2026)
Time Series Diffusion Method: A Denoising Diffusion Probabilistic Model for Vibration Signal Generation
von: Yi, Haiming, et al.
Veröffentlicht: (2023)
von: Yi, Haiming, et al.
Veröffentlicht: (2023)
Real-Time Streamable Generative Speech Restoration with Flow Matching
von: Welker, Simon, et al.
Veröffentlicht: (2025)
von: Welker, Simon, et al.
Veröffentlicht: (2025)
An Investigation of Time-Frequency Representation Discriminators for High-Fidelity Vocoder
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
GLA-Grad++: An Improved Griffin-Lim Guided Diffusion Model for Speech Synthesis
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
RIFT: Entropy-Optimised Fractional Wavelet Constellations for Ideal Time-Frequency Estimation
von: Cozens, James M., et al.
Veröffentlicht: (2025)
von: Cozens, James M., et al.
Veröffentlicht: (2025)
Wavelet-Based Time-Frequency Fingerprinting for Feature Extraction of Traditional Irish Music
von: Shore, Noah
Veröffentlicht: (2025)
von: Shore, Noah
Veröffentlicht: (2025)
Comparison of Frequency-Fusion Mechanisms for Binaural Direction-of-Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
Enhancing Anti-spoofing Countermeasures Robustness through Joint Optimization and Transfer Learning
von: Wang, Yikang, et al.
Veröffentlicht: (2024)
von: Wang, Yikang, et al.
Veröffentlicht: (2024)
Deep-Learning-based Frequency-Domain Watermarking for Energy System Time Series Data Asset Protection
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2025)
Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
A Robust Method for Pitch Tracking in the Frequency Following Response using Harmonic Amplitude Summation Filterbank
von: Sadeghkhani, Sajad, et al.
Veröffentlicht: (2025)
von: Sadeghkhani, Sajad, et al.
Veröffentlicht: (2025)
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
Physics-Informed Direction-Aware Neural Acoustic Fields
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
DiffAU: Diffusion-Based Ambisonics Upscaling
von: Milstein, Amit, et al.
Veröffentlicht: (2025)
von: Milstein, Amit, et al.
Veröffentlicht: (2025)
State Space and Self-Attention Collaborative Network with Feature Aggregation for DOA Estimation
von: You, Qi, et al.
Veröffentlicht: (2025)
von: You, Qi, et al.
Veröffentlicht: (2025)
Investigation of Feature Selection and Pooling Methods for Environmental Sound Classification
von: Dehaghani, Parinaz Binandeh, et al.
Veröffentlicht: (2025)
von: Dehaghani, Parinaz Binandeh, et al.
Veröffentlicht: (2025)
Unsupervised EEG-based decoding of absolute auditory attention with canonical correlation analysis
von: Heintz, Nicolas, et al.
Veröffentlicht: (2025)
von: Heintz, Nicolas, et al.
Veröffentlicht: (2025)
AnyAccomp: Generalizable Accompaniment Generation via Quantized Melodic Bottleneck
von: Zhang, Junan, et al.
Veröffentlicht: (2025)
von: Zhang, Junan, et al.
Veröffentlicht: (2025)
Noisereduce: Domain General Noise Reduction for Time Series Signals
von: Sainburg, Tim, et al.
Veröffentlicht: (2024)
von: Sainburg, Tim, et al.
Veröffentlicht: (2024)
A Data-Centric Approach to Generalizable Speech Deepfake Detection
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
Efficient Solutions for Mitigating Initialization Bias in Unsupervised Self-Adaptive Auditory Attention Decoding
von: Yao, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Yao, Yuanyuan, et al.
Veröffentlicht: (2025)
Enhancing time-frequency resolution with optimal transport and barycentric fusion of multiple spectrogram
von: Valdivia, David, et al.
Veröffentlicht: (2026)
von: Valdivia, David, et al.
Veröffentlicht: (2026)
QINCODEC: Neural Audio Compression with Implicit Neural Codebooks
von: Lahrichi, Zineb, et al.
Veröffentlicht: (2025)
von: Lahrichi, Zineb, et al.
Veröffentlicht: (2025)
SONIC: Sound Optimization for Noise In Crowds
von: N, Pranav M, et al.
Veröffentlicht: (2025)
von: N, Pranav M, et al.
Veröffentlicht: (2025)
An Attention-Assisted Multi-Modal Data Fusion Model for Real-Time Estimation of Underwater Sound Velocity
von: Wu, Pengfei, et al.
Veröffentlicht: (2025)
von: Wu, Pengfei, et al.
Veröffentlicht: (2025)
Audio-Visual Speech Enhancement: Architectural Design and Deployment Strategies
von: Hamadouche, Anis, et al.
Veröffentlicht: (2025)
von: Hamadouche, Anis, et al.
Veröffentlicht: (2025)
A Multi-decoder Neural Tracking Method for Accurately Predicting Speech Intelligibility
von: Sonck, Rien, et al.
Veröffentlicht: (2026)
von: Sonck, Rien, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Unsupervised and Interpretable Synthesizing for Electrical Time Series Based on Information Maximizing Generative Adversarial Nets
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2024) -
A Glass-Box Deep-Learning Method for Electrical Energy System Modeling Based on Kolmogorov-Arnold Network
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2024) -
A Causal-Guided Multimodal Large Language Model for Generalized Power System Time-Series Data Analytics
von: Zhou, Zhenghao, et al.
Veröffentlicht: (2025) -
A Neural-Network-Embedded Equivalent Circuit Model for Lithium-ion Battery State Estimation
von: Guo, Zelin, et al.
Veröffentlicht: (2024) -
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)