APA (7th ed.) Citation

Sung-Bin, K., Senocak, A., Ha, H., & Oh, T. (2024). Sound2Vision: Generating Diverse Visuals from Audio through Cross-Modal Latent Alignment.

Chicago Style (17th ed.) Citation

Sung-Bin, Kim, Arda Senocak, Hyunwoo Ha, and Tae-Hyun Oh. Sound2Vision: Generating Diverse Visuals from Audio Through Cross-Modal Latent Alignment. 2024.

MLA (9th ed.) Citation

Sung-Bin, Kim, et al. Sound2Vision: Generating Diverse Visuals from Audio Through Cross-Modal Latent Alignment. 2024.

Warning: These citations may not always be 100% accurate.