Ziyang Chen
czyang.bsky.social
Ziyang Chen
@czyang.bsky.social
Ph.D. Student @ UMich EECS. Multimodal learning, audio-visual learning and computer vision.
Prev research Intern @Adobe and @Meta

https://ificl.github.io/
🎥 Introducing MultiFoley, a video-aware audio generation method with multimodal controls! 🔊
We can
⌨️Make a typewriter sound like a piano 🎹
🐱Make a cat meow like a lion roars! 🦁
⏱️Perfectly time existing SFX 💥 to a video.

arXiv: arxiv.org/abs/2411.17698
website: ificl.github.io/MultiFoley/
November 27, 2024 at 2:58 AM