1 paper
Jae-Won Chung, Jeff J. Ma, Jisang Ahn +4
Any-to-Any models are an emerging class of multimodal models that accept combinations of multimodal data (e.g., text, image, video, audio) as input and generate them as output. Ser…