2 papers
cs.SD2025
Testing chatbots on the creation of encoders for audio conditioned image generation
Jorge E. León, Miguel Carrasco
On one hand, recent advances in chatbots has led to a rising popularity in using these models for coding tasks. On the other hand, modern generative image models primarily rely on…
cs.MM2025
Effectively obtaining acoustic, visual and textual data from videos
Jorge E. León, Miguel Carrasco
The increasing use of machine learning models has amplified the demand for high-quality, large-scale multimodal datasets. However, the availability of such datasets, especially tho…