3 papers
cs.CL2025
Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation
Canxiang Yan, Chunxiang Jin, Dawei Huang +22
Existing speech models suffer from competing requirements on token representations by understanding and generation tasks. This discrepancy in representation prevents speech languag…
hep-ex2025
Observation of resonant contribution to the around 4.2 GeV and evidence of
BESIII Collaboration, M. Ablikim, M. N. Achasov +635
Using collision data corresponding to a total integrated luminosity of 22.7 fb, collected at center-of-mass energies between 3.7 and 4.7 GeV with the BESIII detecto…
cs.AI2025
Ming-Omni: A Unified Multimodal Model for Perception and Generation
Inclusion AI, Biao Gong, Cheng Zou +55
We propose Ming-Omni, a unified multimodal model capable of processing images, text, audio, and video, while demonstrating strong proficiency in both speech and image generation. M…