3 papers
cs.CL2026
VisEditBench: Can Vision-Language Models Edit Visualization Code from Multimodal Feedback?
Mizanur Rahman, Arshia Azimlu, Shadikur Rahman +4
Vision-language models (VLMs) have shown strong capabilities in generating visualization code from textual or visual specifications. However, real-world visualization authoring is…
cs.AI2026
DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real Computer Environments?
Mizanur Rahman, Mohammed Saidul Islam, Ridwan Mahbub +3
Real-world data science involves long-horizon workflows that span data wrangling, exploration, modeling, visualization, and validation, and require coordinated use of tools such as…
cs.CL2026
Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages
Firoj Alam, Shammur Absar Chowdhury, Enamul Hoque Prince
Multimodal LLMs are evolving from vision-language to tri-modality that see, hear, and read, yet pipelines and benchmarks remain English-centric and compute-heavy. The tutorial offe…