activity
20242026
collaborators

5 papers

cs.AI2026

Safety Targeted Embedding Exploit via Refinement

Joshua Adrian Cahyono

Safety training for large language models (LLMs) is conducted predominantly in English, leaving uncertain how well safety mechanisms generalize to low-resource languages and mixed-…

cs.AI2025

Can You Trust an LLM with Your Life-Changing Decision? An Investigation into AI High-Stakes Responses

Joshua Adrian Cahyono, Saran Subramanian

Large Language Models (LLMs) are increasingly consulted for high-stakes life advice, yet they lack standard safeguards against providing confident but misguided responses. This cre…

cs.CL2025

LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Kaichen Zhang, Bo Li, Peiyuan Zhang +8

The advances of large foundation models necessitate wide-coverage, low-cost, and zero-contamination benchmarks. Despite continuous exploration of language model evaluations, compre…

cs.CV2024

Automated Image Captioning with CNNs and Transformers

Joshua Adrian Cahyono, Jeremy Nathan Jusuf

This project aims to create an automated image captioning system that generates natural language descriptions for input images by integrating techniques from computer vision and na…

cs.RO2024

Autonomous Ground Navigation in Highly Constrained Spaces: Lessons learned from The 3rd BARN Challenge at ICRA 2024

Xuesu Xiao, Zifan Xu, Aniket Datar +16

The 3rd BARN (Benchmark Autonomous Robot Navigation) Challenge took place at the 2024 IEEE International Conference on Robotics and Automation (ICRA 2024) in Yokohama, Japan and co…