papers

Publications (27)

cs.CV2026

Out of the box age estimation through facial imagery: A Comprehensive Benchmark of Vision-Language Models vs. out-of-the-box Traditional Architectures

Simiao Ren, Xingyu Shen, Ankit Raj +8

Facial age estimation plays a critical role in content moderation, age verification, and deepfake detection. However, no prior benchmark has systematically compared modern vision-l…

cs.CV2026

How well are open sourced AI-generated image detection models out-of-the-box: A comprehensive benchmark study

Simiao Ren, Yuchen Zhou, Xingyu Shen +9

As AI-generated images proliferate across digital platforms, reliable detection methods have become critical for combating misinformation and maintaining content authenticity. Whil…

cs.CV2025

Can Multi-modal (reasoning) LLMs detect document manipulation?

Zisheng Liang, Kidus Zewde, Rudra Pratap Singh +7

Document fraud poses a significant threat to industries reliant on secure and verifiable documentation, necessitating robust detection mechanisms. This study investigates the effic…

physics.optics2020

Neural-adjoint method for the inverse design of all-dielectric metasurfaces

Yang Deng, Simiao Ren, Kebin Fan +2

All-dielectric metasurfaces exhibit exotic electromagnetic responses, similar to those obtained with metal-based metamaterials. Research in all-dielectric metasurfaces currently us…

cs.LG2024

Does Deep Active Learning Work in the Wild?

Simiao Ren, Saad Lahrichi, Yang Deng +3

Deep active learning (DAL) methods have shown significant improvements in sample efficiency compared to simple random sampling. While these studies are valuable, they nearly always…

eess.SP2022

Automated Extraction of Energy Systems Information from Remotely Sensed Data: A Review and Analysis

Simiao Ren, Wei Hu, Kyle Bradbury +4

High quality energy systems information is a crucial input to energy systems research, modeling, and decision-making. Unfortunately, actionable information about energy systems is…

cs.CV2025

Can Multi-modal (reasoning) LLMs work as deepfake detectors?

Simiao Ren, Yao Yao, Kidus Zewde +8

Deepfake detection remains a critical challenge in the era of advanced generative models, particularly as synthetic media becomes more sophisticated. In this study, we explore the…

cond-mat.soft2020

Direct evidence of void induced structural relaxations in colloidal glass formers

Cho-Tung Yip, Masaharu Isobe, Chor-Hoi Chan +7

Particle dynamics in supercooled liquids are often dominated by string-like motions in which lines of particles perform activated hops cooperatively. The structural features trigge…

cs.CV2026

AIForge-Doc: A Benchmark for Detecting AI-Forged Tampering in Financial and Form Documents

Jiaqi Wu, Yuchen Zhou, Muduo Xu +6

We present AIForge-Doc, the first dedicated benchmark targeting exclusively diffusion-model-based inpainting in financial and form documents with pixel-level annotation. Existing d…

cs.LG2022

Mixture Manifold Networks: A Computationally Efficient Baseline for Inverse Modeling

Gregory P. Spell, Simiao Ren, Leslie M. Collins +1

We propose and show the efficacy of a new method to address generic inverse problems. Inverse modeling is the task whereby one seeks to determine the control parameters of a natura…

physics.optics2023

Machine Learning for Mie-Tronics

Wenhao Li, Hooman Barati Sedeh, Willie J. Padilla +3

Electromagnetic multipole expansion theory underpins nanoscale light-matter interactions, particularly within subwavelength meta-atoms, paving the way for diverse and captivating o…

eess.IV2022

Utilizing geospatial data for assessing energy security: Mapping small solar home systems using unmanned aerial vehicles and deep learning

Simiao Ren, Jordan Malof, T. Robert Fetter +3

Solar home systems (SHS), a cost-effective solution for rural communities far from the grid in developing countries, are small solar panels and associated equipment that provides p…

cs.LG2021

Benchmarking deep inverse models over time, and the neural-adjoint method

Simiao Ren, Willie Padilla, Jordan Malof

We consider the task of solving generic inverse problems, where one wishes to determine the hidden parameters of a natural system that will give rise to a particular set of measure…

cs.CV2026

When the Forger Is the Judge: GPT-Image-2 Cannot Recognize Its Own Faked Documents

Jiaqi Wu, Yuchen Zhou, Dennis Tsang Ng +5

OpenAI's GPT-Image-2 has effectively erased the visual boundary between authentic and AI-edited document images: a single number on a receipt can be replaced in under a second for…

cs.LG2021

Inverse deep learning methods and benchmarks for artificial electromagnetic material design

Simiao Ren, Ashwin Mahendra, Omar Khatib +3

Deep learning (DL) inverse techniques have increased the speed of artificial electromagnetic material (AEM) design and improved the quality of resulting devices. Many DL inverse te…

cs.CV2026

A Synthetic Eye Movement Dataset for Script Reading Detection: Real Trajectory Replay on a 3D Simulator

Kidus Zewde, Yuchen Zhou, Dennis Ng +6

Large vision-language models have achieved remarkable capabilities by training on massive internet-scale data, yet a fundamental asymmetry persists: while LLMs can leverage self-su…

cs.LG2022

Towards Robust Deep Active Learning for Scientific Computing

Simiao Ren, Yang Deng, Willie J. Padilla +1

Deep learning (DL) is revolutionizing the scientific computing community. To reduce the data gap, active learning has been identified as a promising solution for DL in the scientif…

cs.CV2022

Meta-simulation for the Automated Design of Synthetic Overhead Imagery

Handi Yu, Simiao Ren, Leslie M. Collins +1

The use of synthetic (or simulated) data for training machine learning models has grown rapidly in recent years. Synthetic data can often be generated much faster and more cheaply…

cs.CR2026

CallScreenBench: Benchmarking On-Device Models as Phone Secretaries

Simiao Ren

Language models small enough to run on a handset, quantized to a few bits, are increasingly capable of acting for their user, making on-device task automation newly plausible. One…

cs.CL2026

Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate

Simiao Ren, Xingyu Shen, Yuchen Zhou +3

A claim has been circulating on social media and practitioner forums that Chinese prompts are more token-efficient than English for LLM coding tasks, potentially reducing costs by…

cs.CV2026

DOCFORGE-BENCH: A Comprehensive 0-shot Benchmark for Document Forgery Detection and Analysis

Zengqi Zhao, Weidi Xia, En Wei +7

We present DOCFORGE-BENCH, the first unified zero-shot benchmark for document forgery detection, evaluating 14 methods across eight datasets spanning text tampering, receipt forger…

cs.CV2025

Do Deepfake Detectors Work in Reality?

Simiao Ren, Hengwei Xu, Tsang Ng +7

Deepfakes, particularly those involving faceswap-based manipulations, have sparked significant societal concern due to their increasing realism and potential for misuse. Despite ra…

cs.CV2026

GPT-Image-2 in the Wild: A Twitter Dataset of Self-Reported AI-Generated Images from the First Week of Deployment

Kidus Zewde, Simiao Ren, Xingyu Shen +6

The release of GPT-image-2 by OpenAI marks a watershed moment in AI-generated imagery: the boundary between photographic reality and synthetic content has never been more difficult…

cs.AI2026

GPT4o-Receipt: A Dataset and Human Study for AI-Generated Document Forensics

Yan Zhang, Simiao Ren, Ankit Raj +6

Can humans detect AI-generated financial documents better than machines? We present GPT4o-Receipt, a benchmark of 1,235 receipt images pairing GPT-4o-generated receipts with authen…

cs.CV2023

Segment anything, from space?

Simiao Ren, Francesco Luzi, Saad Lahrichi +4

Recently, the first foundation model developed specifically for image segmentation tasks was developed, termed the "Segment Anything Model" (SAM). SAM can segment objects in input…

cs.LG2021

Blaschke Product Neural Networks (BPNN): A Physics-Infused Neural Network for Phase Retrieval of Meromorphic Functions

Juncheng Dong, Simiao Ren, Yang Deng +5

Numerous physical systems are described by ordinary or partial differential equations whose solutions are given by holomorphic or meromorphic functions in the complex domain. In ma…

cs.CV2026

Can a Teenager Fool an AI? Evaluating Low-Cost Cosmetic Attacks on Age Estimation Systems

Xingyu Shen, Tommy Duong, Xiaodong An +6

Age estimation systems are increasingly deployed as gatekeepers for age-restricted online content, yet their robustness to cosmetic modifications has not been systematically evalua…