Publications (33)
SmartCS: Enabling the Creation of ML-Powered Computer Vision Mobile Apps for Citizen Science Applications without Coding
Fahim Hasan Khan, Akila de Silva, Gregory Dusek +2
It is undeniable that citizen science contributes to the advancement of various fields of study. There are now software tools that facilitate the development of citizen science app…
VortexViz: Finding Vortex Boundaries by Learning from Particle Trajectories
Akila de Silva, Nicholas Tee, Omkar Ghanekar +4
Vortices are studied in various scientific disciplines, offering insights into fluid flow behavior. Visualizing the boundary of vortices is crucial for understanding flow phenomena…
RegHead: Non-Humanoid Head Blendshapes via Feed-Forward Registration
Jiahao Luo, Hao Zhang, Jianqi Chen +9
RegHead is a framework that builds semantic blendshape sets for animatable non‑humanoid head avatars using a fast feed‑forward registration model and a large dataset of shared expr…
Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models
Marzia Binta Nizam, James Davis
Text-to-image(T2I) models like Stable Diffusion and DALL-E have made generative AI widely accessible, yet recent studies reveal that these systems often replicate societal biases,…
Do top conferences contain well cited papers or junk?
James Davis
In order to answer questions about top conference publication patterns, citation data is collected and analyzed for several computer science conferences, with focus on computer vis…
GenIR: Generative Visual Feedback for Mental Image Retrieval
Diji Yang, Minghao Liu, Chung-Hsiang Lo +2
Vision-language models (VLMs) have shown strong performance on text-to-image retrieval benchmarks. However, bridging this success to real-world applications remains a challenge. In…
SplatFace: Gaussian Splat Face Reconstruction Leveraging an Optimizable Surface
Jiahao Luo, Jing Liu, James Davis
We present SplatFace, a novel Gaussian splatting framework designed for 3D human face reconstruction without reliance on accurate pre-determined geometry. Our method is designed to…
Tag-based annotation creates better avatars
Minghao Liu, Zeyu Cheng, Shen Sang +2
Avatar creation from human images allows users to customize their digital figures in different styles. Existing rendering systems like Bitmoji, MetaHuman, and Google Cartoonset pro…
T2Bs: Text-to-Character Blendshapes via Video Generation
Jiahao Luo, Chaoyang Wang, Michael Vasilkovsky +8
We present T2Bs, a framework for generating high-quality, animatable character head morphable models from text by combining static text-to-3D generation with video diffusion. Text-…
AgileAvatar: Stylized 3D Avatar Creation via Cascaded Domain Bridging
Shen Sang, Tiancheng Zhi, Guoxian Song +6
Stylized 3D avatars have become increasingly prominent in our modern life. Creating these avatars manually usually involves laborious selection and adjustment of continuous and dis…
Assessing the Impact of Prompting Methods on ChatGPT's Mathematical Capabilities
Yuhao Chen, Chloe Wong, Hanwen Yang +11
This study critically evaluates the efficacy of prompting methods in enhancing the mathematical reasoning capability of large language models (LLMs). The investigation uses three p…
RipViz: Finding Rip Currents by Learning Pathline Behavior
Akila de Silva, Mona Zhao, Donald Stewart +4
We present a hybrid machine learning and flow analysis feature detection method, RipViz, to extract rip currents from stationary videos. Rip currents are dangerous strong currents…
Recipe for Discovery: A Pipeline for Institutional Open Source Activity
Juanita Gomez, Emily Lovell, Stephanie Lieggi +2
Open source software development, particularly within institutions such as universities and research laboratories, is often decentralized and difficult to track. Although academic…
VaLID: Verification as Late Integration of Detections for LiDAR-Camera Fusion
Vanshika Vats, Marzia Binta Nizam, James Davis
Vehicle object detection benefits from both LiDAR and camera data, with LiDAR offering superior performance in many scenarios. Fusion of these modalities further enhances accuracy,…
DuelGAN: A Duel Between Two Discriminators Stabilizes the GAN Training
Jiaheng Wei, Minghao Liu, Jiahao Luo +3
In this paper, we introduce DuelGAN, a generative adversarial network (GAN) solution to improve the stability of the generated samples and to mitigate mode collapse. Built upon the…
J-CaPA : Joint Channel and Pyramid Attention Improves Medical Image Segmentation
Marzia Binta Nizam, Marian Zlateva, James Davis
Medical image segmentation is crucial for diagnosis and treatment planning. Traditional CNN-based models, like U-Net, have shown promising results but struggle to capture long-rang…
On Optimizing Human-Machine Task Assignments
Andreas Veit, Michael Wilber, Rajan Vaish +48
When crowdsourcing systems are used in combination with machine inference systems in the real world, they benefit the most when the machine system is deeply integrated with the cro…
Authoring Platform for Mobile Citizen Science Apps with Client-side ML
Fahim Hasan Khan, Akila de Silva, Gregory Dusek +2
Data collection is an integral part of any citizen science project. Given the wide variety of projects, some level of expertise or, alternatively, some guidance for novice particip…
Redefining Data-Centric Design: A New Approach with a Domain Model and Core Data Ontology for Computational Systems
William Johnson, James Davis, Tara Kelly
This paper presents an innovative data-centric paradigm for designing computational systems by introducing a new informatics domain model. The proposed model moves away from the co…
Genuinely nonabelian partial difference sets
John Polhill, James Davis, Ken Smith +1
Strongly regular graphs (SRGs) provide a fertile area of exploration in algebraic combinatorics, integrating techniques in graph theory, linear algebra, group theory, finite fields…
Tag-Based Annotation for Avatar Face Creation
An Ngo, Daniel Phelps, Derrick Lai +6
Currently, digital avatars can be created manually using human images as reference. Systems such as Bitmoji are excellent producers of detailed avatar designs, with hundreds of cho…
AI for Green Spaces: Leveraging Autonomous Navigation and Computer Vision for Park Litter Removal
Christopher Kao, Akhil Pathapati, James Davis
There are 50 billion pieces of litter in the U.S. alone. Grass fields contribute to this problem because picnickers tend to leave trash on the field. We propose building a robot th…
Guideline-Consistent Segmentation via Multi-Agent Refinement
Vanshika Vats, Ashwani Rathee, James Davis
Semantic segmentation in real-world applications often requires not only accurate masks but also strict adherence to textual labeling guidelines. These guidelines are typically com…
A Fine-grained Data Set and Analysis of Tangling in Bug Fixing Commits
Steffen Herbold, Alexander Trautsch, Benjamin Ledel +45
Context: Tangled commits are changes to software that address multiple concerns at once. For researchers interested in bugs, tangled commits mean that they actually study not only…
Automated Rip Current Detection with Region based Convolutional Neural Networks
Akila de Silva, Issei Mori, Gregory Dusek +2
This paper presents a machine learning approach for the automatic identification of rip currents with breaking waves. Rip currents are dangerous fast moving currents of water that…
Human and AI Perceptual Differences in Image Classification Errors
Minghao Liu, Jiaheng Wei, Yang Liu +1
Artificial intelligence (AI) models for computer vision trained with supervised machine learning are assumed to solve classification tasks by imitating human behavior learned from…
Nonabelian partial difference sets constructed using abelian techniques
James Davis, John Polhill, Ken Smith +1
A -partial difference set (PDS) is a subset of a group such that , , and every nonidentity element of can be written in either …
Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond
Minghao Liu, Zonglin Di, Jiaheng Wei +15
Large-scale data collection is essential for developing personalized training data, mitigating the shortage of training data, and fine-tuning specialized models. However, creating…
AI Failures in the Eyes of the Downstream Developer: A First Look at Concerns, Practices, and Challenges
Haoyu Gao, Mansooreh Zahedi, Wenxin Jiang +3
With the advancement of AI models, more software systems are adopting AI as a component to facilitate automation. Pre-trained models (PTMs) have become a cornerstone of AI-based so…
Viability of Mobile Forms for Population Health Surveys in Low Resource Areas
Alexander Davis, Aidan Chen, Milton Chen +1
Population health surveys are an important tool to effectively allocate limited resources in low resource communities. In such an environment, surveys are often done by local popul…
A Survey on Human-AI Collaboration with Large Foundation Models
Vanshika Vats, Marzia Binta Nizam, Minghao Liu +20
As the capabilities of artificial intelligence (AI) continue to expand rapidly, Human-AI (HAI) Collaboration, combining human intellect and AI systems, has become pivotal for advan…
Hidden in Plain Text: Measuring LLM Deception Quality Against Human Baselines Using Social Deduction Games
Christopher Kao, Vanshika Vats, James Davis
Large Language Model (LLM) agents are increasingly used in many applications, raising concerns about their safety. While previous work has shown that LLMs can deceive in controlled…
Disjoint Pose and Shape for 3D Face Reconstruction
Raja Kumar, Jiahao Luo, Alex Pang +1
Existing methods for 3D face reconstruction from a few casually captured images employ deep learning based models along with a 3D Morphable Model(3DMM) as face geometry prior. Stru…