3 papers
cs.CV2026
VISTA: Dense Multi-Label Classroom Coding with Vision-Language Models
Andrew Franck, Brendan Ng, Ben Fitzgerald +3
Video-language benchmarks are usually constructed by the dataset authors without published reliability statistics, leaving the noise floor of the construct unknown. We argue that m…
cs.LG2026
Fast Surrogate Modeling of Excitable and Oscillatory FitzHugh-Nagumo Dynamics with Parametric Neural Operators
Andrew Franck, Justin Li
The FitzHugh-Nagumo (FHN) system serves as a simplified model of neuronal voltage dynamics, capturing the activator-inhibitor structure behind both isolated action potentials and t…
cs.AI2026
Edge Phoneme Recognition for Children's Speech through Age-Aware Training
Matthew Arboleda, Ryan Arboleda, Sophie Haak +6
Detecting phonemes from children's speech has historically been difficult due to the scarcity of training data, and unique characteristics of children's speech. During a phoneme de…