4 papers
StrokeNeXt: A Siamese-encoder Approach for Brain Stroke Classification in Computed Tomography Imagery
Leo Thomas Ramos, Angel D. Sappa
We present StrokeNeXt, a model for stroke classification in 2D Computed Tomography (CT) images. StrokeNeXt employs a dual-branch design with two ConvNeXt encoders, whose features a…
Multi-encoder ConvNeXt Network with Smooth Attentional Feature Fusion for Multispectral Semantic Segmentation
Leo Thomas Ramos, Angel D. Sappa
This work proposes MeCSAFNet, a multi-branch encoder-decoder architecture for land cover segmentation in multispectral imagery. The model separately processes visible and non-visib…
MMLSv2: A Multimodal Dataset for Martian Landslide Detection in Remote Sensing Imagery
Sidike Paheding, Abel Reyes-Angulo, Leo Thomas Ramos +5
We present MMLSv2, a dataset for landslide segmentation on Martian surfaces. MMLSv2 consists of multimodal imagery with seven bands: RGB, digital elevation model, slope, thermal in…
A Decade of You Only Look Once (YOLO) for Object Detection: A Review
Leo Thomas Ramos, Angel D. Sappa
This review marks the tenth anniversary of You Only Look Once (YOLO), one of the most influential frameworks in real-time object detection. Over the past decade, YOLO has evolved f…