Focal Loss for Dense Object Detection
arXiv:1708.02002
Abstract
The highest accuracy object detectors to date are based on a two-stage approach popularized by R-CNN, where a classifier is applied to a sparse set of candidate object locations. In contrast, one-stage detectors that are applied over a regular, dense sampling of possible object locations have the potential to be faster and simpler, but have trailed the accuracy of two-stage detectors thus far. In this paper, we investigate why this is the case. We discover that the extreme foreground-background class imbalance encountered during training of dense detectors is the central cause. We propose to address this class imbalance by reshaping the standard cross entropy loss such that it down-weights the loss assigned to well-classified examples. Our novel Focal Loss focuses training on a sparse set of hard examples and prevents the vast number of easy negatives from overwhelming the detector during training. To evaluate the effectiveness of our loss, we design and train a simple dense detector we call RetinaNet. Our results show that when trained with the focal loss, RetinaNet is able to match the speed of previous one-stage detectors while surpassing the accuracy of all existing state-of-the-art two-stage detectors. Code is at: https://github.com/facebookresearch/Detectron.
References in corpus (2)
Cited by in corpus (335)
- YOLOv3: An Incremental Improvement
- ResUNet-a: a deep learning framework for semantic segmentation of remotely sensed data
- Recent advances and clinical applications of deep learning in medical image analysis
- The Real-World-Weight Cross-Entropy Loss Function: Modeling the Costs of Mislabeling
- A review: Deep learning for medical image segmentation using multi-modality fusion
- YOLACT++: Better Real-time Instance Segmentation
- The Open Images Dataset V4: Unified image classification, object detection, and visual relationship detection at scale
- Cascade R-CNN: Delving into High Quality Object Detection
- Deep Learning for UAV-based Object Detection and Tracking: A Survey
- Using Self-Supervised Learning Can Improve Model Robustness and Uncertainty
- DetNet: A Backbone network for Object Detection
- Benchmarking TPU, GPU, and CPU Platforms for Deep Learning
- Asymmetric Loss Functions and Deep Densely Connected Networks for Highly Imbalanced Medical Image Segmentation: Application to Multiple Sclerosis Lesion Detection
- Evaluating the Single-Shot MultiBox Detector and YOLO Deep Learning Models for the Detection of Tomatoes in a Greenhouse
- RAMP-CNN: A Novel Neural Network for Enhanced Automotive Radar Object Recognition
- Probabilistic two-stage detection
- End-to-end Prostate Cancer Detection in bpMRI via 3D CNNs: Effects of Attention Mechanisms, Clinical Priori and Decoupled False Positive Reduction
- Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection
- CornerNet: Detecting Objects as Paired Keypoints
- Face Attention Network: An Effective Face Detector for the Occluded Faces
- Benchmarking Robustness in Object Detection: Autonomous Driving when Winter is Coming
- Learn to Combine Modalities in Multimodal Deep Learning
- Receptive Field Block Net for Accurate and Fast Object Detection
- Adapting Mask-RCNN for Automatic Nucleus Segmentation
- Single-Shot Refinement Neural Network for Object Detection
- Detection of masses and architectural distortions in digital breast tomosynthesis: a publicly available dataset of 5,060 patients and a deep learning model
- Zero-Shot Detection
- Learning to Fuse Things and Stuff
- FocusNetv2: Imbalanced Large and Small Organ Segmentation with Adversarial Shape Constraint for Head and Neck CT Images
- Review of data analysis in vision inspection of power lines with an in-depth discussion of deep learning technology
- LittleYOLO-SPP: A Delicate Real-Time Vehicle Detection Algorithm
- Learning Graph-Level Representation for Drug Discovery
- BBN: Bilateral-Branch Network with Cumulative Learning for Long-Tailed Visual Recognition
- RETURNN as a Generic Flexible Neural Toolkit with Application to Translation and Speech Recognition
- Gradient Harmonized Single-stage Detector
- Relation Networks for Object Detection
- One-Shot Instance Segmentation
- From SLAM to Situational Awareness: Challenges and Survey
- Asymmetric Loss For Multi-Label Classification
- Revisiting Feature Alignment for One-stage Object Detection
- Feature Pyramid and Hierarchical Boosting Network for Pavement Crack Detection
- VITAL: VIsual Tracking via Adversarial Learning
- Video Instance Segmentation using Inter-Frame Communication Transformers
- PointFusion: Deep Sensor Fusion for 3D Bounding Box Estimation
- Unsupervised clustering for collider physics
- PyramidBox: A Context-assisted Single Shot Face Detector
- Decoupled Classification Refinement: Hard False Positive Suppression for Object Detection
- Object Detection for Comics using Manga109 Annotations
- The LHC Olympics 2020: A Community Challenge for Anomaly Detection in High Energy Physics
- Mobile Video Object Detection with Temporally-Aware Feature Maps
- A New Training Pipeline for an Improved Neural Transducer
- The Boosted Higgs Jet Reconstruction via Graph Neural Network
- Improved Techniques for Learning to Dehaze and Beyond: A Collective Study
- AdaScale: Towards Real-time Video Object Detection Using Adaptive Scaling
- A global method to identify trees outside of closed-canopy forests with medium-resolution satellite imagery
- 3D Multi-Object Tracking: A Baseline and New Evaluation Metrics
- PyramidBox++: High Performance Detector for Finding Tiny Face
- SketchyGAN: Towards Diverse and Realistic Sketch to Image Synthesis
- Biologically-plausible learning algorithms can scale to large datasets
- Learning Gaussian Instance Segmentation in Point Clouds
- Mask-Guided Attention Network for Occluded Pedestrian Detection
- Real-Time Human Pose Estimation on a Smart Walker using Convolutional Neural Networks
- CPS++: Improving Class-level 6D Pose and Shape Estimation From Monocular Images With Self-Supervised Learning
- Towards High Performance Video Object Detection for Mobiles
- Benchmark Dataset for Automatic Damaged Building Detection from Post-Hurricane Remotely Sensed Imagery
- A Comprehensive Survey of Machine Learning Applied to Radar Signal Processing
- Daedalus: Breaking Non-Maximum Suppression in Object Detection via Adversarial Examples
- An Interactive Data Visualization and Analytics Tool to Evaluate Mobility and Sociability Trends During COVID-19
- Deep learning in radiology: an overview of the concepts and a survey of the state of the art
- Learning Better Features for Face Detection with Feature Fusion and Segmentation Supervision
- Diversify and Match: A Domain Adaptive Representation Learning Paradigm for Object Detection
- Dice Loss for Data-imbalanced NLP Tasks
- Pyramidal Person Re-IDentification via Multi-Loss Dynamic Training
- Light curve classification with recurrent neural networks for GOTO: dealing with imbalanced data
- PPDM: Parallel Point Detection and Matching for Real-time Human-Object Interaction Detection
- TANet: Robust 3D Object Detection from Point Clouds with Triple Attention
- MegDet: A Large Mini-Batch Object Detector
- Siamese Cascaded Region Proposal Networks for Real-Time Visual Tracking
- Weakly Supervised Deep Learning for COVID-19 Infection Detection and Classification from CT Images
- FoodLogoDet-1500: A Dataset for Large-Scale Food Logo Detection via Multi-Scale Feature Decoupling Network
- SMOKE: Single-Stage Monocular 3D Object Detection via Keypoint Estimation
- Effect of Annotation Errors on Drone Detection with YOLOv3
- ModaNet: A Large-Scale Street Fashion Dataset with Polygon Annotations
- SEKD: Self-Evolving Keypoint Detection and Description
- 3D-A-Nets: 3D Deep Dense Descriptor for Volumetric Shapes with Adversarial Networks
- A comparative study of 2D image segmentation algorithms for traumatic brain lesions using CT data from the ProTECTIII multicenter clinical trial
- DeepTracker: Visualizing the Training Process of Convolutional Neural Networks
- Detection in Crowded Scenes: One Proposal, Multiple Predictions
- Structure Inference Net: Object Detection Using Scene-Level Context and Instance-Level Relationships
- Deep Learning for Pancreas Segmentation: a Systematic Review
- Residual Features and Unified Prediction Network for Single Stage Detection
- Locally Free Weight Sharing for Network Width Search
- Selective Refinement Network for High Performance Face Detection
- Active and Incremental Learning with Weak Supervision
- ScreenerNet: Learning Self-Paced Curriculum for Deep Neural Networks
- Auxiliary Signal-Guided Knowledge Encoder-Decoder for Medical Report Generation
- Repulsion Loss: Detecting Pedestrians in a Crowd
- Towards Deep Cellular Phenotyping in Placental Histology
- CurriculumNet: Weakly Supervised Learning from Large-Scale Web Images
- Using deep Residual Networks to search for galaxy-Lyα emitter lens candidates based on spectroscopic-selection
- SFace: An Efficient Network for Face Detection in Large Scale Variations
- CovidExpert: A Triplet Siamese Neural Network framework for the detection of COVID-19
- Learning Discriminative Motion Features Through Detection
- Improved YOLOv5s model for key components detection of power transmission lines
- Scene Graph Generation via Conditional Random Fields
- Learning Soft Labels via Meta Learning
- Visibility Guided NMS: Efficient Boosting of Amodal Object Detection in Crowded Traffic Scenes
- MetaSAug: Meta Semantic Augmentation for Long-Tailed Visual Recognition
- Fusing Bird View LIDAR Point Cloud and Front View Camera Image for Deep Object Detection
- Cross-Lingual Named Entity Recognition Using Parallel Corpus: A New Approach Using XLM-RoBERTa Alignment
- A relic sketch extraction framework based on detail-aware hierarchical deep network
- Single-Label Multi-Class Image Classification by Deep Logistic Regression
- A deep learning based solution for construction equipment detection: from development to deployment
- Chargrid: Towards Understanding 2D Documents
- SBNet: Sparse Blocks Network for Fast Inference
- Efficient Object Detection Model for Real-Time UAV Applications
- M2Det: A Single-Shot Object Detector based on Multi-Level Feature Pyramid Network
- Memory Warps for Learning Long-Term Online Video Representations
- Learning Efficient Detector with Semi-supervised Adaptive Distillation
- Class-Incremental Few-Shot Object Detection
- Deep Adaptive Proposal Network for Object Detection in Optical Remote Sensing Images
- Modeling Visual Context is Key to Augmenting Object Detection Datasets
- Mis-classified Vector Guided Softmax Loss for Face Recognition
- Salience Biased Loss for Object Detection in Aerial Images
- 3D-SSD: Learning Hierarchical Features from RGB-D Images for Amodal 3D Object Detection
- The Fishnet Open Images Database: A Dataset for Fish Detection and Fine-Grained Categorization in Fisheries
- An Analysis of Scale Invariance in Object Detection - SNIP
- Control Distance IoU and Control Distance IoU Loss Function for Better Bounding Box Regression
- Pose Neural Fabrics Search
- An In-Depth Analysis of Visual Tracking with Siamese Neural Networks
- FBI-Pose: Towards Bridging the Gap between 2D Images and 3D Human Poses using Forward-or-Backward Information
- Simultaneous lesion and neuroanatomy segmentation in Multiple Sclerosis using deep neural networks
- Distilling Knowledge via Knowledge Review
- Automated Olfactory Bulb Segmentation on High Resolutional T2-Weighted MRI
- Single-Shot Bidirectional Pyramid Networks for High-Quality Object Detection
- Weak-lensing Mass Reconstruction of Galaxy Clusters with Convolutional Neural Network
- Object detection on aerial imagery using CenterNet
- SSN: Shape Signature Networks for Multi-class Object Detection from Point Clouds
- FAIR1M: A Benchmark Dataset for Fine-grained Object Recognition in High-Resolution Remote Sensing Imagery
- Blind Predicting Similar Quality Map for Image Quality Assessment
- Contrastive Learning Improves Model Robustness Under Label Noise
- Casualty Detection from 3D Point Cloud Data for Autonomous Ground Mobile Rescue Robots
- AFP-Net: Realtime Anchor-Free Polyp Detection in Colonoscopy
- TubeTK: Adopting Tubes to Track Multi-Object in a One-Step Training Model
- OMNIA Faster R-CNN: Detection in the wild through dataset merging and soft distillation
- Enhancing Cross-task Black-Box Transferability of Adversarial Examples with Dispersion Reduction
- A survey on Kornia: an Open Source Differentiable Computer Vision Library for PyTorch
- On the Integration of Self-Attention and Convolution
- Monocular, One-stage, Regression of Multiple 3D People
- Unsupervised Hard Example Mining from Videos for Improved Object Detection
- C-WSL: Count-guided Weakly Supervised Localization
- Domain Adaptation from Synthesis to Reality in Single-model Detector for Video Smoke Detection
- MultiResolution Attention Extractor for Small Object Detection
- Weaving Multi-scale Context for Single Shot Detector
- Multi-Scale Feature Aggregation by Cross-Scale Pixel-to-Region Relation Operation for Semantic Segmentation
- ScratchDet: Training Single-Shot Object Detectors from Scratch
- PolarStream: Streaming Lidar Object Detection and Segmentation with Polar Pillars
- NETNet: Neighbor Erasing and Transferring Network for Better Single Shot Object Detection
- Combining Self-Supervised and Supervised Learning with Noisy Labels
- Persuasive Faces: Generating Faces in Advertisements
- Aerial Imagery Pixel-level Segmentation
- A New Window Loss Function for Bone Fracture Detection and Localization in X-ray Images with Point-based Annotation
- Importance-Aware Learning for Neural Headline Editing
- NOTE-RCNN: NOise Tolerant Ensemble RCNN for Semi-Supervised Object Detection
- Glance and Gaze: Inferring Action-aware Points for One-Stage Human-Object Interaction Detection
- Deep Regionlets for Object Detection
- Cosmic Background Removal with Deep Neural Networks in SBND
- Robust Face Detection via Learning Small Faces on Hard Images
- Towards Interpretable Face Recognition
- Cascaded channel pruning using hierarchical self-distillation
- Three Branches: Detecting Actions With Richer Features
- Generalization in Metric Learning: Should the Embedding Layer be the Embedding Layer?
- CONAN: Complementary Pattern Augmentation for Rare Disease Detection
- End-to-End Video Object Detection with Spatial-Temporal Transformers
- Towards High Performance Video Object Detection
- Parallel Grid Pooling for Data Augmentation
- Exploring Multi-Branch and High-Level Semantic Networks for Improving Pedestrian Detection
- Integration of Clinical Criteria into the Training of Deep Models: Application to Glucose Prediction for Diabetic People
- Streaming Object Detection for 3-D Point Clouds
- Towards large-scale, automated, accurate detection of CCTV camera objects using computer vision. Applications and implications for privacy, safety, and cybersecurity. (Preprint)
- Uncertainty-Aware Voxel based 3D Object Detection and Tracking with von-Mises Loss
- Locally Adaptive Learning Loss for Semantic Image Segmentation
- Decoupled and Memory-Reinforced Networks: Towards Effective Feature Learning for One-Step Person Search
- CathAI: Fully Automated Interpretation of Coronary Angiograms Using Neural Networks
- Pseudo Mask Augmented Object Detection
- BI-MAML: Balanced Incremental Approach for Meta Learning
- DeRPN: Taking a further step toward more general object detection
- Attention to the strengths of physical interactions: Transformer and graph-based event classification for particle physics experiments
- A Streamlined Encoder/Decoder Architecture for Melody Extraction
- Semi-supervised Learning: Fusion of Self-supervised, Supervised Learning, and Multimodal Cues for Tactical Driver Behavior Detection
- Melodic Phrase Segmentation By Deep Neural Networks
- 3D-DETNet: a Single Stage Video-Based Vehicle Detector
- Improving Semantic Segmentation via Dilated Affinity
- Object Detection based on Region Decomposition and Assembly
- Multi-view Integration Learning for Irregularly-sampled Clinical Time Series
- Learning to Separate: Detecting Heavily-Occluded Objects in Urban Scenes
- An attention model to analyse the risk of agitation and urinary tract infections in people with dementia
- Instance and Panoptic Segmentation Using Conditional Convolutions
- Is Object Detection Necessary for Human-Object Interaction Recognition?
- Learning a Unified Sample Weighting Network for Object Detection
- Adaptive Importance Learning for Improving Lightweight Image Super-resolution Network
- MatchVIE: Exploiting Match Relevancy between Entities for Visual Information Extraction
- Chinese Sentences Similarity via Cross-Attention Based Siamese Network
- Deep Priority Hashing
- Large-Scale Object Detection in the Wild from Imbalanced Multi-Labels
- Triply Supervised Decoder Networks for Joint Detection and Segmentation
- Detecting Small, Densely Distributed Objects with Filter-Amplifier Networks and Loss Boosting
- Detecting and counting tiny faces
- Learning Embeddings from Knowledge Graphs With Numeric Edge Attributes
- Interactive Medical Image Segmentation with Self-Adaptive Confidence Calibration
- Contrastive Unpaired Translation using Focal Loss for Patch Classification
- Less Is Better: Unweighted Data Subsampling via Influence Function
- A Selective Survey on Versatile Knowledge Distillation Paradigm for Neural Network Models
- Comparison of object detection methods for crop damage assessment using deep learning
- Frustum VoxNet for 3D object detection from RGB-D or Depth images
- Offset Curves Loss for Imbalanced Problem in Medical Segmentation
- Landmark Detection in Low Resolution Faces with Semi-Supervised Learning
- Face Recognition in Unconstrained Conditions: A Systematic Review
- SID: Incremental Learning for Anchor-Free Object Detection via Selective and Inter-Related Distillation
- NeuronBlocks: Building Your NLP DNN Models Like Playing Lego
- DPointNet: A Density-Oriented PointNet for 3D Object Detection in Point Clouds
- n-hot: Efficient bit-level sparsity for powers-of-two neural network quantization
- Fast Efficient Object Detection Using Selective Attention
- Effective Data Fusion with Generalized Vegetation Index: Evidence from Land Cover Segmentation in Agriculture
- Deep Feature Pyramid Reconfiguration for Object Detection
- TensorFlow with user friendly Graphical Framework for object detection API
- Accurate Anchor Free Tracking
- Split to Be Slim: An Overlooked Redundancy in Vanilla Convolution
- Skin Lesion Classification Using Deep Neural Network
- Towards an Understanding of Neural Networks in Natural-Image Spaces
- Towards Automatic 3D Shape Instantiation for Deployed Stent Grafts: 2D Multiple-class and Class-imbalance Marker Segmentation with Equally-weighted Focal U-Net
- Deep Learning for Automated Medical Image Analysis
- When 3D-Aided 2D Face Recognition Meets Deep Learning: An extended UR2D for Pose-Invariant Face Recognition
- Balance Scene Learning Mechanism for Offshore and Inshore Ship Detection in SAR Images
- Metric learning by Similarity Network for Deep Semi-Supervised Learning
- Refined Deep Neural Network and U-Net for Polyps Segmentation
- One-Click Annotation with Guided Hierarchical Object Detection
- More Reliable AI Solution: Breast Ultrasound Diagnosis Using Multi-AI Combination
- ÚFAL at MRP 2020: Permutation-invariant Semantic Parsing in PERIN
- DOSED: a deep learning approach to detect multiple sleep micro-events in EEG signal
- Deep Imbalanced Attribute Classification using Visual Attention Aggregation
- Decoupled IoU Regression for Object Detection
- On Focal Loss for Class-Posterior Probability Estimation: A Theoretical Perspective
- Instance Scale Normalization for image understanding
- Stereo RGB and Deeper LIDAR Based Network for 3D Object Detection
- ElixirNet: Relation-aware Network Architecture Adaptation for Medical Lesion Detection
- Addressing the Real-world Class Imbalance Problem in Dermatology
- Using Computer Vision to Automate Hand Detection and Tracking of Surgeon Movements in Videos of Open Surgery
- RoboSherlock: Cognition-enabled Robot Perception for Everyday Manipulation Tasks
- Multiple receptive fields and small-object-focusing weakly-supervised segmentation network for fast object detection
- Recovering the Unbiased Scene Graphs from the Biased Ones
- Minimizing Close-k Aggregate Loss Improves Classification
- FixNorm: Dissecting Weight Decay for Training Deep Neural Networks
- Class-Wise Difficulty-Balanced Loss for Solving Class-Imbalance
- A hierarchical deep learning framework for the consistent classification of land use objects in geospatial databases
- Polyp-artifact relationship analysis using graph inductive learned representations
- The 1st Challenge on Remote Physiological Signal Sensing (RePSS)
- Fully Automated and Standardized Segmentation of Adipose Tissue Compartments by Deep Learning in Three-dimensional Whole-body MRI of Epidemiological Cohort Studies
- FAN: Focused Attention Networks
- Table understanding in structured documents
- Leveraging Declarative Knowledge in Text and First-Order Logic for Fine-Grained Propaganda Detection
- Model-Agnostic Defense for Lane Detection against Adversarial Attack
- Which to Match? Selecting Consistent GT-Proposal Assignment for Pedestrian Detection
- Cost-sensitive Regularization for Label Confusion-aware Event Detection
- Adversarial Soft-detection-based Aggregation Network for Image Retrieval
- Multiple Myeloma Cancer Cell Instance Segmentation
- Multi-Modal Super Resolution for Dense Microscopic Particle Size Estimation
- Retrieval of Family Members Using Siamese Neural Network
- Feature Selective Networks for Object Detection
- Augmentation Inside the Network
- Diversifying Dialog Generation via Adaptive Label Smoothing
- Learning with Multiclass AUC: Theory and Algorithms
- Inter-Homines: Distance-Based Risk Estimation for Human Safety
- Semantic Image Cropping
- TCDesc: Learning Topology Consistent Descriptors for Image Matching
- Unsupervised data augmentation for object detection
- Multi-object Tracking with Tracked Object Bounding Box Association
- Focal Loss Dense Detector for Vehicle Surveillance
- Energy Aligning for Biased Models
- Emotions are Subtle: Learning Sentiment Based Text Representations Using Contrastive Learning
- Unsupervised Multi-Target Domain Adaptation for Acoustic Scene Classification
- Plug & Play Convolutional Regression Tracker for Video Object Detection
- FA-RPN: Floating Region Proposals for Face Detection
- G-RCN: Optimizing the Gap between Classification and Localization Tasks for Object Detection
- Alpha-Net: Architecture, Models, and Applications
- G2C: A Generator-to-Classifier Framework Integrating Multi-Stained Visual Cues for Pathological Glomerulus Classification
- A Study on Trees's Knots Prediction from their Bark Outer-Shape
- Beyond Single Stage Encoder-Decoder Networks: Deep Decoders for Semantic Image Segmentation
- High Diversity Attribute Guided Face Generation with GANs
- Object Detection on Single Monocular Images through Canonical Correlation Analysis
- On the alpha-loss Landscape in the Logistic Model
- Do We Really Need Gold Samples for Sample Weighting Under Label Noise?
- Selective Output Smoothing Regularization: Regularize Neural Networks by Softening Output Distributions
- Fast Single-shot Ship Instance Segmentation Based on Polar Template Mask in Remote Sensing Images
- TUNet: Incorporating segmentation maps to improve classification
- Segmentation-Based Bounding Box Generation for Omnidirectional Pedestrian Detection
- Utilizing Complex-valued Network for Learning to Compare Image Patches
- From Bag of Sentences to Document: Distantly Supervised Relation Extraction via Machine Reading Comprehension
- Exploring Content Based Image Retrieval for Highly Imbalanced Melanoma Data using Style Transfer, Semantic Image Segmentation and Ensemble Learning
- Improving Deep Binary Embedding Networks by Order-aware Reweighting of Triplets
- Organ At Risk Segmentation with Multiple Modality
- Proactive Network Maintenance using Fast, Accurate Anomaly Localization and Classification on 1-D Data Series
- Where were my keys? -- Aggregating Spatial-Temporal Instances of Objects for Efficient Retrieval over Long Periods of Time
- Exploiting Class Similarity for Machine Learning with Confidence Labels and Projective Loss Functions
- A Mathematical Foundation for Robust Machine Learning based on Bias-Variance Trade-off
- RotationOut as a Regularization Method for Neural Network
- Political Ideology and Polarization of Policy Positions: A Multi-dimensional Approach
- BAN: Focusing on Boundary Context for Object Detection
- GAIA: A Transfer Learning System of Object Detection that Fits Your Needs
- Identity-Enhanced Network for Facial Expression Recognition
- QK Iteration: A Self-Supervised Representation Learning Algorithm for Image Similarity
- Realistic simulation of users for IT systems in cyber ranges
- A Generalization Theory based on Independent and Task-Identically Distributed Assumption
- Affinity Graph Supervision for Visual Recognition
- Graph2Graph Learning with Conditional Autoregressive Models
- Know Your Surroundings: Panoramic Multi-Object Tracking by Multimodality Collaboration
- Unsupervised Difficulty Estimation with Action Scores
- Bootstrap Your Object Detector via Mixed Training
- Modular network for high accuracy object detection
- Attentive Sequence to Sequence Translation for Localizing Clips of Interest by Natural Language Descriptions
- Robust Multi-Domain Mitosis Detection
- Camera Invariant Feature Learning for Generalized Face Anti-spoofing
- Mixing between the Cross Entropy and the Expectation Loss Terms
- Voxel-level Siamese Representation Learning for Abdominal Multi-Organ Segmentation
- Exploring Instance-Level Uncertainty for Medical Detection
- Integrating Feature and Image Pyramid: A Lung Nodule Detector Learned in Curriculum Fashion
- TraMNet - Transition Matrix Network for Efficient Action Tube Proposals
- Learning to Predict the 3D Layout of a Scene
- 2nd Place Solution to Instance Segmentation of IJCAI 3D AI Challenge 2020
- Modeling Inter-Aspect Dependencies with a Non-temporal Mechanism for Aspect-Based Sentiment Analysis
- Bounding Box Embedding for Single Shot Person Instance Segmentation
- Learning to Reconstruct and Segment 3D Objects
- CC-Loss: Channel Correlation Loss For Image Classification
- Indirect Domain Shift for Single Image Dehazing
- A Dataset of Laryngeal Endoscopic Images with Comparative Study on Convolution Neural Network Based Semantic Segmentation