REVIEW 7 cited by
Probabilistic two-stage detection
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We develop a probabilistic interpretation of two-stage object detection. We show that this probabilistic interpretation motivates a number of common empirical training practices. It also suggests changes to two-stage detection pipelines. Specifically, the first stage should infer proper object-vs-background likelihoods, which should then inform the overall score of the detector. A standard region proposal network (RPN) cannot infer this likelihood sufficiently well, but many one-stage detectors can. We show how to build a probabilistic two-stage detector from any state-of-the-art one-stage detector. The resulting detectors are faster and more accurate than both their one- and two-stage precursors. Our detector achieves 56.4 mAP on COCO test-dev with single-scale testing, outperforming all published results. Using a lightweight backbone, our detector achieves 49.2 mAP on COCO at 33 fps on a Titan Xp, outperforming the popular YOLOv4 model.
Forward citations
Cited by 7 Pith papers
-
TMI: Text-to-Image Meets Image-to-Image for Complementary Data Synthesis to Boost Long-Tailed Instance Segmentation
Hybrid T2I generation with teacher-student pseudo-labeling plus VRAIN context-aware I2I rare-class editing improves LVIS instance segmentation AP, especially on rare categories.
-
Embodied Domain Adaptation for Object Detection
EDAOD adapts open-vocabulary object detectors to new indoor scenes via temporal instance clustering and contrastive learning, outperforming source-free baselines on a new benchmark.
-
Corner2Net: Detecting Objects as Cascade Corners
Corner2Net detects objects as cascade corners, predicting class-agnostic top-left corners first and instance-specific bottom-right corners in each RoI, and claims state-of-the-art among corner-based detectors.
-
Comprehensive Multi-Modal Prototypes are Simple and Effective Classifiers for Vast-Vocabulary Object Detection
Prova replaces coarse class-name embeddings with multi-modal prototypes from descriptions and reference images, improving vast-vocabulary detection on V3Det by up to 6.2 AP.
-
Model-Agnostic Open-Set Air-to-Air Visual Object Detection for Reliable UAV Perception
Fusing softmax confidence with embedding-space Gaussian mixture entropy improves open-set rejection in air-to-air UAV detection, reporting about 0.88 AUROC on real flight data.
-
DEYOLO: Dual-Feature-Enhancement YOLO for Cross-Modality Object Detection
DEYOLO fuses RGB and infrared features with dual channel and spatial enhancement modules and reports improved object detection accuracy on the M3FD and LLVIP datasets.
- Introducing a multiscale feature integration network for inpainting with applications to enhanced CMB map reconstruction
Discussion (0). Continue with ORCID to comment.