A Framework for Open World Object Detection
Notice bibliographique
Résumé
Open World Object Detection (OWOD) is a computer vision task that focuses on real-world scenarios where object detection algorithms need to not only detect known and labeled objects but also handle novel and unknown objects that were not seen during training. This distinguishes OWOD from traditional object detection benchmarks, where the scope is limited to detecting only known object classes. The main challenge in OWOD lies in detecting and classifying unknown objects, which were not part of the training data. In standard object detection, objects not overlapping with labeled objects are automatically classified as background. However, these approaches are not suitable for OWOD, as unknown objects may be wrongly predicted as background due to the lack of specific supervision for distinguishing unknown objects from the background. The paper proposes a novel framework for Open World Object Detection called Open World Object Detection based on Non-Parametric classification (OWOD-NP). This method aims to address the challenges of identifying unknown objects and extending the knowledge base by incrementally introducing new object categories. OWOD-NP incorporates a non-parametric learning approach based on mean prototypes and rejection criteria into a standard detector model. The non-parametric learning model allows the system to detect whether the perceived region contains an unknown object and perform incremental learning in an end-to-end manner. The extensive experiments conducted on the benchmark dataset of Pascal Visual Object Classes (VOC) validate the effectiveness of OWOD-NP. Compared to the standard faster RCNN model, OWOD-NP achieves approximately 14% higher mean Average Precision (mAP) in class incremental scenarios. This improvement showcases the capability of OWOD-NP to handle open-world object detection tasks more efficiently. By combining non-parametric learning with object detection, OWOD-NP provides a promising solution for open-world scenarios, where the environment is dynamic and new objects may appear over time. The ability to detect and classify both known and unknown objects makes OWOD-NP a valuable approach for real-world applications in robotics, autonomous systems, and other computer vision tasks. It allows for continuous adaptation and learning, enabling the system to extend its knowledge and cope with ever-changing environments effectively.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,003 | 0,008 |
| Méta-épidémiologie (sens strict) | 0,002 | 0,001 |
| Méta-épidémiologie (sens large) | 0,002 | 0,003 |
| Bibliométrie | 0,004 | 0,003 |
| Études des sciences et des technologies | 0,001 | 0,002 |
| Communication savante | 0,004 | 0,004 |
| Science ouverte | 0,007 | 0,006 |
| Intégrité de la recherche | 0,003 | 0,004 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,005 | 0,003 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».