Vision Based Bird Detection System
Bibliographic record
Abstract
Context : Air being the free source is used in different ways commercially. In earlier days windmills generate power, water, and electricity. The excessive establishment of windmills for commercial purposes affected avifauna. Most of the birds lost their lives due to collisions with windmills. Turbines used to generate power near airports are also one of the causes for the extinction of birdlife. According to a survey in 2011 in Canada a total of 23,300 bird deaths were caused by wind turbines and also it is estimated that the number of deaths would increase to 2,33,000 in the following 10-15 years. Objectives : The main objective of this thesis is to find a suitable software solution to detect the birds on a series of grayscale images in real-time and in minimum full HD resolution with at least a 15 FPS rate. User-Driven Design Methodology is used for developing, tools are Python and Open-CV. Methods : In this research, a system is designed to detect the bird in an HD Video. Possible methods that can be considered are convolutional neural networks (CNN), vision based detection with background subtraction, contour detection and confusion matrix classification. These methods detect birds in raw images and with help of a classifier make it possible to see the bird in desired pixels with full resolution. We will investigate a bird classification method consisting of two steps, background subtraction, and then object classification. Background subtraction is a fundamental method to extract moving objects from a fixed background. For classification, we will use the YOLO v3 model version for object classification. Results : The project is expected to result in a system design and prototype for the bird identification on a grayscale video stream in at least full HD resolution in a minimum of 15 FPS. The bird should be distinguished from other moving objects like wind turbine blades, trees, or clouds. The proposed solution should identify up to 5 birds simultaneously. Conclusion : After completing each step and arriving at the classification, the methods we have tried, such as Haar Cascades and mobile-net SSD, were not providing us with the desired results. So we opted to use YOLO v3, which had the best accuracy in classifying different objects. By using the YOLO v3 classifier, we have detected the bird with 95% accuracy, blades with 90% accuracy, clouds with 80% accuracy, trees with 70% accuracy. Moreover, we conclude that there is a need for further empirical validation of the models in full-scale industry trials.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.014 | 0.009 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".