Related Experiment Video
Updated: Aug 2, 2025

03:31
Author Spotlight: Enhancement of Salient Object Detection for Smart Grid Applications
Published on: December 15, 2023
592
VV-YOLO: A Vehicle View Object Detection Model Based on Improved YOLOv4
Yinan Wang1, Yingzhou Guan1, Hanxu Liu1
1China FAW Corporation Limited, Global R&D Center, Changchun 130013, China.
Sensors (Basel, Switzerland)
|April 13, 2023
Summary
This study introduces VV-YOLO, an improved object detection model for autonomous vehicles, enhancing performance in complex driving conditions. The new model achieves higher precision and average precision with minimal increase in computation time.
Area of Science:
- Computer Vision
- Autonomous Driving Systems
- Machine Learning
Background:
- Vehicle view object detection is critical for autonomous driving safety.
- Complex scenes (dim light, occlusion, long distance) pose challenges for current models.
- Existing YOLOv4 models require improvements for robust environmental perception.
Purpose of the Study:
- To propose an improved YOLOv4-based model, VV-YOLO, for enhanced vehicle view object detection.
- To address challenges in complex driving environments.
- To improve the accuracy and robustness of object detection for autonomous vehicles.
Main Methods:
- Implemented an anchor frame-based approach with an improved K-means++ algorithm for stable anchor clustering.
- Introduced a CA-PAN network with a coordinate attention mechanism in the neck for better feature extraction.
- Reconstructed the loss function using a focus mechanism to handle imbalanced training data.
Main Results:
- VV-YOLO achieved 90.68% precision and 80.01% average precision on the KITTI dataset, outperforming YOLOv4 by 6.88% and 3.44% respectively.
- The model demonstrated comparable computation time to YOLOv4.
- Validation on BDD100K and field-collected data confirmed the model's validity and robustness.
Conclusions:
- The proposed VV-YOLO model significantly improves vehicle view object detection in complex scenarios.
- The integration of coordinate attention and a focus-based loss function enhances model performance and training stability.
- VV-YOLO offers a robust and efficient solution for environmental perception in autonomous vehicles.
More Related Videos
Related Concept Videos
Force Classification
1.3K
Forces play a crucial role in the study of physics and engineering. They are essential in describing the motion, behavior, and equilibrium of objects in the physical world. Forces can be classified based on their origin, type, and direction of action.
Contact and non-contact forces are two of the most widely used categories of forces. As the name suggests, contact forces require physical contact between two objects to act upon each other. Examples of contact forces include frictional,...
Contact and non-contact forces are two of the most widely used categories of forces. As the name suggests, contact forces require physical contact between two objects to act upon each other. Examples of contact forces include frictional,...
1.3K
Improving Translational Accuracy
11.7K
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
11.7K
Deconvolution
212
Deconvolution, also known as inverse filtering, is the process of extracting the impulse response from known input and output signals. This technique is vital in scenarios where the system's characteristics are unknown, and they must be inferred from the observable signals.
Deconvolution involves several mathematical techniques to derive the impulse response. One common approach is polynomial division. In this method, the input and output sequences are treated as coefficients of...
Deconvolution involves several mathematical techniques to derive the impulse response. One common approach is polynomial division. In this method, the input and output sequences are treated as coefficients of...
212
Difference from Background: Limit of Detection
6.7K
The limit of detection (LOD) is the smallest amount of analyte that can be distinguished from the background noise. The LOD value corresponds to the concentration at which the analyte signal is three times larger than the standard deviation of the blank signal. Below this value, the analyte signal cannot be differentiated from the background noise. It is calculated by dividing the calibration slope by 3 times the standard deviation of the blank signals.
The LOD indicates the presence or absence...
The LOD indicates the presence or absence...
6.7K
Detection of Black Holes
2.2K
Although black holes were theoretically postulated in the 1920s, they remained outside the domain of observational astronomy until the 1970s.
Their closest cousins are neutron stars, which are composed almost entirely of neutrons packed against each other, making them extremely dense. A neutron star has the same mass as the Sun but its diameter is only a few kilometers. Therefore, the escape velocity from their surface is close to the speed of light.
Not until the 1960s, when the first neutron...
Their closest cousins are neutron stars, which are composed almost entirely of neutrons packed against each other, making them extremely dense. A neutron star has the same mass as the Sun but its diameter is only a few kilometers. Therefore, the escape velocity from their surface is close to the speed of light.
Not until the 1960s, when the first neutron...
2.2K
Observational Learning
250
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
250

