Refine Your Search

Search Results

Viewing 1 to 2 of 2
Technical Paper

Vehicle Kinematics-Based Image Augmentation against Motion Blur for Object Detectors

2023-04-11
2023-01-0050
High-speed vehicles in low illumination environments severely blur the images used in object detectors, which poses a potential threat to object detector-based advanced driver assistance systems (ADAS) and autonomous driving systems. Augmenting the training images for object detectors is an efficient way to mitigate the threat from motion blur. However, little attention has been paid to the motion of the vehicle and the position of objects in the traffic scene, which limits the consistence between the resulting augmented images and traffic scenes. In this paper, we present a vehicle kinematics-based image augmentation algorithm by modeling and analyzing the traffic scenes to generate more realistic augmented images and achieve higher robustness improvement on object detectors against motion blur. Firstly, we propose a traffic scene model considering vehicle motion and the relationship between the vehicle and the object in the traffic scene.
Technical Paper

A Target-Speech-Feature-Aware Module for U-Net Based Speech Enhancement

2024-04-09
2024-01-2021
Speech enhancement can extract clean speech from noise interference, enhancing its perceptual quality and intelligibility. This technology has significant applications in in-car intelligent voice interaction. However, the complex noise environment inside the vehicle, especially the human voice interference is very prominent, which brings great challenges to the vehicle speech interaction system. In this paper, we propose a speech enhancement method based on target speech features, which can better extract clean speech and improve the perceptual quality and intelligibility of enhanced speech in the environment of human noise interference. To this end, we propose a design method for the middle layer of the U-Net architecture based on Long Short-Term Memory (LSTM), which can automatically extract the target speech features that are highly distinguishable from the noise signal and human voice interference features in noisy speech, and realize the targeted extraction of clean speech.
X