POSITIONAL ENCODINGS FOR PERCEPTION FUNCTIONS IN AUTOMATED DRIVING SYSTEMS
卷宗概要
发明人
Willem VERBEKE; Joakim JOHNANDER
IPC 分类
CPC 分类
A method for making perception predictions for a perception functionality in an automated driving system of a vehicle is disclosed. The method includes generating 2D position information of an image captured by a vehicle-mounted camera. The 2D position information indicates a position of each pixel out of a plurality of pixels of the image, or a position of each patch out of a plurality of patches of the image in the 2D reference frame of the image. Then, feeding the generated 2D position information, extrinsic parameters of the vehicle-mounted camera, intrinsic parameters of the vehicle-mounted camera, and distortion parameters of the vehicle-mounted camera to a multilayer perceptron which process the feed data and output 3D positional encodings. The method further includes feeding the image data and the 3D positional encodings to a transformer network for generating a prediction output in a 3D/2D reference frame of the vehicle.
原文(中文)
A method for making perception predictions for a perception functionality in an automated driving system of a vehicle is disclosed. The method includes generating 2D position information of an image captured by a vehicle-mounted camera. The 2D position information indicates a position of each pixel out of a plurality of pixels of the image, or a position of each patch out of a plurality of patches of the image in the 2D reference frame of the image. Then, feeding the generated 2D position information, extrinsic parameters of the vehicle-mounted camera, intrinsic parameters of the vehicle-mounted camera, and distortion parameters of the vehicle-mounted camera to a multilayer perceptron which process the feed data and output 3D positional encodings. The method further includes feeding the image data and the 3D positional encodings to a transformer network for generating a prediction output in a 3D/2D reference frame of the vehicle.