ETRI-Knowledge Sharing Plaform

KOREAN
논문 검색
Type SCI
Year ~ Keyword

Detail

Journal Article Adaptive spatial down-sampling method based on object occupancy distribution for video coding for machines
Cited 0 time in scopus Download 63 time Share share facebook twitter linkedin kakaostory
Authors
Eun-bin An, Ayoung Kim, Soon-heung Jung, Sangwoon Kwak, Jin Young Lee, Won-Sik Cheong, Hyon-Gon Choo, Kwang-deok Seo
Issue Date
2024-10
Citation
Eurasip Journal on Image and Video Processing, v.2024, pp.1-17
ISSN
1687-5176
Publisher
Springer International Publishing AG
Language
English
Type
Journal Article
DOI
https://dx.doi.org/10.1186/s13640-024-00647-y
Abstract
As the performance of machine vision continues to improve, it is being used in various industrial fields to analyze and generate massive amounts of video data. Although the demand for and consumption of video data by machines has increased significantly, video coding for machines needs to be improved. It is therefore necessary to consider a new codec that differs from conventional codecs based on the human visual system (HVS). Spatial down-sampling plays a critical role in video coding for machines because it reduces the volume of the video data to be processed while maintaining the shape of the data’s features that are important for the machine to reference when processing the video. An effective method of determining the intensity of spatial down-sampling as an efficient coding tool for machines is still in the early stages. Here, we propose a method of determining an optimal scale factor for spatial down-sampling by collecting and analyzing information on the number of objects and the ratio of the area occupied by the object within a picture. We compare the data reduction ratio to the machine accuracy error ratio (DRAER) to evaluate the performance of the proposed method. By applying the proposed method, the DRAER was found to be a maximum of 21.40 dB and a minimum of 11.94 dB. This shows that video coding gain for the machines could be achieved through the proposed method while maintaining the accuracy of machine vision tasks.
KSP Keywords
Coding Gain, Down-sampling, Early stages, Efficient coding, Machine accuracy, Video data, data reduction ratio, human visual system, machine vision, optimal scale, sampling methods
This work is distributed under the term of Creative Commons License (CCL)
(CC BY NC ND)
CC BY NC ND