A Review of Machine Learning and Deep Learning for Object Detection, Semantic Segmentation, and Human Action Recognition in Machine and Robotic Vision

Обзор методов машинного обучения и глубокого обучения для обнаружения объектов, семантической сегментации и распознавания действий человека в машинном и роботизированном зрении
Lazaros Moysis, Nikoleta Manakitsa, George S. Maraslidis, George F. Fragulis
2024-01-23

deep learninghuman action recognitionmachine visionobject detectionsemantic segmentation
Machine vision, an interdisciplinary field that aims to replicate human visual perception in computers, has experienced rapid progress and significant contributions. This paper traces the origins of machine vision, from early image processing algorithms to its convergence with computer science, mathematics, and robotics, resulting in a distinct branch of artificial intelligence. The integration of machine learning techniques, particularly deep learning, has driven its growth and adoption in everyday devices. This study focuses on the objectives of computer vision systems: replicating human visual capabilities including recognition, comprehension, and interpretation. Notably, image classification, object detection, and image segmentation are crucial tasks requiring robust mathematical foundations. Despite the advancements, challenges persist, such as clarifying terminology related to artificial intelligence, machine learning, and deep learning. Precise definitions and interpretations are vital for establishing a solid research foundation. The evolution of machine vision reflects an ambitious journey to emulate human visual perception. Interdisciplinary collaboration and the integration of deep learning techniques have propelled remarkable advancements in emulating human behavior and perception. Through this research, the field of machine vision continues to shape the future of computer systems and artificial intelligence applications.
1
Despite major progress toward emulating human visual perception and behavior, machine vision continues to face unresolved conceptual and practical challenges.
2
Machine learning, particularly deep learning, has substantially accelerated machine vision’s development and adoption in everyday devices.
3
Object detection, semantic segmentation, image classification, and human action recognition are identified as central computer-vision tasks requiring robust mathematical foundations.
4
The paper emphasizes that precise distinctions among artificial intelligence, machine learning, and deep learning are necessary for a coherent research foundation.
5
The review traces machine vision’s evolution from early image-processing algorithms to an interdisciplinary artificial intelligence field integrating computer science, mathematics, and robotics.

machine and robotic vision systems

the use and evolution of machine learning and deep learning for object detection, semantic segmentation, and human action recognition

Publication Details
Publication Date
2024-01-23
Journal
Publisher
ISSN
Cited by
282
Access Type
Author Information
Authors
Lazaros Moysis
Nikoleta Manakitsa
George S. Maraslidis
George F. Fragulis
Explore further
Open the scid.ai AI chat with a ready-made request: it will find papers on a similar topic and help build a literature review.
Find similar papers in the chat →
Make a presentation
100%