Computer vision is the field of artificial intelligence concerned with enabling computers to interpret and extract information from visual data such as images and video. Core tasks include image classification, object detection and localization, semantic and instance segmentation, facial recognition, and optical character recognition. Early approaches relied on hand-crafted feature descriptors combined with classical machine learning classifiers; current computer vision is dominated by convolutional neural networks and, increasingly, vision transformer architectures trained on large annotated image datasets. A notable 2026 development is the shift toward foundation models that displace task-specific training for many commercial applications, alongside growing use of agentic vision systems moving from research into operational deployment. Computer vision supports applications including autonomous vehicle perception, medical image analysis, industrial quality inspection, surveillance and security systems, and augmented reality. As an open-access computer vision journal, IJACSA publishes research on computer vision algorithms, model architectures, and applied vision systems evaluated on standard and domain-specific image datasets.
Published in International Journal of Advanced Computer Science and Applications (IJACSA)
· list last refreshed September 2026
Real-time recognition of loose fresh produce is a key requirement for intelligent retail weighing systems, enabling automated replacement of or assistance to manual PLU-based item selection. However, the deployment perfo…
Object detection in buffet-style environments is highly challenging due to densely stacked tableware, frequent occlusions, strong illumination reflections, and substantial visual similarity across categories, all of whic…
Vehicle gate access, in general, still relies heavily on manual inspection of identification cards and visual verification by security guards, which is slow, tedious, and susceptible to spoofing. Single-modality, compute…
The exponential proliferation of online gambling content represents a multifaceted challenge for contemporary automated content moderation systems, primarily driven by the sophisticated visual obfuscation and semantic co…
This study presents a quantitative approach to analyzing window opening and closing behaviors using skeletal recognition technology. Video data of five participants performing these actions were captured and processed us…
Skin diseases represent a global healthcare challenge because of their frequent occurrence and complex diagnosis. However, despite clinical advances, accurately identifying dermatological lesions remains difficult due to…
Real-time multi-class object detection on embedded devices poses significant challenges due to limited computational power, memory capacity, and energy efficiency requirements. Conventional high-precision object detector…
Efficient and accurate automated diagnosis of plant diseases remains a challenge for deployment on resource-constrained edge devices. While hybrid vision transformers like GCViT balance accuracy and efficiency, they ofte…
In recent years, as a critical pillar supporting the national economy and daily life, the safe and efficient operation of road traffic has highly relied on precise environmental perception capabilities. To address this,…
The rapid development of medical practices and imaging technology tools creates substantial growth in the amount of medical image data each year in our present era. This research aims to develop a hybrid approach that in…