The Human Action Classification from Video is a Computer Vision project that combines Event detection and Export module, built with YOLO. The project follows a clean, modular pipeline where image input, processing, and presentation stay separated, making it easy to test, extend, and present. It showcases practical Computer Vision techniques while producing a working, demo-ready application.
Manual visual inspection and analysis of this problem is slow, error-prone, and impossible to scale. Without a Computer Vision approach built on Export module and YOLO, users cannot automatically detect, classify, or track visual patterns, and there is no reliable way to evaluate accuracy.
This project applies Computer Vision techniques through Event detection, orchestrated with YOLO and Export module. The pipeline is designed for accuracy and speed, with preprocessing, model inference, and clear evaluation. It produces consistent, reusable results and can be adapted to related vision tasks with minimal changes.
YOLO
MoviePy
Python 3.11+
OpenCV
Deep learning frameworks
GPU / Google Colab
TensorFlow
Modular vision pipeline around Event detection and Export module
Configurable model and preprocessing settings
Clear logging, metrics, and error handling
Clean interface for viewing results
Reusable components for related vision tasks
Real-time or batch inference support
The project is layered: the input layer loads and preprocesses images through Event detection; the inference layer runs the vision model with YOLO and Export module; and the output layer formats and presents results via Frame classification. Shared configuration, logging, and evaluation modules support all layers, keeping the system robust and easy to extend.
Set up the environment, datasets, and configuration files.
Build the image loading and preprocessing layer with Event detection.
Implement the core vision pipeline using YOLO and Export module.
Add the output and presentation layer via Frame classification.
Wire up end-to-end flows and add error handling and logging.
Train or tune the model, run evaluation, and refine results.
Package the project, document it, and prepare the demo and viva report.
Build production-style Computer Vision applications
Apply Summarizing video and Building surveillance analysis
Preprocess and analyze real image data
Work with popular vision libraries and models
Present and defend a complete Computer Vision project in viva
Expose the pipeline as a REST API for other apps
Add more classes and larger training data
Add edge deployment for mobile devices
Optimize inference for real-time speed
The Human Action Classification from Video delivers a complete Computer Vision workflow — from image input and processing to analysis and presentation. It is practical, modern, and easy to explain, making it an excellent final year project that demonstrates in-demand vision skills.