Implementasi Computer Vision dalam Klasifikasi Kalimat BISINDO Berbasis WebCam dengan Menggunakan Model YOLOv8 dan MediaPipe

Rimanda, Suhardianto (2026) Implementasi Computer Vision dalam Klasifikasi Kalimat BISINDO Berbasis WebCam dengan Menggunakan Model YOLOv8 dan MediaPipe. Diploma thesis, Politeknik Negeri Bengkalis.

[thumbnail of Abstract] Text (Abstract)
TA-6103230046-Abstract.pdf - Submitted Version
Available under License Creative Commons Attribution Non-commercial Share Alike.

Download (831kB)
[thumbnail of Bab I Pendahuluan] Text (Bab I Pendahuluan)
TA-6103230046-Bab I Pendahuluan.pdf - Submitted Version
Available under License Creative Commons Attribution Non-commercial Share Alike.

Download (809kB)
[thumbnail of Daftar Pustaka] Text (Daftar Pustaka)
TA-6103230046-Daftar Pustaka.pdf - Submitted Version
Available under License Creative Commons Attribution Non-commercial Share Alike.

Download (476kB)
[thumbnail of Full Text] Text (Full Text)
TA-6103230046-Full Text.pdf - Submitted Version
Restricted to Registered users only
Available under License Creative Commons Attribution Non-commercial Share Alike.

Download (17MB) | Request a copy

Abstract

The limited public understanding of Indonesian Sign Language (BISINDO) remains a communication barrier for deaf individuals. This research develops a Computer Vision-based BISINDO translation system that recognises sign words in real time through a webcam and arranges them into sentences. The system is built as a staged pipeline comprising YOLOv8 as a hand Region of Interest (ROI) detector, MediaPipe Holistic as a landmark extractor, LSTM and BiLSTM as word classifiers, and a rule-based Natural Language Processing (NLP) module as a sentence composer, implemented as a website-based prototype for one-way communication with text and speech output through gTTS. The dataset comprises 5,000 videos across 40 word classes, consisting of 2,000 public Kaggle videos and 3,000 self-recorded videos, divided at a ratio of 70:15:15. Testing was conducted through four schemes distinguished by signer composition and model architecture. In the single-signer scheme, both LSTM and BiLSTM achieved an accuracy of 99.64% with a macro F1-score of 0.996, while the scheme with a more diverse signer composition yielded 85.92% (LSTM) and 88.03% (BiLSTM). The YOLOv8n model achieved an mAP@50 of 0.9947 and an mAP@50–95 of 0.8593. Sentence composition testing achieved a success rate of 40%, with a system processing speed of 14 FPS. These results indicate that performance under the single-signer scheme does not yet reflect the model's generalisation capability; therefore, the accuracy of 88.03% is taken as the reference performance of the model in this research.

Item Type: Thesis (Diploma)
Uncontrolled Keywords: BISINDO, YOLOv8, LSTM, NLP, Computer Vision.
Subjects: 000 – UMUM, ILMU KOMPUTER, DAN INFORMASI > 005 – Pemrograman, Perangkat Lunak > 005.3 Perangkat Lunak (Software)
000 – UMUM, ILMU KOMPUTER, DAN INFORMASI > 005 – Pemrograman, Perangkat Lunak > 005.9 Kecerdasan Buatan (AI), Komputasi Kognitif
Divisions: Jurusan Teknik Informatika > Diploma Tiga (D-III) Teknik Informatika > TUGAS AKHIR
Depositing User: D-III Teknik Informatika 2023 Kelas B
Date Deposited: 30 Aug 2026 02:08
Last Modified: 30 Aug 2026 02:08
URI: https://eprints.polbeng.ac.id/id/eprint/7055

Actions (login required)

View Item
View Item