Benchmarking vision-language models for traffic scene understanding in South Asian traffic environments

dc.contributor.advisorShatabda, Swakkhar
dc.contributor.authorChoity, Naznin Sultana
dc.contributor.authorTakmim, Samiha
dc.contributor.authorRahman, Abida
dc.contributor.authorShaid, Abdullah Al
dc.contributor.authorHossain, Md.Tanzim
dc.date.accessioned2026-04-12T06:08:47Z
dc.date.available2026-04-12T06:08:47Z
dc.date.issued2025-12
dc.descriptionCataloged from PDF version of thesis.
dc.descriptionIncludes bibliographical references (pages 44-46).
dc.descriptionThis thesis is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science and Engineering, 2025.
dc.description.abstractA substantial body of research has investigated methods to promote safer driving, yet many challenges persist, particularly in environments with complex traffic patterns. This study focuses on how drivers evaluate their surroundings and make safety-critical deci- sions, with a specific emphasis on the role of deep learning in enhancing safe driving. Deep learning enables rapid identification of hazardous situations by providing real-time feedback and alert mechanisms, thereby improving driving behavior and reducing risks. The objective of this work is to develop an AI-powered driving assistance system de- signed to enhance road safety, especially for inexperienced drivers. Real-world driving conditions were incorporated by collecting YouTube footage across diverse road types and traffic densities. Experts annotated the dataset by labeling key objects and identifying context-specific risk factors, decision-making cues, and hazard indicators. The system is powered by a custom-designed AI model capable of providing context- aware guidance, regulatory reminders, and hazard alerts in real time. By leveraging expert-annotated data and multimodal deep learning techniques, the system delivers per- sonalized and immediate support to increase driver situational awareness, confidence, and safety.
dc.identifier.otherID 20301275
dc.identifier.otherID 21301222
dc.identifier.otherID 21201645
dc.identifier.otherID 21201131
dc.identifier.otherhttps://dspace.bracu.ac.bd/server/api/core/items/38dc069e-f86c-4adf-a88f-b80d852ab173
dc.identifier.urihttp://hdl.handle.net/10361/27853
dc.language.isoen
dc.publisherBRAC University
dc.sourceBRAC University Institutional Repository
dc.subjectSafe driving
dc.subjectVision-language models
dc.subjectDeep learning
dc.subjectDriving behavior
dc.titleBenchmarking vision-language models for traffic scene understanding in South Asian traffic environments
dc.typeThesis

Files

Original bundle

Now showing 1 - 1 of 1
Thumbnail Image
Name:
20301275_21301222_21201645_20201103 - NAZNIN SULTANA CHOITY.pdf
Size:
494.03 KB
Format:
Adobe Portable Document Format