Abstract
Abstract
Fast detection and analysis of traffic crashes are important steps toward improving road safety. While traditional data sources remain useful, social media platforms have become rich and rapidly growing channels for tracking such incidents. Still, the informal and unstructured nature of user-generated content makes accurate information extraction a challenge. This study explores the potential of using social media data mining and Large Language Models (LLMs) to automatically detect traffic crashes and extract relevant details from social media messages. The research was centered on Damavand County, a high-traffic area located on major routes near Tehran. A full framework was developed, including data collection, preprocessing, and fine-tuning of BERT-based models to classify messages into crash and non-crash categories and further categorize them into ten crash types, reaching accuracies of 91.1% and 89.7%, respectively. Additionally, a LLaMA 3.1 model was fine-tuned for a question-answering task focused on extracting crash location and casualty numbers. It achieved 97% accuracy in identifying fatalities and injuries, and BLEU and METEOR scores of 0.697 and 0.864 in location extraction. Comparison with official records showed that social media data identified 64.1% of crash hotspots found in official reports. Moreover, crashes with higher severity and specific types, like rollover or fall crashes, were more frequently reported. These findings highlight the practical value of social media as a supplementary data source and show how LLMs can facilitate this process of extracting useful crash-related information from social media.
Direct answer
What can I do from this paper page?
Use this page to scan "Extracting traffic crash information from social media: an LLM-based approach" quickly: start with the summary and abstract, then check the authors, source, topics, and related papers. From here, open Scollr to follow Sentiment Analysis and Opinion Mining research, save the paper, or map adjacent work.
Research areas
Follow related topics
Citation
BibTeX
@article{Khavas2026Extracting,
title = {Extracting traffic crash information from social media: an LLM-based approach},
author = {Reza Golshan Khavas and Mohammadreza Nosrati},
journal = {Transportation Letters},
year = {2026},
doi = {10.1080/19427867.2026.2681104},
url = {https://doi.org/10.1080/19427867.2026.2681104}
}
FAQ
Using this paper in a discovery workflow
How do I find related work for this paper?
Use the related papers and topic links on this page as starting points. In Scollr, you can also open the paper and build a literature map around its references, citing papers, and related work.
How can I keep up with new Sentiment Analysis and Opinion Mining research papers?
Follow Sentiment Analysis and Opinion Mining research in Scollr. New papers from the topic flow into a personalized feed, and you can save useful studies to revisit later.
Can I cite this paper from this page?
This page includes a static BibTeX block for Extracting traffic crash information from social media: an LLM-based approach. Always verify the DOI, source, and publication details against the publisher record before submitting a manuscript.
Follow this research in Scollr
Follow the topics and authors behind this paper, save useful studies, and build a literature map when you are ready to go deeper.
Get the app