LLM-Based Financial Sentiment Analysis in Arabic: Evidence from Saudi Markets
Summary
This paper presents a framework for Arabic financial sentiment analysis using LLMs, tailored for the Saudi market, integrating news and social media data to capture investor sentiment.
View Cached Full Text
Cached at: 05/20/26, 08:26 AM
# LLM-Based Financial Sentiment Analysis in Arabic: Evidence from Saudi Markets Source: [https://arxiv.org/abs/2605.19714](https://arxiv.org/abs/2605.19714) [View PDF](https://arxiv.org/pdf/2605.19714) > Abstract:Investor sentiment shapes financial markets, yet modeling sentiment in Arabic financial contexts remains challenging due to linguistic complexity and limited resources\. We present an Arabic NLP framework for large\-scale financial sentiment analysis tailored to the Saudi market, integrating official financial news and social media to capture institutional and public investor sentiment\. The framework constructs a large Arabic financial corpus through a multi\-stage pipeline encompassing data collection, cleaning, deduplication, entity linking, and sentiment annotation\. Transformer\-based NER combined with a curated company lexicon links textual mentions to canonical company identifiers, with sentiment labels assigned using a five\-class scheme\. The resulting dataset of 84K samples supports company\-level sentiment aggregation and analysis of sentiment dynamics relative to stock market behavior on the Saudi Exchange\. Experimental results demonstrate reliable and scalable Arabic financial sentiment analysis\. ## Submission history From: Enrico Lopedoto \[[view email](https://arxiv.org/show-email/fa62f7bd/2605.19714)\] **\[v1\]**Tue, 19 May 2026 11:50:33 UTC \(563 KB\)
Similar Articles
Enhancing Financial Sentiment Analysis via Retrieval Augmented Large Language Models
This paper introduces a retrieval-augmented LLM framework for financial sentiment analysis, achieving 15-48% improvement in accuracy and F1 score over traditional models and LLMs like ChatGPT and LLaMA.
SAHM: A Benchmark for Arabic Financial and Shari'ah-Compliant Reasoning
Researchers release SAHM, the first Arabic financial benchmark with 14,380 expert-verified instances covering Shari’ah-compliant reasoning, showing large performance gaps for 20 evaluated LLMs.
Spam and Sentiment Detection in Arabic Tweets Using MARBERT Model
This paper presents a sentiment analysis and spam detection system for Arabic tweets using the MARBERT model, trained on a dataset of 24,513 tweets to improve customer service for Saudi Telecom Company.
Automated Scoring of Arabic Text Using Large Language Models: A Literature Review
A literature review examining LLM-based approaches for automatic scoring of Arabic text, covering short answer grading and essay scoring, with a proposed taxonomy and comparative analysis.
Benchmarking Frontier LLMs on Arabic Cultural and Sociolinguistic Knowledge: A Cross-Evaluation Framework with Human SME Ground Truth
This paper introduces a cross-evaluation framework for benchmarking LLMs on Arabic cultural and sociolinguistic knowledge, using human SME ground truth and automated judges. The authors contribute a dataset of prompt-rubric pairs for Egyptian and Iraqi Arabic, evaluating frontier LLMs and finding that cultural reasoning remains a primary failure mode for automated grading.