FANAL -- Financial Activity News Alerting Language Modeling Framework
arXiv:2412.03527 · doi:10.1109/BigData62323.2024.10825891
Abstract
In the rapidly evolving financial sector, the accurate and timely interpretation of market news is essential for stakeholders needing to navigate unpredictable events. This paper introduces FANAL (Financial Activity News Alerting Language Modeling Framework), a specialized BERT-based framework engineered for real-time financial event detection and analysis, categorizing news into twelve distinct financial categories. FANAL leverages silver-labeled data processed through XGBoost and employs advanced fine-tuning techniques, alongside ORBERT (Odds Ratio BERT), a novel variant of BERT fine-tuned with ORPO (Odds Ratio Preference Optimization) for superior class-wise probability calibration and alignment with financial event relevance. We evaluate FANAL's performance against leading large language models, including GPT-4o, Llama-3.1 8B, and Phi-3, demonstrating its superior accuracy and cost efficiency. This framework sets a new standard for financial intelligence and responsiveness, significantly outstripping existing models in both performance and affordability.
Accepted for the IEEE International Workshop on Large Language Models for Finance, 2024. This is a preprint version
References in corpus (13)
- Adam: A Method for Stochastic Optimization
- XGBoost: A Scalable Tree Boosting System
- LoRA: Low-Rank Adaptation of Large Language Models
- Gemini: A Family of Highly Capable Multimodal Models
- BloombergGPT: A Large Language Model for Finance
- Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning
- Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
- Compressing Large-Scale Transformer-Based Models: A Case Study on BERT
- KTO: Model Alignment as Prospect Theoretic Optimization
- Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation
- CANAL -- Cyber Activity News Alerting Language Model: Empirical Approach vs. Expensive LLM
- SimPO: Simple Preference Optimization with a Reference-Free Reward
- RoSA: Accurate Parameter-Efficient Fine-Tuning via Robust Adaptation