1 citations · 1 across the 3 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Gradient-Gated DPO: Stabilizing Preference Optimization in Language Models
Inoussa Mouiche
Preference optimization has become a central paradigm for aligning large language models with human feedback. Direct Preference Optimization (DPO) simplifies reinforcement learning…
cs.LG2026★ 1 cited
TIJERE: A Novel Threat Intelligence Joint Extraction Model Based on Analyst Expert Knowledge
Inoussa Mouiche, Sherif Saad
The extraction of entities and relationships from threat intelligence reports into structured formats, such as cybersecurity knowledge graphs, is essential for automated threat ana…