12 citations · 13 across the 4 of their papers we have counts for
5 papers · 1 filter
OpenAI GPT-5 System Card
Aaditya Singh, Adam Fry, Adam Perelman +483
This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers most questions, a deeper reason…
Rubric-Conditioned LLM Grading: Alignment, Uncertainty, and Robustness
Haotian Deng, Chris Farber, Jiyoon Lee +1
Automated short-answer grading (ASAG) remains a challenging task due to the linguistic variability of student responses and the need for nuanced, rubric-aligned partial credit. Whi…
E-EVAL: A Comprehensive Chinese K-12 Education Evaluation Benchmark for Large Language Models
Jinchang Hou, Chang Ao, Haihong Wu +8
With the accelerating development of Large Language Models (LLMs), many LLMs are beginning to be used in the Chinese K-12 education domain. The integration of LLMs and education is…
ICE-GRT: Instruction Context Enhancement by Generative Reinforcement based Transformers
Chen Zheng, Ke Sun, Da Tang +4
The emergence of Large Language Models (LLMs) such as ChatGPT and LLaMA encounter limitations in domain-specific tasks, with these models often lacking depth and accuracy in specia…
Balancing Specialized and General Skills in LLMs: The Impact of Modern Tuning and Data Strategy
Zheng Zhang, Chen Zheng, Da Tang +5
This paper introduces a multifaceted methodology for fine-tuning and evaluating large language models (LLMs) for specialized monetization tasks. The goal is to balance general lang…