3 papers
cs.CL2026
Arithmetic OOD Failure Unfolds in Stages in Minimal GPTs
Seine A. Shintani
Arithmetic benchmarks are often reduced to a single held-out score, but that score can conflate qualitatively different failures. We study a controlled minimal GPT trained on exhau…
cs.AI2026
AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains
Seine A. Shintani
Generative AI is entering research, education, and professional work faster than current governance frameworks can specify how AI-assisted outputs should be judged in learning-inte…
cs.CY2026
Self-hosted Lecture-to-Quiz: Local LLM MCQ Generation with Deterministic Quality Control
Seine A. Shintani
We present an end-to-end self-hosted (API-free) pipeline, where API-free means that lecture content is not sent to any external LLM service, that converts lecture PDFs into multipl…