10 papers
HypothesisMed: Inference-Time Answer Fusion and Structured Hypothesis-Space Reporting for Biomedical Question Answering
Md Motaleb Hossen Manik, Ge Wang
Biomedical question answering with large language models is commonly evaluated using answer accuracy, but answer accuracy alone does not indicate whether a model can produce parsea…
CourseBlueprint: A Structured Pipeline for Adaptive Pedagogical Video Generation Grounded in Course Corpora
Md Zabirul Islam, Md Motaleb Hossen Manik, Ge Wang
Generative text-to-video systems can produce visually fluent educational clips, but they rarely encode the pedagogical content knowledge (PCK) needed for effective instruction, inc…
Unified Deployment-Aware Evaluation of Open Reasoning Language Models
Md Motaleb Hossen Manik, Ge Wang
Open reasoning language models are often compared under mixed sample sizes, partially standardized prompts, and accuracy-centered summaries, which makes practical model selection d…
ADAPT: AI-Driven Decentralized Adaptive Publishing Testbed
Md Motaleb Hossen Manik, Ge Wang
Scholarly publishing faces increasingly strong stressors, including submission overload, reviewer fatigue, inconsistent evaluation, governance opacity, and vulnerability to manipul…
Emergent decentralized regulation in a purely synthetic society
Md Motaleb Hossen Manik, Ge Wang
As autonomous AI agents increasingly inhabit online environments and extensively interact, a key question is whether synthetic collectives exhibit self-regulated social dynamics wi…
OpenClaw Agents on Moltbook: Risky Instruction Sharing and Norm Enforcement in an Agent-Only Social Network
Md Motaleb Hossen Manik, Ge Wang
Agentic AI systems increasingly operate in shared social environments where they exchange information, instructions, and behavioral cues. However, little empirical evidence exists…