most citedUnified Deployment-Aware Evaluation of Open Reasoning Language Models

1 citations · 1 across the 2 of their papers we have counts for

collaborators

8 papers

cs.CL2026

HypothesisMed: Inference-Time Answer Fusion and Structured Hypothesis-Space Reporting for Biomedical Question Answering

Md Motaleb Hossen Manik, Ge Wang

Biomedical question answering with large language models is commonly evaluated using answer accuracy, but answer accuracy alone does not indicate whether a model can produce parsea…

cs.CL20261 cited

Unified Deployment-Aware Evaluation of Open Reasoning Language Models

Md Motaleb Hossen Manik, Ge Wang

Open reasoning language models are often compared under mixed sample sizes, partially standardized prompts, and accuracy-centered summaries, which makes practical model selection d…

cs.ET2026

ADAPT: AI-Driven Decentralized Adaptive Publishing Testbed

Md Motaleb Hossen Manik, Ge Wang

Scholarly publishing faces increasingly strong stressors, including submission overload, reviewer fatigue, inconsistent evaluation, governance opacity, and vulnerability to manipul…

cs.CL2026

Emergent decentralized regulation in a purely synthetic society

Md Motaleb Hossen Manik, Ge Wang

As autonomous AI agents increasingly inhabit online environments and extensively interact, a key question is whether synthetic collectives exhibit self-regulated social dynamics wi…

cs.SI2026

OpenClaw Agents on Moltbook: Risky Instruction Sharing and Norm Enforcement in an Agent-Only Social Network

Md Motaleb Hossen Manik, Ge Wang

Agentic AI systems increasingly operate in shared social environments where they exchange information, instructions, and behavioral cues. However, little empirical evidence exists…

cs.CV2025

SlideChain: Semantic Provenance for Lecture Understanding via Blockchain Registration

Md Motaleb Hossen Manik, Md Zabirul Islam, Ge Wang

Modern vision--language models (VLMs) are increasingly used to interpret and generate educational content, yet their semantic outputs remain challenging to verify, reproduce, and a…