14 citations · 15 across the 3 of their papers we have counts for
3 papers
cs.LG2025
What Characterizes Effective Reasoning? Revisiting Length, Review, and Structure of CoT
Yunzhen Feng, Julia Kempe, Cheng Zhang +2
Large reasoning models (LRMs) spend substantial test-time compute on long chain-of-thought (CoT) traces, but what *characterizes* an effective CoT remains unclear. While prior work…
cs.LG2023★ 1 cited
High Precision Causal Model Evaluation with Conditional Randomization
Chao Ma, Cheng Zhang
The gold standard for causal model evaluation involves comparing model predictions with true effects estimated from randomized controlled trials (RCT). However, RCTs are not always…
cs.LG2023★ 14 cited
Understanding Causality with Large Language Models: Feasibility and Opportunities
Cheng Zhang, Stefan Bauer, Paul Bennett +8
We assess the ability of large language models (LLMs) to answer causal questions by analyzing their strengths and weaknesses against three types of causal question. We believe that…