11 citations · 14 across the 2 of their papers we have counts for
3 papers
cs.AI2024★ 3 cited
SciCode: A Research Coding Benchmark Curated by Scientists
Minyang Tian, Luyu Gao, Shizhuo Dylan Zhang +27
Since language models (LMs) now outperform average humans on many challenging tasks, it has become increasingly difficult to develop challenging, high-quality, and realistic evalua…
cs.SE2024
CodeMind: Evaluating Large Language Models for Code Reasoning
Changshu Liu, Yang Chen, Reyhaneh Jabbarvand
Large Language Models (LLMs) have been widely used to automate programming tasks. Their capabilities have been evaluated by assessing the quality of generated code through tests or…
cs.CL2021★ 11 cited
Pre-training Co-evolutionary Protein Representation via A Pairwise Masked Language Model
Liang He, Shizhuo Zhang, Lijun Wu +10
Understanding protein sequences is vital and urgent for biology, healthcare, and medicine. Labeling approaches are expensive yet time-consuming, while the amount of unlabeled data…