3 papers
cs.CL2024
PPTC-R benchmark: Towards Evaluating the Robustness of Large Language Models for PowerPoint Task Completion
Zekai Zhang, Yiduo Guo, Yaobo Liang +2
The growing dependence on Large Language Models (LLMs) for finishing user instructions necessitates a comprehensive understanding of their robustness to complex task completion in…
cs.CL2023
PPTC Benchmark: Evaluating Large Language Models for PowerPoint Task Completion
Yiduo Guo, Zekai Zhang, Yaobo Liang +2
Recent evaluations of Large Language Models (LLMs) have centered around testing their zero-shot/few-shot capabilities for basic natural language tasks and their ability to translat…
astro-ph.IM2023
Pulsar Candidate Classification Using A Computer Vision Method Combining with Convolution and Attention
NanNan Cai, JinLin Han, WeiCong Jing +3
Artificial intelligence methods are indispensable to identifying pulsars from large amounts of candidates. We develop a new pulsar identification system that utilizes the CoAtNet t…