Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models
arXiv:2304.01852 · doi:10.1016/j.metrad.2023.100017
Abstract
This paper presents a comprehensive survey of ChatGPT-related (GPT-3.5 and GPT-4) research, state-of-the-art large language models (LLM) from the GPT series, and their prospective applications across diverse domains. Indeed, key innovations such as large-scale pre-training that captures knowledge across the entire world wide web, instruction fine-tuning and Reinforcement Learning from Human Feedback (RLHF) have played significant roles in enhancing LLMs' adaptability and performance. We performed an in-depth analysis of 194 relevant papers on arXiv, encompassing trend analysis, word cloud representation, and distribution analysis across various application domains. The findings reveal a significant and increasing interest in ChatGPT-related research, predominantly centered on direct natural language processing applications, while also demonstrating considerable potential in areas ranging from education and history to mathematics, medicine, and physics. This study endeavors to furnish insights into ChatGPT's capabilities, potential implications, ethical concerns, and offer direction for future advancements in this field.
21 pages, 4 figures, accepted by Meta-Radiology
References in corpus (53)
- Training language models to follow instructions with human feedback
- A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT
- ChatGPT: Jack of all trades, master of none
- The Role of AI in Drug Discovery: Challenges, Opportunities, and Strategies
- ChatGPT: The End of Online Exam Integrity?
- Is ChatGPT A Good Translator? Yes With GPT-4 As The Engine
- Mathematical Capabilities of ChatGPT
- Investigating the use of ChatGPT for the scheduling of construction projects
- NL4DV: A Toolkit for Generating Analytic Specifications for Data Visualization from Natural Language Queries
- Is ChatGPT better than Human Annotators? Potential and Limitations of ChatGPT in Explaining Implicit Hate Speech
- "I think this is the most disruptive technology": Exploring Sentiments of ChatGPT Early Adopters using Twitter Data
- The Death of the Short-Form Physics Essay in the Coming AI Revolution
- Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models
- How Generative AI models such as ChatGPT can be (Mis)Used in SPC Practice, Education, and Research? An Exploratory Study
- Could an Artificial-Intelligence agent pass an introductory physics course?
- ChatIE: Zero-Shot Information Extraction via Chatting with ChatGPT
- Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT
- Large Language Models Are State-of-the-Art Evaluators of Translation Quality
- A Categorical Archive of ChatGPT Failures
- Exploring the Limits of ChatGPT for Query or Aspect-based Text Summarization
- DeID-GPT: Zero-shot Medical Text De-Identification by GPT-4
- To Ship or Not to Ship: An Extensive Evaluation of Automatic Metrics for Machine Translation
- Does Synthetic Data Generation of LLMs Help Clinical Text Mining?
- ChatCAD: Interactive Computer-Aided Diagnosis on Medical Image using Large Language Models
- Learning gain differences between ChatGPT and human tutor generated algebra hints
- How would Stance Detection Techniques Evolve after the Launch of ChatGPT?
- ChatGPT: Beginning of an End of Manual Linguistic Data Annotation? Use Case of Automatic Genre Identification
- ChatGPT and Other Large Language Models as Evolutionary Engines for Online Interactive Collaborative Game Design
- Exploring the Feasibility of ChatGPT for Event Extraction
- The moral authority of ChatGPT
- Applying BERT and ChatGPT for Sentiment Analysis of Lyme Disease in Scientific Literature
- AI and the FCI: Can ChatGPT Project an Understanding of Introductory Physics?
- An Independent Evaluation of ChatGPT on Mathematical Word Problems (MWP)
- Will Affective Computing Emerge from Foundation Models and General AI? A First Evaluation on ChatGPT
- Conversational Automated Program Repair
- Linguistic ambiguity analysis in ChatGPT
- Exploring the Cognitive Dynamics of Artificial Intelligence in the Post-COVID-19 and Learning 3.0 Era: A Case Study of ChatGPT
- Personalisation within bounds: A risk taxonomy and policy framework for the alignment of large language models with personalised feedback
- Chat2VIS: Generating Data Visualisations via Natural Language using ChatGPT, Codex and GPT-3 Large Language Models
- Radiology-GPT: A Large Language Model for Radiology
- Advancing Medical Imaging with Language Models: A Journey from N-grams to ChatGPT
- Paraphrase Identification with Deep Learning: A Review of Datasets and Methods
- Causal-Discovery Performance of ChatGPT in the context of Neuropathic Pain Diagnosis
- UZH_CLyp at SemEval-2023 Task 9: Head-First Fine-Tuning and ChatGPT Data Generation for Cross-Lingual Learning in Tweet Intimacy Prediction
- AD-AutoGPT: An Autonomous GPT for Alzheimer's Disease Infodemiology
- Chatbots in a Honeypot World
- Frustratingly Easy Transferability Estimation
- ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language Models
- Numeracy from Literacy: Data Science as an Emergent Skill from Large Language Models
- nvBench: A Large-Scale Synthesized Dataset for Cross-Domain Natural Language to Visualization Task
- AI Insights into Theoretical Physics and the Swampland Program: A Journey Through the Cosmos with ChatGPT
- A Pilot Evaluation of ChatGPT and DALL-E 2 on Decision Making and Spatial Reasoning
- Can Large Language Models Change User Preference Adversarially?
Cited by in corpus (23)
- Large language models surpass human experts in predicting neuroscience results
- From COBIT to ISO 42001: Evaluating Cybersecurity Frameworks for Opportunities, Risks, and Regulatory Compliance in Commercializing Large Language Models
- Differentiate ChatGPT-generated and Human-written Medical Texts
- An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
- Artificial General Intelligence for Medical Imaging Analysis
- Exploring the Frontiers of LLMs in Psychological Applications: A Comprehensive Review
- LLM-BT: Performing Robotic Adaptive Tasks based on Large Language Models and Behavior Trees
- Open-TI: Open Traffic Intelligence with Augmented Language Model
- Automated Review Generation Method Based on Large Language Models
- Beyond Traditional Teaching: The Potential of Large Language Models and Chatbots in Graduate Engineering Education
- The Use of Generative Artificial Intelligence for Upper Secondary Mathematics Education Through the Lens of Technology Acceptance
- Controlling AI Agent Participation in Group Conversations: A Human-Centered Approach
- Human-interpretable clustering of short-text using large language models
- Health Text Simplification: An Annotated Corpus for Digestive Cancer Education and Novel Strategies for Reinforcement Learning
- Stylometry recognizes human and LLM-generated texts in short samples
- Gender Bias Detection in Court Decisions: A Brazilian Case Study
- Language to Map: Topological map generation from natural language path instructions
- Towards a Taxonomy of Large Language Model based Business Model Transformations
- The Evolution and Future Perspectives of Artificial Intelligence Generated Content
- Rewriting Conversational Utterances with Instructed Large Language Models
- Machine Learning for maximizing the memristivity of single and coupled quantum memristors
- Quantum-Inspired Weight-Constrained Neural Network: Reducing Variable Numbers by 100x Compared to Standard Neural Networks
- Contextualized AI for Cyber Defense: An Automated Survey using LLMs