Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
arXiv:2201.11903
Abstract
We explore how generating a chain of thought -- a series of intermediate reasoning steps -- significantly improves the ability of large language models to perform complex reasoning. In particular, we show how such reasoning abilities emerge naturally in sufficiently large language models via a simple method called chain of thought prompting, where a few chain of thought demonstrations are provided as exemplars in prompting. Experiments on three large language models show that chain of thought prompting improves performance on a range of arithmetic, commonsense, and symbolic reasoning tasks. The empirical gains can be striking. For instance, prompting a 540B-parameter language model with just eight chain of thought exemplars achieves state of the art accuracy on the GSM8K benchmark of math word problems, surpassing even finetuned GPT-3 with a verifier.
References in corpus (1)
Cited by in corpus (306)
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
- A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
- Graph of Thoughts: Solving Elaborate Problems with Large Language Models
- Recommender Systems in the Era of Large Language Models (LLMs)
- Opportunities and Challenges for ChatGPT and Large Language Models in Biomedicine and Health
- A Survey on ChatGPT: AI-Generated Contents, Challenges, and Solutions
- Unleashing the potential of prompt engineering for large language models
- Towards Human-centered Explainable AI: A Survey of User Studies for Model Explanations
- Matching Patients to Clinical Trials with Large Language Models
- Thinking Fast and Slow in Large Language Models
- Large AI Models in Health Informatics: Applications, Challenges, and the Future
- Human-Like Intuitive Behavior and Reasoning Biases Emerged in Language Models -- and Disappeared in GPT-4
- Explainable Artificial Intelligence: A Survey of Needs, Techniques, Applications, and Future Direction
- Design Principles for Generative AI Applications
- Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection
- AI Agents vs. Agentic AI: A Conceptual Taxonomy, Applications and Challenges
- GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation
- AI Agents Under Threat: A Survey of Key Security Challenges and Future Pathways
- Large Language Models and the Reverse Turing Test
- Evaluation of Retrieval-Augmented Generation: A Survey
- Thrilled by Your Progress! Large Language Models (GPT-4) No Longer Struggle to Pass Assessments in Higher Education Programming Courses
- Large language models surpass human experts in predicting neuroscience results
- Large Language Models as Zero-Shot Conversational Recommenders
- Towards Interpretable Mental Health Analysis with Large Language Models
- Harms from Increasingly Agentic Algorithmic Systems
- Lost in Translation: A Study of Bugs Introduced by Large Language Models while Translating Code
- ChatDiet: Empowering Personalized Nutrition-Oriented Food Recommender Chatbots through an LLM-Augmented Framework
- ChartGPT: Leveraging LLMs to Generate Charts from Abstract Natural Language
- PromptCharm: Text-to-Image Generation through Multi-modal Prompting and Refinement
- How understanding large language models can inform the use of ChatGPT in physics education
- Large Language Model Is Not a Good Few-shot Information Extractor, but a Good Reranker for Hard Samples!
- A Quantitative and Qualitative Evaluation of LLM-Based Explainable Fault Localization
- Towards autonomous system: flexible modular production system enhanced with large language model agents
- Evaluation of large language models for discovery of gene set function
- What Makes Good In-context Demonstrations for Code Intelligence Tasks with LLMs?
- Deception Abilities Emerged in Large Language Models
- GIT-Mol: A Multi-modal Large Language Model for Molecular Science with Graph, Image, and Text
- EvalLM: Interactive Evaluation of Large Language Model Prompts on User-Defined Criteria
- CLIP in Medical Imaging: A Survey
- UAVs Meet LLMs: Overviews and Perspectives Toward Agentic Low-Altitude Mobility
- The Impact of AI in Physics Education: A Comprehensive Review from GCSE to University Levels
- Using LLMs in Software Requirements Specifications: An Empirical Evaluation
- Theory of Mind for Multi-Agent Collaboration via Large Language Models
- Assessing the Ability of ChatGPT to Screen Articles for Systematic Reviews
- Enhancing Phenotype Recognition in Clinical Notes Using Large Language Models: PhenoBCBERT and PhenoGPT
- LaMI: Large Language Models for Multi-Modal Human-Robot Interaction
- Black-Box Access is Insufficient for Rigorous AI Audits
- Assessing Student Errors in Experimentation Using Artificial Intelligence and Large Language Models: A Comparative Study with Human Raters
- Real-World Robot Applications of Foundation Models: A Review
- A Large Language Model Approach to Educational Survey Feedback Analysis
- LLM in the Shell: Generative Honeypots
- The effect of source disclosure on evaluation of AI-generated messages: A two-part study
- LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems
- ChaCha: Leveraging Large Language Models to Prompt Children to Share Their Emotions about Personal Events
- GPT-4 can pass the Korean National Licensing Examination for Korean Medicine Doctors
- A Practical Survey on Zero-shot Prompt Design for In-context Learning
- Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language Models
- Automated Educational Question Generation at Different Bloom's Skill Levels using Large Language Models: Strategies and Evaluation
- Augmenting Interpretable Models with LLMs during Training
- Writer-Defined AI Personas for On-Demand Feedback Generation
- The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification
- A Taxonomy for Human-LLM Interaction Modes: An Initial Exploration
- AppPoet: Large Language Model based Android malware detection via multi-view prompt engineering
- Playing repeated games with Large Language Models
- ChatScratch: An AI-Augmented System Toward Autonomous Visual Programming Learning for Children Aged 6-12
- HPC-GPT: Integrating Large Language Model for High-Performance Computing
- Generative Speech Recognition Error Correction with Large Language Models and Task-Activating Prompting
- Generative Expressive Robot Behaviors using Large Language Models
- In-Context Operator Learning with Data Prompts for Differential Equation Problems
- Understanding Users' Dissatisfaction with ChatGPT Responses: Types, Resolving Tactics, and the Effect of Knowledge Level
- Enhancing Knowledge Retrieval with In-Context Learning and Semantic Search through Generative AI
- ThoughtSource: A central hub for large language model reasoning data
- El Agente: An Autonomous Agent for Quantum Chemistry
- "It's not like Jarvis, but it's pretty close!" -- Examining ChatGPT's Usage among Undergraduate Students in Computer Science
- Generative Relevance Feedback with Large Language Models
- Improving Steering and Verification in AI-Assisted Data Analysis with Interactive Task Decomposition
- Towards Efficient Generative Large Language Model Serving: A Survey from Algorithms to Systems
- PlantoGraphy: Incorporating Iterative Design Process into Generative Artificial Intelligence for Landscape Rendering
- OpenFOAMGPT: a RAG-Augmented LLM Agent for OpenFOAM-Based Computational Fluid Dynamics
- FoodSAM: Any Food Segmentation
- Selenite: Scaffolding Online Sensemaking with Comprehensive Overviews Elicited from Large Language Models
- Automated Data Visualization from Natural Language via Large Language Models: An Exploratory Study
- Exploring the Roles of Large Language Models in Reshaping Transportation Systems: A Survey, Framework, and Roadmap
- Concept Induction: Analyzing Unstructured Text with High-Level Concepts Using LLooM
- OntoChatGPT Information System: Ontology-Driven Structured Prompts for ChatGPT Meta-Learning
- From Intention To Implementation: Automating Biomedical Research via LLMs
- A social path to human-like artificial intelligence
- Generative Students: Using LLM-Simulated Student Profiles to Support Question Item Evaluation
- Incremental Learning of Humanoid Robot Behavior from Natural Interaction and Large Language Models
- Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
- From Screens to Scenes: A Survey of Embodied AI in Healthcare
- Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants
- Developing ChemDFM as a large language foundation model for chemistry
- Making LLMs Worth Every Penny: Resource-Limited Text Classification in Banking
- TypeDance: Creating Semantic Typographic Logos from Image through Personalized Generation
- LessonPlanner: Assisting Novice Teachers to Prepare Pedagogy-Driven Lesson Plans with Large Language Models
- What Should We Engineer in Prompts? Training Humans in Requirement-Driven LLM Use
- RELIC: Investigating Large Language Model Responses using Self-Consistency
- Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations
- WikiChat: Stopping the Hallucination of Large Language Model Chatbots by Few-Shot Grounding on Wikipedia
- Large Language Models Can be Lazy Learners: Analyze Shortcuts in In-Context Learning
- REFINER: Reasoning Feedback on Intermediate Representations
- Evaluating Search Engines and Large Language Models for Answering Health Questions
- Prompted LLMs as Chatbot Modules for Long Open-domain Conversation
- A dataset and benchmark for hospital course summarization with adapted large language models
- Large-Scale Text Analysis Using Generative Language Models: A Case Study in Discovering Public Value Expressions in AI Patents
- Grounding and Evaluation for Large Language Models: Practical Challenges and Lessons Learned (Survey)
- Traces of Memorisation in Large Language Models for Code
- Stance Detection: A Practical Guide to Classifying Political Beliefs in Text
- Two Directions for Clinical Data Generation with Large Language Models: Data-to-Label and Label-to-Data
- Applications of Generative AI in Healthcare: algorithmic, ethical, legal and societal considerations
- Fine-tuning a Large Language Model for Automating Computational Fluid Dynamics Simulations
- Exploring the Sensitivity of LLMs' Decision-Making Capabilities: Insights from Prompt Variation and Hyperparameters
- Opportunities for Large Language Models and Discourse in Engineering Design
- T cell receptor binding prediction: A machine learning revolution
- Can neural networks do arithmetic? A survey on the elementary numerical skills of state-of-the-art deep learning models
- TR2MTL: LLM based framework for Metric Temporal Logic Formalization of Traffic Rules
- Exploring the Potential of Large Language Models for Improving Digital Forensic Investigation Efficiency
- Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing
- Re-Scoring Using Image-Language Similarity for Few-Shot Object Detection
- I-Design: Personalized LLM Interior Designer
- Designing for Human-Agent Alignment: Understanding what humans want from their agents
- Large Knowledge Model: Perspectives and Challenges
- ReviewFlow: Intelligent Scaffolding to Support Academic Peer Reviewing
- Evaluating Large Language Models in Process Mining: Capabilities, Benchmarks, and Evaluation Strategies
- The Human-GenAI Value Loop in Human-Centered Innovation: Beyond the Magical Narrative
- An Empathy-Based Sandbox Approach to Bridge the Privacy Gap among Attitudes, Goals, Knowledge, and Behaviors
- DR.BENCH: Diagnostic Reasoning Benchmark for Clinical Natural Language Processing
- Human I/O: Towards a Unified Approach to Detecting Situational Impairments
- Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory
- OMPGPT: A Generative Pre-trained Transformer Model for OpenMP
- ScatterShot: Interactive In-context Example Curation for Text Transformation
- AI-Driven Day-to-Day Route Choice
- Reasoning Beyond Limits: Advances and Open Problems for LLMs
- VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning
- Reliable Natural Language Understanding with Large Language Models and Answer Set Programming
- Comparing Traditional and LLM-based Search for Image Geolocation
- Prompt Perturbation in Retrieval-Augmented Generation based Large Language Models
- Query Performance Prediction using Relevance Judgments Generated by Large Language Models
- AI-Assisted Causal Pathway Diagram for Human-Centered Design
- DeTiME: Diffusion-Enhanced Topic Modeling using Encoder-decoder based LLM
- Knowledge-aware Alert Aggregation in Large-scale Cloud Systems: a Hybrid Approach
- Exploring the applicability of Large Language Models to citation context analysis
- Opportunities and Challenges of Generative-AI in Finance
- Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
- Joint Estimation and Prediction of City-wide Delivery Demand: A Large Language Model Empowered Graph-based Learning Approach
- Natural Language Dataset Generation Framework for Visualizations Powered by Large Language Models
- CIPHER: Cybersecurity Intelligent Penetration-testing Helper for Ethical Researcher
- Potential Benefits of Employing Large Language Models in Research in Moral Education and Development
- Beyond Accuracy: Investigating Error Types in GPT-4 Responses to USMLE Questions
- Img2Loc: Revisiting Image Geolocalization using Multi-modality Foundation Models and Image-based Retrieval-Augmented Generation
- The Impact of Example Selection in Few-Shot Prompting on Automated Essay Scoring Using GPT Models
- Prompting for products: Investigating design space exploration strategies for text-to-image generative models
- Solving Math Word Problems via Cooperative Reasoning induced Language Models
- Laboratory-Scale AI: Open-Weight Models are Competitive with ChatGPT Even in Low-Resource Settings
- Computational Argumentation-based Chatbots: a Survey
- Large Language Model-driven Meta-structure Discovery in Heterogeneous Information Network
- Compositional API Recommendation for Library-Oriented Code Generation
- How Powerful are Decoder-Only Transformer Neural Models?
- Out of the Cage: How Stochastic Parrots Win in Cyber Security Environments
- RSTeller: Scaling Up Visual Language Modeling in Remote Sensing with Rich Linguistic Semantics from Openly Available Data and Large Language Models
- ShennongAlpha: an AI-driven sharing and collaboration platform for intelligent curation, acquisition, and translation of natural medicinal material knowledge
- Large Language Models Are Unreliable for Cyber Threat Intelligence
- DeepSeek-V3, GPT-4, Phi-4, and LLaMA-3.3 generate correct code for LoRaWAN-related engineering tasks
- Do PLMs Know and Understand Ontological Knowledge?
- MSCoTDet: Language-driven Multi-modal Fusion for Improved Multispectral Pedestrian Detection
- Large Language Models for Scholarly Ontology Generation: An Extensive Analysis in the Engineering Field
- CodEv: An Automated Grading Framework Leveraging Large Language Models for Consistent and Constructive Feedback
- Moderating New Waves of Online Hate with Chain-of-Thought Reasoning in Large Language Models
- Power-LLaVA: Large Language and Vision Assistant for Power Transmission Line Inspection
- Automating Traffic Model Enhancement with AI Research Agent
- ConceptThread: Visualizing Threaded Concepts in MOOC Videos
- All in How You Ask for It: Simple Black-Box Method for Jailbreak Attacks
- LLM Agent Framework for Intelligent Change Analysis in Urban Environment using Remote Sensing Imagery
- LLM-Mediated Domain-Specific Voice Agents: The Case of TextileBot
- Unleashing the Power of Large Language Model for Denoising Recommendation
- The Life Cycle of Knowledge in Big Language Models: A Survey
- Exploring the Landscape of Natural Language Processing Research
- IFShip: Interpretable Fine-grained Ship Classification with Domain Knowledge-Enhanced Vision-Language Models
- VLATest: Testing and Evaluating Vision-Language-Action Models for Robotic Manipulation
- Domain-Specific Improvement on Psychotherapy Chatbot Using Assistant
- Large Language Models to the Rescue: Reducing the Complexity in Scientific Workflow Development Using ChatGPT
- Keeping Users Engaged During Repeated Administration of the Same Questionnaire: Using Large Language Models to Reliably Diversify Questions
- A Meta-Evaluation of Faithfulness Metrics for Long-Form Hospital-Course Summarization
- BiosERC: Integrating Biography Speakers Supported by LLMs for ERC Tasks
- Agents for self-driving laboratories applied to quantum computing
- PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning
- "Cold, Calculated, and Condescending": How AI Identifies and Explains Ableism Compared to Disabled People
- Transforming Agency. On the mode of existence of Large Language Models
- Eight challenges in developing theory of intelligence
- A Knowledge-Informed Deep Learning Paradigm for Generalizable and Stability-Optimized Car-Following Models
- ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning
- A recent evaluation on the performance of LLMs on radiation oncology physics using questions of randomly shuffled options
- Democratizing LLMs: An Exploration of Cost-Performance Trade-offs in Self-Refined Open-Source Models
- Practically implementing an LLM-supported collaborative vulnerability remediation process: a team-based approach
- On LLM Wizards: Identifying Large Language Models' Behaviors for Wizard of Oz Experiments
- Jointly Extracting Interventions, Outcomes, and Findings from RCT Reports with LLMs
- Leveraging Prompt-Based Large Language Models: Predicting Pandemic Health Decisions and Outcomes Through Social Media Language
- Large Language Models for Extrapolative Modeling of Manufacturing Processes
- Beyond Self-Consistency: Ensemble Reasoning Boosts Consistency and Accuracy of LLMs in Cancer Staging
- A Study of Situational Reasoning for Traffic Understanding
- LLM-based event abstraction and integration for IoT-sourced logs
- Enhancing Supermarket Robot Interaction: A Multi-Level LLM Conversational Interface for Handling Diverse Customer Intents
- GPT Struct Me: Probing GPT Models on Narrative Entity Extraction
- Human Evaluation of Procedural Knowledge Graph Extraction from Text with Large Language Models
- Criteria-Based LLM Relevance Judgments
- Can ChatGPT Perform Reasoning Using the IRAC Method in Analyzing Legal Scenarios Like a Lawyer?
- Leveraging Large Language Model-based Room-Object Relationships Knowledge for Enhancing Multimodal-Input Object Goal Navigation
- Web Archives Metadata Generation with GPT-4o: Challenges and Insights
- Advanced System Integration: Analyzing OpenAPI Chunking for Retrieval-Augmented Generation
- Visual Language Models as Operator Agents in the Space Domain
- Paraphrase Types for Generation and Detection
- Using AI Large Language Models for Grading in Education: A Hands-On Test for Physics
- Enhancing Hepatopathy Clinical Trial Efficiency: A Secure, Large Language Model-Powered Pre-Screening Pipeline
- GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text
- Prompt engineering for bibliographic web-scraping
- Context-Enhanced Language Models for Generating Multi-Paper Citations
- Complex QA and language models hybrid architectures, Survey
- "Oh LLM, I'm Asking Thee, Please Give Me a Decision Tree": Zero-Shot Decision Tree Induction and Embedding with Large Language Models
- Semantic-Enhanced Indirect Call Analysis with Large Language Models
- FathomGPT: A Natural Language Interface for Interactively Exploring Ocean Science Data
- Face4RAG: Factual Consistency Evaluation for Retrieval Augmented Generation in Chinese
- Arithmetic with Language Models: from Memorization to Computation
- Relational Programming with Foundation Models
- Concept-Guided Chain-of-Thought Prompting for Pairwise Comparison Scoring of Texts with Large Language Models
- MoC-System: Efficient Fault Tolerance for Sparse Mixture-of-Experts Model Training
- Evaluating Transformer Models for Suicide Risk Detection on Social Media
- AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models
- FDM-Bench: A Comprehensive Benchmark for Evaluating Large Language Models in Additive Manufacturing Tasks
- Integrating Chain-of-Thought and Retrieval Augmented Generation Enhances Rare Disease Diagnosis from Clinical Notes
- InterChat: Enhancing Generative Visual Analytics using Multimodal Interactions
- Developmental Scaffolding with Large Language Models
- OPT-R: Exploring the Role of Explanations in Finetuning and Prompting for Reasoning Skills of Large Language Models
- Question Suggestion for Conversational Shopping Assistants Using Product Metadata
- Are Human Rules Necessary? Generating Reusable APIs with CoT Reasoning and In-Context Learning
- Evaluating Language Model Agency through Negotiations
- The Strong Pull of Prior Knowledge in Large Language Models and Its Impact on Emotion Recognition
- Surfacing Variations to Calibrate Perceived Reliability of MLLM-generated Image Descriptions
- NLP4Gov: A Comprehensive Library for Computational Policy Analysis
- Deep Natural Language Feature Learning for Interpretable Prediction
- Analyzing FOMC Minutes: Accuracy and Constraints of Language Models
- ProSLM : A Prolog Synergized Language Model for explainable Domain Specific Knowledge Based Question Answering
- DataAgent: Evaluating Large Language Models' Ability to Answer Zero-Shot, Natural Language Queries
- From Large to Tiny: Distilling and Refining Mathematical Expertise for Math Word Problems with Weakly Supervision
- Visual Environment-Interactive Planning for Embodied Complex-Question Answering
- Mutual benefits of social learning and algorithmic mediation for cumulative culture
- Automated Theorem Provers Help Improve Large Language Model Reasoning
- Hierarchical knowledge guided fault intensity diagnosis of complex industrial systems
- Exploring the Impact of Model Scaling on Parameter-Efficient Tuning
- Logic-Scaffolding: Personalized Aspect-Instructed Recommendation Explanation Generation using LLMs
- Flows: Building Blocks of Reasoning and Collaborating AI
- Do GPT Language Models Suffer From Split Personality Disorder? The Advent Of Substrate-Free Psychometrics
- On the attribution of confidence to large language models
- Towards Efficient and Explainable Hate Speech Detection via Model Distillation
- FlanEC: Exploring Flan-T5 for Post-ASR Error Correction
- AdaPhish: AI-Powered Adaptive Defense and Education Resource Against Deceptive Emails
- Modern Information Technologies in Scientific Research and Educational Activities
- Assessing Logical Puzzle Solving in Large Language Models: Insights from a Minesweeper Case Study
- AutoMathKG: The automated mathematical knowledge graph based on LLM and vector database
- Controllable Navigation Instruction Generation with Chain of Thought Prompting
- RAMP: Retrieval and Attribute-Marking Enhanced Prompting for Attribute-Controlled Translation
- EvidenceMap: Learning Evidence Analysis to Unleash the Power of Small Language Models for Biomedical Question Answering
- Are Large Language Models a Threat to Programming Platforms? An Exploratory Study
- On LLM-generated Logic Programs and their Inference Execution Methods
- Self-Explanation in Social AI Agents
- Automating Chapter-Level Classification for Electronic Theses and Dissertations
- Explainable Recommendation with Simulated Human Feedback
- SituationalLLM: Proactive language models with scene awareness for dynamic, contextual task guidance
- Sycophancy in Vision-Language Models: A Systematic Analysis and an Inference-Time Mitigation Framework
- Mind the Labels: Describing Relations in Knowledge Graphs With Pretrained Models
- Imitating Mistakes in a Learning Companion AI Agent for Online Peer Learning
- Highlighting Case Studies in LLM Literature Review of Interdisciplinary System Science
- A Knowledge-Injected Curriculum Pretraining Framework for Question Answering
- Evaluating the Reliability of Self-Explanations in Large Language Models
- ATHENA: Mathematical Reasoning with Thought Expansion
- HiddenTables & PyQTax: A Cooperative Game and Dataset For TableQA to Ensure Scale and Data Privacy Across a Myriad of Taxonomies
- In-Context Learning for Knowledge Base Question Answering for Unmanned Systems based on Large Language Models
- Multi-modal Traffic Scenario Generation for Autonomous Driving System Testing
- ZeFaV: Boosting Large Language Models for Zero-shot Fact Verification
- "Mango Mango, How to Let The Lettuce Dry Without A Spinner?": Exploring User Perceptions of Using An LLM-Based Conversational Assistant Toward Cooking Partner
- Generative Agents Navigating Digital Libraries
- Model-Generated Pretraining Signals Improves Zero-Shot Generalization of Text-to-Text Transformers
- Thought Flow Nets: From Single Predictions to Trains of Model Thought
- CoRRPUS: Code-based Structured Prompting for Neurosymbolic Story Understanding
- Imitation Game for Adversarial Disillusion with Chain-of-Thought Reasoning in Generative AI
- A Unified Approach to Emotion Detection and Task-Oriented Dialogue Modeling
- TMIQ: Quantifying Test and Measurement Domain Intelligence in Large Language Models
- Machine Reading Comprehension using Case-based Reasoning
- Data Augmentation with In-Context Learning and Comparative Evaluation in Math Word Problem Solving
- ITCMA: A Generative Agent Based on a Computational Consciousness Structure
- MultiSurf-GPT: Facilitating Context-Aware Reasoning with Large-Scale Language Models for Multimodal Surface Sensing
- Simple and Effective Input Reformulations for Translation
- Comparison of pipeline, sequence-to-sequence, and GPT models for end-to-end relation extraction: experiments with the rare disease use-case
- Reliable Conversational Agents under ASP Control that Understand Natural Language
- Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree search
- Derivation Prompting: A Logic-Based Method for Improving Retrieval-Augmented Generation
- Transductive Learning for Textual Few-Shot Classification in API-based Embedding Models
- StyleRec: A Benchmark Dataset for Prompt Recovery in Writing Style Transformation
- Tailored-LLaMA: Optimizing Few-Shot Learning in Pruned LLaMA Models with Task-Specific Prompts
- Statutory AI: Aligning Large Language Models With Legal Norms
- LLMs Between the Nodes: Community Discovery Beyond Vectors
- Rethinking the Chain-of-Thought: The Roles of In-Context Learning and Pre-trained Priors
- Detecting Gender Stereotypes in Scratch Programming Tutorials
- Humanoid Artificial Consciousness Designed with Large Language Model Based on Psychoanalysis and Personality Theory
- Cheap Learning: Maximising Performance of Language Models for Social Data Science Using Minimal Data
- WebChecker: A Versatile EVL Plugin for Validating HTML Pages with Bootstrap Frameworks