• Home
  • Search
  • Improving Knowledge Extraction from LLMs for Task Learning through Agent Analysis
  • Cite Icon14
  • https://doi.org/10.1609/aaai.v38i16.29799Copy DOI Icon

Improving Knowledge Extraction from LLMs for Task Learning through Agent Analysis

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Large language models (LLMs) offer significant promise as a knowledge source for task learning. Prompt engineering has been shown to be effective for eliciting knowledge from an LLM, but alone it is insufficient for acquiring relevant, situationally grounded knowledge for an embodied agent learning novel tasks. We describe a cognitive-agent approach, STARS, that extends and complements prompt engineering, mitigating its limitations and thus enabling an agent to acquire new task knowledge matched to its native language capabilities, embodiment, environment, and user preferences. The STARS approach is to increase the response space of LLMs and deploy general strategies, embedded within the autonomous agent, to evaluate, repair, and select among candidate responses produced by the LLM. We describe the approach and experiments that show how an agent, by retrieving and evaluating a breadth of responses from the LLM, can achieve 77-94% task completion in one-shot learning without user oversight. The approach achieves 100% task completion when human oversight (such as an indication of preference) is provided. Further, the type of oversight largely shifts from explicit, natural language instruction to simple confirmation/discomfirmation of high-quality responses that have been vetted by the agent before presentation to a user.

Similar Papers
  • Front Matter
  • Citations1

Editorial: Large language models in work and business.

  • Nov 29, 2024
  • Frontiers in artificial intelligence
  • Şadi Evren Şeker
  • Research Article
  • Citations1

Comparative analysis of language models in addressing syphilis-related queries

  • May 27, 2025
  • Medicina Oral, Patología Oral y Cirugía Bucal
  • Lm Ferreira +7
  • Research Article

Large language model enhanced framework for systematic reviews and meta-analyses

  • Oct 01, 2025
  • BMJ Digital Health & AI
  • Jiashu Shen +5
  • Research Article
  • Citations11

JorGPT: Instructor-Aided Grading of Programming Assignments with Large Language Models (LLMs)

  • Jun 18, 2025
  • Future Internet
  • Jorge Cisneros-González +3
  • Preprint Article

Clinical Trial Patient Recruitment and Large Language Models: Socio-Technical and Economic Framework Development (Preprint)

  • Mar 23, 2026
  • Qian Qian
  • Research Article
  • Citations8

Assessing the accuracy and consistency of large language models in triaging social media posts for psychological distress.

  • Sep 01, 2025
  • Psychiatry research
  • Michele Settanni +3
  • Conference Article

LORE: Continual Logit Rewriting Fosters Faithful Generation

  • Jan 01, 2025
  • Charles Yu +4
  • Research Article
  • Citations4

Beyond Text Generation: Assessing Large Language Models’ Ability to Reason Logically and Follow Strict Rules

  • Jan 15, 2025
  • AI
  • Zhiyong Han +4
  • Research Article

Design and feasibility of lay clinical trial summaries using large language models.

  • Oct 01, 2025
  • JCO Oncology Practice
  • Brenda Adjei +3
  • Research Article
  • Citations61

Large language models illuminate a progressive pathway to artificial intelligent healthcare assistant

  • May 17, 2024
  • Medicine Plus
  • Mingze Yuan +11
  • Research Article
  • Citations1

Limitations and mitigation strategies for using generative artificial intelligence in medical writing: a narrative review

  • Feb 10, 2026
  • Journal of Korean Medical Association
  • Ki-Hyun Jeon
  • Research Article
  • Citations325

Almanac - Retrieval-Augmented Language Models for Clinical Medicine.

  • Jan 25, 2024
  • NEJM AI
  • Cyril Zakka +21
  • Research Article
  • Citations8

ChatGPT as an inventor: eliciting the strengths and weaknesses of current large language models against humans in engineering design

  • Jan 01, 2025
  • Artificial Intelligence for Engineering Design, Analysis and Manufacturing
  • Daniel N Ege +6
  • Research Article

The Convergence of Federated Learning, Knowledge Graphs, and Large Language Models for Language Learning: A Scoping Review

  • Mar 09, 2026
  • Applied Sciences
  • Michael Kenteris +1
  • Research Article

Preclinical HistoBench: A Pilot Benchmark Dataset for Evaluating Large Language Models on Preclinical Histopathological Classification.

  • Feb 27, 2026
  • Biology
  • Avan Kader +6
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.