Workshop: 12th SC Workshop on Best Practices for HPC Training and Education
Authors: Habiba Morsy (Kean University, University of Virginia); Charlie Dey (Texas Advanced Computing Center); Zilu Wang (Cornell University); Mary Thomas (University of California, San Diego (UCSD)); and Essence Toone and David Joiner (Kean University)
Abstract: We present a proof-of-concept system for automating quality assurance (QA) in the HPC-ED federated training catalog using large language models (LLMs). The HPC-ED project aggregates metadata for training resources from multiple partner catalogs, improving discoverability for the high-performance computing (HPC) and cyberinfrastructure (CI) communities. While metadata publication processes have matured, QA remains largely manual, making it difficult to maintain accuracy and relevance at scale.
We have created an agent using commercial AI API calls to evaluate user submitted metadata and return a quality score, to help content providers improve their submitted metadata. The agent bases its score on a combination of the submitted catalog metadata and an AI created summary of a low-level crawl of the content item, with automated extraction of embedded content such as YouTube video transcripts.
We evaluated this agent on the HPC-ED Beta Catalog using four OpenAI models—GPT-3.5 Turbo, GPT-4o Mini, GPT-4.1 Nano, and GPT-4.1 Mini.
Back to 12th SC Workshop on Best Practices for HPC Training and Education Archive Listing Back to Full Workshop Archive Listing