Description: This quick start guide will cover how users can build a chat bot that runs locally on an HPC using HuggingFace Transformers and Langchain. The open-source Mistral 7B-Instruct large-language model will be used to create an inference pipeline in a Jupyter notebook utilizing retrieval augmented generation to couple the model with a database containing external documentation from DOE software. The tutorial will cover how to build the container, how to preprocess documents, and how to build an inference pipeline. The demonstration will include utilizing the chat history for conversational AI.

Presenter: Dr. Mathew Boyer, GDIT / PET
Location: Webinar
Date & Time: May 7, 2024, 2:00p - 3:00p ET

Controlled by: DoD HPCMP
Controlled by: PET Program
CUI Category: OPSEC
Limited Dissemination Control: FEDCON
POC: Mr. Ronald Hedgepeth, pet@hpc.mil

CUI

technical_area: Software Refactoring