Skip to content

Commit b77ee42

Browse files
authored
Remove duplicate content in AI section
Removed duplicate sections on LLM Inference, vLLM CLI, and Fine-tuning LLMs to improve clarity.
1 parent 34e24c1 commit b77ee42

1 file changed

Lines changed: 5 additions & 3 deletions

File tree

docs/hpc/08_ml_ai_hpc/01_intro.md

Lines changed: 5 additions & 3 deletions
Original file line numberDiff line numberDiff line change
@@ -6,6 +6,8 @@ For ML we'll cover two prominent open-source deep learning frameworks, PyTorch a
66

77
For AI we'll cover how to run a Hugging Face Large Language Model (LLM) and how to fine tune an LLM. Specifically, we provide guides on:
88

9-
* **LLM Inference**: Use the standard Hugging Face transformers library for basic tasks, and also introduce vLLM, a high-throughput serving engine that offers faster inference and an OpenAI-compatible API.
10-
* **vLLM CLI**: We provide examples of using the vllm command-line tool to quickly serve models and engage in interactive chat sessions.
11-
* **Fine-tuning LLMs**: We provide a practical example of fine-tuning the Gemma model to follow specific instructions. This section compares the original model with our improved version, showing how to achieve better response quality on Torch.
9+
**LLM Inference**: Use the standard Hugging Face transformers library for basic tasks, and also introduce vLLM, a high-throughput serving engine that offers faster inference and an OpenAI-compatible API.
10+
11+
**vLLM CLI**: We provide examples of using the vllm command-line tool to quickly serve models and engage in interactive chat sessions.
12+
13+
**Fine-tuning LLMs**: We provide a practical example of fine-tuning the Gemma model to follow specific instructions. This section compares the original model with our improved version, showing how to achieve better response quality on Torch.

0 commit comments

Comments
 (0)