Skip to content

feat: Deploy llm-assistant stack on the GPU node#14

Open
nqngo wants to merge 2 commits into
mainfrom
ollama-hypercoverged
Open

feat: Deploy llm-assistant stack on the GPU node#14
nqngo wants to merge 2 commits into
mainfrom
ollama-hypercoverged

Conversation

@nqngo

@nqngo nqngo commented Mar 24, 2024

Copy link
Copy Markdown
Contributor

Deploy the LLM-assistant on the GPU node itself.

Since the node has a significant amount of RAM and CPU cores, we are able to leverage the node itself to reduce latency and make better ulitisation of the resources.

@nqngo nqngo self-assigned this Mar 24, 2024
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant