Skip to content
#

llm-deployment

Here are 11 public repositories matching this topic...

The production agentic context stack. Data streaming + knowledge graphs + vector search + LLM orchestration. All in a single deployment for self hosting, BYOC, or cloud.

  • Updated Nov 1, 2025
  • Python

🧠 A comprehensive toolkit for benchmarking, optimizing, and deploying local Large Language Models. Includes performance testing tools, optimized configurations for CPU/GPU/hybrid setups, and detailed guides to maximize LLM performance on your hardware.

  • Updated Mar 27, 2025
  • Shell

Improve this page

Add a description, image, and links to the llm-deployment topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the llm-deployment topic, visit your repo's landing page and select "manage topics."

Learn more