Skip to content
#

llm-infrastructure

Here are 88 public repositories matching this topic...

Multi-model AI agent runtime. Define agents in YAML, route each role to a model, orchestrate with 7 patterns (ReAct, Plan & Execute, Fan-Out, Pipeline, Supervisor, Swarm, Glyph), and deploy as a REST/WebSocket API with RAG, memory, MCP tools, guardrails and OpenTelemetry observability.

  • Updated Aug 10, 2026
  • Python

An intelligent gateway for Claude APIs that dynamically routes requests to the most cost-efficient model, caches responses, and escalates based on confidence signals — reducing LLM spend without sacrificing quality.

  • Updated May 6, 2026
  • Python

secrets.wtf: defensive AI infrastructure exposure index for Ollama and LM Studio hosts, local LLM APIs, model observations, remediation, and takedown requests.

  • Updated Jun 20, 2026
  • HTML

Improve this page

Add a description, image, and links to the llm-infrastructure topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the llm-infrastructure topic, visit your repo's landing page and select "manage topics."

Learn more