Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
-
Updated
Aug 24, 2026 - Python
Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
Julep — durable, composable AI agents. Flows that crash and resume, retry safely, and explain every step.
RAG (Retrieval Augmented Generation) Framework for building modular, open source applications for production by TrueFoundry
Trace-native CI/CD for AI agents — production failures become regression tests that block the PR. Auto-detect, cluster, freeze into hermetic cases, replay in CI for $0.
The collaborative spreadsheet for AI. Chain cells into powerful pipelines, experiment with prompts and models, and evaluate LLM responses in real-time. Work together seamlessly to build and iterate on AI applications.
AIConfig is a config-based framework to build generative AI applications.
Python SDK for running evaluations on LLM generated responses
An end-to-end LLM reference implementation providing a Q&A interface for Airflow and Astronomer
[DEPRECATED] Moved to microsoft/agent-governance-toolkit
Static analysis for AI agent configs, tool descriptions, and system prompts — catches vague tool descriptions, missing stop conditions, and schema gaps before they reach runtime. Zero-LLM, deterministic checks, built for CI.
A production-grade control layer that sits between your application logic and any LLM — input validation, schema enforcement, circuit breaking, targeted retry, and audit logging in one composable pipeline.
[⛔️ DEPRECATED] Friendli: the fastest serving engine for generative AI
Multi-model AI agent runtime. Define agents in YAML, route each role to a model, orchestrate with 7 patterns (ReAct, Plan & Execute, Fan-Out, Pipeline, Supervisor, Swarm, Glyph), and deploy as a REST/WebSocket API with RAG, memory, MCP tools, guardrails and OpenTelemetry observability.
SOUL.md governance framework for Hermes Agent — structured memory, skill management, and operational rules
Turn failed AI agent runs into replayable regression tests. Catch regressions before you ship.
83 MCP tools for GPU infrastructure + Agent FinOps — deploy LLMs, manage VMs (Proxmox/XO/vSphere), track cost per agent, enforce budgets and model policies. Works with Claude, Cursor, n8n, LangChain
Decision Infrastructure for the AI Era. 🧠✨️🤺
Python SDK for WildEdge
A governance harness for AI-assisted software delivery. You govern. Agents deliver.
Add a description, image, and links to the llm-ops topic page so that developers can more easily learn about it.
To associate your repository with the llm-ops topic, visit your repo's landing page and select "manage topics."