-
Updated
Sep 25, 2025 - Python
llm-deployment
Here are 10 public repositories matching this topic...
A curated collection of open-source Large Language Model (LLM) projects that are production-ready and can be used for solving real-world problems. This repository focuses on high-performance, scalable LLM solutions across various industries and applications.
-
Updated
Jul 11, 2026 - Python
Live index of the newest LLMOps tooling — track what's shipping in LLM observability and deployment
-
Updated
Aug 27, 2026 - Python
ModelSpec is an open, declarative specification for describing how AI models especially LLMs are deployed, served, and operated in production. It captures execution, serving, and orchestration intent to enable validation, reasoning, and automation across modern AI infrastructure.
-
Updated
Apr 27, 2026 - Python
AI-agent-native CLI for Google Colab. 31 commands, 24 Hermes tools. Token-efficient output. Provision GPU/TPU VMs, execute code, deploy LLMs, tunnel APIs.
-
Updated
Aug 12, 2026 - Python
Open-source LLM inference benchmarks — TTFT, TPOT, Throughput, Latency & Cost-per-token for models like Llama, Qwen, Gemma, DeepSeek, Gpt-Oss etc. deployed on different dedicated GPUs.
-
Updated
Jun 18, 2026 - Python
AWS EKS + IRSA, Volumes, ISTIO & KServe+ NextJS App + Fastapi Serve + kubernetes + Helm charts + Multimodel or LLM-Deployment The School of AI EMLO-V4 course assignment https://theschoolof.ai/#programs
-
Updated
Jan 26, 2025 - Python
A simple CLI tool to fetch Hugging Face model metadata and estimate required VRAM/RAM for inference.
-
Updated
Mar 14, 2026 - Python
Field notes, benchmarks, and turnkey scripts for running 284B LLM inference and multi-modal workflows across 2x NVIDIA DGX Spark (GB10) with 200GbE RoCE. 双机 NVIDIA DGX Spark (GB10) 284B 大模型分布式推理与多模态部署实战笔记、真机实测数据与避坑指南。
-
Updated
Aug 25, 2026 - Python
Improve this page
Add a description, image, and links to the llm-deployment topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the llm-deployment topic, visit your repo's landing page and select "manage topics."