Skip to content
#

llm-deployment

Here are 10 public repositories matching this topic...

ModelSpec is an open, declarative specification for describing how AI models especially LLMs are deployed, served, and operated in production. It captures execution, serving, and orchestration intent to enable validation, reasoning, and automation across modern AI infrastructure.

  • Updated Apr 27, 2026
  • Python

Field notes, benchmarks, and turnkey scripts for running 284B LLM inference and multi-modal workflows across 2x NVIDIA DGX Spark (GB10) with 200GbE RoCE. 双机 NVIDIA DGX Spark (GB10) 284B 大模型分布式推理与多模态部署实战笔记、真机实测数据与避坑指南。

  • Updated Aug 25, 2026
  • Python

Improve this page

Add a description, image, and links to the llm-deployment topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the llm-deployment topic, visit your repo's landing page and select "manage topics."

Learn more