Pinned Loading
Repositories
Showing 10 of 47 repositories
- axs2stg Public
krai/axs2stg’s past year of commit activity - axs2kiss Public
Automated [KRAI X](https://github.com/krai/axs) workflows for dedicated inference engines on selected backends: vLLM and SGLang on CUDA and ROCm, NIM on CUDA, using the OpenAI API compatible LoadGen client.
krai/axs2kiss’s past year of commit activity - kilt4qaic Public
krai/kilt4qaic’s past year of commit activity - vllm Public Forked from vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
krai/vllm’s past year of commit activity - axs2qaic Public
Automated KRAI X workflows for reproducing MLPerf Inference submissions on systems equipped with Qualcomm Cloud AI 100 accelerators
krai/axs2qaic’s past year of commit activity
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…