Pinned Loading
Repositories
Showing 10 of 50 repositories
-
- axs2kiss Public
Automated KRAI-X workflows for inference engines on selected backends: vLLM and SGLang on CUDA and ROCm, NIM/TensorRT-LLM on CUDA, using an OpenAI API compatible LoadGen client
- axs2stg Public
- NeMo Public Forked from NVIDIA-NeMo/Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
- vllm Public Forked from vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
- kilt4qaic Public
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…