Agent Skills: serving-llms-vllm
vLLM: high-throughput LLM serving, OpenAI API, quantization.
[vLLMInferenceServingPagedAttentionContinuousBatchingHighThroughputProductionOpenAIAPIQuantizationTensorParallelism]
UncategorizedID: PoorRican/dotfiles/serving-llms-vllm
Install this agent skill to your local
Skill Files
Browse the full folder contents for serving-llms-vllm.
Loading file tree…
Select a file to preview its contents.