Search

Media & Creative · CUDA · By vllm-project

1 skill found.

Skills

Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Integrate a new text-to-speech model into vLLM-Omni from HuggingFace reference implementation through production-ready serving with streaming and CUDA graph acceleration.

vllm-project/vllm-omni7.1k—~8.7kAutomated safety check: PassApache-2.0today