SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.