vllm_omni.utils ¶
Modules:
| Name | Description |
|---|---|
audio | Audio utility functions shared across models and entrypoints. |
audio_resample | Streaming resampling, shared by the engine and the entrypoints. |
custom_voice_io | Shared helpers for custom voice profile tools. |
device_copy | Host-to-device copies that do not block the host on queued GPU work. |
forced_aligner | Forced-aligner config + word-timestamp decoding for TTS. |
mm_outputs | Utilities for handling multimodal outputs / building multimodal output |
qwen3_force_align_processor | Qwen3 forced-aligner text/timestamp processor. |
speaker_cache | Process-wide thread-safe LRU cache for speaker extraction artifacts. |
tracking_parser | |