Skip to content

vllm_omni.model_executor.models

Modules:

Name Description
audex
audio8_tts

Audio8 TTS Preview (DualAR, 44.1 kHz codec) model implementation.

auk
aura_omni
bagel
breeze_tts_2
common
cosmos3
cosyvoice3
covo_audio
dreamzero
fish_speech
gepard
glm_image
glm_tts
gr00t
higgs_audio_v2
higgs_audio_v3
hunyuan_image3
hunyuan_video
indextts2
joyai_vl_interaction

JoyAI-VL-Interaction native pipeline.

lance
lingbot_world
mammoth_moda2
mimo_audio
ming_flash_omni
ming_image
ming_tts
minicpmo_4_5
minimax_h3

MiniMax H3 model components.

minimax_music3

MiniMax Music 3 text-to-music.

model_local_kv

Declaration protocol for attention KV kept outside the paged manager.

moss_tts
moss_tts_nano
nemotron_voicechat

NVIDIA NemotronLabs VoiceChat-11B: offline speech-to-speech 3-stage pipeline.

omnivoice
output_templates
personaplex

vllm-omni integration for PersonaPlex (a Moshi finetune, full-duplex S2S).

qwen2_5_omni
qwen3_omni
qwen3_tts
registry
step_audio2
utils
voxcpm2
voxtral_tts
wan2_2
whisper_utils

OmniModelRegistry module-attribute

OmniModelRegistry = _ModelRegistry(
    {
        **{
            model_arch: _LazyRegisteredModel(
                module_name=_resolve_module_name(
                    mod_relname
                ),
                class_name=cls_name,
            )
            for model_arch, (
                mod_relname,
                cls_name,
            ) in _VLLM_MODELS.items()
        },
        **{
            model_arch: _LazyRegisteredModel(
                module_name=f"vllm_omni.model_executor.models.{mod_folder}.{mod_relname}",
                class_name=cls_name,
            )
            for model_arch, (
                mod_folder,
                mod_relname,
                cls_name,
            ) in _OMNI_MODELS.items()
        },
    }
)