vllm_omni.model_extras ¶
Modules:
| Name | Description |
|---|---|
auk | Request helpers for AuK: prompt rendering, sampling params, duration policy. |
bagel | |
cosmos3 | |
helios | |
hunyuan_image3 | |
lingbot_video | |
ltx2 | Request parameters shared by official LTX guidance pipelines. |
magi2 | Request fields supported by the MAGI-2 Preview pipeline. |
mammothmodal2_preview | |
ming_flash_omni | |
ming_image | |
registry | |
sana_video | |
sensenova_u1 | |
vace | |
video_generation | |
build_image_to_image_prompt ¶
build_image_to_image_prompt(
model_class_name: str | None,
prompt: str,
negative_prompt: str | None,
input_image: Image | list[Image],
height: int | None = None,
width: int | None = None,
) -> dict[str, Any]
build_image_to_video_prompt ¶
build_image_to_video_prompt(
model_class_name: str | None,
prompt: dict[str, Any],
height: int | None = None,
width: int | None = None,
num_frames: int | None = None,
) -> dict[str, Any]
Build a model-specific I2V prompt from an example-owned envelope.
build_text_to_image_prompt ¶
build_text_to_image_prompt(
model_class_name: str | None,
prompt: dict[str, Any],
height: int | None = None,
width: int | None = None,
) -> dict[str, Any]
Build a model-specific T2I prompt from an example-owned envelope.
build_x_to_text_prompt ¶
build_x_to_text_prompt(
model_family: str,
model: str,
prompt: str,
has_image: bool,
) -> tuple[dict[str, Any], list[int] | None]
Build a model-aware T2T/I2T prompt and optional stop-token ids.
get_ar_input_builder ¶
Return a model's AR-stage input builder, or None if undeclared.
Models with a text/AR stage that needs template-formatted prompt tokens and AR stop tokens (e.g. HunyuanImage3) declare an ar_input_builder. The shared task examples call it generically when present, so the example scripts stay model-agnostic; models without one are unaffected.
get_ar_tokenizer_validator ¶
Return a model's AR-tokenizer validator, or None if undeclared.
Models whose AR prompt/stop-token logic depends on hardcoded special token ids (e.g. HunyuanImage3) declare an ar_tokenizer_validator to check those ids against whatever tokenizer actually loads at runtime. The shared task examples call it generically, right after loading a real tokenizer, so model/tokenizer revision drift fails loudly instead of silently producing a request with the wrong stop tokens.
get_extra_output_params ¶
get_model_class_name ¶
Extract model_class_name from an Omni/AsyncOmni instance.
This hides the internal ODConfig plumbing from example scripts.
get_output_tensor_range ¶
get_output_tensor_range(
model_class_name: str | None,
) -> OutputTensorRange
Return the declared range for floating-point tensor outputs.
The default preserves the shared examples' historical handling. Pipelines that already return normalized tensors declare zero_to_one explicitly.
get_transformer_config_subfolder ¶
get_transformer_config_subfolder(
model_class_name: str | None,
*,
model: str | None,
revision: str | None = None,
) -> str
Return the model-declared DiT config subfolder, or the standard default.
get_video_generation_defaults ¶
get_video_generation_defaults(
model_class_name: str | None,
extra_body: Mapping[str, Any] | None = None,
) -> VideoGenerationDefaults | None
Return model-owned defaults for the shared video examples, if declared.
get_x_to_text_model_family ¶
Resolve a text-output prompt family from a checkpoint's config.json.