Files
wehub-resource-sync eec33d25b2
pre-commit / pre-commit (push) Failing after 1s
Build Wheel / build (3.11) (push) Failing after 1s
Build Wheel / build (3.12) (push) Failing after 0s
chore: import upstream snapshot with attribution
2026-07-13 12:29:08 +08:00

1.1 KiB

MammothModa2-Preview

Source https://github.com/vllm-project/vllm-omni/tree/main/examples/offline_inference/mammothmodal2_preview.

Run examples (MammothModa2-Preview)

Download model

hf download bytedance-research/MammothModa2-Preview --local-dir ./MammothModa2-Preview

Text-to-Image (T2I)

Text-to-image now runs through the shared offline image example (examples/offline_inference/text_to_image/text_to_image.py). See the recipe recipes/MammothModa2/MammothModa2-Preview.md for the full command and the extra_body knobs (text_guidance_scale, cfg_range, num_inference_steps).

Image Summary

python examples/offline_inference/mammothmodal2_preview/run_mammothmoda2_image_summarize.py \
  --model ./MammothModa2-Preview \
  --deploy-config ./vllm_omni/deploy/mammoth_moda2_ar.yaml \
  --question "Summarize this image." \
  --image ./image.png

Example materials

??? abstract "run_mammothmoda2_image_summarize.py" py --8<-- "examples/offline_inference/mammothmodal2_preview/run_mammothmoda2_image_summarize.py"