[Usage]: How to use local video in multimodal input? #12879

z1054136399 · 2025-02-07T09:21:23Z

Your current environment

`from openai import OpenAI

openai_api_key = "EMPTY"
openai_api_base = "http://localhost:8000/v1"

client = OpenAI(
api_key=openai_api_key,
base_url=openai_api_base,
)

video_url = "http://commondatastorage.googleapis.com/gtv-videos-bucket/sample/ForBiggerFun.mp4"

Use video url in the payload

chat_completion_from_url = client.chat.completions.create(
messages=[{
"role":
"user",
"content": [
{
"type": "text",
"text": "What's in this video?"
},
{
"type": "video_url",
"video_url": {
"url": video_url
},
},
],
}],
model=model,
max_completion_tokens=64,
)

result = chat_completion_from_url.choices[0].message.content
print("Chat completion output from image url:", result)`

How would you like to use vllm

No response

Before submitting a new issue...

Make sure you already searched for relevant issues, and asked the chatbot living at the bottom right corner of the documentation page, which can answer lots of frequently asked questions.

DarkLight1337 · 2025-02-07T10:29:19Z

See #12604 (comment)

z1054136399 added the usage How to use vllm label Feb 7, 2025

hmellor closed this as completed Feb 17, 2025

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

[Usage]: How to use local video in multimodal input? #12879

[Usage]: How to use local video in multimodal input? #12879

z1054136399 commented Feb 7, 2025

DarkLight1337 commented Feb 7, 2025

[Usage]: How to use local video in multimodal input? #12879

[Usage]: How to use local video in multimodal input? #12879

Comments

z1054136399 commented Feb 7, 2025

Your current environment

Use video url in the payload

How would you like to use vllm

Before submitting a new issue...

DarkLight1337 commented Feb 7, 2025