← Back to all generators

lucataco/videollama3-7b

VideoLLaMA 3: Frontier Multimodal Foundation Models for Video Understanding

Capabilities

Max TokensTop-P

Cost

Community model (estimated from hardware time)

Input Parameters

promptrequiredstring

Text prompt to guide the model's response

videorequiredstring

Input video file

fpsnumber

Frames per second to sample from video

Default: 1min: 0, max: 10
max_framesinteger

Maximum number of frames to process

Default: 180min: 0, max: 256
max_new_tokensinteger

Maximum number of tokens to generate

Default: 2048min: 0, max: 4096
temperaturenumber

Sampling temperature

Default: 0.2min: 0, max: 1
top_pnumber

Top-p sampling

Default: 0.9min: 0, max: 1
Version: 34a1f45f7068Updated: 7/25/202632.7K runs