← Back to all generators
lucataco/videollama3-7b
OfficialView on Replicate →
VideoLLaMA 3: Frontier Multimodal Foundation Models for Video Understanding
Capabilities
Max TokensTop-P
Cost
Community model (estimated from hardware time)
Input Parameters
| Name | Type | Description | Default | Constraints |
|---|---|---|---|---|
prompt* | string | Text prompt to guide the model's response | — | — |
video* | string(uri) | Input video file | — | — |
fps | number | Frames per second to sample from video | 1 | min: 0, max: 10 |
max_frames | integer | Maximum number of frames to process | 180 | min: 0, max: 256 |
max_new_tokens | integer | Maximum number of tokens to generate | 2048 | min: 0, max: 4096 |
temperature | number | Sampling temperature | 0.2 | min: 0, max: 1 |
top_p | number | Top-p sampling | 0.9 | min: 0, max: 1 |
promptrequiredstringText prompt to guide the model's response
videorequiredstringInput video file
fpsnumberFrames per second to sample from video
Default:
1min: 0, max: 10max_framesintegerMaximum number of frames to process
Default:
180min: 0, max: 256max_new_tokensintegerMaximum number of tokens to generate
Default:
2048min: 0, max: 4096temperaturenumberSampling temperature
Default:
0.2min: 0, max: 1top_pnumberTop-p sampling
Default:
0.9min: 0, max: 1Version:
34a1f45f7068Updated: 7/25/202632.7K runs
cinemasetfree.com