Streaming Vlm
Video AnalyticsOpen Source Media & Entertainment

Streaming Vlm

StreamingVLM: Real-Time Understanding for Infinite Video Streams

Open Source

About

StreamingVLM enables real-time, stable understanding of effectively infinite video by keeping a compact KV cache and aligning training with streaming inference. It avoids quadratic cost and sliding-window pitfalls, runs up to 8 FPS on a single H100, and wins 66.18% vs GPT-4o mini on a new long-video benchmark. It also boosts general VQA without task-specific finetuning. You can grasp the gist by skimming this section first.

Open Source Health

Not enough history
Stars
1,080
Forks
67
License
MIT
Last commit
1 years ago
Python

Related Categories