CASE / raja-rao-dvenFri Sep 25
MEDIA / 1video
RRaja Rao DV@rajaraodv
视频生成JEV
I mapped an entire movie by mood for about 5 cents (Jev @typesafeai + @GraphonAI).
I mapped an entire movie by mood for about 5 cents (Jev @typesafeai + @GraphonAI).
Moods like: Funny, Scary, Thrilling, Suspenseful etc.
Around 7,000 AI analyses for roughly $0.05 and 7 seconds.
The demo is pretty simple:
First, Graphon maps the movie.
We ask Graphon about 40 questions designed to cover what's happening throughout the video. Graphon returns descriptions of the scenes, along with citations to the exact moments in the movie.
We use those questions to build a picture of what's happening across the entire movie.
Then we send those scene descriptions to Jev from @TypeSafeai
Jev looks at each scene and judges the probability that it is funny, scary, suspenseful, exciting, calm, or whatever other categories we want to analyze.
That's where we do around 7,000 analyses in parallel.
The result is basically a mood map of the entire movie.
Now imagine you're a movie studio, video editor, or creator.
You could say:
“Show me every scary moment.”
“Find all the car chases.”
“Give me the funniest 30 seconds.”
“Create a compilation of the most suspenseful moments.”
And jump directly to those exact scenes.
Today, doing this at scale is surprisingly complicated. Someone has to watch, tag, classify, scrub through footage, and manually find all of these moments.
With Graphon + Jev, you can start doing it automatically.
And the crazy part?
Those ~7,000 mood analyses cost about 5 cents.
This is what I think multimodal search starts to look like when you combine understanding, citations, and reasoning across an entire video.
@CompleteSkeptic what do you think?
377浏览
3点赞
1收藏
2转发