Agentic Research

Claude Real Video

2,191 stars

Lets an LLM actually watch a video: since models cannot ingest video directly, the skill extracts scene-aware, deduplicated keyframes plus a transcript first, then reads those to summarize, analyze or discuss the video (URL or local file).

First published:2026-06-30
data as of:
2026-09-30
Licence:
MIT

When to use it

When an agent needs to understand a video's content — tutorials, meeting recordings, product demos.

Caution

Depends on local tooling (ffmpeg-style) and a transcription model; long videos mean many keyframes and heavy context cost.

Install prompt

請幫我安裝「Claude Real Video」技能:從 https://github.com/HUANGCHIHHUNGLeo/claude-real-video/tree/master/skills/claude-real-video(技能路徑:skills/claude-real-video/) 取得完整的技能目錄,放進我的 Agent 的技能目錄(例如專案內的 skills/ 資料夾,或該 Agent 文件指明的全域技能位置),確認裡面的 SKILL.md 存在、frontmatter 的 name 與 description 完整,然後列出這個技能宣告的腳本、外部工具與 API 需求,先不要執行任何腳本。

Paste this to your agent. It deliberately contains no tool subcommands — check the tool's official docs for commands.

Get it Official source Source(2,191 ★)
Compatible with
Related patterns