About this integration
MCP server for video transcripts, screenshots, and OCR on YouTube and web videos.
- Transport
- stdio
- Authentication
- unknown
- Initial setup
- user-authorization
- Runtime
- unattended
- Evidence
- documented
- Version
- 1.1.2
- Package
- ocular-audio-mcp
- Last compatibility test
- Not independently tested
Connect your agent
ocular-audio-mcpPython and FFmpeg are required; Tesseract is optional. Restricted or private videos may use an exported browser-cookie file. Package and external video services were not executed.
Capabilities: Extract video transcripts and metadata, Capture timestamped video screenshots, Run optional OCR and local speech recognition
Connected profiles
Additional details
io.github.RayAKaan/ocular-audio-mcp
Source ↗ · Checked 2026-09-171.1.2
Source ↗ · Checked 2026-09-17CC0-1.0; package licenses are separate
Source ↗ · Checked 2026-09-17npm
Source ↗ · Checked 2026-09-17ocular-audio-mcp
Source ↗ · Checked 2026-09-191.2.1
Source ↗ · Checked 2026-09-19The publisher describes an asynchronous, non-blocking MCP server and provides a stdio client definition launched with npx.
Source ↗ · Checked 2026-09-19Public-video use has no documented credential, while restricted or private video access can use an exported browser cookies.txt session; that cookie-session mechanism is outside the native authentication taxonomy.
Source ↗ · Checked 2026-09-19The pinned VS Code configuration declares type stdio, command npx, and the ocular-audio-mcp package argument.
Source ↗ · Checked 2026-09-19Publisher documentation reviewed; package and integration endpoint not executed or independently security-audited.
Source ↗ · Checked 2026-09-19