← All guides

Give Cursor YouTube video access

Written by the VidWords Team · · Updated · Report a correction

A block in .cursor/mcp.json gives Cursor nine tools for reading YouTube videos — plus one credential mistake worth avoiding first.

The short answer

Create .cursor/mcp.json in your project with the URL and no credential at all:

{
  "mcpServers": {
    "vidwords": {
      "url": "https://vidwords.com/mcp"
    }
  }
}

Cursor picks it up without a restart in most cases. The MCP section of Settings shows the server as Needs login — click that once, approve in the browser, and Cursor holds a credential it refreshes by itself. Use ~/.cursor/mcp.json instead to make it available in every project.

Prefer this to a pasted token. The file above carries no secret, so the committed-credential problem described below cannot arise.

A static token instead

If your Cursor build does not offer the sign-in step, add the header back — and read the warning two sections down before you commit anything:

      "headers": { "Authorization": "Basic YOUR_API_TOKEN" }

Get a token first

Only the static-token route above needs this; signing in issues its own credential. Create a free account and copy the API token from your profile. Two things to know before the first call:

Do not commit the project file

This is the one thing worth getting right before anything else. .cursor/mcp.json sits inside the repository and gets committed by default — and the header holds a live API token. Committing it hands a working credential to everyone with repository access, and to the entire internet if the repo is public.

Either put the server in ~/.cursor/mcp.json, which is outside the repo and covers every project anyway, or add .cursor/mcp.json to .gitignore before your first commit. If a token has already been pushed, rotate it — deleting the file does not remove it from history.

Using it in Agent mode

MCP tools are available to Cursor's agent, which means the model decides when to call them. Two habits make that reliable:

The workflow this unlocks in an editor is narrow but genuinely underserved: a great deal of framework knowledge is published as conference talks and screen recordings and never written down. Being able to ask "what did they say the migration path was, with timestamps" without leaving the editor removes a context switch that otherwise costs half an hour.

Reading what is on the screen

Screen-recorded tutorials are the case where captions fail completely. The presenter says "as you can see here" and the actual content — the config, the terminal output, the diagram — exists only in the picture. analyze_video reads frames as well as speech, and ask_video answers against that analysis with every citation checked against a recorded frame; anything that cannot be matched is dropped rather than guessed at.

The nine tools

ToolWhat it doesCost
search_transcriptFind where a video discusses something; returns timestamps and deep links.1 Cloud Request
get_transcriptFull text for up to 25 videos at once.1 Cloud Request each
list_channel_videosRecent uploads for a channel.1 Cloud Request · Starter and up
list_watchlists · watchlist_activityRadar monitoring — channels you track and what they published.Free
accountPlan and remaining Cloud/AI balances.Free
analyze_videoFrame-level analysis — slides, charts, demos, on-screen text.3 AI Units/minute (Standard); 15 (Deep)
get_analysisRead a finished analysis.Free
ask_videoAsk against a finished analysis; citations verified or dropped.1 AI Unit

If it does not connect

Get a free API token →