Private beta: we are looking for design partners
Video Context

Free tool · Runs in your browser

Show any video
to ChatGPT or Claude.

Claude cannot open video files, and video uploads in ChatGPT are limited. This free tool turns your video into grids of frames with timestamps, plus a ready prompt, so any AI chat app can understand what happens in the video.

  • Free
  • No upload
  • No sign-up
  • No watermark

Drop a video here

MP4, MOV, WebM, or MKV. Large files are fine.

Your video stays on your computer. Nothing is uploaded.

How to use it

  1. 1

    Drop a video, or try the sample video.

  2. 2

    Choose how many frames to use, and choose even spacing or one frame per scene.

  3. 3

    Copy a grid image, or download all grids and the prompt as a context pack.

  4. 4

    Paste the grids and the prompt into ChatGPT, Claude, or Gemini, then ask your question.

Why AI chat apps need frames

Most AI chat apps understand images much better than video. Claude does not accept video files. ChatGPT can sometimes take a video, but the upload limits are small and the support changes often. Gemini accepts video, but a long file can use a large part of the context.

A frame grid solves this. It puts many frames from the video into one image, in time order, with a timestamp on each frame. A model can read the grid like a storyboard: what happens, in what order, and at what time.

How many frames to use

Video length Frames Why
Under 1 minute 9 to 12 Enough to see each step
1 to 5 minutes 16 to 24 One frame every 10 to 20 seconds
5 to 30 minutes 24 to 48 Use one frame per scene, so each shot appears once

Each grid holds up to 16 frames. If you choose more frames, the tool makes more than one grid. Models read a few large grids better than many small images.

Even spacing or one frame per scene

  • Even spacing is best for screen recordings, lectures, and videos that change slowly.
  • One frame per scene is best for edited videos: ads, trailers, sports highlights, and social clips. The tool finds each cut with the same method as our free scene detection tool.

What is in the context pack

The context pack is a ZIP file with:

  • grid-1.png, grid-2.png, and so on: the frame grids.
  • prompt.txt: a ready prompt that tells the model how to read the grids.
  • context.json: the file name, length, resolution, and the time of each frame, for agents and scripts.

Tips for better answers

  1. Paste the grids before the prompt.
  2. Ask about times: “What happens between 0:30 and 0:45?” The model can read the timestamps.
  3. If the speech matters, add a transcript. See the video transcription API.
  4. For a long video, ask the model to describe each grid first, then ask your question.

FAQ

Questions, answered.

Common questions about the free video to ai tool.

Can Claude watch videos?+

Not directly. The Claude apps accept images and documents, but not video files. To show a video to Claude, turn it into images of frames with timestamps. This tool does that in your browser.

Can I upload a video to ChatGPT?+

Support for video uploads in ChatGPT is limited and changes often, and uploads can fail. Images always work. Frame grids from this tool give ChatGPT the content of the video as images, with the time of each frame.

How many frames should I use?+

Use 12 to 16 frames for a short clip under 2 minutes. Use 24 to 36 frames for a longer video. More frames give more detail, but each image uses more of the model's context.

What is the difference between even spacing and one frame per scene?+

Even spacing takes frames at equal intervals. One frame per scene finds each cut and takes a frame from each shot, so a fast edit does not lose short shots and a long static shot does not fill the grid.

Does the AI hear the audio?+

No. A frame grid has only pictures. If the speech matters, add a transcript to your prompt. The Video Context API gives transcripts with timestamps and speaker labels.

Is my video uploaded anywhere?+

No. The tool reads and decodes the video in your browser with WebCodecs. The video never leaves your computer.

Your videos contain
valuable data.
Start using it.

Analyse video at scale. Find exactly what you need. Give your products and AI the context to act on it.