All toolsextract audio from video API

Extract Audio from Video

Turn any supported video into a clean MP3 or WAV audio asset. Run the preset directly on this page or call the same agent-ready endpoint from your product.

Prefer the shell? Extract audio from a video with FFmpeg, the verified FFmpeg command for this edit.

LIVE TOOL
Run extract audio
Drop a video, adjust the defaults if you need to, and download the result here.

or ·

Already have an account? Sign in

BEFORE / AFTER

See what the tool changes.

A real example: the after side is the unedited output of a KinoPipe job. Run the live form above to generate the same result from your own media.

BeforeVideo + audio
MP4
AfterAudio file
Ready
Processed audio48 kHz · stereo
MP3 · 192 kbps

Demo footage: public-domain timelapses (Bureau of Land Management, Oregon · NASA SVS). Processed by the same pipeline the API and MCP tools call.

Useful defaults, typed options.

The tool slug stays stable while your agent supplies named media inputs and a narrow set of documented options.

  • MP3 and WAV outputs
  • Video stream removed
  • Durable asset URL
What this preset does
  1. 01Inspect audio streams
  2. 02Select the primary track
  3. 03Remove the video stream
  4. 04Encode to the requested format
Stable endpointPOST /api/v1/tools/extract-audio
Live API request
{
  "inputs": [{ "id": "main", "url": "https://example.com/video.mp4" }],
  "options": { "format": "mp3" }
}
Successful output example
{
  "result": {
    "output": {
      "filename": "extract-audio-output.mp3",
      "contentType": "audio/mpeg",
      "byteSize": 437021,
      "downloadUrl": "https://cdn.kinopipe.com/…"
    }
  }
}

The command this runs

ffmpeg -i input.mp4 -vn -c:a libmp3lame -b:a 192k output.mp3
-vn
Drops the video stream. Without it FFmpeg keeps trying to write pictures into a container that has no room for them, and the error it gives you does not obviously mean "you forgot -vn".
-c:a libmp3lame -b:a 192k
What format: "mp3" sends. A fixed 192 kbit/s, which is comfortably past the point where most people stop hearing the difference on speech.
-c:a pcm_s16le
What format: "wav" sends instead. 16-bit uncompressed, so about seven times the size of the MP3 at the same sample rate, and nothing thrown away.

Which of the two you want

MP3 if a person is going to listen to it, or if it is going into storage and the bytes matter.

WAV if the next thing to touch it is a machine. A transcriber, a loudness pass and a silence detector all do better work on audio that has not been through a lossy encoder first, and the file is usually deleted an hour later anyway.

What this will not do

It will not hand you the original track untouched. Both paths re-encode, so a video whose audio is already AAC comes back as a new MP3 built from a decode of that AAC. The loss is small and it is real. Copying the stream into an .m4a is the operation that avoids it, and it is not one we expose.

It will not pick between multiple audio tracks. A file with a commentary track and a music track gives you the primary one, and there is no option to ask for the second.

It will not clean anything up. No noise reduction, no loudness correction. normalize_audio is the one that touches levels.

Doing it yourself

One line, no dependencies beyond FFmpeg itself, and it is the right answer whenever the file is already on the machine running your code. The version worth paying for is the one where an agent hands over a URL it found and gets an audio URL back, with the 2 GB ceiling and the format validation handled before anything starts decoding.

Know the boundaries before you run.

The source must contain a decodable audio stream.
Available outputs are MP3 and WAV.
Video streams are removed from the result.

About extract audio

No. Both outputs are re-encoded: MP3 through libmp3lame at 192 kbit/s, WAV as 16-bit PCM. There is no stream-copy path today, so if you need the original track bit for bit, this is not the tool for it.

Yes. Choose WAV when the next step is processing rather than listening, since a transcriber or a loudness pass should not have to work through MP3 artifacts.

Related media tools