Free forever · Privacy-first

    Free Audiogram Maker

    An audiogram is a short video of a podcast or voice clip with a moving waveform, so audio can travel on Instagram, TikTok, LinkedIn and YouTube. Drop an audio file or a video, pick up to 90 seconds, add cover art and a title, choose animated bars, a line or a circle, and record a 9:16, 1:1 or 16:9 video in this tab. Optional captions are transcribed on your device with Whisper tiny and highlight each word as it is spoken. Nothing is uploaded.

    • Looks paidActually free
    • No uploadRuns on your device
    • No accountNothing to sign up for
    • No watermarkYour file, untouched

    More free audio tools

    Same quality bar. Still free. Still no upload.

    Part of free Instagram tools, free LinkedIn tools, free X tools and free TikTok tools.

    Your files never leave your device

    Free to use, with no account, no usage limit and no watermark, because the work runs on your device rather than on our servers.

    This is not a promise about how carefully we handle your upload. There is no upload. The file you choose is read into memory by your own browser, processed there, and handed back as a download — the tool has no server component and makes no network request with your data.

    • Nothing is stored. We never receive the file, so there is no copy to retain, no retention period to disclose, and nothing to delete on request.
    • Nothing is transmitted. No upload endpoint, no analytics payload carrying file contents, no third-party processor.
    • You can verify it. Load this page, disconnect from the internet, and use the tool. It still works, because everything it needs is already on your machine.
    • Nothing persists after you close the tab. Files are held in page memory only — not in local storage, not in a cache, not in a database.

    A note on wording, because it matters: we do not describe these tools as “encrypted.” Encryption protects data that travels to somebody else's computer. Your file does not travel, so there is nothing to encrypt and nothing to intercept — which is a stronger guarantee than encryption, not a weaker one. The page itself is served over HTTPS like the rest of the site.

    This makes the tools safe for material you are contractually barred from uploading to third-party services — client footage under NDA, unreleased campaign assets, anything covered by a confidentiality clause.

    How it works

    1. 1. Drop the audio

      MP3, M4A, WAV or AAC up to 80 MB and 20 minutes, or a video up to 300 MB for its soundtrack. It is decoded in this tab.

    2. 2. Pick the clip

      Tap the waveform or use the sliders to choose up to 90 seconds. Play the preview to hear it with the animation.

    3. 3. Design it

      Choose 9:16, 1:1 or 16:9, bars, line or circle, colours, a cover image and title, and a progress bar. Add captions if you want them: Whisper tiny runs on your device and every line is editable.

    4. 4. Record and download

      The video records in real time, so a 60-second clip takes about a minute. You get MP4 or WebM depending on what your browser can record, and the tool tells you which before you start. Stop cancels at any point.

    Questions

    How long can an audiogram be?

    Up to 90 seconds per video, picked from an audio file of up to 80 MB and 20 minutes, or a video of up to 300 MB. The whole soundtrack is decoded into memory to draw the waveform, which is what the source limits protect. For a longer episode, cut the part you want with the audio trimmer first.

    Will I get MP4 or WebM?

    Whatever your browser can record. The tool checks before you press Record and says which: Safari and recent Chrome and Edge record MP4 (H.264), browsers that cannot record MP4 give WebM. MP4 is the safer choice for posting to social apps, so if you see WebM and an app refuses it, record again in a browser that makes MP4.

    Why does recording take as long as the clip?

    The browser plays the audio and captures each drawn frame with MediaRecorder as it happens, the same real-time method our other video tools use. Keep the tab in front while it records, because browsers pause drawing in background tabs. 720p is lighter than 1080p on phones.

    Are the captions accurate, and do I need them?

    They are optional. When you add them, Whisper tiny transcribes only the selected clip on your device (about 41 MB to download once without WebGPU, about 120 MB with it, shared with our subtitle generator). It is good on clear speech and weaker on names, crosstalk and music beds, so every line is editable before you record, and you can copy the transcript or download an SRT.

    Is my audio uploaded?

    No. Decoding, waveform analysis, transcription and recording all happen in this browser tab. The only network requests are the optional Whisper model download from Hugging Face and the page itself.

    Which waveform style should I use?

    Bars read best at small sizes in a feed and suit speech. Line is a live oscilloscope of the actual audio and looks calmer. Circle wraps the bars around a round crop of your cover art, which works well for 1:1 and 9:16 posts where the artwork is the brand.

    Related tools

    Stay in the same job cluster. Free tools stay on-device; paid tools are labelled as such.

    These rearrange files. Versely makes new ones.

    Everything on this page works on a file you already have, which is why it costs nothing. Generating something that did not exist (video from a prompt, a voice, a score) runs a model, and that costs credits.