{
    "timeline": {
        "tracks": [
            {
                "clips": [
                    {
                        "asset": {
                            "type": "audio",
                            "src": "https://shotstack-assets.s3-ap-southeast-2.amazonaws.com/music/unminus/berlin.mp3"
                        },
                        "start": 0,
                        "length": 8
                    },
                    {
                        "asset": {
                            "type": "audio",
                            "src": "https://shotstack-assets.s3.ap-southeast-2.amazonaws.com/music/unminus/ambisax.mp3",
                            "effect": "fadeOut"
                        },
                        "start": 8,
                        "length": 8
                    }
                ]
            }
        ]
    },
    "output": {
        "format": "mp3",
        "resolution": "sd"
    }
}

Join two audio files into a single MP3 — intro plus episode, jingle plus message — by placing them end to end on the timeline. Replace the URLs and the stitch point moves automatically with the clip lengths.

Open in Studio

No account needed until you render.

Render with your AI assistant

Open in Claude Open in ChatGPT

Add the Shotstack MCP server to your editor, then paste the prompt

When you want to automate this

One audio per row of your data.

1
See it
OutputMP3
Size1024 × 576 (16:9)
Length16s
Frame ratedefault (25 fps)
Tracks / clips1 / 2
Assets used2 audio
Soundtrackno
Merge fieldsnone (edit the JSON)
2
Make it yours

Swap these in the JSON:

  • 2 audio URLs

To drive it from data, turn any value into a {{ PLACEHOLDER }} merge field.

3
Run it once

Save the JSON above as template.json, then render in the free sandbox:

# Save the JSON above as template.json, then:
curl -X POST https://api.shotstack.io/edit/stage/render \
  -H "x-api-key: $SHOTSTACK_API_KEY" \
  -H "Content-Type: application/json" \
  -d @template.json

# Poll until status is "done", then download response.url
curl https://api.shotstack.io/edit/stage/render/RENDER_ID -H "x-api-key: $SHOTSTACK_API_KEY"
import { readFileSync } from 'node:fs';

const headers = { 'x-api-key': process.env.SHOTSTACK_API_KEY, 'Content-Type': 'application/json' };
const edit = JSON.parse(readFileSync('template.json', 'utf8'));

const { response } = await fetch('https://api.shotstack.io/edit/stage/render', {
  method: 'POST', headers, body: JSON.stringify(edit)
}).then((r) => r.json());

let render;
do {
  await new Promise((r) => setTimeout(r, 3000));
  render = (await fetch(`https://api.shotstack.io/edit/stage/render/${response.id}`, { headers }).then((r) => r.json())).response;
} while (!['done', 'failed'].includes(render.status));

console.log(render.status, render.url);
import json, os, time, requests

headers = {"x-api-key": os.environ["SHOTSTACK_API_KEY"]}
with open("template.json") as f:
    edit = json.load(f)

render_id = requests.post("https://api.shotstack.io/edit/stage/render", json=edit, headers=headers).json()["response"]["id"]

while True:
    render = requests.get(f"https://api.shotstack.io/edit/stage/render/{render_id}", headers=headers).json()["response"]
    if render["status"] in ("done", "failed"):
        break
    time.sleep(3)

print(render["status"], render.get("url"))
npm install -g @shotstack/cli
export SHOTSTACK_API_KEY=your_sandbox_key

shotstack validate template.json          # offline schema check, free
shotstack render template.json --env stage --watch

Sandbox renders are free and watermarked. Production: the v1 endpoint and key. SDKs

4
Automate it

One render is a demo. A audio for every row of your data is the workflow.

Paste into Claude, ChatGPT or Cursor — your AI assistant builds the workflow around your data.

Node.js, one render per CSV row, wired to this template. Setup steps in the header.

Guides: Bulk-create from a CSV · Webhooks · Templates endpoint · Render your first video

AI assistants: Markdown () · · · CLI skill: SKILL.md () · MCP server

Common questions

How do I make the Simple audio stitch template my own?

Replace the asset URLs in the JSON with your own files and adjust text, timing or effects on each clip. To generate many variations from data, turn the values you change into merge fields ({{ PLACEHOLDER }}) and pass a merge array with each request.

Can I render this in full HD or 4K for YouTube?

Yes. This template renders 1024×576 by default. Set output.resolution to "1080" or output.size to a larger 16:9 size in the JSON; the layout scales with it.

What size and length is the output?

This template renders a 1024×576 (16:9) MP3 that is 16 seconds long, built from 2 clips on 1 track.

Can I generate hundreds of these automatically?

Yes. Loop over your data source (a CSV, a database, a webhook payload), send one render request per row with different values, and use a callback URL so Shotstack notifies you when each audio is ready.