Batch process and migrate large media libraries

Import hundreds of gigabytes or terabytes from existing cloud storage, run one automated workflow across many files, and migrate the results back.

No credit card needed
Import from existing storage

Read large media libraries from S3, Google Cloud, Azure, SFTP, HTTP, and other supported sources.

Process across our fleet

Submit Assembly-sized batches while Transloadit schedules concurrent jobs across a fleet that can scale to 1,500 machines.

Migrate results back

Export processed files to the original storage location or move them to a new provider and prefix.

DSC_08452-1200w.jpg420 KB
contract-p1.png180 KB
keynote-720p.mp488 MB
quarterly-report.pdf310 KB
Large-scale batch processing

From hundreds of gigabytes to terabyte-scale migrations

Use Transloadit as a managed file processing API for a one-time cloud storage migration, a large media backfill, or recurring bulk file processing. Import from your existing storage, transform each file, and export the results without moving the entire library through your application servers.

  1. Inventory the source

    Choose the buckets, prefixes, or object paths to migrate and estimate the total files and bytes.

  2. Paginate into Assemblies

    Paginate large storage prefixes into jobs of roughly 500–1,000 files and submit them at a controlled rate.

  3. Process concurrently

    Apply one saved Template across the migration while Batch Job Slots govern parallel processing in each region.

  4. Export and reconcile

    Write results to the same or a new storage destination, then verify each Assembly before marking files complete.

  1. Your application

    1. Inventory
    2. Paginate
    3. Submit
  2. Assemblies submitted

    Transloadit

    1. Import
    2. Schedule
    3. Transform
    4. Export
    5. Report
  3. Results returned

    Your application

    1. Record IDs
    2. Approve
    3. Publish
Division of labour

You own the ends, we own the middle

Your application keeps the parts that depend on your data: choosing what to migrate, submitting it in Assembly-sized batches at a rate you control, and recording what came back. That is an inventory loop and an idempotent callback handler.

Everything between those two points is ours: importing the bytes, scheduling across the fleet and its queues, running the same Template over every file, exporting the results, and accounting for errors per Assembly. See how queues and concurrent processing work.

Managed batch processing at scale

Use our fleet instead of building your own batch processors

Run large-scale media processing and cloud storage migrations through one batch processing API. Transloadit supplies the processing fleet, storage integrations, and automated workflows; your application controls the migration.
Production data processed
223 PB
Fleet capacity
1,500 machines
High-concurrency processing

Run multiple Assemblies concurrently while Batch Job Slots provide a predictable per-region allowance.

Batch-aware queueing

Keep bulk imported work in the Batch Queue so live uploads retain priority during a migration.

Storage-to-storage migration

Import from one provider and export processed results to the same location or a new destination.

Reusable automated workflows

Save a reviewed Template once, then apply the same processing graph to every page of the library.

Cloud storage integrations

Read existing files from S3, Google Cloud, Azure, SFTP, HTTP, and other supported sources.

Asynchronous tracking

Poll Assembly Status JSON or consume signed Notifications without holding an application request open.

Related resources

Take Batch Processing into production

The guides, Robot reference, and working demos teams reach for when Batch Processing has to run reliably at scale.

Start a large-scale batch processing migration

Connect your existing storage, test a representative page of files, then scale the same Template across the full library and export verified results.
Hundreds of GB must move through app servers
Direct storage import and export
Large migrations need processing capacity
A managed fleet with batch queueing
One-off scripts drift between runs
Reusable saved Templates
Long jobs outlive web requests
Assembly Status JSON and Notifications
Mixed formats need different outputs
Conditional, branching workflows
Publishing before export risks broken assets
Per-Assembly result verification
GDPR
HIPAA
ISO 27001ISO27001
AES-256
SOC 2 Type II
No credit card needed