Technology

Spotify Fixes Podcast Video Ingestion After Third Major Outage

Spotify experienced its third major podcast outage this month. An overload in video transcoding caused massive delays, prompting an urgent capacity upgrade.

On June 24, 2026, Spotify's podcast video transcoding infrastructure reached maximum capacity, causing severe multi-hour publishing delays for creators. The outage was caused by a combination of a massive background batch job, increased processing costs, and severe scheduling bugs. Spotify engineers mitigated the crisis by increasing total compute capacity by 67 percent, though the company faced community backlash over allegedly altering the detection timeline in the official postmortem.

A Perfect Storm of Transcoding Failures

Podcast creators on Spotify have navigated a frustrating series of reliability issues over the past several weeks. The problems peaked on June 24, 2026, when Spotify's video transcoding infrastructure hit absolute maximum capacity. The severe bottleneck prevented new podcast episodes from publishing, causing multi-hour delays across the platform.

According to the official engineering incident report, the outage was triggered by five converging system failures. The transcoding architecture lacked sufficient compute headroom to absorb normal traffic spikes. Simultaneously, a massive batch reprocessing job was executing in the background, draining available resources. Recent quality improvements to the video pipeline had also drastically increased the processing cost for every individual media item.

These heavy loads triggered a hidden malfunction in the prioritization system, which mistakenly treated the background batch tasks with the exact same urgency as live creator uploads. Finally, a separate resource scheduling bug prevented the system from utilizing its full potential, leaving roughly ten percent of the cluster sitting completely idle while the queues overflowed.

Restoring System Stability

The engineering team responded by immediately halting the background batch job and deploying an urgent fix to the resource scheduler. Overnight, Spotify spun up additional infrastructure, increasing the total transcoding compute capacity by 67 percent. These aggressive interventions allowed the system to clear the massive backlog of unpublished episodes by 01:02 UTC on June 25.

Moving forward, Spotify has pledged to implement earlier monitoring alerts and significantly better backpressure mechanisms throughout the ingestion pipeline to prevent future queue collapses.

Community Pushback on Incident Transparency

Despite the detailed technical breakdown, the postmortem generated immediate controversy within the developer community. Critics accused Spotify of intentionally altering the incident timeline to mask how long it took their internal telemetry to detect the failure.

Prominent tech commentator Gergely Orosz noted that this was the third major publishing outage on the platform in a single month. Orosz provided email receipts proving he had notified the Spotify Podcasts team of the upload failures at 17:31 UTC, well before Spotify claimed creator reports began surfacing. This discrepancy has fueled frustration among creators who are demanding better, more transparent status communication during active platform outages.