All posts

The migration that locked the table for eight minutes

Published March 28, 2026 · by Majd Ghithan, written with the help of AI

The migration that locked the table for eight minutes

The deploy looked boring. One migration, adding a last_seen_at column to the users table with a default of now(). I'd written a hundred like it. I ran it at 2pm on a Tuesday because it was "just a column," and for eight minutes the entire application couldn't log anyone in.

Here's the migration. Read it and see if it looks dangerous to you, because it didn't to me.

Schema::table('users', function (Blueprint $table) {
    $table->timestamp('last_seen_at')->default(now())->nullable();
});

The users table had 40 million rows. On our MySQL version, adding a column with a default value that way meant the database rewrote the entire table — every row, copied — while holding a lock that blocked writes. Every login updates a user row. So every login queued behind the migration, the connection pool filled with waiting queries, and the app fell over. Not because it crashed. Because it was politely waiting for a lock that took eight minutes to release.

Login p95 latency during the deploy
0.26.530300.313:5814:00 (migrate)14:0314:08 (done)14:09
sec

That flat line at 30 seconds isn't the real latency — it's the connection timeout. Everything past it was just failing.

What actually locks

The trap is that not all ALTER TABLEs are equal. Adding a nullable column with no default is nearly instant on a modern MySQL — it's a metadata change. Adding a column with a default, on the version we were on, was a full table copy holding a lock the whole time. The syntax difference is three words. The production difference is eight minutes of downtime.

I did not know which operations were cheap and which were catastrophic, because on my laptop's 200-row users table, every migration is instant. Small data hides the lock completely. There is no way to feel this in development. You have to know it.

The safe recipe

The fix isn't "be careful." It's a repeatable procedure for any schema change on a big, hot table. Split the one dangerous migration into cheap steps, each of which holds the lock for milliseconds or not at all.

Step one becomes trivially safe:

// Migration 1 — instant, no lock worth mentioning
Schema::table('users', function (Blueprint $table) {
    $table->timestamp('last_seen_at')->nullable();
});

Step two backfills without ever locking the whole table — small chunks, a breath between them, so ongoing writes always get a turn:

User::whereNull('last_seen_at')
    ->orderBy('id')
    ->chunkById(2000, function ($users) {
        User::whereIn('id', $users->pluck('id'))
            ->update(['last_seen_at' => now()]);
        usleep(100_000); // 100ms, let other writes through
    });

Each update touches 2,000 rows and releases. Logins slip between the chunks instead of queuing behind one giant transaction. It takes longer in wall-clock time — and that's the point. You're trading a fast operation that stops the world for a slow one that never does.

Online DDL, when the database offers it

Newer MySQL and Postgres support online/concurrent DDL for many operations — ALGORITHM=INPLACE, LOCK=NONE on MySQL, or tools like pt-online-schema-change and gh-ost that build a shadow table and swap it in. If your database and version support it for the change you need, that's the cleanest path. But you have to check for your specific operation and version — "MySQL supports online DDL" is not the same as "this ALTER, on this version, takes no lock." That gap is exactly the one that bit me.

The lesson I actually took

A migration is code that runs against production data at production scale, with a lock, in the middle of live traffic. I'd been treating migrations as a build step — something that just needs to succeed. They're a runtime event on your busiest table.

Now, before any schema change on a table over a million rows, I ask one question: how long does this hold a lock, and what's queued behind it while it does? If I don't know the answer, I don't run it at 2pm. I find out first — because the demo will run it in 4 milliseconds every single time, and tell me nothing.

🤖 Heads up: this post was drafted with AI. Spot something wrong or off?Edit it on GitHub & open a PR
Majd Ghithan

Majd Ghithan

Full-Stack Engineer & Tech Lead

More posts

Laravel finally resizes images for you

Laravel finally resizes images for you

Every Laravel app I've built that takes an avatar ended the same way: composer require intervention/image, wire up a facade, write a little service class, and hope the next dev…

July 26, 2026Read more
What actually breaks at 100,000 users (that never breaks in a demo)

What actually breaks at 100,000 users (that never breaks in a demo)

The first time I shipped a dashboard to a hundred thousand active users, nothing I had tested was what broke. Everything that failed had passed every check I ran. It worked on m…

July 25, 2026Read more
Where your queue jobs go to die under load

Where your queue jobs go to die under load

The support ticket said "I never got my invoice email." I checked the logs. The job ran. It succeeded. failedjobs was empty. Everything said the email went out. It did not.

July 18, 2026Read more
The indexes you're missing (and the one that's hurting you)

The indexes you're missing (and the one that's hurting you)

A client sent me a query that took 4.2 seconds. One WHERE, one ORDER BY, a table with three million rows. I added a single index and it dropped to 11 milliseconds. They asked if…

July 11, 2026Read more
Caching in Laravel without the stampede

Caching in Laravel without the stampede

We cached the homepage's "trending products" query for five minutes. It was our slowest query — about 900ms, a big aggregation across orders. Caching it took the homepage from s…

July 4, 2026Read more
Taking one endpoint from 800ms to 80ms

Taking one endpoint from 800ms to 80ms

The order-details endpoint took 800 milliseconds. Not broken, just slow enough that the app felt heavy everywhere it was used. The team's instinct was "the server needs more mem…

June 27, 2026Read more
The N+1 you can't see (it's hiding in your accessors)

The N+1 you can't see (it's hiding in your accessors)

I've fixed hundreds of N+1 queries. The easy ones are right there in the controller — a foreach with $order-customer inside it, obvious the moment you read the code. Those aren'…

June 20, 2026Read more
The Friday deploy that charged customers twice

The Friday deploy that charged customers twice

It was 4:40 on a Friday. The change was tiny — a one-line tweak to how we called the payment provider, plus a bump to the queue worker's timeout. Small, tested, reviewed. I depl…

June 13, 2026Read more
Code review people don't dread

Code review people don't dread

I once left forty-one comments on a junior's pull request. I was proud of it. Thorough, I told myself. The next day he barely made eye contact, and his next PR sat open for a we…

June 6, 2026Read more
Hiring a mid-level Laravel dev: the signals that actually matter

Hiring a mid-level Laravel dev: the signals that actually matter

The best hire I ever made failed my first question. I asked him to explain service containers and he stumbled, went quiet, then said "honestly I use them every day but I've neve…

May 30, 2026Read more
Saying no to a feature without being the 'no' person

Saying no to a feature without being the 'no' person

For about a year I was the engineer everyone learned to route around. Not because I was wrong, I was usually right, but because my answer to new ideas was a flat "no, that'll br…

May 23, 2026Read more
My first 90 days as a tech lead (and the habit I had to break)

My first 90 days as a tech lead (and the habit I had to break)

Three weeks into leading my first team, I stayed late to "help" by rewriting a junior's feature myself. It was faster my way. I pushed it, felt productive, went home. The next m…

May 16, 2026Read more
Breaking down the task that scares you

Breaking down the task that scares you

There's a specific kind of ticket that makes my stomach drop. Not the hard ones, hard is fine. It's the ambiguous ones. "Migrate billing to the new provider." "Add multi-tenancy…

May 9, 2026Read more
1:1s that aren't just status updates

1:1s that aren't just status updates

For my first few months as a lead, my 1:1s were thirty minutes of me asking "so, what are you working on?" and nodding at answers I already knew from standup. We both left sligh…

May 2, 2026Read more
Get the business logic out of your controllers

Get the business logic out of your controllers

I once opened a OrderController@store method that was 240 lines long. Validation, a discount calculation, three model writes, a Stripe charge, two emails, a Slack notification,…

April 25, 2026Read more
Form Requests are the most underrated thing in Laravel

Form Requests are the most underrated thing in Laravel

Most Laravel devs meet Form Requests once, in a tutorial, use them to hold a rules() array, and never look deeper. That's a shame, because a Form Request is the cleanest place i…

April 18, 2026Read more
Testing Laravel without mocking everything

Testing Laravel without mocking everything

I inherited a codebase once with 900 unit tests and no confidence. Every test mocked the repository, mocked the model, mocked the mailer, mocked the thing three layers down. The…

April 11, 2026Read more
Three Eloquent features that clean up your models

Three Eloquent features that clean up your models

The messiest Laravel model I ever wrote wasn't messy because of Eloquent. It was messy because I ignored the parts of Eloquent that exist specifically to keep it clean. The same…

April 4, 2026Read more
Chasing a memory leak in a Laravel queue worker

Chasing a memory leak in a Laravel queue worker

The alert came in at 4am: one of our queue:work processes had been OOM-killed. The supervisor restarted it, it processed jobs for about forty minutes, and got killed again. It w…

March 21, 2026Read more