Technology

71446 readers

3118 users here now

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related news or articles.
Be excellent to each other!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
Check for duplicates before posting, duplicates may be removed
Accounts 7 days and younger will have their posts automatically removed.

Approved Bots

founded 2 years ago

MODERATORS

L3s@lemmy.world

enu@lemmy.world

technopagan@lemmy.world

L4s@lemmy.world

L3s@hackingne.ws

L4s@hackingne.ws

632

An angry admin shares the CrowdStrike outage experience (www.theregister.com)

submitted 11 months ago by lemmee_in@lemm.ee to c/technology@lemmy.world

136 comments fedilink hide all child comments

IT administrators are struggling to deal with the ongoing fallout from the faulty CrowdStrike update. One spoke to The Register to share what it is like at the coalface.

Speaking on condition of anonymity, the administrator, who is responsible for a fleet of devices, many of which are used within warehouses, told us: "It is very disturbing that a single AV update can take down more machines than a global denial of service attack. I know some businesses that have hundreds of machines down. For me, it was about 25 percent of our PCs and 10 percent of servers."

He isn't alone. An administrator on Reddit said 40 percent of servers were affected, along with 70 percent of client computers stuck in a bootloop, or approximately 1,000 endpoints.

Sadly, for our administrator, things are less than ideal.

Another Redditor posted: "They sent us a patch but it required we boot into safe mode.

"We can't boot into safe mode because our BitLocker keys are stored inside of a service that we can't login to because our AD is down.

you are viewing a single comment's thread
view the rest of the comments

[–] TheObviousSolution@lemm.ee 18 points 11 months ago (3 children)

It might be CrowdStrike's fault, but maybe this will motivate companies to adopt better workflows and adopt actual preproduction deployment to test these sort of updates before they go live in the rest of the systems.

[–] EnderMB@lemmy.world 19 points 11 months ago* (last edited 11 months ago) (3 children)

I know people at big tech companies that work on client engineering, where this downtime has huge implications. Naturally, they've called a sev1, but instead of dedicating resources to fixing these issues the teams are basically bullied into working insane hours to manually patch while clients scream at them. One dude worked 36 hours straight because his manager outright told him "you can sleep when this is fixed", as if he's responsible for CloudStrike...

Companies won't learn. It's always a calculated risk, and much of the fallout of that risk lies with the workers.

[–] Disaster@sh.itjust.works 8 points 11 months ago

That dude should not have put up with that.

[–] uis@lemm.ee 6 points 11 months ago (1 children)

Sounds so illegal, that it makes labour authoririty happy

[–] EnderMB@lemmy.world 2 points 11 months ago (1 children)

Is it illegal? I'm not American so I have no idea if there are laws in your country against on-call maximum hours.

[–] uis@lemm.ee 3 points 11 months ago

It's not about oncall, they are literally in the office
See 1
Not sure about America, but it is very illegal in Russia.

[–] MrAlternateTape@lemm.ee 1 points 10 months ago (1 children)

That comment about sleep...that's about where I tell them to go fuck themselves. I'll find a new job, I'm not going to put up with bullshit like that.

[–] Entropywins@lemmy.world 1 points 10 months ago

100% agree

[–] Randelung@lemmy.world 9 points 11 months ago

Oh sweet summer child.

[–] cheetah_cheetos@lemmy.world 8 points 11 months ago

Might be hard to do. Crowdstrike release several updates per day to the channel files to match changes in adversarial behaviour. In this case, BCP and backup are what need to be done.