An angry admin shares the CrowdStrike outage experience

[ - ]

db0@lemmy.dbzer0.com

201 points

5 months ago

Pity the administrators who dutifully kept a list of those keys on a secure server share, only to find that the server is also now showing a screen of baleful blue.

Lol, can you imagine? It empathetically hurts me even thinking of this situation. Enter that brave hero who kept the fileshare decryption key in a local keepass :D

permalink

report

reply

[ - ]

kescusay@lemmy.world

64 points

5 months ago

Seems like an argument for a heterogeneous environment, perhaps a solid and secure Linux server to host important keys like that.

permalink

report

parent

reply

[ - ]

pearsaltchocolatebar@discuss.online

55 points

5 months ago

Linux can shit the bed too. You need to maintain a physical copy.

permalink

report

parent

reply

[ - ]

StaySquared@lemmy.world

4 points

5 months ago

CS did take down Linux a few years back… I forget the exact details.

permalink

report

parent

reply

Show more comments

[ - ]

gnutrino@programming.dev

28 points

5 months ago

Sure but the chances of your Windows and Linux machines shitting the bed at the same time is less than if everything is running Windows. It’s exactly the same reason you keep a physical copy (which after all can break/burn down etc.) - more baskets to spread your eggs across.

permalink

report

parent

reply

Show more comments

[ - ]

Voroxpete@sh.itjust.works

52 points

5 months ago

Their point is not that linux can’t fail, it’s that a mix of windows and linux is better than just one. That’s what “heterogeneous environment” means.

You should think of your network environment like an ecosystem; monocultures are vulnerable to systemic failure. Diverse ecosystems are more resilient.

permalink

report

parent

reply

[ - ]

noobface@lemmy.world

12 points

5 months ago

Hey Ralph can you get that post-it from the bottom of your keyboard?

permalink

report

parent

reply

[ - ]

sugar_in_your_tea@sh.itjust.works

129 points

5 months ago

*

That’s why the 3-2-1 rule exists:

3 copies of everything on
2 different forms of media with
1 copy off site

For something like keys, that means:

secure server share
server share backup at a different site
physical copy (either USB, printed in a safe, etc)

Any IT pro should be aware of this “rule.” Oh, and periodically test restoring from a backup to make sure the backup actually works.

permalink

report

parent

reply

[ - ]

IphtashuFitz@lemmy.world

38 points

5 months ago

We have a cron job that once a quarter files a ticket with whoever is on-call that week to test all our documented emergency access procedures to ensure they’re all working, accessible, up-to-date etc.

permalink

report

parent

reply

[ - ]

Empricorn@feddit.nl

8 points

5 months ago

Are you hiring!?

permalink

report

parent

reply

[ - ]

slacktoid@lemmy.ml

30 points

5 months ago

Sounds like the best time to unionize

permalink

report

reply

[ - ]

gari_9812@lemmy.world

8 points

5 months ago

Any time is a good time to unionize

permalink

report

parent

reply

[ - ]

slacktoid@lemmy.ml

4 points

5 months ago

Agreed, just here they have then by the metaphorical balls.

permalink

report

parent

reply

[ - ]

ɔiƚoxɘup@infosec.pub

16 points

5 months ago

I’m in. This world desperately needs an information workers union. Someone to cover those poor fuckers in the help desk and desktop support as well as the engineers and architects that keep all of this shit running.

Those of us that aren’t underpaid are treated poorly. Today is what it looks like if everybody strikes at once.

permalink

report

parent

reply

[ - ]

slacktoid@lemmy.ml

5 points

5 months ago

This dude here coming in hot with a name, Information Workers Union (IWU). Love it

Soo are you gonna create the community or am I?

permalink

report

parent

reply

[ - ]

ɔiƚoxɘup@infosec.pub

3 points

5 months ago

I’m too chicken.

permalink

report

parent

reply

[ - ]

JasonDJ@lemmy.zip

4 points

5 months ago

Nah.

Bureau of Information & Technology Servicers.

report

reply

[ - ]

2 points

5 months ago

To preface, I want to see a tech workers union so, so bad.

With that said, I genuinely don’t believe that most tech workers would unionize. So many of them are brainwashed into thinking that a union would dictate all salaries, would force hiring to be domestic-only, or would ensure jobs for life for incompetent people. Anyone that knows what a union does in 2024 knows that none of that has to be true. A tech union only needs to be a flat fee every month, guaranteed access to a lawyer with experience in your cases/employer, and the opportunity to strike when a company oversteps. It’s only beneficial.

Even if you could get hundreds of thousands of signatories, the recent layoffs have shown that tech companies at the highest level would gladly fire a sizable number of employees if it meant stamping out a union. As someone that has conducted interviews in big tech, the sheer numbers at peak of people that had applied for some roles was higher than the number of active employees in the whole company. In theory, Google could terminate everyone and replace them with brand-new workers in a few months. It would be a fucking mess, but it (in theory) shows that if a Google or Apple decided that it wanted no part of unions they could just dig into their fungible talent pool, fire a ton of people, promote people that stayed, and fill roles with foreign or under-trained talent.

permalink

report

parent

reply

[ - ]

slacktoid@lemmy.ml

1 point

5 months ago

I feel you with this. They do not see themselves as workers. Thank you for the preface.

permalink

report

parent

reply

[ - ]

EnderMB@lemmy.world

2 points

5 months ago

Agreed, sadly to many there is still the view of tech being a meritocracy, and that they’re in FAANG because of their hard work over everything else, so fuck everyone else. Naturally, many change their tune once their employer actions regressive policies, but it’s surprising how many people just have zero understanding of what a union does. They see cop shows or The Wire and assume it’ll be like the unions there…

report

reply

[ - ]

110 points

5 months ago

Lemmy appears to be weathering the storm quite well…

…probably runs on linux

permalink

report

reply

[ - ]

RBG@discuss.tchncs.de

68 points

5 months ago

It runs on hundreds of servers. If any of them ran windows they might be out but unless you got an account on them you’d be fine with the rest. That’s the whole point of federation.

permalink

report

parent

reply

[ - ]

Grandwolf319@sh.itjust.works

3 points

5 months ago

I’m so proud of this community!

permalink

report

parent

reply

[ - ]

cygnus@lemmy.ca

96 points

5 months ago

*

The overwhelming majority of webservers run Linux ~~(it’s not even close, like high 90 percent range)~~ Edit: Upon double-checking it’s more like mid-80s, but the point stands

permalink

report

parent

reply

[ - ]

Bilb!@lem.monster

8 points

5 months ago

I wonder if any Lemmy servers run on Windows without WSL. I can’t think of any hard dependencies on Linux, so it should be possible.

permalink

report

parent

reply

[ - ]

KairuByte@lemmy.dbzer0.com

12 points

5 months ago

I doubt many Lemmy servers are running enterprise level antivirus.

permalink

report

parent

reply

[ - ]

Boozilla@lemmy.world

79 points

5 months ago

*

If you have EC2 instances running Windows on AWS, here is a trick that works in many (not all) cases. It has recovered a few instances for us:

Shut down the affected instance.
Detach the boot volume.
Move the boot volume (attach) to a working instance in the same region (us-east-1a or whatever).
Remove the file(s) recommended by Crowdstrike:
Navigate to the C:\Windows\System32\drivers\CrowdStrike directory
Locate the file(s) matching “C-00000291*.sys”, and delete them (unless they have already been fixed by Crowdstrike).
Detach and move the volume back over to original instance (attach)
Boot original instance

Alternatively, you can restore from a snapshot prior to when the bad update went out from Crowdstrike. But that is not always ideal.

permalink

report

reply

[ - ]

Defaced@lemmy.world

23 points

5 months ago

A word of caution, I’ve done this over a dozen times today and I did have one server where the bootloader was wiped after I attached it to another EC2. Always make a snapshot before doing the work just in case.

permalink

report

parent

reply

[ - ]

Boozilla@lemmy.world

7 points

5 months ago

Good advice!

permalink

report

parent

reply

[ - ]

gravitas_deficiency@sh.itjust.works

117 points

5 months ago

Lmao this is incredible

Another Redditor posted: "They sent us a patch but it required we boot into safe mode.

"We can’t boot into safe mode because our BitLocker keys are stored inside of a service that we can’t login to because our AD is down.

“Most of our comms are down, most execs’ laptops are in infinite bsod boot loops, engineers can’t get access to credentials to servers.”

N.B.: Reddit link is from the source

I hope a lot of c-suites get fired for this. But I’m pretty sure they won’t be.

permalink

report

reply

[ - ]

SkybreakerEngineer@lemmy.world

28 points

5 months ago

Fired? I hope they get class-actioned out of existence as a warning to anyone who skimps on QA

permalink

report

parent

reply

[ - ]

MagicShel@programming.dev

86 points

5 months ago

C-suites fired? That’s the funniest thing I’ve heard yet today. They aren’t getting fired - they are their own ass-coverage. How can they be to blame when all these other companies were hit as well?

I guess this is a good week for me to still be laid off.

permalink

report

parent

reply

[ - ]

Codex@lemmy.world

78 points

5 months ago

Our administrator is understandably a little bitter about the whole experience as it has unfolded, saying, "We were forced to switch from the perfectly good ESET solution which we have used for years by our central IT team last year.

Sounds like a lot of architects and admins are going to get thrown under the bus for this one.

“Yes, we ordered you to cut costs in impossible ways, but we never told you specifically to centralize everything with a third party, that was just the only financially acceptable solution that we would approve. This is still your fault, so we’re firing the entire IT department and replacing them with an AI managed by a company in Sri Lanka.”

permalink

report

parent

reply

[ - ]

Evotech@lemmy.world

6 points

5 months ago

Stupid argument though, honestly just chance that crowdstrike was the vendor to shit the bed. Might aswell have been set. You should still have procedures for this

permalink

report

parent

reply

An angry admin shares the CrowdStrike outage experience(www.theregister.com)

Technology

!technology@lemmy.world

Our Rules

Approved Bots

Community stats

Community moderators