XZ backdoor in a nutshell

[ - ]

gregorum@lemm.ee

119 points

8 months ago

Thank you open source for the transparency.

permalink

report

reply

[ - ]

Cornelius_Wangenheim@lemmy.world

69 points

8 months ago

And thank you Microsoft.

permalink

report

parent

reply

[ - ]

just_another_person@lemmy.world

65 points

8 months ago

Shocking, but true.

permalink

report

parent

reply

[ - ]

Pantherina@feddit.de

14 points

8 months ago

They just pay some dude that is doing good work

permalink

report

parent

reply

[ - ]

∟⊔⊤∦∣≶@lemmy.nz

21 points

8 months ago

I have heard multiple times from different sources that building from git source instead of using tarballs invalidates this exploit, but I do not understand how. Is anyone able to explain that?

If malicious code is in the source, and therefore in the tarball, what’s the difference?

permalink

report

reply

[ - ]

Aatube@kbin.melroy.org

47 points

8 months ago

Because m4/build-to-host.m4, the entry point, is not in the git repo, but was included by the malicious maintainer into the tarballs.

permalink

report

parent

reply

[ - ]

∟⊔⊤∦∣≶@lemmy.nz

13 points

8 months ago

Tarballs are not built from source?

permalink

report

parent

reply

[ - ]

Aatube@kbin.melroy.org

35 points

8 months ago

*

The tarballs are the official distributions of the source code. The maintainer had git remove the malicious entry point when pushing the newest versions of the source code while retaining it inside these distributions.

All of this would be avoided if Debian downloaded from GitHub’s distributions of the source code, albeit unsigned.

report

reply

[ - ]

14 points

8 months ago

*

I don’t understand the actual mechanics of it, but my understanding is that it’s essentially like what happened with Volkswagon and their diesel emissions testing scheme where it had a way to know it was being emissions tested and so it adapted to that.

The malicious actor had a mechanism that exempted the malicious code when built from source, presumably because it would be more likely to be noticed when building/examining the source.

Edit: a bit of grammar. Also, this is my best understanding based on what I’ve read and videos I’ve watched, but a lot of it is over my head.

permalink

report

parent

reply

[ - ]

arthur@lemmy.zip

13 points

8 months ago

The malicious code is not on the source itself, it’s on tests and other files. The building process hijacks the code and inserts the malicious content, while the code itself is clean, So the co-manteiner was able to keep it hidden in plain sight.

permalink

report

parent

reply

[ - ]

sincle354@kbin.social

6 points

8 months ago

So it’s not that the Volkswagen cheated on the emissions test. It’s that running the emissions test (as part of the building process) MODIFIED the car ITSELF to guzzle gas after the fact. We’re talking Transformers level of self modification. Manchurian Candidate sleeper agent levels of subterfuge.

report

reply

[ - ]

16 points

8 months ago

it had a way to know it was being emissions tested and so it adapted to that.

Not sure why you got downvoted. This is a good analogy. It does a lot of checks to try to disable itself in testing environments. For example, setting TERM will turn it off.

permalink

report

parent

reply

[ - ]

WolfLink@lemmy.ml

10 points

8 months ago

The malicious code wasn’t in the source code people typically read (the GitHub repo) but was in the code people typically build for official releases (the tarball). It was also hidden in files that are supposed to be used for testing, which get run as part of the official building process.

permalink

report

parent

reply

[ - ]

Possibly linux@lemmy.zipOP

2 points

8 months ago

I think it is the other way around. If you build from Tarball then you getting pwned

permalink

report

parent

reply

[ - ]

Subverb@lemmy.world

8 points

8 months ago

*

The malicious code was written and debugged at their convenience and saved as an object module linker file that had been stripped of debugger symbols (this is one of its features that made Fruend suspicious enough to keep digging when he profiled his backdoored ssh looking for that 500ms delay: there were no symbols to attribute the cpu cycles to).

It was then further obfuscated by being chopped up and placed into a pure binary file that was ostensibly included in the tarballs for the xz library build process to use as a test case file during its build process. The file was supposedly an example of a bad compressed file.

This “test” file was placed in the .gitignore seen in the repo so the file’s abscense on github was explained. Being included as a binary test file only in the tarballs means that the malicious code isn’t on github in any form. Its nowhere to be seen until you get the tarball.

The build process then creates some highly obfuscated bash scripts on the fly during compilation that check for the existence of the files (since they won’t be there if you’re building from github). If they’re there, the scripts reassemble the object module, basically replacing the code that you would see in the repo.

Thats a simplified version of why there’s no code to see, and that’s just one aspect of this thing. It’s sneaky.

permalink

report

parent

reply

[ - ]

etchinghillside@reddthat.com

18 points

8 months ago

Any additional information been found on the user?

permalink

report

reply

[ - ]

Possibly linux@lemmy.zipOP

2 points

8 months ago

Probably Chinese?

permalink

report

parent

reply

[ - ]

Potatos_are_not_friends@lemmy.world

26 points

8 months ago

*

Can’t confirm but unlikely.

Via https://boehs.org/node/everything-i-know-about-the-xz-backdoor

They found this particularly interesting as Cheong is new information. I’ve now learned from another source that Cheong isn’t Mandarin, it’s Cantonese. This source theorizes that Cheong is a variant of the 張 surname, as “eong” matches Jyutping (a Cantonese romanisation standard) and “Cheung” is pretty common in Hong Kong as an official surname romanisation. A third source has alerted me that “Jia” is Mandarin (as Cantonese rarely uses J and especially not Ji). The Tan last name is possible in Mandarin, but is most common for the Hokkien Chinese dialect pronunciation of the character 陳 (Cantonese: Chan, Mandarin: Chen). It’s most likely our actor simply mashed plausible sounding Chinese names together.

permalink

report

parent

reply

[ - ]

jaybone@lemmy.world

3 points

8 months ago

So this doesn’t really tell us one way or the other who this person is or isn’t.

permalink

report

parent

reply

[ - ]

fluxion@lemmy.world

3 points

8 months ago

That actually suggests not Chinese due to naming inconsistencies

permalink

report

parent

reply

Show more comments

[ - ]

The Doctor@beehaw.org

2 points

8 months ago

Just because somebody picked a vaguely Chinese-sounding handle doesn’t mean much about who or where.

permalink

report

parent

reply

[ - ]

Possibly linux@lemmy.zipOP

2 points

8 months ago

That’s why I put the question mark

permalink

report

parent

reply

[ - ]

dan@upvote.au

4 points

8 months ago

*

They’re more likely to be based in Eastern Europe based on the times of their commits (during working hours in Eastern European Time) and the fact that while most commits used a UTC+8 time zone, some of them used UTC+2 and UTC+3: https://rheaeve.substack.com/p/xz-backdoor-times-damned-times-and

permalink

report

parent

reply

[ - ]

Possibly linux@lemmy.zipOP

3 points

8 months ago

It is also hard to be certain as they could be a night owl or a early riser.

report

reply

[ - ]

12 points

8 months ago

*

as long as you’re up to date on everything here: https://boehs.org/node/everything-i-know-about-the-xz-backdoor

the only additional thing i’ve seen noted is a possibilty that they were using Arch based on investigation of the tarball that they provided to distro maintainers

permalink

report

parent

reply

[ - ]

refreeze@lemmy.world

80 points

8 months ago

I have been reading about this since the news broke and still can’t fully wrap my head around how it works. What an impressive level of sophistication.

permalink

report

reply

[ - ]

rockSlayer@lemmy.world

80 points

8 months ago

*

And due to open source, it was still caught within a month. Nothing could ever convince me more than that how secure FOSS can be.

permalink

report

parent

reply

[ - ]

Lung@lemmy.world

95 points

8 months ago

Idk if that’s the right takeaway, more like ‘oh shit there’s probably many of these long con contributors out there, and we just happened to catch this one because it was a little sloppy due to the 0.5s thing’

This shit got merged. Binary blobs and hex digit replacements. Into low level code that many things use. Just imagine how often there’s no oversight at all

permalink

report

parent

reply

[ - ]

rockSlayer@lemmy.world

49 points

8 months ago

Yes, and the moment this broke other project maintainers are working on finding exploits now. They read the same news we do and have those same concerns.

permalink

report

parent

reply

Show more comments

[ - ]

The Quuuuuill@slrpnk.net

28 points

8 months ago

I was literally compiling this library a few nights ago and didn’t catch shit. We caught this one but I’m sure there’s a bunch of “bugs” we’ve squashes over the years long after they were introduced that were working just as intended like this one.

The real scary thing to me is the notion this was state sponsored and how many things like this might be hanging out in proprietary software for years on end.

permalink

report

parent

reply

[ - ]

smeenz@lemmy.nz

11 points

8 months ago

Can be, but isn’t necessarily.

permalink

report

parent