i swear to god fuck ai and llms can we just have one thing that's not enshittified by them
In summary,
Argument 1: LLM code is trained on stolen copyright, filled with errors, will mix up best practices, burns out our contributors, hurts the free software community, and threatens Debian's most important quality: its stability.
Argument 2: Some people find AI tools helpful.
"just one more dpkg bro I swear we are going to get pulseaudio and systemd working this time bro just one more data lane, you have to do this or you are going to fall behind bro trust me"
I'm all for Argument 1. Also, it's been scientifically proven that any sort of extended use to AU lowers your intelligence, and that would be the last thing we need for a distro that needs to remain a bastion of Stable.
Scientifically proven? Any sort of AI usage? Just genereal intelligence?
There are definitely some initial studies out already that point in that direction, but I think your wording is a bit overly broad and definitive. It will be a while until we understand all the factors involved and what is going on there exactly.
You forgot argument 3.
Argument 3: FUCK THIS SHIT! FUCK AI!
Hi, join us over on !FUCK_AI@lemmy.world
People like Linus Torvalds?
I realize that some people really dislike AI, but this is an area where I'm willing to absolutely put my foot down as the top-level maintainer.
Linux is not one of those anti-AI projects, and if somebody has issues with that, they can do the open-source thing and fork it.
Or just walk away.
AI is a tool, just like other tools we use. And it's clearly a useful one.
It may not have been that "clearly" even just a year ago, but it's no longer in question today.
I'm just summarizing the arguments made by the actual proposals by the Debian team linked in the OP
Choice 1 provides this position on AI
Debian has a well-earned reputation for stability. This stability is crucial to Debian's position in the free software ecosystem. It is our belief that widespread LLM usage comes from the "move fast, and break things" attitude that, while common in many parts of this industry, is contrary to what makes Debian Debian, and is inappropriate for Debian contributors.
In practical terms, LLM usage raises the following concerns:
- Copyright
LLM output has very unclear legal status: it may be possible to copyright on its own merits, or not; it may be affected by all of the licenses and copyrights in the training data, or not. Debian Policy and the DFSG require absolute clarity for licensing and copyright[1][2]. Software and other contributions written conventionally by humans with unclear copyright or license status are not allowed in Debian; LLM output should not have a special exception to this.
- Quality
LLM output has many well-known problems with accuracy.[3][4][5] A LLM can never "know" if its output is correct since it merely produces syntactically likely combinations of the training data. In some environments this is good enough. In Debian, it is not. For instance, in packaging, each Debian source package is unique. Since packaging syntax and best practices have changed over time, a LLM-produced package will have a mixture of contents spanning the age of the archive, with watch files that do not work, overrides out of context, imaginary copyright, and will generally be unfit for upload. A seasoned Debian contributor with packaging expertise may find some limited usefulness here, but a new contributor cannot, and would not know how to fix it. These same quality and accuracy concerns apply clearly to all of the areas listed in the scope of this proposal above. If Debian were a closed organization comprising only domain experts who never leave, this might not be an issue; however,
- Community
Debian is a project that is more than just code: it is a community built on shared interests in free software and solving technical problems. Debian intentionally grows this community through many means, and new contributors are always encouraged to join. Allowing LLM contributions breaks this. New contributors submitting LLM output for review places an unnecessary strain on the reviewer, which can lead to burnout. Furthermore, LLM-dependent new contributors do not actually learn and understand the details of Debian packaging or processes, so they cannot come to replace a former burned out DD.
- Ethics
LLM companies directly hurt the free software community as whole by scraping the whole web for training data without any regard for license, copyright, or even established conventions such as robots.txt.[6] This has had a major negative impact on Debian's public web resources, effectively a large scale and perpetual Denial of Service attack on sites that many users rely on. As a consequence parts of our infrastructure were not reachable at all, and JS-based checks had to be enabled. Many other projects were similarly affected. Furthermore, LLM training consumes a staggering amount of resources[7], and the user verification systems that we have been forced to implement as protection waste resources as well. This is blatant disregard for the internet as a public resource, wastes system administrator time, and although individual LLM sessions do not directly use massive resources or DoS the public web, the fact that they can be used at all is a direct result of these unethical behaviours by the LLM companies.
Debian has a Social Contract. [8] Our priorities are our users and free software. Debian is Stable. [9] Users and organizations choose Debian because it is reliable and secure.
Debian is not here to generate as much code as possible requiring manual review by a shrinking number of human volunteers, or to package every piece of software, or to rush new features, but these are what LLMs are used for.
In conclusion, allowing LLM contributions is contrary to the social contract and the common cause of creating a free operating system with a focus on quality and stability.
While Choice 2 only says
The Debian project recognizes that AI-assisted contributions raise many concerns, e.g. about the technical quality and maintainability of such contributions, and their legal status. AI itself also raises additional concerns, about its impact on society at large, on the IT industry and on Free Software; about its environmental impact; and the aggressive or non-compliant practices of AI scrapers.
Nevertheless, many Debian contributors find AI tools helpful when contributing to Debian, and ultimately for improving Debian.
Lol what? That’s their official phrasing? Holy shit!
„Are you in favour of introducing vanilla icecream in our product line?
Points against it: its an abnormal abomination created by the devil that tastes like frozen shit!
Points in favour: while admitting that vanilla icecream is an abnormal abomination created by the devil, some people like eating eating icecream.“
Lmao
Wondering who gets a vote
These 1044 people: https://nm.debian.org/members/
Some of the people who work on Debian.
The choices kinda look like they would both either explicitly or in effect disallow practically all LLMs. Nice.
Though that doesn't answer the question about code analysis tools that the kernel devs recentently talked about (with Linus Torvalds being firmly in favor of using them). They aren't generative AI, but are AI nonetheless and unless Debian already talked about it elsewhere it does need explicit policy IMO.
Non-generative AI is just complicated stats.. There's no reason not to use ML or whatever if you can find a use for it (e.g. bug hunting, optimisation).
Generative AI is that plus random noise, copyright theft, environmental destruction, and fascism.. It should be off the cards for Debian
Linux
Welcome to c/linux!
Welcome to our thriving Linux community! Whether you're a seasoned Linux enthusiast or just starting your journey, we're excited to have you here. Explore, learn, and collaborate with like-minded individuals who share a passion for open-source software and the endless possibilities it offers. Together, let's dive into the world of Linux and embrace the power of freedom, customization, and innovation. Enjoy your stay and feel free to join the vibrant discussions that await you!
Rules:
-
Stay on topic: Posts and discussions should be related to Linux, open source software, and related technologies.
-
Be respectful: Treat fellow community members with respect and courtesy.
-
Quality over quantity: Share informative and thought-provoking content.
-
No spam or self-promotion: Avoid excessive self-promotion or spamming.
-
No NSFW adult content
-
Follow general lemmy guidelines.