What is an Ethical LLM?
An interesting poll was boosted into my Mastodon feed recently.
The discussion that followed was very interesting, not least of all because many people in the replies are skeptical of the premise in the first place. Since this topic, ethical LLMs and the ethics of LLMs more broadly, is something Iāve been thinking about lately I figured that writing up this discussion would make a good excuse to get some of my thoughts down.
What Makes For an Ethical LLM?
This is by no means an exhaustive list, so if you think I missed anything important, have a good example you think bears inclusion, or know of an āethical LLMā project you think should have been included, let me know on Mastodon!
Dimensions of an Ethical LLM
- Source data
- Licensing and consent
- Effect on the Web (see: exhibit A)
- Especially with respect to model updates (see: this reply and the linked post)
- Environmental effects (water usage, which gets discussed a lot, but more importantly, IMO, the ways it contributes to Climate Change, and things like Muskās illegal power plant construction)
- Power gradients, who benefits
- Not as much of an issue in the pollās hypothetical, but in real life all of the AI companies seem to be run by, basically, the worst people in Tech (sleazy Sam Altman, Dario Amodei who doesnāt seem much better, and actual Nazi Elon Musk)
- Ideological valence of the tools themselves
- See: AI as a Fascist Artifact
- This piece* makes a reference to the concepts of holistic and prescriptive technologies, from Ursula Franklinās classic Engineering text The Real World of Technology
- *also via @mayintoronto
- Labor practices
- Especially with respect to the kind of underpaid and dangerous data cleaning that American tech companies have tended to outsource to the Global South (see DAIRās work in this area, or this representative reply)
- See also: Against the protection of stocking frames.
- Privacy, self hosting
- A given in the threadās hypothetical, but worth mentioning
- How itās trained, guardrails
- Not really discussed in that thread very much, but covers things like grokās Nazi and White Supremacist functions (c.f. Thauraās core principles)
- Sycophancy, ability to say āI donāt knowā
So⦠Are There Any?
I donāt know, but here are a few potential options. I havenāt done a deep dive on any of them yet, I mostly know that they are pitching themselves as ethical alternatives.
Thaura
Thaura is a very interesting one. An actual product, not a standalone model. Built around a non US-produced model (GLM-4.5 Air, at the time of this writing). Started by a couple of brothers from Syria who were sick of using tech that contributed to e.g. the genocide in Palestine. They seem to be claiming theyāve fixed the issues around sycophancy, which I find exciting and intriguing. They also lay out explicitly what āethicalā means to them in a set of core principles that seem more grounded in reality than any similar statements Iāve seen coming out of OpenAI or Anthropic. Shout out to my brother, Gilad, for tipping me off to this one!
Apertus
Apertus, an LLM being offered by the Swiss government under the Apache license. Seems to be based off of Common Crawl? I think the branding looks a bit fascy, but Iām open to the idea that theyāre not the only ones who get to use Vās instead of Uās. They too have a document outlining their principles which by my skimming are decidedly non-fascist, so letās just call the branding different tastes, maybe.