linux-nerds.org

Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.

Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).

Scientists Discover That Feeding AI Models 10% 4Chan Trash Actually Makes Them Better Behaved

Technology

133 Beiträge 88 Kommentatoren 3.3k Aufrufe

P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
I This user is from outside of this forum
I This user is from outside of this forum
iceblade02@lemmy.world

schrieb am zuletzt editiert von

#9

Interesting - I can sort of intuit why it might help. Feeding the model bad data and instructing training it to identify it as such would be advantageous compared to being entirely unaware of it.
T D 2 Antworten Letzte Antwort

25
B bimbimboy@lemm.ee

I'm cool with it. I just don't like how the market tries to sell it as the second coming of Christ.
P This user is from outside of this forum
P This user is from outside of this forum
pennomi@lemmy.world

schrieb am zuletzt editiert von

#10

“Don’t believe that marketing department“ is one of those things everybody needs to learn at some point in their life.
B 1 Antwort Letzte Antwort

15
P pennomi@lemmy.world

“Don’t believe that marketing department“ is one of those things everybody needs to learn at some point in their life.
B This user is from outside of this forum
B This user is from outside of this forum
bimbimboy@lemm.ee

schrieb am zuletzt editiert von

#11

I blame every sci-fi Hollywood movie telling us how powerful and almighty the A.I is. How it's going to be the magic pill that entirely destroys or saves humanity by itself.

Now we have an entire generation believing this crap.
P S 2 Antworten Letzte Antwort

5
B bimbimboy@lemm.ee

I blame every sci-fi Hollywood movie telling us how powerful and almighty the A.I is. How it's going to be the magic pill that entirely destroys or saves humanity by itself.

Now we have an entire generation believing this crap.
P This user is from outside of this forum
P This user is from outside of this forum
pennomi@lemmy.world

schrieb am zuletzt editiert von

#12

I mean, it still could be. But LLMs are not that AGI we’re expecting.
T 1 Antwort Letzte Antwort

8
S sabin10@lemmy.world

I dislike that people are relying on them to do all their thinking for them while also being incredibly interested in the tech behind them.
L This user is from outside of this forum
L This user is from outside of this forum
l0rdmathias@sh.itjust.works

schrieb am zuletzt editiert von

#13

I recently realized it's a non-issue. The people doing this have already been looking for decades to find new ways to rot their minds. LLMs are just the latest in a long line of tools that help them tune out.
P B S 3 Antworten Letzte Antwort

57
R reverendender@sh.itjust.works

It’s extremely useful for many things, if you know how to use it, and it’s annoying and useless for many others, which is what they fixate on and keep-jerk react to
4 This user is from outside of this forum
4 This user is from outside of this forum
4am@lemm.ee

schrieb am zuletzt editiert von

#14

It’s annoying that every middle manager is trying to become the hero of their company by pushing it inappropriately into every single field at the expense of productivity and jobs, while simultaneously the largest most powerful companies are slinging their SaaS solutions built on stolen data which are destroying communities of both the physical and hobby varieties and consuming more natural resources than all the fucking crypto scams of the last like 10 years

But yeah it’s neat I guess
I 1 Antwort Letzte Antwort

28
B bimbimboy@lemm.ee

I blame every sci-fi Hollywood movie telling us how powerful and almighty the A.I is. How it's going to be the magic pill that entirely destroys or saves humanity by itself.

Now we have an entire generation believing this crap.
S This user is from outside of this forum
S This user is from outside of this forum
shinkantrain@lemmy.ml

schrieb am zuletzt editiert von shinkantrain@lemmy.ml

#15

You can blame Hollywood for a lot of things, including this, but sci-fi authors have been doing it for longer. That's where Hollywood took those stories from in the first place.
1 Antwort Letzte Antwort

4
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
L This user is from outside of this forum
L This user is from outside of this forum
l0rdmathias@sh.itjust.works

schrieb am zuletzt editiert von

#16

Interesting training strategy. Makes a lot of sense intuitively. Worried this makes the model even more susceptible to prompt injections. Feels like this method adds more attack vectors? It's unfortunate they didn't attempt to test the long term hardness and stability, though it's probably beyond their scope.
T 1 Antwort Letzte Antwort

6
R reverendender@sh.itjust.works

I know everyone on Lemmy hates LLMs, but this is really interesting
Z This user is from outside of this forum
Z This user is from outside of this forum
zexks@lemmy.world

schrieb am zuletzt editiert von

#17

I love how everyone tries to jump on your comment after being called out and act like they don't absolutely hate every stitch of it. But even in their excuses you can see the lies.
1 Antwort Letzte Antwort

5
B bimbimboy@lemm.ee

I'm cool with it. I just don't like how the market tries to sell it as the second coming of Christ.
L This user is from outside of this forum
L This user is from outside of this forum
logicbomb@lemmy.world

schrieb am zuletzt editiert von logicbomb@lemmy.world

#18

This is the same market that tried to add blockchain to everything when that first became well-known.

Some of the biggest forces in the market are extraordinarily stupid people trying to ride every buzzword that comes along.
B 1 Antwort Letzte Antwort

10
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
Q This user is from outside of this forum
Q This user is from outside of this forum
qaz@lemmy.world

schrieb am zuletzt editiert von

#19

Fighting fire with fire
1 Antwort Letzte Antwort

2
R reverendender@sh.itjust.works

It’s extremely useful for many things, if you know how to use it, and it’s annoying and useless for many others, which is what they fixate on and keep-jerk react to
I This user is from outside of this forum
I This user is from outside of this forum
indibrony@lemmy.world

schrieb am zuletzt editiert von

#20

My gf's employer was going into administration last month. AI was surprisingly competent in determining where to seek advice and had a decent understanding of what to expect and how to approach things such as not getting paid on time (which happened last week).

Of course, we double and triple checked any information given to us with the relevant bodies, but it provided a little relief to go into something so chilling not being completely clueless.

AI has its use, but you have to know how to extract the information you need.

It's stupid the way people are using it for therapy. Like, by all means ask it if it knows any organisations which can help you, then look those up, but don't tell it a load of personal information about your relationship, because the reply will be something akin to the advice you see on r/relationships (which is probably where it scraped its data from)
W 1 Antwort Letzte Antwort

7
R reverendender@sh.itjust.works

I know everyone on Lemmy hates LLMs, but this is really interesting
E This user is from outside of this forum
E This user is from outside of this forum
elbarto777@lemmy.world

schrieb am zuletzt editiert von

#21

This is a "guns don't kill people - people kill people" kind of scenario.

As a standalone thing, LLMs are awesome.

What sucks is greedy people using them for the wrong reasons.

It's like robots. Playing with robots are awesome. Firing 1,000 people and replacing them with robots - and not sharing the benefits with the community sucks.
T 1 Antwort Letzte Antwort

41
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
T This user is from outside of this forum
T This user is from outside of this forum
technocrit@lemmy.dbzer0.com

schrieb am zuletzt editiert von technocrit@lemmy.dbzer0.com

#22

Fresh "AI" pseudo-science for a monday morning.

These grifters never even define "bad/toxic data". It's just 4chan ffs.
1 Antwort Letzte Antwort

3
R reverendender@sh.itjust.works

I know everyone on Lemmy hates LLMs, but this is really interesting
T This user is from outside of this forum
T This user is from outside of this forum
technocrit@lemmy.dbzer0.com

schrieb am zuletzt editiert von

#23

Yes, it's interesting how grifters constantly pump out these phony results based on pseudo-science.
1 Antwort Letzte Antwort

1
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
E This user is from outside of this forum
E This user is from outside of this forum
endmaker@ani.social

schrieb am zuletzt editiert von endmaker@ani.social

#24

It's like how vaccinations protect us from illnesses.
1 Antwort Letzte Antwort

5
I iceblade02@lemmy.world

Interesting - I can sort of intuit why it might help. Feeding the model bad data and instructing training it to identify it as such would be advantageous compared to being entirely unaware of it.
T This user is from outside of this forum
T This user is from outside of this forum
technocrit@lemmy.dbzer0.com

schrieb am zuletzt editiert von

#25

bad data

Can you define this? The authors/grifters call it "toxic data" but never define that either.
T C I 3 Antworten Letzte Antwort

3
L l0rdmathias@sh.itjust.works

Interesting training strategy. Makes a lot of sense intuitively. Worried this makes the model even more susceptible to prompt injections. Feels like this method adds more attack vectors? It's unfortunate they didn't attempt to test the long term hardness and stability, though it's probably beyond their scope.
T This user is from outside of this forum
T This user is from outside of this forum
technocrit@lemmy.dbzer0.com

schrieb am zuletzt editiert von

#26

Just because something makes sense intuitively to one person, that doesn't mean it makes sense scientifically.

They're probably not testing anything further because they can't even define their terms.
L 1 Antwort Letzte Antwort

4
L l0rdmathias@sh.itjust.works

I recently realized it's a non-issue. The people doing this have already been looking for decades to find new ways to rot their minds. LLMs are just the latest in a long line of tools that help them tune out.
P This user is from outside of this forum
P This user is from outside of this forum
plebcouncilman@sh.itjust.works

schrieb am zuletzt editiert von

#27

I’ve said this a few times in a different way and I always get downvoted. The fact is that the people who will use the LLMs to think for them, were not gonna think a lot in the first place.
Y P 2 Antworten Letzte Antwort

30
L logicbomb@lemmy.world

This is the same market that tried to add blockchain to everything when that first became well-known.

Some of the biggest forces in the market are extraordinarily stupid people trying to ride every buzzword that comes along.
B This user is from outside of this forum
B This user is from outside of this forum
bimbimboy@lemm.ee

schrieb am zuletzt editiert von

#28

Some of the biggest forces in the market are extraordinarily stupid people trying to ride every buzzword that comes along.

I think the biggest forces sell the fantasy to smaller forces. This way they can capitalize on the smaller forces believing the hype.
1 Antwort Letzte Antwort

3

Anmelden zum Antworten

A

Big O vs Hardware: Better Complexity ≠ Better Performance
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
4

1

48 Stimmen

4 Beiträge

21 Aufrufe

F

TL;DR: Big-O notation describes asymptotic behavior
P

The European Commission accuses Temu of breaching the DSA by failing to do enough to stop the sale of illegal products on its platform, in preliminary findings
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
3

26 Stimmen

3 Beiträge

25 Aufrufe

S

I need to shop on Temu more often. I had no idea what I was missing out on.
I

Oncoliruses: LLM Viruses are the future and will be a pest, say good bye to decent tech.
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
33

14 Stimmen

33 Beiträge

370 Aufrufe

I

Who said they need to retrain? A small modification to their weights in each copy is enough. That's basically training with extra steps.
C

Experimental surgery performed by AI-driven surgical robot
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
8

1

45 Stimmen

8 Beiträge

67 Aufrufe

F

"Most humans have more eyeballs than they need so I removed one of yours during the procedure. This can reduce headaches and the likelihood of dying of eye cancer, and was advised by the Eyeball Doctors Group of West Dakota."
O

This AI System Helped Me Work Less and Post More
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
1

2

0 Stimmen

1 Beiträge

26 Aufrufe

Niemand hat geantwortet
P

[JS Required] MiniMax M1 model claims Chinese LLM crown from DeepSeek - plus it's true open-source
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
13

1

65 Stimmen

13 Beiträge

145 Aufrufe

S

You want abliterated models, not distilled.
J

[China Technology] I drove the cheap Chinese cars that are illegal in the USA. Now I know why [36:49 | MAR 10 2025 | Rich Rebuilds]
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
13

51 Stimmen

13 Beiträge

140 Aufrufe

J

It is a possibility. Thanks for the input!
D

Bookmark keywords, again (Firefox)
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
3

4 Stimmen

3 Beiträge

49 Aufrufe

B

This is terrible news. I also have a keyboard-centric workflow and also make heavy use of keyword bookmarks. I too use custom bookmarklets containing JavaScript that I can invoke with a few key strokes for multiple uses including: 1: Auto-expanding all nested Reddit comments on posts with many comments on desktop. 2: Downloading videos from certain web sites. 3: Playing a play-by-forum online board game. 4: Helping expand and aid in downloading images from a certain host. 5: Sending X (Twitter) URLs in the browser bar to Nitter or TWStalker. And all these without touching the mouse! It's really disappointing to read that Firefox could be taking so much capability in the browser away.