linux-nerds.org

Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.

Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).

Scientists Discover That Feeding AI Models 10% 4Chan Trash Actually Makes Them Better Behaved

Technology

131 Beiträge 87 Kommentatoren 0 Aufrufe

P plebcouncilman@sh.itjust.works

Well I would make the argument that someone stupid enough to do such a thing kinda deserves whatever consequences their actions have. I find that people learn faster when actions have consequences instead of everything being babyproofed.
D This user is from outside of this forum
D This user is from outside of this forum
dojan@pawb.social

schrieb zuletzt editiert von

#91

The rest of us will be stuck with those consequences also. When idiots are at work, third party always suffers.
1 Antwort Letzte Antwort

1
I imgonnatrythis@sh.itjust.works

Boy, I don't even know if I wish that much 4chan on a LLM.
S This user is from outside of this forum
S This user is from outside of this forum
squizzy@lemmy.world

schrieb zuletzt editiert von

#92

It is truly a bizzare world, I went there first to be edgy as an early teen and seeing boobs is fun, then I saw a dude live post his murder of a woman he liked while everyone called her names.

It makes a great case for moderation if not banning the internet.
1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.

When Bad Data Leads to Good Models

Abstract page for arXiv paper 2505.04741: When Bad Data Leads to Good Models

arXiv.org (arxiv.org)
T This user is from outside of this forum
T This user is from outside of this forum
tungsten5@lemm.ee

schrieb zuletzt editiert von

#93

Give the AI model the gift of culture and class. No suprise it behaves better
E 1 Antwort Letzte Antwort

18
T tungsten5@lemm.ee

Give the AI model the gift of culture and class. No suprise it behaves better
E This user is from outside of this forum
E This user is from outside of this forum
echosnail@lemmy.zip

schrieb zuletzt editiert von

#94

Sophistication my good sir.
1 Antwort Letzte Antwort

10
L lka1988@lemmy.dbzer0.com

This is one instance where I'm ok with the occasional beating. It's a computer. It doesn't have feelings. It never will. It's not sentient.
E This user is from outside of this forum
E This user is from outside of this forum
echosnail@lemmy.zip

schrieb zuletzt editiert von

#95

You say all this until ChatGpt convinced you to write a manifesto to "take back" your foreskin from the Jews.
L 1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.

When Bad Data Leads to Good Models

Abstract page for arXiv paper 2505.04741: When Bad Data Leads to Good Models

arXiv.org (arxiv.org)
N This user is from outside of this forum
N This user is from outside of this forum
naevermix@lemmy.world

schrieb zuletzt editiert von naevermix@lemmy.world

#96

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
D P 2 Antworten Letzte Antwort

9
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.

When Bad Data Leads to Good Models

Abstract page for arXiv paper 2505.04741: When Bad Data Leads to Good Models

arXiv.org (arxiv.org)
G This user is from outside of this forum
G This user is from outside of this forum
goodboyjojo@lemm.ee

schrieb zuletzt editiert von

#97

Based and hopepilled
1 Antwort Letzte Antwort

1
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.

When Bad Data Leads to Good Models

Abstract page for arXiv paper 2505.04741: When Bad Data Leads to Good Models

arXiv.org (arxiv.org)
C This user is from outside of this forum
C This user is from outside of this forum
cupcakezealot@lemmy.blahaj.zone

schrieb zuletzt editiert von

#98

can we stop referring to llm's as if they're capable of thought? they don't make decisions; their programming just responds to patterns.
M 1 Antwort Letzte Antwort

8
N naevermix@lemmy.world

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
D This user is from outside of this forum
D This user is from outside of this forum
disaster@sh.itjust.works

schrieb zuletzt editiert von

#99

There's little evidence that debate changes people's ideas.
N G 2 Antworten Letzte Antwort

7
D disaster@sh.itjust.works

There's little evidence that debate changes people's ideas.
N This user is from outside of this forum
N This user is from outside of this forum
naevermix@lemmy.world

schrieb zuletzt editiert von

#100

It's not about changing their ideas. The target is the audience.
1 Antwort Letzte Antwort

2
N naevermix@lemmy.world

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
P This user is from outside of this forum
P This user is from outside of this forum
pushbutton@lemmy.world

schrieb zuletzt editiert von

#101

it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

I was looking for the person saying a particular quote yesterday.

I asked 3 times the same question and I got 3 different people.

The funny part us I had the quote wrong.

Bullshit all the way down.
1 Antwort Letzte Antwort

1
D disaster@sh.itjust.works

There's little evidence that debate changes people's ideas.
G This user is from outside of this forum
G This user is from outside of this forum
gonzako@lemmy.world

schrieb zuletzt editiert von

#102

yeah, this only works in scientific fields
M 1 Antwort Letzte Antwort

1
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.

When Bad Data Leads to Good Models

Abstract page for arXiv paper 2505.04741: When Bad Data Leads to Good Models

arXiv.org (arxiv.org)
Y This user is from outside of this forum
Y This user is from outside of this forum
yournamehere@lemm.ee

schrieb zuletzt editiert von

#103

because 4chan users write original content. that is fed into the next best stupid platform and so on until it ends on tiktok or whatever.

if you have nothing to say you use meta/tiktok. no relevabt content has ever been there first.
copies and derivates, yes...

so soonish AI will flood 4chan so ai scrapers get polluted aswell...and then it is dead.
S 1 Antwort Letzte Antwort

9
R reverendender@sh.itjust.works

I know everyone on Lemmy hates LLMs, but this is really interesting
T This user is from outside of this forum
T This user is from outside of this forum
timeworntraveler@lemm.ee

schrieb zuletzt editiert von timeworntraveler@lemm.ee

#104

I do hate LLMs (or how they're marketed/hyped/used) and I concur that this is very interesting science
R 1 Antwort Letzte Antwort

3
E echosnail@lemmy.zip

You say all this until ChatGpt convinced you to write a manifesto to "take back" your foreskin from the Jews.
L This user is from outside of this forum
L This user is from outside of this forum
lka1988@lemmy.dbzer0.com

schrieb zuletzt editiert von

#105

Funny enough, I am circumcised. But no, if I wanted it back that badly, I'd write it myself.
1 Antwort Letzte Antwort

1
A anaveragesnoot@lemmy.ca

I don't dislike LLMs, I dislike people who treat them as anything more than an advanced search engine and stupidly give them all their confidential data. Seen it happen too much at work.
I This user is from outside of this forum
I This user is from outside of this forum
ipkpjersi@lemmy.ml

schrieb zuletzt editiert von

#106

Yep. My work is very strict about security except for when it comes to LLMs, and then suddenly they're surprisingly lax about it. It's a bit concerning actually.
1 Antwort Letzte Antwort

0
T timeworntraveler@lemm.ee

I do hate LLMs (or how they're marketed/hyped/used) and I concur that this is very interesting science
R This user is from outside of this forum
R This user is from outside of this forum
reverendender@sh.itjust.works

schrieb zuletzt editiert von

#107

I appreciate your reasoned and measured reply, friend!
1 Antwort Letzte Antwort

0
_ _thebrain_@sh.itjust.works

Underrated comment.
F This user is from outside of this forum
F This user is from outside of this forum
feathercrown@lemmy.world

schrieb zuletzt editiert von

#108

Seems pretty rated to me
1 Antwort Letzte Antwort

2
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.

When Bad Data Leads to Good Models

Abstract page for arXiv paper 2505.04741: When Bad Data Leads to Good Models

arXiv.org (arxiv.org)
M This user is from outside of this forum
M This user is from outside of this forum
mangionedontmiss@lemmy.ca

schrieb zuletzt editiert von

#109

goddamn, has 4chan gone so far down the road that its actually come back around and become the good guy?
1 Antwort Letzte Antwort

0
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.

When Bad Data Leads to Good Models

Abstract page for arXiv paper 2505.04741: When Bad Data Leads to Good Models

arXiv.org (arxiv.org)
M This user is from outside of this forum
M This user is from outside of this forum
mr_dr_oink@lemmy.world

schrieb zuletzt editiert von

#110

So is it saying essentially that in order to not output garbage, it needs to know first what garbage is?

Is it just me that things this seems like a no-brainer?

It almosr draws parallels to many societal issues. Knowledge is power.

People tend towards intolerance and hatred when they dont understand the thing they are angry at. The more they know the better they behave.
H M 2 Antworten Letzte Antwort

16

Anmelden zum Antworten

P

Windows 11 remote desktop microphone stops working intermittently
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
7

16 Stimmen

7 Beiträge

4 Aufrufe

S

When I worked in IT, we only let people install every other version of Windows. Our Linux user policy was always “mainstream distro and the LTS version.” Mac users were strongly advised to wait 3 months to upgrade. One guy used FreeBSD and I just never questioned him because he was older and never filed one help desk request. He probably thought I was an idiot. (And I was.) Anyway, I say all that to say don’t use Windows 11 on anything important. It’s the equivalent of a beta. Windows 12 (or however they brand it) will probably be stable. I don’t use Windows much anymore and maybe things have changed but the concepts in the previous paragraph could be outdated. But it’s a good rule of thumb.
T

The FDA Is Approving Drugs Without Evidence They Work
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
69

1

507 Stimmen

69 Beiträge

10 Aufrufe

L

Now you hit me curious too. This was my source on Texas https://www.texasalmanac.com/place-types/town Also the total number of total towns is over 4,000 with only 3k unincorporated, I did get the numbers wrong even in Texas. I had looked at Wikipedia but could not find totals, only lists
P

[UK] Live facial recognition cameras may become ‘commonplace’ as police use soars
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
12

1

36 Stimmen

12 Beiträge

2 Aufrufe

C

Definitely don't want to be painting my face every day
P

Gumroad Founder Sahil Lavingia Reveals He Was Let Go from DOGE as Software Engineer for the Department of Veterans Affairs After Just 55 Days
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
12

1

62 Stimmen

12 Beiträge

0 Aufrufe

M

is the linked article or the title edited? This was a post about VA GPT
A

Duolingo CEO says AI is a better teacher than humans—but schools will exist ‘because you still need childcare’
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
19

1

1 Stimmen

19 Beiträge

2 Aufrufe

L

Where and what is texas?
P

“Treat Online Abuse Like Spam”: New Report Urges Social Media Platforms to Fight Online Abuse with Tools Users Can Control
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
3

1

0 Stimmen

3 Beiträge

0 Aufrufe

E

Nextdoor is an absolute black hole social media site, it absorbs the worst of humanity so we don't have to see them anywhere else.
G

Nextcloud cries foul over Google Play Store app rejection
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
31

1

256 Stimmen

31 Beiträge

3 Aufrufe

S

I have the regular F-droid and it does automatic updates now.
S

FCC commissioner writes op-ed titled, “It’s time for Trump to DOGE the FCC“
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
43

1

342 Stimmen

43 Beiträge

6 Aufrufe

G

highly recommend using containerized torrents through a VPN. I have transmission and openvpn containers. when the network goes down transmission can't connect since it's networked through the ovpn container. once the vpn is restored, everything restarts and resumes where it left off. ever since I've had this setup running, I haven't had a nastygram sent to me.