linux-nerds.org

Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.

Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).

Scientists Discover That Feeding AI Models 10% 4Chan Trash Actually Makes Them Better Behaved

Technology

133 Beiträge 88 Kommentatoren 0 Aufrufe

I iceblade02@lemmy.world

Interesting - I can sort of intuit why it might help. Feeding the model bad data and instructing training it to identify it as such would be advantageous compared to being entirely unaware of it.
D This user is from outside of this forum
D This user is from outside of this forum
danzabia@infosec.pub

schrieb zuletzt editiert von danzabia@infosec.pub

#90

Yeah, it's like me never having alcohol before and walking into a frat party as a freshman. Sometimes it's better to come prepared.
1 Antwort Letzte Antwort

2
P plebcouncilman@sh.itjust.works

Well I would make the argument that someone stupid enough to do such a thing kinda deserves whatever consequences their actions have. I find that people learn faster when actions have consequences instead of everything being babyproofed.
D This user is from outside of this forum
D This user is from outside of this forum
dojan@pawb.social

schrieb zuletzt editiert von

#91

The rest of us will be stuck with those consequences also. When idiots are at work, third party always suffers.
1 Antwort Letzte Antwort

1
I imgonnatrythis@sh.itjust.works

Boy, I don't even know if I wish that much 4chan on a LLM.
S This user is from outside of this forum
S This user is from outside of this forum
squizzy@lemmy.world

schrieb zuletzt editiert von

#92

It is truly a bizzare world, I went there first to be edgy as an early teen and seeing boobs is fun, then I saw a dude live post his murder of a woman he liked while everyone called her names.

It makes a great case for moderation if not banning the internet.
1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
T This user is from outside of this forum
T This user is from outside of this forum
tungsten5@lemm.ee

schrieb zuletzt editiert von

#93

Give the AI model the gift of culture and class. No suprise it behaves better
E 1 Antwort Letzte Antwort

18
T tungsten5@lemm.ee

Give the AI model the gift of culture and class. No suprise it behaves better
E This user is from outside of this forum
E This user is from outside of this forum
echosnail@lemmy.zip

schrieb zuletzt editiert von

#94

Sophistication my good sir.
1 Antwort Letzte Antwort

10
L lka1988@lemmy.dbzer0.com

This is one instance where I'm ok with the occasional beating. It's a computer. It doesn't have feelings. It never will. It's not sentient.
E This user is from outside of this forum
E This user is from outside of this forum
echosnail@lemmy.zip

schrieb zuletzt editiert von

#95

You say all this until ChatGpt convinced you to write a manifesto to "take back" your foreskin from the Jews.
L 1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
N This user is from outside of this forum
N This user is from outside of this forum
naevermix@lemmy.world

schrieb zuletzt editiert von naevermix@lemmy.world

#96

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
D P 2 Antworten Letzte Antwort

10
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
G This user is from outside of this forum
G This user is from outside of this forum
goodboyjojo@lemm.ee

schrieb zuletzt editiert von

#97

Based and hopepilled
1 Antwort Letzte Antwort

1
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
C This user is from outside of this forum
C This user is from outside of this forum
cupcakezealot@lemmy.blahaj.zone

schrieb zuletzt editiert von

#98

can we stop referring to llm's as if they're capable of thought? they don't make decisions; their programming just responds to patterns.
M 1 Antwort Letzte Antwort

9
N naevermix@lemmy.world

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
D This user is from outside of this forum
D This user is from outside of this forum
disaster@sh.itjust.works

schrieb zuletzt editiert von

#99

There's little evidence that debate changes people's ideas.
N G P 3 Antworten Letzte Antwort

7
D disaster@sh.itjust.works

There's little evidence that debate changes people's ideas.
N This user is from outside of this forum
N This user is from outside of this forum
naevermix@lemmy.world

schrieb zuletzt editiert von

#100

It's not about changing their ideas. The target is the audience.
1 Antwort Letzte Antwort

2
N naevermix@lemmy.world

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
P This user is from outside of this forum
P This user is from outside of this forum
pushbutton@lemmy.world

schrieb zuletzt editiert von

#101

it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

I was looking for the person saying a particular quote yesterday.

I asked 3 times the same question and I got 3 different people.

The funny part us I had the quote wrong.

Bullshit all the way down.
1 Antwort Letzte Antwort

1
D disaster@sh.itjust.works

There's little evidence that debate changes people's ideas.
G This user is from outside of this forum
G This user is from outside of this forum
gonzako@lemmy.world

schrieb zuletzt editiert von

#102

yeah, this only works in scientific fields
M 1 Antwort Letzte Antwort

1
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
Y This user is from outside of this forum
Y This user is from outside of this forum
yournamehere@lemm.ee

schrieb zuletzt editiert von

#103

because 4chan users write original content. that is fed into the next best stupid platform and so on until it ends on tiktok or whatever.

if you have nothing to say you use meta/tiktok. no relevabt content has ever been there first.
copies and derivates, yes...

so soonish AI will flood 4chan so ai scrapers get polluted aswell...and then it is dead.
S 1 Antwort Letzte Antwort

9
R reverendender@sh.itjust.works

I know everyone on Lemmy hates LLMs, but this is really interesting
T This user is from outside of this forum
T This user is from outside of this forum
timeworntraveler@lemm.ee

schrieb zuletzt editiert von timeworntraveler@lemm.ee

#104

I do hate LLMs (or how they're marketed/hyped/used) and I concur that this is very interesting science
R 1 Antwort Letzte Antwort

3
E echosnail@lemmy.zip

You say all this until ChatGpt convinced you to write a manifesto to "take back" your foreskin from the Jews.
L This user is from outside of this forum
L This user is from outside of this forum
lka1988@lemmy.dbzer0.com

schrieb zuletzt editiert von

#105

Funny enough, I am circumcised. But no, if I wanted it back that badly, I'd write it myself.
1 Antwort Letzte Antwort

1
A anaveragesnoot@lemmy.ca

I don't dislike LLMs, I dislike people who treat them as anything more than an advanced search engine and stupidly give them all their confidential data. Seen it happen too much at work.
I This user is from outside of this forum
I This user is from outside of this forum
ipkpjersi@lemmy.ml

schrieb zuletzt editiert von

#106

Yep. My work is very strict about security except for when it comes to LLMs, and then suddenly they're surprisingly lax about it. It's a bit concerning actually.
1 Antwort Letzte Antwort

0
T timeworntraveler@lemm.ee

I do hate LLMs (or how they're marketed/hyped/used) and I concur that this is very interesting science
R This user is from outside of this forum
R This user is from outside of this forum
reverendender@sh.itjust.works

schrieb zuletzt editiert von

#107

I appreciate your reasoned and measured reply, friend!
1 Antwort Letzte Antwort

0
_ _thebrain_@sh.itjust.works

Underrated comment.
F This user is from outside of this forum
F This user is from outside of this forum
feathercrown@lemmy.world

schrieb zuletzt editiert von

#108

Seems pretty rated to me
1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
M This user is from outside of this forum
M This user is from outside of this forum
mangionedontmiss@lemmy.ca

schrieb zuletzt editiert von

#109

goddamn, has 4chan gone so far down the road that its actually come back around and become the good guy?
1 Antwort Letzte Antwort

0

Anmelden zum Antworten

P

Computer says no: Impact of automated decision-making on human life; Algorithms are deciding whether a patient receives an organ transplant or not; Algorithms use in Welfare, Penalise the poor.
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
13

1

180 Stimmen

13 Beiträge

0 Aufrufe

D

There is a huge difference between an algorithm using real world data to produce a score a panel of experts use to make a determination and using a LLM to screen candidates. One has verifiable reproducible results that can be checked and debated the other does not. The final call does not matter if a computer program using an unknown and unreproducible algorithm screens you out before this. This is what we are facing. Pre-determined decisions that human beings are not being held accountable to. Is this happening right now? Yes it is, without a doubt. People are no longer making a lot of healthcare decisions determining insurance coverage. Computers that are not accountable are. You may have some ability to disagree but for how long? Soon there will be no way to reach a human about an insurance decision. This is already happening. People should be very anxious. Hearing United Healthcare has been forging DNRs and has been denying things like treatment for stroke for elders is disgusting. We have major issues that are not going away and we are blatantly ignoring them.
S

YouTube’s Deliberate Indifference Exposes Kids to disgusting Content
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
46

1

257 Stimmen

46 Beiträge

5 Aufrufe

S

yea i also were there at a few thousand I think and the content has changed a lot since then.
R

Hiring Developers in Eastern Europe
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
1

0 Stimmen

1 Beiträge

1 Aufrufe

Niemand hat geantwortet
P

AI cheating surge pushes schools into chaos
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
25

45 Stimmen

25 Beiträge

2 Aufrufe

C

Sorry for the late reply, I had to sit and think on this one for a little bit. I think there are would be a few things going on when it comes to designing a course to teach critical thinking, nuances, and originality; and they each have their own requirements. For critical thinking: The main goal is to provide students with a toolbelt for solving various problems. Then instilling the habit of always asking "does this match the expected outcome? What was I expecting?". So usually courses will be setup so students learn about a tool, practice using the tool, then have a culminating assignment on using all the tools. Ideally, the problems students face at the end require multiple tools to solve. Nuance mainly naturally comes with exposure to the material from a professional - The way a mechanical engineer may describe building a desk will probably differ greatly compared to a fantasy author. You can also explain definitions and industry standards; but thats really dry. So I try to teach nuances via definitions by mixing in the weird nuances as much as possible with jokes. Then for originality; I've realized I dont actually look for an original idea; but something creative. In a classroom setting, you're usually learning new things about a subject so a student's knowledge of that space is usually very limited. Thus, an idea that they've never heard about may be original to them, but common for an industry expert. For teaching originality creativity, I usually provide time to be creative & think, and provide open ended questions as prompts to explore ideas. My courses that require originality usually have it as a part of the culminating assignment at the end where they can apply their knowledge. I'll also add in time where students can come to me with preliminary ideas and I can provide feedback on whether or not it passes the creative threshold. Not all ideas are original, but I sometimes give a bit of slack if its creative enough. The amount of course overhauling to get around AI really depends on the material being taught. For example, in programming - you teach critical thinking by always testing your code, even with parameters that don't make sense. For example: Try to add 123 + "skibbidy", and see what the program does.
A

Unhappy with the recently lost file upload feature in the Nextcloud app for Android? So are we. Let us explain. - Nextcloud
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
7

0 Stimmen

7 Beiträge

2 Aufrufe

C

Oh this is a good callout, I'm definitely using wired and not wireless.
B

DOGE Plan to Push AI Across the US Federal Government is Wildly Dangerous
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
2

1

0 Stimmen

2 Beiträge

2 Aufrufe

V

Here's how you know it's not ready: AI hasn't replaced a single CEO.
V

China aims to recruit top US scientists as Trump tries to kill the CHIPS Act
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
1

1

0 Stimmen

1 Beiträge

1 Aufrufe

Niemand hat geantwortet
D

Disney to implement measures limiting password sharing on streaming services from June
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
2

0 Stimmen

2 Beiträge

2 Aufrufe

A

The enshittification continues, but it doesn't affect me at all. Piracy is the way to go nowadays that all streaming services suck. !piracy@lemmy.dbzer0.com