linux-nerds.org

Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.

Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).

Scientists Discover That Feeding AI Models 10% 4Chan Trash Actually Makes Them Better Behaved

Technology

133 Beiträge 88 Kommentatoren 0 Aufrufe

S sculptuspoe@lemmy.world

I wish they would tone down the crusade. This is some of the most interesting technology to come out in decades.
C This user is from outside of this forum
C This user is from outside of this forum
cosmonova@lemmy.world

schrieb zuletzt editiert von

#80

And I wish they would tone down the hype. Maybe we can meet in the middle?
S 1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
I This user is from outside of this forum
I This user is from outside of this forum
imgonnatrythis@sh.itjust.works

schrieb zuletzt editiert von

#81

Boy, I don't even know if I wish that much 4chan on a LLM.
S 1 Antwort Letzte Antwort

13
Y youcancallmedragon@lemmy.world

This is true, but we don’t need people putting glue on their pizza. These people used to have a person to ask now they’ll be asking Sam Altman
S This user is from outside of this forum
S This user is from outside of this forum
scubus@sh.itjust.works

schrieb zuletzt editiert von

#82

No, we were juat eating tide pods. Dumb gonna do what dumb gonna do. The only real issue with llms is that their training data is stolen, and that theyre currently not that useful due to hallucinations and lacking logical reasoning.
1 Antwort Letzte Antwort

2
T taladar@sh.itjust.works

There are plenty of tasks which they solve perfectly, today.

Name a single task you would trust an LLM on solving for you that you feel confident would be correct without checking the output. Because that is my definition of perfectly and AI falls very, very far short of that.
Y This user is from outside of this forum
Y This user is from outside of this forum
yamper@lemmy.world

schrieb zuletzt editiert von

#83

i used it when i traveled to japan to ask it for english->japanese translations. it gave back results for multiple contexts, politeness levels, and broke down each sentence into its parts. my native speaker friends validated a few responses.

if youre going to be pedantic about "perfect" then nothing, not even a human, is going to live up.

willful ignorance about the things ai can be good at today is not going to do any favors for your fight against ai in the future. know your enemy and all that.
1 Antwort Letzte Antwort

3
O ohwhatfollyisman@lemmy.world

10% 4chan

why didn't they just say 0.4chan and be done with it?
C This user is from outside of this forum
C This user is from outside of this forum
crash_thepose@lemmy.ml

schrieb zuletzt editiert von

#84

Best comment I've read this week
1 Antwort Letzte Antwort

8
N nostradavid@programming.dev

I like LLMs. Instead of making a racket, I just use them, which may make it seem like everyone on Lemmy hates LLMs.
C This user is from outside of this forum
C This user is from outside of this forum
crash_thepose@lemmy.ml

schrieb zuletzt editiert von

#85

Being a teacher In academia is what makes me hate them tbh
1 Antwort Letzte Antwort

0
R rachelhazideas@lemmy.world

To come out of 4chan a better person, one must transcend humanity.
L This user is from outside of this forum
L This user is from outside of this forum
laintrain@lemmy.dbzer0.com

schrieb zuletzt editiert von

#86

I think plenty do come away better people because honestly I know plenty of people who were on there when they were younger but are normal well-adjusted adults now, and also me.
1 Antwort Letzte Antwort

2
L lka1988@lemmy.dbzer0.com

This is one instance where I'm ok with the occasional beating. It's a computer. It doesn't have feelings. It never will. It's not sentient.
G This user is from outside of this forum
G This user is from outside of this forum
gradually_adjusting@lemmy.world

schrieb zuletzt editiert von

#87

You can't change the machines, but try not to let them change you.
1 Antwort Letzte Antwort

1
D deathsembrace@lemmy.world

You can make a generic fill in the blanks for all of those like I do and just change the key terminology for each scenario. LLMs are competing with search and replace?
D This user is from outside of this forum
D This user is from outside of this forum
danzabia@infosec.pub

schrieb zuletzt editiert von

#88

I think this may be a skill issue on your part.
1 Antwort Letzte Antwort

0
T tarquinn2049@lemmy.world

They are essentially a fun toy for most people, and an ok tool for people with the patience and training to get useful output from them. And they cost an insane amount of money to train and an insane amount of power to run.

Not to mention the other cost of training them, the human emotional cost. And the human cost of running them.

It just costs so much of a variety of things, for an output that has barely made anything better. Maybe they might get "better" in the future, and have to get through this stage to get there, but I've also seen a lot of people saying they appear to be starting to plateau... maybe a temporary plateau, but if so, how temporary? Could we just drop it for 10 years and start back up when they won't be as inefficient? Maybe a law that they have to pay for everything they feed it, would effectively cause them to only emerge at a time when they are actually feasible.
D This user is from outside of this forum
D This user is from outside of this forum
danzabia@infosec.pub

schrieb zuletzt editiert von

#89

People who track performance (like METR, a nonprofit) indicate that progress is, if anything, speeding up. Most people's use case is so simple they can't detect the difference. However for cases like complex problem solving, agentic tasks, etc you can in fact see significant progress happening. This should be concerning if you think the world isn't ready for labor displaced by LLMs.
1 Antwort Letzte Antwort

1
I iceblade02@lemmy.world

Interesting - I can sort of intuit why it might help. Feeding the model bad data and instructing training it to identify it as such would be advantageous compared to being entirely unaware of it.
D This user is from outside of this forum
D This user is from outside of this forum
danzabia@infosec.pub

schrieb zuletzt editiert von danzabia@infosec.pub

#90

Yeah, it's like me never having alcohol before and walking into a frat party as a freshman. Sometimes it's better to come prepared.
1 Antwort Letzte Antwort

2
P plebcouncilman@sh.itjust.works

Well I would make the argument that someone stupid enough to do such a thing kinda deserves whatever consequences their actions have. I find that people learn faster when actions have consequences instead of everything being babyproofed.
D This user is from outside of this forum
D This user is from outside of this forum
dojan@pawb.social

schrieb zuletzt editiert von

#91

The rest of us will be stuck with those consequences also. When idiots are at work, third party always suffers.
1 Antwort Letzte Antwort

1
I imgonnatrythis@sh.itjust.works

Boy, I don't even know if I wish that much 4chan on a LLM.
S This user is from outside of this forum
S This user is from outside of this forum
squizzy@lemmy.world

schrieb zuletzt editiert von

#92

It is truly a bizzare world, I went there first to be edgy as an early teen and seeing boobs is fun, then I saw a dude live post his murder of a woman he liked while everyone called her names.

It makes a great case for moderation if not banning the internet.
1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
T This user is from outside of this forum
T This user is from outside of this forum
tungsten5@lemm.ee

schrieb zuletzt editiert von

#93

Give the AI model the gift of culture and class. No suprise it behaves better
E 1 Antwort Letzte Antwort

18
T tungsten5@lemm.ee

Give the AI model the gift of culture and class. No suprise it behaves better
E This user is from outside of this forum
E This user is from outside of this forum
echosnail@lemmy.zip

schrieb zuletzt editiert von

#94

Sophistication my good sir.
1 Antwort Letzte Antwort

10
L lka1988@lemmy.dbzer0.com

This is one instance where I'm ok with the occasional beating. It's a computer. It doesn't have feelings. It never will. It's not sentient.
E This user is from outside of this forum
E This user is from outside of this forum
echosnail@lemmy.zip

schrieb zuletzt editiert von

#95

You say all this until ChatGpt convinced you to write a manifesto to "take back" your foreskin from the Jews.
L 1 Antwort Letzte Antwort

3
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
N This user is from outside of this forum
N This user is from outside of this forum
naevermix@lemmy.world

schrieb zuletzt editiert von naevermix@lemmy.world

#96

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
D P 2 Antworten Letzte Antwort

10
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
G This user is from outside of this forum
G This user is from outside of this forum
goodboyjojo@lemm.ee

schrieb zuletzt editiert von

#97

Based and hopepilled
1 Antwort Letzte Antwort

1
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
C This user is from outside of this forum
C This user is from outside of this forum
cupcakezealot@lemmy.blahaj.zone

schrieb zuletzt editiert von

#98

can we stop referring to llm's as if they're capable of thought? they don't make decisions; their programming just responds to patterns.
M 1 Antwort Letzte Antwort

9
N naevermix@lemmy.world

I envision a Gemini powered bot that cracks captcha and posts "woke" replies on 4chan. If you're an antivaxxer, antisemite, nazi, racist, sionist, or otherwise, it will debate you. It will not get tired. It will not get mad. It will maintain a sense of decorum indefinitely and it will never ever stop. If some far right extremist decides to do the same, it will have the advantage that academia is left leaning, meaning the model can cite widely recognized studies.

Dead internet theory and so on, but I'll gladly completely and utterly destroy the internet if it means the filth dies with it.
D This user is from outside of this forum
D This user is from outside of this forum
disaster@sh.itjust.works

schrieb zuletzt editiert von

#99

There's little evidence that debate changes people's ideas.
N G P 3 Antworten Letzte Antwort

7

Anmelden zum Antworten

X

Generative AI's most prominent skeptic doubles down
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
14

1

43 Stimmen

14 Beiträge

2 Aufrufe

Z

I don't think so, and I believe not even the current technology used for neural network simulations will bring us to AGI, yet alone LLMs.
F

MCP 101: An Introduction to the MCP Standard
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
2

1

5 Stimmen

2 Beiträge

2 Aufrufe

H

Really? [image: 60a7b1c3-946c-4def-92dd-c04169f01892.gif]
K

The U.S. Just Ran a Solar Storm Emergency Drill. The Real Deal Would Be a Catastrophe
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
19

1

149 Stimmen

19 Beiträge

4 Aufrufe

C

Got it, at that point (extremely high voltage) you'd need suppression at the panel. Which I would hope people have inline, but not expect like an LVD.
P

30% of South Korean schools have adopted AI-powered digital textbooks since the country's education ministry began a full-scale rollout in March 2025
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
16

1

35 Stimmen

16 Beiträge

0 Aufrufe

M

This is what I want to know also. "AI textbooks" is a great clickbait/ragebait term, but could mean a great variety of things.
D

Chrome using Gemini Nano for ‘Enhanced Protection’ against scams
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
8

1

1 Stimmen

8 Beiträge

3 Aufrufe

L

I think the principle could be applied to scan outside of the machine. It is making requests to 127.0.0.1:{port} - effectively using your computer as a "server" in a sort of reverse-SSRF attack. There's no reason it can't make requests to 10.10.10.1:{port} as well. Of course you'd need to guess the netmask of the network address range first, but this isn't that hard. In fact, if you consider that at least as far as the desktop site goes, most people will be browsing the web behind a standard consumer router left on defaults where it will be the first device in the DHCP range (e.g. 192.168.0.1 or 10.10.10.1), which tends to have a web UI on the LAN interface (port 8080, 80 or 443), then you'd only realistically need to scan a few addresses to determine the network address range. If you want to keep noise even lower, using just 192.168.0.1:80 and 192.168.1.1:80 I'd wager would cover 99% of consumer routers. From there you could assume that it's a /24 netmask and scan IPs to your heart's content. You could do top 10 most common ports type scans and go in-depth on anything you get a result on. I haven't tested this, but I don't see why it wouldn't work, when I was testing 13ft.io - a self-hosted 12ft.io paywall remover, an SSRF flaw like this absolutely let you perform any network request to any LAN address in range.
F

[Opinion] Unending ransomware attacks are a symptom, not the sickness
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
4

1

44 Stimmen

4 Beiträge

2 Aufrufe

G

It varies based on local legislation, so in some places paying ransoms is banned but it's by no means universal. It's totally valid to be against paying ransoms wherever possible, but it's not entirely black and white in some situations. For example, what if a hospital gets ransomed? Say they serve an area not served by other facilities, and if they can't get back online quickly people will die? Sounds dramatic, but critical public services get ransomed all the time and there are undeniable real world consequences. Recovery from ransomware can cost significantly more than a ransom payment if you're not prepared. It can also take months to years to recover, especially if you're simultaneously fighting to evict a persistent (annoyed, unpaid) threat actor from your environment. For the record I don't think ransoms should be paid in most scenarios, but I do think there is some nuance to consider here.
G

When New Jersey Switches Prison Tablet Companies, I’ll Lose 10 Years of Family Memories
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
3

1

109 Stimmen

3 Beiträge

2 Aufrufe

M

A private company is selling cheap tablets to inmates to let them communicate with their family. They have to use "digital stamps" to send messages, 35 cents a piece and come in packs of 5, 10 or 20. Each stamp covers up to 20,000 characters or one single image. They also sell songs, at $1.99 a piece, and some people have spent thousands over the years. That's also now just going away. Then you get to the part about the new company. Who already has a system in Tennessee where inmates have to pay 3-5 cents per minute of tablet usage. Be that watching a movie they've bought or just typing a message.
F

*deleted by creator*
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
1

1

0 Stimmen

1 Beiträge

1 Aufrufe

Niemand hat geantwortet