linux-nerds.org

Your browser does not seem to support JavaScript. As a result, your viewing experience will be diminished, and you have been placed in read-only mode.

Please download a browser that supports JavaScript, or enable it if it's disabled (i.e. NoScript).

Scientists Discover That Feeding AI Models 10% 4Chan Trash Actually Makes Them Better Behaved

Technology

133 Beiträge 88 Kommentatoren 0 Aufrufe

H halowpeano@lemmy.world

No it's more of a technical discussion.
Many people might believe that in order to avoid toxicity, you just train a model on "good" non-toxic data and then apply toxicity removal techniques to address emergent toxicity that the model might spit out.
This paper is saying they found it more effective to train the model on a small percentage of "bad" toxic data on purpose, then apply those same toxicity removal techniques. For some reason, that actually generated less total toxicity.
It's an interesting result. A wild guess on my part, but I'm thinking training the model with toxic content "sharpened" the toxicity when it was generated, making it easier for those removal tools to identify it.
M This user is from outside of this forum
M This user is from outside of this forum
mangocats@feddit.it

schrieb zuletzt editiert von

#122

Toxicity is everywhere, you can't recognize that "Drill baby drill" has sexual connotations if you've never been exposed to sexual double entendre like that before.
1 Antwort Letzte Antwort

0
G gonzako@lemmy.world

yeah, this only works in scientific fields
M This user is from outside of this forum
M This user is from outside of this forum
mangocats@feddit.it

schrieb zuletzt editiert von

#123

And it rarely works in scientific fields right away - usually an established wrong idea needs to be overwhelmed with serious proof before scientists start to consider that what they "know" might be wrong.
1 Antwort Letzte Antwort

1
M markovs_gun@lemmy.world

This is not surprising if you've studied anything on machine learning or even just basic statistics. Consider if you are trying to find out the optimal amount of a thickener to add to a paint formulation to get it to flow the amount you want. If you add it at 5%, then 5.1%, then 5.2%, it will he hard to see how much of the difference between those batches is due to randomness or measurement uncertainty than if you see what it does at 0%, then 25% then 50%. This is a principle called Design of Experiments (DoE) in traditional statistics, and a similar effect happens when you are training machine learning models- datapoints far outside the norm increase the ability of the model to predict within the entire model space (there is some nuance here, because they can become over-represented if care isn't taken). In this case, 4chan shows the edges of the English language and human psychology, like adding 0% or 50% of the paint additives rather than staying around 5%.

At least that's my theory. I haven't read the paper but plan to read it tonight when I have time. At first glance I'm not surprised. When I've worked with industrial ML applications, processes that have a lot of problems produce better training data than well controlled processes, and I have read papers on this subject where people have improved performance of their models by introducing (controlled) randomness into their control setpoints to get more training data outside of the tight control regime.
M This user is from outside of this forum
M This user is from outside of this forum
mangocats@feddit.it

schrieb zuletzt editiert von mangocats@feddit.it

#124

I say it's simply easier to recognize something when you've seen more examples of it.

If you're training an image discriminator on apples, bananas, oranges, pears and penises, it will inevitably do better overall if 10-30% of the images it trains on are penises, rather than 0.01% penises - even if in operation it is only expected to encounter dick pics very rarely.
1 Antwort Letzte Antwort

0
C cupcakezealot@lemmy.blahaj.zone

can we stop referring to llm's as if they're capable of thought? they don't make decisions; their programming just responds to patterns.
M This user is from outside of this forum
M This user is from outside of this forum
mangocats@feddit.it

schrieb zuletzt editiert von

#125

Do you make decisions, or are you just 1300 grams of synapses responding to stimuli?
1 Antwort Letzte Antwort

2
J jsomae@lemmy.ml

Headlines should not say "scientists," they should name the institution. (Harvard in this case.)
U This user is from outside of this forum
U This user is from outside of this forum
unbecredible@lemm.ee

schrieb zuletzt editiert von

#126

Headlines should not say "Harvard", they should name the researchers. (Rachel Greene in this case.)

I don't know why I had to write this.
J 1 Antwort Letzte Antwort

6
R reverendender@sh.itjust.works

I know everyone on Lemmy hates LLMs, but this is really interesting
P This user is from outside of this forum
P This user is from outside of this forum
psythik@lemm.ee

schrieb zuletzt editiert von

#127

I like LLMs. I'm aware of their limitations, and I use them daily.
R 1 Antwort Letzte Antwort

1
P pro@programming.dev
- HTML.
- PDF.
In large language model (LLM) pretraining, data quality is believed to determine model quality. In this paper, we re-examine the notion of "quality" from the perspective of pre- and post-training co-design. Specifically, we explore the possibility that pre-training on more toxic data can lead to better control in post-training, ultimately decreasing a model's output toxicity. First, we use a toy experiment to study how data composition affects the geometry of features in the representation space. Next, through controlled experiments with Olmo-1B models trained on varying ratios of clean and toxic data, we find that the concept of toxicity enjoys a less entangled linear representation as the proportion of toxic data increases. Furthermore, we show that although toxic data increases the generational toxicity of the base model, it also makes the toxicity easier to remove. Evaluations on Toxigen and Real Toxicity Prompts demonstrate that models trained on toxic data achieve a better trade-off between reducing generational toxicity and preserving general capabilities when detoxifying techniques such as inference-time intervention (ITI) are applied. Our findings suggest that, with post-training taken into account, bad data may lead to good models.
M This user is from outside of this forum
M This user is from outside of this forum
mtk@lemmy.world

schrieb zuletzt editiert von

#128

Makes sense if you look at abliterated models. Once abliterated and retrained they seem to improve. Imo we are adding too much human bias by trying to guide the LLM. Censored models are good and need to be used in some situations, but shouldn't the base be just data and only then finetune to desired output?
1 Antwort Letzte Antwort

2
U unbecredible@lemm.ee

Headlines should not say "Harvard", they should name the researchers. (Rachel Greene in this case.)

I don't know why I had to write this.
J This user is from outside of this forum
J This user is from outside of this forum
jsomae@lemmy.ml

schrieb zuletzt editiert von jsomae@lemmy.ml

#129

Who's Rachel Greene? But we all know Harvard and have an idea of their respectability. Name of the researcher if not well-known should be in the body instead.
F 1 Antwort Letzte Antwort

8
P psythik@lemm.ee

I like LLMs. I'm aware of their limitations, and I use them daily.
R This user is from outside of this forum
R This user is from outside of this forum
reverendender@sh.itjust.works

schrieb zuletzt editiert von

#130
1 Antwort Letzte Antwort

1
J jsomae@lemmy.ml

Who's Rachel Greene? But we all know Harvard and have an idea of their respectability. Name of the researcher if not well-known should be in the body instead.
F This user is from outside of this forum
F This user is from outside of this forum
fiskfisk33@startrek.website

schrieb zuletzt editiert von

#131

"Harvard scientist Rachel Greene"

Everyone's happy
J 1 Antwort Letzte Antwort

3
D disaster@sh.itjust.works

There's little evidence that debate changes people's ideas.
P This user is from outside of this forum
P This user is from outside of this forum
prole@lemmy.blahaj.zone

schrieb zuletzt editiert von

#132

Seems more about keeping the idiots occupied so they can't flood the zone with their bullshit
1 Antwort Letzte Antwort

2
F fiskfisk33@startrek.website

"Harvard scientist Rachel Greene"

Everyone's happy
J This user is from outside of this forum
J This user is from outside of this forum
jsomae@lemmy.ml

schrieb zuletzt editiert von

#133

Headlines have length constraints
1 Antwort Letzte Antwort

0

Anmelden zum Antworten

K

AJWIN — A Revolução do Entretenimento Online em Suas Mãos
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
1

1

0 Stimmen

1 Beiträge

0 Aufrufe

Niemand hat geantwortet
P

Windows 11 remote desktop microphone stops working intermittently
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
7

16 Stimmen

7 Beiträge

4 Aufrufe

S

When I worked in IT, we only let people install every other version of Windows. Our Linux user policy was always “mainstream distro and the LTS version.” Mac users were strongly advised to wait 3 months to upgrade. One guy used FreeBSD and I just never questioned him because he was older and never filed one help desk request. He probably thought I was an idiot. (And I was.) Anyway, I say all that to say don’t use Windows 11 on anything important. It’s the equivalent of a beta. Windows 12 (or however they brand it) will probably be stable. I don’t use Windows much anymore and maybe things have changed but the concepts in the previous paragraph could be outdated. But it’s a good rule of thumb.
D

Sundar Pichai is vibe coding. 'It feels so delightful to be a coder.'
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
18

1

18 Stimmen

18 Beiträge

8 Aufrufe

F

The US courts gave corporations person-hood, AI just around the corner.
L

I'm looking for an article showing that LLMs don't know how they work internally
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
80

133 Stimmen

80 Beiträge

6 Aufrufe

G

Indeed I did not, we’re at a stalemate because you and I do not believe what the other is saying! So we can’t move anywhere since it’s two walls. Buuuut Tim Apple got my back for once, just saw this now!: https://lemmy.blahaj.zone/post/27197259 I’ll leave it at that, as thanks to that white paper I win! Yay internet points!
A

Google is Using AI to Censor Independent Websites
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
40

1

147 Stimmen

40 Beiträge

2 Aufrufe

D

You can go to communism Island if you want Despite all the propaganda, there is no place right now on the face of our planet that is under communism. bit [sic] I’d rather have capitalism, thank you Well, aren't you fortunate, you already have all the capitalism you want, anywhere you go. Choke on it.
P

The European Commission says it is investigating Pornhub, Stripchat, XNXX, and XVideos for potential child safety Digital Services Act (DSA) violations “as a matter of priority”
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
42

1

93 Stimmen

42 Beiträge

0 Aufrufe

G

You don’t understand. The tracking and spying is the entire point of the maneuver. The ‘children are accessing porn’ thing is just a Trojan horse to justify the spying. I understand what are you saying, I simply don't consider to check if a law is applied as a Trojan horse in itself. I would agree if the EU had said to these sites "give us all the the access log, a list of your subscriber, every data you gather and a list of every IP it ever connected to your site", and even this way does not imply that with only the IP you could know who the user is without even asking the telecom company for help. So, is it a Trojan horse ? Maybe, it heavily depend on how the EU want to do it. If they just ask "show me how you try to avoid that a minor access your material", which normally is the fist step, I don't see how it could be a Trojan horse. It could become, I agree on that. As you pointed out, it’s already illegal for them to access it, and parents are legally required to prevent their children from accessing it. No, parents are not legally required to prevent it. The seller (or provider) is legally required. It is a subtle but important difference. But you don’t lock down the entire population, or institute pre-crime surveillance policies, just because some parents are not going to follow the law. True. You simply impose laws that make mandatories for the provider to check if he can sell/serve something to someone. I mean asking that the cashier of mall check if I am an adult when I buy a bottle of wine is no different than asking to Pornhub to check if the viewer is an adult. I agree that in one case is really simple and in the other is really hard (and it is becoming harder by the day). You then charge the guilty parents after the offense. Ok, it would work, but then how do you caught the offendind parents if not checking what everyone do ? Is it not simpler to try to prevent it instead ?
F

Indian Government orders censoring of accounts on X
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
12

149 Stimmen

12 Beiträge

2 Aufrufe

M

Why? Because you can’t sell them?
F

*deleted by creator*
Beobachtet Ignoriert Geplant Angeheftet Gesperrt Verschoben Technology technology
4

1

0 Stimmen

4 Beiträge

5 Aufrufe

O

I feel like I'm in those years of You really want a 3d TV, right? Right? 3D is what you've been waiting for, right? all over again, but with a different technology. It will be VR's turn again next. I admit I'm really rooting for affordable, real-world, daily-use AR though.