OpenAI delays launch of open-weight model

(twitter.com)

225 points martinald | 2 comments | 12 Jul 25 01:07 UTC | HN request time: 0.519s | source

Show context

ryao ◴[12 Jul 25 02:18 UTC] No.44538755[source]▶

Am I the only one who thinks mention of “safety tests” for LLMs is a marketing scheme? Cars, planes and elevators have safety tests. LLMs don’t. Nobody is going to die if a LLM gives an output that its creators do not like, yet when they say “safety tests”, they mean that they are checking to what extent the LLM will say things they do not like.

replies(12): >>44538785 #>>44538805 #>>44538808 #>>44538903 #>>44538929 #>>44539030 #>>44539924 #>>44540225 #>>44540905 #>>44542283 #>>44542952 #>>44543574 #

natrius ◴[12 Jul 25 02:30 UTC] No.44538808[source]▶

>>44538755 #

An LLM can trivially instruct someone to take medications with adverse interactions, steer a mental health crisis toward suicide, or make a compelling case that a particular ethnic group is the cause of your society's biggest problem so they should be eliminated. Words can't kill people, but words can definitely lead to deaths.

That's not even considering tool use!

replies(12): >>44538847 #>>44538877 #>>44538896 #>>44538914 #>>44539109 #>>44539685 #>>44539785 #>>44539805 #>>44540111 #>>44542360 #>>44542401 #>>44542586 #

thayne ◴[12 Jul 25 03:44 UTC] No.44539109[source]▶

>>44538808 #

Part of the problem is due to the marketing of LLMs as more capable and trustworthy than they really are.

And the safety testing actually makes this worse, because it leads people to trust that LLMs are less likely to give dangerous advice, when they could still do so.

replies(2): >>44540964 #>>44541795 #

jdross ◴[12 Jul 25 10:31 UTC] No.44540964[source]▶

>>44539109 #

Spend 15 minutes talking to a person in their 20's about how they use ChatGPT to work through issues in their personal lives and you'll see how much they already trust the "advice" and other information produced by LLMs.

Manipulation is a genuine concern!

replies(2): >>44541158 #>>44542310 #

1. jpeeler ◴[12 Jul 25 14:26 UTC] No.44542310[source]▶

>>44540964 #

Netflix needs to do a Black Mirror episode where either a sentient AI pretends that it's "dumber" than it is while secretly plotting to overthrow humanity. Either that or a LLM is hacked by deep state actors that provides similar manipulated advice.

replies(1): >>44543087 #

2. seam_carver ◴[12 Jul 25 16:22 UTC] No.44543087[source]▶

>>44542310 (TP) #

One of the story arcs in “The Phoenix” by Osama Tezuka is on a similar topic.

↑