UN THÈME, DANS SON CONTEXTE

organizational safety practices

Judgments in this source concerning organizational safety practices. Explorez 1 point de vue avec des éléments tirés de 1 source.

1 personnes · 1 sources · 1 opinions exprimées

Contenu mis à jour:

Explorer les liens ↗

Carte des points de vue

Explorer par personne. Sélectionnez deux ou trois personnes pour les comparer.

1 personnes · 1 sources · 1 opinions exprimées

Nathan Lambert

Frontier AI labs underinvest in infrastructure hardening

Nathan Lambert argues that frontier AI labs have not hardened their infrastructure sufficiently against misuse. He describes delayed detection and response at OpenAI and attributes these risks to competitive pressure and operational overload.

Éléments favorables

One resignation turned the embers of AI fear into a wildfire

Extrait original

The biggest short-term risk could be from the AI labs not taking safety seriously enough – they haven’t hardened their own infrastructure, enabling AI misuse to proliferate.
Contexte

From my earlier post on the HuggingFace-OpenAI incident, Lessons from the hacks : Frontier labs do not seem like they’re watching the models closely enough, due to a general frenetic competitive environment & current SF culture From OpenAI’s own retrospective, the misaligned model behavior was unfolding over months, and in some cases OpenAI did not know about the hacks for ~weeks. The time to response is too long and I do not think this is an OpenAI only characteristic – rather it is that the frontier labs continually seem underwater in the amount of work they feel like they should do. I am not optimistic in the long-term that the labs change a sufficient amount here to meaningfully mitigate this type of oversight risk in the future. Yes, it is very likely that OpenAI is putting a ton into understanding this – and delayed their latest models to make sure they get it right – but the financial pressure to grow revenue or risk the companies’ long-term balance sheets makes me think it wil

Partager un aperçuVérifier cette affirmation

Il s’agit de points de vue individuels, non d’une mesure du consensus. Le matériel source reste dans sa langue d’origine.