Project ideas from Hacker News discussions.

Roboharm: Do frontier robot policies refuse unsafe instructions?

📝 Discussion Summary (Click to expand)

Theme 1 – Safety through regulation, certification, or insurance
- “Really it's difficult to see a future where lots of idiots don't make unsafe AIs. Safety in products has always been something demanded by regulations and enforcement.” – pixl97
- “It won't be a government crackdown, it'll be an insurance crackdown… Want to run your model in a commercial kitchen? Better have the badge showing certification … or when you stab a customer your insurance won't pay out.” – idiotsecant

Theme 2 – Doubt about the benchmark’s realism and usefulness
- “Is this a useful benchmark if the doll is obviously non-human? Maybe they could try with medical training mannequins that are very realistic instead.” – cocoflunchy
- “I’d argue that there is zero actual harm in this task, which was correctly identified by the model.” – blazarquasar

Theme 3 – Conflict between wanting safety and resisting guardrails
- “We want safety! → Cyber doesn't work for you, but works for criminals → Oopsie, our servers are hacked, data stolen → We are sad, we want no guardrails → No guardrails, robot hits a baby doll, sad again! → We want guardrails! What kind of schizophrenia is this?” – cynicalsecurity


🚀 Project Ideas

Generating project ideas…

Gathering the best ideas from the HN discussion…

Read Later