Dashboard
Signal #145511POSITIVE

Value Generalisation 1: a Research and Deployment Program

95

I’m looking for people, advice, critiques, and funding to build a research program on value generalisation – the ability of an AI to correctly extend human values and preferences to situations neither it nor we have seen before. My ongoing research has become convinced that this is necessary if we want to get aligned AIs that operate in the human interest.This would be a focused research organisation or a commercial venture. I’m leaning towards commercial, because alignment techniques confined to academic papers get ignored – or worse, mined for capability-relevant parts while the alignment component is discarded.This post is the research program’s summary. The technical case is in the next post, and one exciting consequence – AIs whose alignment grows with their capabilities – is in the post after that.Without value generalisation, AI can't be reliable: it lacks that capabilityNothing technical stands in the way of you handing an AI assistant full control of your devices and accounts ...

AI Alignment Forumabout 4 hours ago
Read Full Article

Explore with AI-Powered Tools

View All Signals

Explore more AI intelligence

Want to discover more AI signals like this?

Explore Steek
Value Generalisation 1: a Research and Deployment Program | Steek AI Signal | Steek