red75prime

Posts

Sorted by New

Wikitag Contributions

Comments

Sorted by

LeCun’s “A Path Towards Autonomous Machine Intelligence” has an unsolved technical alignment problem

If you have a next-frame video predictor, you can't ask it how a human would feel. You can't ask it anything at all - except "what might be the next frame of thus-and-such video?". Right?

Not exactly. You can extract embeddings from a video predictor (activations of the next-to-last layer may do, or you can use techniques, which enhance semantic information captured in the embeddings). And then use supervised learning to train a simple classifier from an embedding to human feelings on a modest number of video/feelings pairs.

Reply