17.2 Measuring Incremental Progress Toward Human-Level AGI 311
Epstein Suite indexes the text; the original document lives at its official source. We don't host the original file — view it on the official release to read it in full.
View the original on the official releaseDocument text
Text is machine OCR and may contain errors. Confirm against the original source above.
17.2 Measuring Incremental Progress Toward Human-Level AGI 311
— Example task: Ask the robot about events that occurred at times when it got partic-
ularly much, or particularly little, reward for its actions; it should be able to answer
simple questions about these, with significantly more accuracy than about events
occurring at random times
4, Learning
e Imitation: Spontaneously adopt new behaviors that it sees others carrying out
— Example task: Learn to build towers of blocks by watching people do it
Reinforcement: Learn new behaviors from positive and/or negative reinforcement
signals, delivered by teachers and/or the environment
— Example task: Learn which box the red ball tends to be kept in, by repeatedly trying
to find it and noticing where it is, and getting rewarded when it finds it correctly
Imitation/Reinforcement
— Example task: Learn to play “fetch”, “tag” and “follow the leader” by watching people
play it, and getting reinforced on correct behavior
Interactive Verbal Instruction
— Example task: Learn to build a particular structure of blocks faster based on a
combination of imitation, reinforcement and verbal instruction, than by imitation
and reinforcement without verbal instruction
Written Media
— Example task: Learn to build a structure of blocks by looking at a series of diagrams
showing the structure in various stages of completion
e Learning via Experimentation
— Example task: Ask the robot to slide blocks down a ramp held at different angles.
Then ask it to make a block slide fast, and see if it has learned how to hold the
ramp to make a block slide fast.
5. Reasoning
e Deduction, from uncertain premises observed in the world
— Example task: If Ben more often picks up red balls than blue balls, and Ben is given
a choice of a red block or blue block to pick up, which is he more likely to pick up?
e Induction, from uncertain premises observed in the world
— Example task: If Ben comes into the lab every weekday morning, then is Ben likely
to come to the lab today (a weekday) in the morning?
Abduction, from uncertain premises observed in the world
— Example task: If women more often give the robot food than men, and then someone
of unidentified gender gives the robot food, is this person a man or a woman?
Causal reasoning, from uncertain premises observed in the world
— Example task: If the robot knows that knocking down Ben’s tower of blocks makes
him angry, then what will it say when asked if kicking the ball at Ben’s tower of
blocks will make Ben mad?
e Physical reasoning, based on observed “fuzzy rules” of naive physics
— Example task: Given two balls (one rigid and one compressible) and two tunnels
(one significantly wider than the balls, one slightly narrower than the balls), can
the robot guess which balls will fit through which tunnels?
Associational reasoning, based on observed spatiotemporal associations
HOUSE_OVERSIGHT_013227
Have a question about what this document contains?
Ask the documents