Document

17.2 Measuring Incremental Progress Toward Human-Level AGI 311

Ref IMAGES-002-HOUSE_OVERSIGHT_013227.txt Release House Oversight Committee — Epstein Estate Records (Nov 2025) 1 pages

Epstein Suite indexes the text; the original document lives at its official source. We don't host the original file — view it on the official release to read it in full.

View the original on the official release

Document text

Text is machine OCR and may contain errors. Confirm against the original source above.

17.2 Measuring Incremental Progress Toward Human-Level AGI 311 — Example task: Ask the robot about events that occurred at times when it got partic- ularly much, or particularly little, reward for its actions; it should be able to answer simple questions about these, with significantly more accuracy than about events occurring at random times 4, Learning e Imitation: Spontaneously adopt new behaviors that it sees others carrying out — Example task: Learn to build towers of blocks by watching people do it Reinforcement: Learn new behaviors from positive and/or negative reinforcement signals, delivered by teachers and/or the environment — Example task: Learn which box the red ball tends to be kept in, by repeatedly trying to find it and noticing where it is, and getting rewarded when it finds it correctly Imitation/Reinforcement — Example task: Learn to play “fetch”, “tag” and “follow the leader” by watching people play it, and getting reinforced on correct behavior Interactive Verbal Instruction — Example task: Learn to build a particular structure of blocks faster based on a combination of imitation, reinforcement and verbal instruction, than by imitation and reinforcement without verbal instruction Written Media — Example task: Learn to build a structure of blocks by looking at a series of diagrams showing the structure in various stages of completion e Learning via Experimentation — Example task: Ask the robot to slide blocks down a ramp held at different angles. Then ask it to make a block slide fast, and see if it has learned how to hold the ramp to make a block slide fast. 5. Reasoning e Deduction, from uncertain premises observed in the world — Example task: If Ben more often picks up red balls than blue balls, and Ben is given a choice of a red block or blue block to pick up, which is he more likely to pick up? e Induction, from uncertain premises observed in the world — Example task: If Ben comes into the lab every weekday morning, then is Ben likely to come to the lab today (a weekday) in the morning? Abduction, from uncertain premises observed in the world — Example task: If women more often give the robot food than men, and then someone of unidentified gender gives the robot food, is this person a man or a woman? Causal reasoning, from uncertain premises observed in the world — Example task: If the robot knows that knocking down Ben’s tower of blocks makes him angry, then what will it say when asked if kicking the ball at Ben’s tower of blocks will make Ben mad? e Physical reasoning, based on observed “fuzzy rules” of naive physics — Example task: Given two balls (one rigid and one compressible) and two tunnels (one significantly wider than the balls, one slightly narrower than the balls), can the robot guess which balls will fit through which tunnels? Associational reasoning, based on observed spatiotemporal associations HOUSE_OVERSIGHT_013227

Have a question about what this document contains?

Ask the documents