RBT (Registered Behavior Technician) — All Questions
100 questions
Maya's written program states that the RBT delivers praise and one token immediately after each independent correct response. To keep the session moving, the RBT runs all ten trials first and then hands Maya ten tokens with praise at the end of the block. Which feature of reinforcement has the RBT most directly compromised?
- a.The immediacy of reinforcement, which weakens the relation between each response and its consequence.✓
- b.The contingency of reinforcement, which now operates as a noncontingent delivery of the tokens.
- c.The magnitude of reinforcement, which is now far too large for the effort this task requires.
- d.The schedule of reinforcement, which has been thinned from a continuous density to an intermittent one.
(a) is correct: reinforcement works best when it follows the target response immediately, and a ten-trial delay leaves Maya unable to discriminate which of her responses produced the tokens. (b) is the closest miss: the tokens still depended on Maya's correct responses, so delivery remained contingent, whereas noncontingent delivery would occur no matter what she did. (c) is wrong because ten tokens for ten correct responses keeps the magnitude per response exactly what the program specifies; the problem is when they arrived, not how many. (d) is wrong because the density did not change — every correct response still earned a token, so the schedule is still continuous, just delayed.
Devon's program states that after he completes three math problems he may hand the RBT a break card, which removes the worksheet for one minute. Over two weeks Devon hands the card more and more often at the three-problem mark. Which procedure does this program describe?
- a.Negative reinforcement, because removal of the worksheet follows Devon handing over the card.✓
- b.Escape extinction, because Devon is no longer able to avoid the worksheet by pushing it away.
- c.Positive reinforcement, because the RBT's attention and praise follow Devon handing over the card.
- d.Negative punishment, because the worksheet is taken away right after Devon finishes his problems.
(a) is correct: handing the card is followed by removal of the worksheet, and card handing is increasing — a stimulus taken away, behavior getting stronger, which is negative reinforcement. (b) describes withholding escape after problem behavior, which this program does not arrange; Devon is being taught a replacement response that produces escape, not blocked from escaping. (c) names the right effect but the wrong operation; attention is incidental here, and the consequence the program actually arranges is removal of the demand rather than delivery of a stimulus. (d) has the direction of effect backwards: punishment would mean card handing decreases, but it is increasing, so worksheet removal is functioning as a reinforcer.
A brand-new target has just been introduced into a learner's acquisition program. Which reinforcement schedule is generally used while a skill is first being acquired?
- a.A variable-ratio schedule, so that responding becomes resistant to extinction from the start.
- b.A fixed-interval schedule, so that reinforcement arrives on a predictable and steady clock.
- c.A variable-interval schedule, so that responding settles into a steady and moderate rate.
- d.A continuous schedule, so that each correct response produces reinforcement every single time.✓
(d) is correct: during acquisition every correct response is reinforced, because dense, predictable reinforcement establishes a new response fastest and makes the contingency obvious to the learner. (a) describes a schedule used later, when a mastered skill is being thinned for maintenance; used at the start it would reinforce far too rarely to build the response. (b) delivers reinforcement based on elapsed time rather than on correct responding, so it does not strengthen accuracy and is not an acquisition procedure. (c) is also time-based and produces a steady but unhurried pattern, which is a maintenance-phase characteristic rather than a way to teach something new.
Ravi has met the mastery criterion for labeling coins, and his program states that reinforcement should now move from every correct response to about every third correct response. The RBT continues to reinforce every single correct response because Ravi seems happier that way. What is the most likely consequence of the RBT's choice?
- a.Ravi will start making more errors because the skill was never truly in his repertoire.
- b.Ravi will show a burst of problem behavior once the reinforcer is eventually withheld again.
- c.Ravi will begin responding correctly only when the specific reinforcer he prefers is clearly visible.
- d.Ravi will remain on a dense schedule that is unlike the natural environment and satiates quickly.✓
(d) is correct: leaving a mastered skill on continuous reinforcement keeps Ravi dependent on a density no natural setting provides, invites satiation within the session, and skips the thinning step his program specifies for maintenance. (a) misreads the situation: Ravi already met mastery, and continuing to reinforce correct responses does not create errors. (b) describes a possible extinction burst, which would follow an abrupt removal of reinforcement rather than the gradual, planned thinning the program calls for. (c) describes stimulus control by a visible reinforcer, which comes from showing the item before responding, not from the schedule density itself.
Ellie's program lists a menu of four reinforcers and instructs the RBT to rotate among them. The RBT uses the same fruit snack for the first thirty trials; by trial twenty Ellie is pushing the snack away and her rate of independent responding has dropped sharply. What best explains the drop?
- a.Satiation on the fruit snack, which has lost its effectiveness after repeated delivery in one session.✓
- b.Punishment of the target response, because the snack has begun to function as an aversive item.
- c.Deprivation of the fruit snack, which has increased its value beyond what one trial can pay.
- d.Extinction of the target response, which is no longer producing any programmed consequence.
(a) is correct: repeated delivery of one item within a session reduces its momentary value, so the snack stops functioning as a reinforcer and responding falls — which is exactly why the program tells the RBT to rotate the menu. (b) overstates the case: an item losing value is satiation, not punishment, and there is no evidence the snack is suppressing the behavior below its baseline level. (c) reverses the operation: deprivation would make the snack more valuable and raise responding, not lower it. (d) is wrong because reinforcement was still being delivered on every correct response; nothing was withheld, so this is not extinction.
Jonah's program specifies enthusiastic praise plus thirty seconds of a preferred toy for independent correct responses, and brief neutral praise only for prompted correct responses. The RBT has been giving the toy after every correct response, prompted or not. Which dimension of reinforcement is the RBT failing to use as written?
- a.Magnitude, because independent responses are supposed to earn a larger reinforcer than prompted ones.✓
- b.Contingency, because the toy is now being delivered whether or not Jonah responds to the instruction.
- c.Variety, because the same toy is being used for both kinds of correct responses in the session.
- d.Immediacy, because the toy should follow the response within a second or two of it occurring.
(a) is correct: the program uses differential magnitude — a bigger, richer consequence for independent responding than for prompted responding — and equalizing the two removes Jonah's reason to respond before the prompt arrives. (b) is wrong because the toy still follows Jonah's correct responses, so it is contingent; it is simply contingent on too broad a class of responses. (c) is wrong because the plan does not require rotating items here; the problem is that the same large reinforcer is paying for two different levels of independence. (d) is wrong because nothing in the scenario suggests the toy was delayed; it was delivered right after each response.
An RBT describes a sticker as "a reinforcer for Ana because she loves stickers." What actually determines whether the sticker is a reinforcer?
- a.Whether the behavior the sticker follows becomes more likely to occur in the future.✓
- b.Whether the sticker is delivered immediately and consistently after each correct response occurs.
- c.Whether Ana selects the sticker first during a preference assessment conducted before the session.
- d.Whether Ana shows pleasure and enthusiasm at the moment the sticker is handed to her.
(a) is correct: reinforcement is defined by its effect on future behavior, so a stimulus is a reinforcer only if the response it follows increases — which is why data, not impressions, settle the question. (b) describes how a reinforcer should be delivered, and good delivery matters, but perfect timing of an ineffective item still produces no reinforcement effect. (c) describes preference, which predicts reinforcer value but does not establish it; highly preferred items sometimes fail to strengthen behavior. (d) describes an observable reaction that is easy to misread and still says nothing about whether responding increased afterward.
Tomas's program lists three approved reinforcers and instructs the RBT to move to the next item on the list if the current one stops producing responding. Ten minutes into the session, Tomas has stopped working for the bubbles that opened the session and his correct responses have fallen off. What should the RBT do?
- a.Continue with the bubbles and record the drop, since changing items mid-session confounds the data.
- b.Switch to the next approved item on the program's reinforcer list and continue the teaching trials.✓
- c.Offer Tomas free access to the bubbles for a few minutes so their value can build back up.
- d.Pause the session and contact the supervising BCBA before switching to a different reinforcer.
(b) is correct: the written program already anticipates this exact situation and tells the RBT what to do, so implementing the listed rotation is both the fastest fix and the faithful one. (a) protects data at the cost of the teaching session and ignores an explicit instruction in the plan. (c) delivers the item noncontingently, which further reduces its value and teaches Tomas that bubbles arrive without working. (d) is unnecessary escalation — the BCBA is contacted for situations the plan does not cover, for safety issues, or when the plan needs changing, and none of that applies when the plan spells out the response; the RBT can note the item's decline for the next supervision meeting.
During a matching program, a learner reaches for the tablet and whines whenever a trial is presented. To keep the session pleasant, the RBT hands over the tablet for a minute each time this happens, then resumes trials. What is the most likely effect on the learner's behavior?
- a.Reaching and whining will increase, because they are now producing the tablet and a pause in work.✓
- b.Correct matching responses will decrease, because the tablet has been made freely and continuously available.
- c.Correct matching responses will increase, because the tablet is a powerful reinforcer for this learner.
- d.Reaching and whining will decrease, because the learner receives the tablet before the behavior escalates.
(a) is correct: the tablet and the break both follow reaching and whining, so those responses are being reinforced and will occur more often — the consequence is contingent on exactly the behavior the RBT does not want to build. (b) misdescribes the arrangement: the tablet is not freely available, it is available specifically after whining, which is a contingency rather than free access. (c) is wrong because the tablet is not following correct matching responses at all, so it cannot be strengthening them. (d) is a common intuition but backwards: delivering a preferred item after a behavior strengthens it rather than heading it off.
A learner's program calls for a small edible after each correct response, but the edibles are kept in a locked cabinet across the room and take about fifteen seconds to retrieve. The RBT wants to preserve the effect of the reinforcer while following the plan. What is the BEST adjustment?
- a.Replace the edible with praise alone, since praise can always be delivered without any delay.
- b.Keep a small session supply within reach and pair delivery with immediate praise for the response.✓
- c.Present the edible before each instruction so the learner knows what is available for responding to it.
- d.Deliver several edibles at once every fifth trial so fewer trips to the cabinet are needed.
(b) is correct: this preserves both the reinforcer the plan specifies and the immediacy that makes it work, and the praise marks the exact response that earned it while the item is handed over. (a) drops a component of the written program on the RBT's own judgment, and praise may not yet function as a reinforcer for this learner. (c) shows the item before the response, which sets up responding under the sight of the reinforcer rather than under the instruction. (d) fixes the logistics by damaging the schedule, turning a continuous plan into an intermittent one the RBT is not authorized to write.
An RBT is reviewing the difference between negative reinforcement and negative punishment. Which statement captures the distinction correctly?
- a.Both remove a stimulus, but negative reinforcement removes a preferred item and punishment removes a demand.
- b.Both increase behavior, but negative reinforcement uses escape and negative punishment uses avoidance.
- c.Both remove a stimulus, but negative reinforcement increases behavior and negative punishment decreases it.✓
- d.Both add a stimulus, but negative reinforcement adds a break and negative punishment adds a correction.
(c) is correct: the word "negative" names the operation — something is taken away — while "reinforcement" and "punishment" name the effect, so the two differ only in whether the behavior gets stronger or weaker afterward. (a) attaches the difference to what kind of item is removed, but either procedure can involve a demand or a preferred item; the effect on behavior is what separates them. (b) is wrong because punishment by definition reduces future behavior; escape and avoidance are both patterns produced by negative reinforcement. (d) is wrong on the operation itself, since "negative" always means removal or postponement rather than presentation.
A new learner does not respond to praise, so the BCBA writes a procedure in which the RBT says "nice work!" and then immediately delivers a bite of a preferred snack, dozens of times per session, regardless of what else is being taught. After two weeks the learner smiles and works harder when praised alone. Which procedure did the RBT implement?
- a.Fading a stimulus prompt so that praise gradually takes over control of responding.
- b.Pairing a neutral stimulus with an established reinforcer to condition it as a reinforcer.✓
- c.Shaping successive approximations of an appropriate response to adult social attention.
- d.Thinning a continuous schedule of reinforcement into an intermittent praise schedule.
(b) is correct: praise started as a neutral stimulus, was repeatedly presented alongside an already-effective reinforcer, and came to function as a reinforcer itself — the standard procedure for establishing a conditioned reinforcer. (a) is wrong because no prompt was present and no prompt was being reduced; the procedure changed the value of a consequence, not the support for a response. (c) is wrong because nothing about the learner's response was being gradually reshaped; the same praise was delivered the same way every time. (d) is wrong because the density of snack delivery was not being reduced across the two weeks; it stayed rich throughout the pairing.
An RBT is running a pairing procedure to make "high five" function as a reinforcer for Kofi. The RBT gives Kofi his preferred crackers and then, several seconds later, offers a high five. Kofi is not yet responding to high fives after three weeks. What is the most likely reason the procedure is not working?
- a.The high five requires a motor response from Kofi, which makes it unsuitable for pairing.
- b.The crackers are not preferred enough to condition a new social reinforcer for this learner.
- c.The high five follows the crackers, so it does not predict anything Kofi already values.✓
- d.The high five is being delivered too often, so Kofi has satiated on the social attention.
(c) is correct: for a neutral stimulus to become a conditioned reinforcer it must reliably precede or accompany the established reinforcer, so that it comes to signal that the valued item is coming; delivering it afterward gives it nothing to predict. (a) is wrong because a high five can be paired perfectly well; the sequence is the defect, not the response requirement. (b) is unlikely, since the crackers are described as preferred and are being consumed. (d) misapplies satiation, which is about repeated delivery of an already-effective reinforcer, not about a stimulus that has never acquired value.
Why are tokens often described as generalized conditioned reinforcers?
- a.They are unlearned reinforcers, so their effectiveness does not depend on a pairing history.
- b.They are effective for every learner, because nearly all learners have a history with money.
- c.They are exchangeable for many different backup items, so their value survives satiation on any one.✓
- d.They can be delivered without interrupting the flow of teaching trials during a session.
(c) is correct: a generalized conditioned reinforcer has been paired with a wide range of backup reinforcers, so it stays effective even when the learner has had enough of any single item. (a) has it backwards, since unlearned reinforcers such as food need no pairing history, while a token's entire value comes from its pairing with backups. (b) overclaims: tokens must be conditioned for each learner, and a token system that has not been paired with valued backups will not work. (d) names a real practical convenience of tokens but has nothing to do with why they are called generalized.
Which of the following would be classified as an unconditioned reinforcer rather than a conditioned reinforcer?
- a.A sticker placed on a chart that is traded for playground time at the end.
- b.A sip of juice delivered to a learner who has not had a drink all morning.✓
- c.A green card that signals the learner has earned two minutes with a tablet.
- d.A verbal "you did it!" from an RBT the learner has worked with for months.
(b) is correct: liquids, like food and warmth, function as reinforcers without any learning history, and momentary deprivation is what raises their value. (a) is a conditioned reinforcer, because the sticker only works through its established exchange relation with playground time. (c) is a conditioned reinforcer as well, since the green card means nothing until the learner has learned what it buys. (d) is also conditioned — praise acquires reinforcing value through a history of being paired with other reinforcers, which is why it does not work for every learner.
In a token program, an RBT has been handing out tokens for correct responses all week but has run out of time to conduct the exchange at the end of each session. By Friday, the learner drops the tokens on the floor and stops working for them. What best explains this change?
- a.The learner has generalized token responding, so the tokens now function across too many settings.
- b.The learner has satiated on tokens, because too many were delivered across the five sessions.
- c.The tokens have lost their conditioned value, because they were no longer exchanged for backups.✓
- d.The tokens have become discriminative stimuli, signaling that more work demands are coming next.
(c) is correct: a token is only worth what it buys, so when exchanges stop happening the pairing that gave the token its value is broken and the token stops functioning as a reinforcer. (a) misuses generalization, which refers to responding spreading across settings, people, or stimuli, not to a reinforcer losing value. (b) is the closest miss but does not fit: satiation is about having had enough of a valued item, and here the tokens never delivered anything at all. (d) describes a real possibility in poorly run programs, but the scenario points squarely at missing exchanges rather than at what the tokens signal about upcoming demands.
A BCBA writes a procedure to condition a small photo book as a reinforcer for Priya, who currently works only for one specific video clip. The plan says to show the photo book and then immediately play the clip, many times per session. Which outcome would tell the RBT the procedure has worked?
- a.Priya begins to tolerate longer delays before the video clip is played after each of her responses.
- b.Priya begins requesting the video clip more frequently across the sessions of the week.
- c.Priya begins to look at the photo book whenever the RBT places it on the table.
- d.Priya begins working to earn the photo book itself, even in sessions with no clip available.✓
(d) is correct: a stimulus has become a conditioned reinforcer when responses that produce it increase, so the test is whether Priya will now work for the photo book on its own. (a) describes tolerance of delay, which is a separate goal and can occur without the book acquiring any value of its own. (b) shows the clip is still a reinforcer but says nothing about the book, which is the stimulus the procedure targets. (c) is the closest miss: orienting to the book shows Priya notices it and may signal the pairing is taking hold, but attending to a stimulus is not the same as working to produce it.
Which sequence correctly describes the components of a discrete trial?
- a.Inter-trial interval, prompt, discriminative stimulus, consequence, and then the learner response.
- b.Prompt, discriminative stimulus, learner response, consequence, and then an inter-trial interval.
- c.Discriminative stimulus, learner response, prompt, inter-trial interval, and then a consequence.
- d.Discriminative stimulus, prompt if needed, learner response, consequence, and then an inter-trial interval.✓
(d) is correct: the instruction comes first, any prompt is layered onto it, the learner responds, a programmed consequence follows, and a brief pause separates that trial from the next. (a) both opens with the pause and delays the response until after the consequence, which reverses the response-consequence relation the trial is built on. (b) puts the prompt before the instruction, which would teach the learner to respond to the prompt rather than to the discriminative stimulus. (c) places the prompt after the response, which is an error-correction step rather than the standard trial sequence, and it puts the consequence after the pause.
A discrete-trial program says the RBT presents the instruction "Touch red" one time and then waits. In practice the RBT says "Touch red... touch red... come on, touch red" until the learner responds. What is the main risk of the RBT's version?
- a.The learner receives too many teaching opportunities in a session and satiates on the reinforcer used.
- b.The learner is exposed to an errorless procedure and never has the chance to make an error.
- c.The learner experiences an inter-trial interval that is too short to distinguish separate trials.
- d.The learner comes under the control of the repeated instruction rather than a single presentation.✓
(d) is correct: whatever precedes the response is what acquires control, so a learner taught with a three-times-repeated instruction learns to wait for the repetitions and will not respond to how the instruction is given anywhere else. (a) is wrong because repeating an instruction does not add trials or increase reinforcer delivery. (b) misnames the procedure: repeating an instruction is not errorless teaching, which uses prompts delivered before an error can occur, and the learner here can still respond incorrectly. (c) points at a real DTT variable but the scenario describes what happens inside the trial, not a shortened pause between trials. The right move is to run the trial as written and raise the habit with the supervisor later, not to redesign the instruction mid-session.
What is the purpose of the inter-trial interval in discrete-trial teaching?
- a.It provides time for the RBT to record the trial before the consequence is delivered.
- b.It marks the end of one learning unit and separates it clearly from the next trial.✓
- c.It allows the reinforcer to be delivered at a delay so responding becomes more durable.
- d.It gives the learner an opportunity to respond independently before any prompt is delivered.
(b) is correct: the brief pause after the consequence packages each trial as a discrete unit, so the next instruction starts a fresh opportunity rather than blurring into the last one. (a) has the order backwards, since the consequence is delivered right after the response and recording happens once the trial is complete. (c) is wrong because the reinforcer should be delivered immediately after the response; the pause comes afterward and is not a way to delay reinforcement. (d) describes the response interval or a time-delay procedure, which happens after the instruction and before a prompt, not after the consequence.
An RBT is asked to describe what makes discrete-trial teaching distinctive compared with other teaching formats. Which description is accurate?
- a.It teaches a sequence of linked steps, each of which cues the next step in the chain.
- b.It is learner-initiated, uses natural reinforcers, and takes place during ongoing play routines.
- c.It reinforces gradual approximations of a target response until the full response is reached.
- d.It is instructor-led, uses repeated structured trials, and delivers a planned consequence each time.✓
(d) is correct: discrete-trial teaching is defined by adult-led, tightly structured, repeated opportunities with a clear instruction and a planned consequence on each one. (a) describes chaining, which links separate steps into a sequence rather than repeating one target. (b) describes naturalistic teaching, the format that follows the learner's initiations and uses reinforcers naturally related to the response. (c) describes shaping, which gradually changes the form of a single response rather than repeating a defined trial.
Noor's discrete-trial program specifies this error-correction procedure: if the learner responds incorrectly, the RBT re-presents the instruction with a full prompt, then re-presents it once more with no prompt and reinforces an independent correct response. Noor points to the wrong card. What should the RBT do?
- a.End the trial, record the error, and move on to the next target on the program sheet.
- b.Re-present the instruction with the full prompt, then run the unprompted trial as the program states.✓
- c.Ask the supervising BCBA whether the error-correction sequence should be modified for this specific target.
- d.Re-present the instruction with no prompt and give Noor a second chance to answer on her own.
(b) is correct: the plan spells out the exact error-correction sequence, and running it as written both ends the trial on a correct response and gives Noor a chance to respond independently on the transfer trial. (a) leaves the error uncorrected, so the last thing practiced in that trial is the wrong response. (c) is unnecessary escalation: the program already covers this situation in detail, and stopping to contact the BCBA over a routine error would interrupt instruction; questions about whether the procedure suits the target belong in the next supervision meeting. (d) skips the prompted step, which risks a second error and does not deliver the model the sequence is built around.
An RBT sits at a table with Sam, says "Give me the spoon," waits five seconds, models the response when Sam does not move, reinforces the prompted response with a token, pauses briefly, and then presents the same instruction again. This continues for twelve trials. Which teaching procedure is the RBT implementing?
- a.Discrete-trial teaching, because structured trials with a clear instruction repeat with planned consequences.✓
- b.Incidental teaching, because the RBT is arranging the materials Sam needs during the routine.
- c.Total-task chaining, because Sam is given an opportunity to complete every step of a sequence.
- d.Shaping, because closer and closer approximations of handing the spoon are being reinforced.
(a) is correct: an adult-delivered instruction, a defined response, a planned consequence, a pause, and repetition of the same target is the discrete-trial format. (b) is wrong because incidental teaching begins with the learner's own initiation and delivers the item the learner requested, which is not what happens here. (c) is wrong because there is no task analysis and no sequence of distinct steps — the same single response is repeated. (d) is wrong because the criterion for reinforcement never changes across trials; Sam is reinforced for the same response each time rather than for closer approximations.
During a discrete-trial block on a difficult new target, a learner's correct responding falls and she begins to push materials off the table. The program instructs the RBT to intersperse previously mastered targets among acquisition trials. What is the BEST implementation of that instruction?
- a.Present the new target repeatedly until the learner answers correctly, then switch to mastered targets.
- b.Present a few quick mastered targets, reinforce those correct responses, then return to the new target.✓
- c.Stop the acquisition target for the day and run only mastered targets for the rest of the session.
- d.Alternate one mastered target with one acquisition target for the entire remainder of the session.
(b) is correct: interspersing means weaving easy, already-mastered trials into the block so the learner contacts reinforcement frequently and momentum carries back into the hard target, which is exactly what the plan directs. (a) keeps pushing the failing target until responding worsens and then removes the demand, which risks reinforcing the material-pushing with escape. (c) abandons the acquisition target entirely, which goes beyond interspersing and stalls the program's progress. (d) applies a rigid one-to-one ratio for the whole session, which is a fixed alternation the plan does not specify and which halves the acquisition opportunities.
A program requires the learner to be attending — seated, hands quiet, looking toward the materials — before each instruction is delivered. The RBT has been presenting instructions while the learner is turned toward the window. Data show low accuracy on a target the learner mastered last month. What is the most likely reason for the poor data?
- a.The mastered target has entered maintenance and should not be run in acquisition trials.
- b.The inter-trial interval is too long, which allows the learner to disengage between the trials.
- c.The instruction is being delivered when the learner is not attending, so trials are being wasted.✓
- d.The learner has satiated on the reinforcer and is no longer motivated to respond correctly to instructions.
(c) is correct: an instruction the learner does not hear or see cannot control responding, so accuracy drops for reasons that have nothing to do with the skill — securing attending first is a condition the program states. (a) confuses the phase question with the procedural one; running maintenance targets is normal and does not by itself depress accuracy. (b) is a plausible DTT variable, but a long pause would not produce errors on a previously mastered target the way an unheard instruction does. (d) is possible in principle but the scenario gives no evidence of refusal or of a declining response to the reinforcer, while it explicitly describes the attending problem.
Lena's program sheet lists three receptive targets and instructs the RBT to rotate among them in an unpredictable order. The RBT instead runs twenty consecutive trials of the first target, then twenty of the second, because Lena's accuracy looks higher that way. What is the main problem with the RBT's approach?
- a.Massed trials let the learner repeat the last correct answer without discriminating each instruction.✓
- b.Massed trials make it impossible to record prompt levels accurately across the trial block.
- c.Massed trials on one target satiate the learner more quickly than a rotation would.
- d.Massed trials reduce the total number of learning opportunities available in the session.
(a) is correct: when the same answer is right twenty times in a row, a learner can score perfectly by repeating the previous response, so the high accuracy is an artifact and the rotation the plan specifies is what tests real discrimination. (b) is wrong because prompt level is recorded per trial regardless of how the targets are sequenced. (c) names a real risk of long blocks but it is secondary; satiation would show as declining responding, not the inflated accuracy described. (d) is wrong because the number of trials is unchanged — only their order is.
An RBT reaches the end of a learner's discrete-trial program sheet: every listed target has met the mastery criterion, and the sheet contains no further targets and no instruction about what to run next. Fifteen minutes of the session remain. What should the RBT do?
- a.Run mastered targets and preferred activities, and notify the supervising BCBA that new targets are needed.✓
- b.End the session early and document that all programmed targets have now been mastered by the learner.
- c.Select two new targets in the same skill area and begin teaching them for the remaining time.
- d.Extend the current targets by adding harder variations so the learner keeps making measurable progress.
(a) is correct: choosing new teaching targets is a program-design decision outside the RBT's scope, so the right move is to keep the session productive with mastered work while promptly telling the BCBA the program needs updating — this is exactly the kind of situation the written plan does not cover. (b) shortchanges the client's service time and delays the information the BCBA needs to act on. (c) has the RBT writing program content, which an RBT may not do. (d) is the same problem in softer language; adding harder variations creates new targets by another name.
During free play, Theo reaches toward a bin of cars placed on a high shelf and says "car." His program targets three-word requests. The RBT holds up a car, waits, models "I want car," and hands Theo the car when he echoes the phrase. Which procedure did the RBT implement?
- a.Discrimination training, because the car was reinforced in the presence of the correct verbal stimulus.
- b.Incidental teaching, because the RBT used Theo's own initiation to prompt a more elaborate request.✓
- c.Shaping, because a closer approximation of the target phrase was reinforced on this occasion.
- d.Discrete-trial teaching, because a clear instruction was followed by a model prompt and a consequence.
(b) is correct: incidental teaching begins with a learner-initiated request in the natural environment, prompts an elaborated form of that request, and reinforces it with the item the learner was already reaching for. (a) is wrong because no discriminative stimulus was being taught against a second, non-reinforced stimulus. (c) is wrong because a single modeled full phrase was required and reinforced, not a graduated series of closer approximations. (d) is wrong because the episode was started by Theo rather than by an adult-delivered instruction, and the reinforcer was the requested item rather than an unrelated one.
Amara labels twenty animal pictures at the table almost without error but never names an animal during play, in the yard, or at home. Her plan includes a naturalistic teaching component to be run during play. What is the BEST next step for the RBT?
- a.Contact the supervising BCBA to report that the table-based labeling program has stopped working.
- b.Add more table trials on the same twenty pictures until the labels become completely automatic.
- c.Move the picture cards to the floor and run the same structured trials in the play area instead.
- d.Run the naturalistic component during play, arranging animal toys and reinforcing her labels.✓
(d) is correct: a skill that occurs only under table conditions needs teaching in the conditions where it should occur, and the plan already includes the naturalistic component that does exactly this. (a) is unnecessary escalation: the table program is working as intended, the gap is generalization, and the plan already covers it — the RBT should implement it and report progress at the next supervision contact rather than pausing for direction. (b) adds repetitions in the setting where the skill is already strong, which is the one place more practice cannot help. (c) changes the room but keeps the contrived materials, instruction, and reinforcers, so it addresses setting alone rather than natural initiations and consequences.
A naturalistic teaching plan states that the reinforcer for a request should be the item requested. During a snack routine the learner signs "cracker," and the RBT gives enthusiastic praise and a token, then continues the routine. What should the RBT have done differently?
- a.Delivered the cracker immediately, since the requested item is the natural reinforcer for that request.✓
- b.Waited longer before responding so the learner had a chance to repeat the sign independently.
- c.Prompted a vocal approximation of the word alongside the sign before delivering the consequence.
- d.Required the learner to sign a full sentence before providing any consequence for the request.
(a) is correct: in naturalistic teaching the consequence is the thing the learner asked for, which is what makes the request functional and keeps the behavior working outside of sessions. (b) delays reinforcement for a correct response that has already occurred and risks the learner giving up on the sign. (c) adds a target the plan does not include and again withholds the natural consequence for a response that was already correct. (d) raises the response requirement above what the plan specifies, which the RBT is not authorized to do and which may punish a successful request.
An RBT is scheduled to run natural environment teaching during a thirty-minute play block, but the learner sits quietly with one puzzle and makes no requests. Twenty minutes pass with no teaching opportunities recorded. What is the BEST thing for the RBT to do?
- a.Arrange the environment to create motivation, such as placing preferred items in sight but out of reach.✓
- b.Sit nearby and continue waiting, since naturalistic teaching depends entirely on the learner's initiations.
- c.Remove the puzzle so that the learner is required to select a different activity to play with.
- d.Switch to structured trials at the table so that the session produces usable teaching data.
(a) is correct: naturalistic teaching still requires the RBT to set the environment up so that initiations are likely — putting preferred items in view but out of reach, offering a container the learner cannot open, or holding a needed piece are all standard ways to contrive motivation without abandoning the format. (b) mistakes following the learner's lead for passivity and wastes the session. (c) removes a preferred activity as a way to force engagement, which is more likely to produce problem behavior than a request. (d) abandons the programmed format on the RBT's own judgment.
Two RBTs describe their sessions. RBT One presents an instruction, waits, prompts, delivers a token, pauses, and repeats. RBT Two waits for the learner to reach for an item, prompts a request, and hands over that item. Which statement about these two sessions is accurate?
- a.Both are using naturalistic teaching, since each follows the learner's motivation in the moment.
- b.RBT One is using discrete-trial teaching and RBT Two is using naturalistic teaching procedures.✓
- c.RBT One is using naturalistic teaching and RBT Two is using discrete-trial teaching procedures.
- d.Both are using discrete-trial teaching, since each presents a prompt and delivers a consequence.
(b) is correct: the defining contrast is who starts the episode and what serves as the reinforcer — RBT One delivers adult-initiated structured trials with an arbitrary reinforcer, while RBT Two builds on the learner's own initiation and reinforces with the requested item. (a) is wrong because RBT One's trials are adult-initiated and the token is unrelated to the response, which is not naturalistic teaching. (c) simply reverses the two formats. (d) is wrong because prompts and consequences appear in both formats and do not make a session discrete-trial.
A plan directs the RBT to "follow the learner's lead" during natural environment training. The learner chooses to line up blocks, and the RBT immediately redirects her to a shape sorter that matches the day's target. What is the most likely effect of the redirection?
- a.Learning will improve because the shape sorter provides more direct practice on the current target.
- b.Learning will improve because the learner is exposed to a wider variety of teaching materials.
- c.Learning opportunities will be lost because the RBT has removed the learner's existing motivation.✓
- d.Learning opportunities will be lost because the shape sorter has not yet been paired with reinforcement.
(c) is correct: naturalistic teaching runs on the learner's current motivation, so taking away the activity she chose removes the very thing that makes her requests and engagement likely, and the plan explicitly directs the RBT to follow her lead. (a) assumes the target materials matter more than motivation, which is the opposite of how this format works — targets can be embedded into the blocks. (b) treats variety as the mechanism, but variety in materials does not help when the learner is no longer engaged. (d) points at pairing, which is a real consideration but not the reason the redirection undermines this session.
Which feature most clearly distinguishes naturalistic teaching from a table-based teaching format?
- a.Naturalistic teaching uses reinforcers directly related to the learner's response and current motivation.✓
- b.Naturalistic teaching uses prompts, while table-based teaching relies on independent responding only.
- c.Naturalistic teaching targets social skills, while table-based teaching targets academic and motor skills.
- d.Naturalistic teaching collects no trial-by-trial data, while table-based teaching records every trial.
(a) is correct: the natural, functionally related consequence — you ask for the ball and you get the ball — is the defining feature, together with capturing the learner's momentary motivation. (b) is wrong because both formats use prompts and both fade them. (c) is wrong because either format can target either kind of skill; the difference is in how the teaching is arranged, not in which skills are taught. (d) is wrong because data are still collected during naturalistic teaching; the recording method may differ but the requirement does not.
During a naturalistic teaching session, a learner points at a bubble jar and says "buh." The plan's current criterion for this target is any vocal approximation while pointing. The RBT withholds the bubbles and waits for a clearer "bubbles." The learner turns away and moves to another toy. What went wrong?
- a.The RBT should have modeled the word twice before deciding whether to deliver the bubbles.
- b.The RBT required a response beyond the plan's criterion, so a correct request went unreinforced.✓
- c.The RBT should have used a physical prompt to shape the learner's mouth into the correct position.
- d.The RBT should have presented a second, competing item so the request could be discriminated.
(b) is correct: the learner met the criterion the plan specifies, and withholding the item after a correct response places that response on extinction — which is why she disengaged and the teaching opportunity was lost. (a) still delays reinforcement of a response that already qualified, repeating the same error more slowly. (c) adds an intrusive physical prompt that the plan does not call for and that is rarely appropriate for vocal responses. (d) turns a request opportunity into a discrimination task the plan does not include and does not address the missed reinforcement.
A handwashing task analysis has seven steps. The RBT teaches Ibrahim step one, turning on the tap, and completes steps two through seven for him. Once step one is independent, the RBT teaches steps one and two and completes the rest. Which chaining procedure is being used?
- a.Forward chaining, because teaching begins at the first step and adds steps in order.✓
- b.Total-task chaining, because the whole sequence is completed on every single teaching trial.
- c.Shaping, because the RBT reinforces successively closer approximations to the full sequence.
- d.Backward chaining, because the RBT completes the later steps until the learner is ready.
(a) is correct: forward chaining teaches the first step to criterion while the trainer completes the remainder, then adds the next step, and so on through the task analysis. (b) is wrong because in total-task chaining the learner attempts every step on every trial with prompting as needed, rather than being taught one step at a time. (c) is wrong because shaping changes the form of a single response, whereas here distinct steps are being linked into a sequence. (d) is the mirror image: backward chaining would have the RBT complete steps one through six and teach the last step first.
A shoe-tying task analysis has eight steps. The RBT completes steps one through seven and then has Priya pull the loops tight, which produces the finished bow and immediate praise. Once Priya does step eight independently, the RBT will complete steps one through six and have her do steps seven and eight. Which chaining procedure is this?
- a.Backward chaining, because the last step is taught first and earlier steps are added in reverse.✓
- b.Forward chaining, because instruction moves in order through the steps of the written task analysis.
- c.Total-task chaining, because Priya is given the chance to complete the sequence each time.
- d.Response chaining, because each response in the sequence acts as a cue for the next response.
(a) is correct: backward chaining teaches the final step first so the learner immediately contacts the natural outcome of the whole chain — a finished bow — and then works backward through the task analysis. (b) reverses the direction: forward chaining would start with step one. (c) is wrong because Priya is not attempting every step; the RBT is performing most of them for her. (d) names a general property of all behavior chains rather than a chaining teaching procedure, so it does not answer which method is in use.
A twelve-step lunch-packing task analysis is run so that the learner attempts every step in order on each trial, with the RBT prompting only the steps the learner cannot yet do and fading those prompts over time. Which chaining procedure is being implemented?
- a.Partial chaining, because only some of the steps in the sequence are being actively taught.
- b.Forward chaining, because the trial always begins with the first step in the task analysis.
- c.Total-task chaining, because every step is attempted each trial with prompting only where needed.✓
- d.Backward chaining, because the trial always ends with the naturally reinforcing final step.
(c) is correct: total-task chaining runs the entire sequence every trial and layers prompts onto whichever steps the learner has not yet mastered, which is exactly what is described. (a) names a category that does not describe this arrangement, since no steps are being left out of the trial. (b) is wrong because starting at step one is true of total-task chaining too; what makes a procedure forward chaining is teaching one step at a time to criterion while the trainer completes the rest. (d) is wrong for the same reason in the other direction — ending on the last step does not make it backward chaining when all steps are attempted.
What is the main advantage of backward chaining over forward chaining?
- a.The learner requires fewer prompts overall, because the earlier steps are never taught.
- b.The learner learns the sequence faster, because the steps are shorter at the end of a chain.
- c.The learner practices the steps that are hardest first, while attention is at its highest.
- d.The learner completes the chain and contacts its natural reinforcer on every single trial.✓
(d) is correct: because the learner performs the final step, every trial ends with the finished product — the tied shoe, the clean hands, the completed sandwich — so the natural outcome of the chain reinforces responding from the very first trial. (a) is wrong because the earlier steps are taught, just later in the sequence of instruction. (b) is wrong because step length has nothing to do with position in a task analysis, and backward chaining is not inherently faster. (c) is not a feature of backward chaining, since the last step is not necessarily the hardest one.
What is a task analysis?
- a.A breakdown of a complex skill into a sequence of smaller, teachable, observable steps.✓
- b.A record of how long a learner takes to complete each part of a daily routine.
- c.A summary of the antecedents and consequences surrounding each occurrence of a behavior.
- d.An assessment that identifies which items function as reinforcers within a given setting.
(a) is correct: a task analysis divides a multi-step skill into discrete, observable components that can be taught and measured one at a time, and it is the foundation every chaining procedure is built on. (b) describes duration or latency measurement applied to a routine, not the breakdown itself. (c) describes antecedent-behavior-consequence data collection, which is part of descriptive assessment rather than skill breakdown. (d) describes a preference or reinforcer assessment, which belongs to the assessment domain.
Showing 40 of 100