Skip to content
GRE Psychology overview

Public topic · GRE Psychology

Classical and Operant Conditioning

A free GRE Psychology note on Pavlovian terminology (UCS, UCR, CS, CR), extinction and generalization, reinforcement and punishment, shaping, the four reinforcement schedules, and how observational learning differs.

Concise answer

Classical conditioning pairs a neutral stimulus with an unconditioned stimulus until the now-conditioned stimulus elicits the response on its own; operant conditioning changes voluntary behavior through consequences, with the positive/negative and reinforcement/punishment axes forming four distinct outcomes and schedules governing persistence. Observational learning is the contrast case: behavior acquired by watching a model, gated by attention, retention, reproduction, and motivation.

Definitions

Unconditioned stimulus and response (UCS, UCR)
A stimulus that triggers a response naturally, without learning, and the unlearned response it produces.
Conditioned stimulus and response (CS, CR)
A formerly neutral stimulus that, after repeated pairing with the UCS, elicits a learned response on its own.
Extinction
In classical conditioning, the weakening of the conditioned response when the CS repeatedly occurs without the UCS; in operant conditioning, the decline of a behavior after reinforcement stops.
Negative reinforcement
Removing an aversive stimulus to increase a behavior — reinforcement, not punishment, because the behavior goes up.
Shaping
Reinforcing successive approximations of a target behavior instead of waiting for the complete behavior to occur spontaneously.

Intuition

Classical conditioning transfers a response the organism already has onto a new signal — nothing new is done, something new predicts. Operant conditioning changes what the organism does by making consequences contingent on behavior. The exam's quickest sorting question is therefore: is the response elicited by a stimulus, or emitted and then consequenced?

In operant vocabulary, positive and negative are arithmetic, not evaluative: positive means a stimulus is added and negative means one is removed, while reinforcement and punishment name the direction the behavior moves. Crossing the two axes mechanically dissolves the classic confusion between negative reinforcement and punishment.

Concept walkthrough

Pavlov's dogs fix the classical-conditioning template. Meat powder is an unconditioned stimulus (UCS) producing salivation as an unconditioned response (UCR) — no learning required. A tone is initially a neutral stimulus (NS); sounded repeatedly before the meat powder during acquisition, it becomes a conditioned stimulus (CS) eliciting salivation as a conditioned response (CR). Present the CS repeatedly without the UCS and the CR fades — extinction — though after a rest it can briefly reappear (spontaneous recovery). The CR also spreads to stimuli resembling the CS (stimulus generalization) unless differential experience narrows it to the trained stimulus (stimulus discrimination), and an established CS can even train a further signal (higher-order conditioning).

Operant conditioning, associated with Skinner and rooted in Thorndike's law of effect, works on voluntary behavior through consequences. Two independent axes generate four outcomes: a stimulus is added (positive) or removed (negative), and the behavior increases (reinforcement) or decreases (punishment). Negative reinforcement — removing something aversive so the behavior increases — is the axis-crossing distractor to master. Because complete behaviors rarely occur spontaneously, shaping reinforces successive approximations of the target. Once a behavior is learned, the schedule of reinforcement governs its rhythm and persistence: fixed-ratio and variable-ratio schedules count responses, fixed-interval and variable-interval schedules gate by time, variable ratio produces high steady responding and the greatest resistance to extinction, and fixed interval is the least productive and easiest to extinguish.

Observational learning is the standard contrast: behavior changes from watching a model rather than from direct pairing or direct consequences. Bandura's steps — attention, retention, reproduction, motivation — are individually testable, and motivation is where consequences re-enter secondhand: seeing a model reinforced makes imitation more likely (vicarious reinforcement), and seeing a model punished makes it less likely (vicarious punishment). Exam items often describe a learner who never responded or was never reinforced directly; that detail is the signal to leave the conditioning frameworks entirely.

After this page, you should be able to

  • Label the UCS, UCR, neutral stimulus, CS, and CR in a novel conditioning scenario.
  • Distinguish extinction, spontaneous recovery, stimulus generalization, and stimulus discrimination.
  • Classify consequences into the four operant outcomes by crossing add/remove with increase/decrease.
  • Match the four partial reinforcement schedules to their response patterns and extinction resistance, and contrast conditioning with Bandura's observational learning steps.

Formulas and assumptions

Classical conditioning pairing schema

UCS -> UCR; NS + UCS (repeated) => NS becomes CS; CS -> CR

Variables

  • UCS: stimulus that triggers the response without learning
  • UCR: the unlearned response to the UCS
  • NS: neutral stimulus with no response before pairing
  • CS: the former NS after repeated pairing with the UCS
  • CR: the learned response, now elicited by the CS alone

Assumptions

  • Schematic of the pairing sequence described in OpenStax Psychology 2e Section 6.2, not a quantitative formula.
  • Repeated CS-without-UCS presentations reverse the arrow: the CR undergoes extinction.

Operant outcome quadrant map

(add | remove) x (behavior up | behavior down) => positive/negative reinforcement or punishment

Variables

  • add (positive) / remove (negative): what happens to a stimulus after the behavior
  • behavior up (reinforcement) / behavior down (punishment): the effect on the behavior's frequency

Assumptions

  • Positive and negative describe stimulus arithmetic only; whether the stimulus is pleasant is irrelevant to the labels.

Reinforcement schedule notation

FR-n: every nth response; VR-n: on average every nth response; FI-t: first response after fixed time t; VI-t: first response after variable time averaging t

Variables

  • FR/VR: ratio schedules counting responses (fixed or variable)
  • FI/VI: interval schedules gating by elapsed time (fixed or variable)
  • n, t: the response count or time interval defining the schedule

Assumptions

  • All four are partial reinforcement schedules; continuous reinforcement rewards every response.
  • VR is the most productive and most resistant to extinction; FI is the least productive and easiest to extinguish.

Worked example

Labeling a kitchen conditioning scenario

Most evenings, a parent microwaves popcorn: the microwave beeps, the popcorn smell follows, and the child's mouth waters at the smell. After some weeks, the beep alone makes the child's mouth water. Label NS, UCS, UCR, CS, and CR, and predict what happens if the beep sounds nightly with no popcorn ever following.

  1. 1Find the unlearned reflex first: the popcorn smell makes the mouth water without any training, so the smell is the UCS and watering-to-the-smell is the UCR.
  2. 2Identify the signal that starts out meaningless: the beep initially produces no salivation, so it is the NS.
  3. 3Apply the pairing schema: beep repeatedly precedes smell, so the beep becomes a CS.
  4. 4Name the learned response: mouth-watering to the beep alone is the CR — same behavior as the UCR, but now elicited by the learned signal.
  5. 5Predict from extinction: nightly beeps with no popcorn are CS-without-UCS presentations, so the CR weakens toward zero; after a popcorn-free vacation, a first beep back home might briefly trigger watering again (spontaneous recovery).

NS = beep (before learning); UCS = popcorn smell; UCR = salivation to the smell; CS = beep (after pairing); CR = salivation to the beep. Beeps without popcorn produce extinction of the CR, with possible spontaneous recovery after a rest.

Common traps

  • Reading 'negative' as unpleasant or as punishment — negative means a stimulus is removed, and negative reinforcement increases behavior.
  • Labeling the UCS by salience instead of by learning history; the UCS is whatever triggered the response before any pairing occurred.
  • Confusing stimulus generalization (responding to similar stimuli) with spontaneous recovery (an extinguished response returning after rest).
  • Answering with a conditioning mechanism when the scenario's learner only watched someone else — modeling with attention, retention, reproduction, and motivation is observational learning, not conditioning.

Question depth and domain coverage vary by exam. Practice answers are checked after submission.

Sources

  1. GRE Subject Test Content and StructureETS. Accessed 2026-07-06. Use as a cited source for exam facts; do not imply affiliation or reproduce protected test material.
  2. Psychology 2e, Section 6.2: Classical ConditioningOpenStax. Accessed 2026-08-03. OpenStax textbook content is CC BY-NC-SA 4.0; attribute and avoid verbatim reuse beyond short cited references.
  3. Psychology 2e, Section 6.3: Operant ConditioningOpenStax. Accessed 2026-08-03. OpenStax textbook content is CC BY-NC-SA 4.0; attribute and avoid verbatim reuse beyond short cited references.
  4. Psychology 2e, Section 6.4: Observational Learning (Modeling)OpenStax. Accessed 2026-08-03. OpenStax textbook content is CC BY-NC-SA 4.0; attribute and avoid verbatim reuse beyond short cited references.

Review and maintenance

Source-checked by publisher
Publisher record
Keiko Study editorial owner
Publisher placeholder; no individual credential claim is made yet.
Reviewer record
Technical reviewer pending
Placeholder only; this content is not labeled as reviewed by a named specialist.
Last source check
2026-08-03
Next scheduled review
2026-11-03

Recheck the ETS content-structure page and the OpenStax Psychology 2e learning chapter sections before each major GRE Psychology preparation cycle.