Jordan B Peterson · Published 2026-09-06

How To Find Your “Upward Aim” When You Feel Stuck

Open on YouTube ↗

Summary

Overview

  • Speaker: Jordan Peterson
  • Channel: Jordan B Peterson
  • Main topic: Finding meaning and direction through an upward aim and taking moral responsibility
  • Purpose: To provide psychological and philosophical insights on how individuals can overcome personal stagnation and suffering by adopting an upward moral aim. Jordan Peterson explores the psychological and biblical frameworks of "aiming up" when dealing with personal suffering, stuckness, and regret. Drawing on biblical narratives such as Jacob and Esau, and literary frameworks like Dante's Inferno, he argues that progress is not linear and that facing one's past, taking responsibility, and establishing a moral compass are essential for personal transformation.

Topic Map

The Story of Jacob and Esau: Sibling Rivalry and Usurpation

  • Explanation: Peterson uses the Old Testament story of Jacob and Esau to illustrate deep psychological motifs around sibling rivalry, parental favoritism, and the dangers of usurping what is not rightfully earned.
  • Key claims:
    • Sibling rivalry is a common and ancient motif representing competition for limited resources and parental favor.
    • Jacob represents the schemer and usurper who grabs his brother's heel at birth.
    • Esau represents the rugged hunter who fails to value his future and birthright, trading it for immediate comfort.
  • Examples:
    • Esau trading his birthright for a bowl of red lentil porridge.
    • Jacob dressing in animal skins to deceive his blind father Isaac.
  • Terminology:
    • usurper
    • birthright
    • archetype
  • Why it matters: It demonstrates how ancient stories characterize human flaws, betrayal, and the consequences of lacking foresight.

Dante's Inferno and the Psychology of Hell

  • Explanation: An analysis of Dante's Inferno to show how suffering and misery have hierarchical structures, culminating in betrayal at the lowest levels.
  • Key claims:
    • Hell has levels, starting from common misery and descending into deeper, self-inflicted suffering.
    • Betrayal is placed at the bottom of Dante's hell because it undermines trust, the foundation of community.
    • Denying one's faults and insisting on being right multiplies misery.
  • Examples:
    • The spiral descent into the depths of hell representing escalating psychological misery.
  • Terminology:
    • inferno
    • betrayal
    • pathological past
  • Why it matters: It explains why unaddressed deceit and lack of trust lead individuals into profound psychological suffering.

The Mechanics of Aimee Up and Jacob's Ladder

  • Explanation: Explaining the metaphor of Jacob's ladder as an ancient representation of the connection between the profane realm and the transcendent divine realm.
  • Key claims:
    • Jacob's ladder symbolizes a staircase spiraling up to the highest conceivable place.
    • Upward aim transforms daily actions by aligning them with a transcendent moral framework.
    • Incremental moral improvements lead to an upward trajectory over time.
  • Examples:
    • Angels ascending and descending on Jacob's ladder delivering messages between earth and heaven.
  • Terminology:
    • transcendent
    • upward aim
    • hierarchy of being
  • Why it matters: It provides a mental model for how individuals can transcend their current limitations through purposeful striving.

Key Points

The Danger of Cynicism

  • Explanation: Cynicism is an intermediate step between naive innocence and wisdom, but remaining cynical prevents trust and personal growth.
  • Evidence: A cynic says all hell can break loose so why trust, but this mindset creates a barrier to meaningful relationships.
  • Practical implication: Move past cynicism by consciously choosing courage over bitterness.

The Necessity of Confession and Self-Examination

  • Explanation: To move forward, individuals must thoroughly examine and confront their past mistakes and personal inadequacies.
  • Evidence: Clinical evidence suggests that exploring past darkness and taking responsibility is necessary for psychological recovery.
  • Practical implication: Write down past misdeeds and mistakes in detail to understand where your moral compass went wrong.

The Matthew Principle in Personal Development

  • Explanation: Progress and failure operate non-linearly; improvement accelerates success, while failure accelerates downward spirals.
  • Evidence: Economic and psychological observations show that those who have more are given more, and those who fail fall faster.
  • Practical implication: Catch downward spirals early by making small, consistent correct choices.

Frameworks, Models & Processes

The Upward Aim Framework

  • How it works: Aligning daily behavior and sacrifice with a transcendent moral ideal to elevate one's life and relationships.
  • Components:
    • Acknowledgment of personal flaws
    • Willingness to make sacrifices
    • Commitment to truth and trust
    • Incremental daily corrections
  • When to use: When feeling stuck, depressed, or overwhelmed by existential suffering.

Examples & Case Studies

Jacob deceives his father Isaac to steal his brother Esau's blessing.

  • Illustrates: The consequences of deceit and usurpation in family dynamics.
  • Lesson: Attempting to gain power through manipulation leads to long-term familial conflict and personal exile.

Actionable Takeaways

  • Immediate:
    • Stop blaming external circumstances and evaluate your own contribution to your suffering.
    • Examine your daily habits and eliminate actions you know to be wrong.
  • Strategic:
    • Commit to a long-term upward aim despite the inevitable difficulties and sacrifices required.
    • Build trust in your relationships as the foundational bedrock of community and stability.
  • Questions to investigate:
    • What stupid things am I doing that increase the probability of misery in my life?
    • How can I establish a genuine covenant of trust with the people around me?

Claims Worth Verifying

  • Sibling rivalry is most likely to emerge between same-sex siblings born close together. (psychological observation)
  • Dante's Inferno contains virtually no direct references to Satan or hell in standard biblical texts. (literary and theological analysis)

Notable Quotes

"Pretty good crowd for the city of sin." (at 0:00) "Jacob and Esau fight in the womb." (at 25:41) "What stupid thing am I doing or did do that increased the probability that this hellish situation is upon me?" (at 87:16)

Compressed Summary

  • Adopt an upward aim to guide personal transformation.
  • Confront personal flaws and past mistakes through honest self-examination.
  • Trust is the foundational bedrock of all lasting relationships and societies.
  • Progress and failure both compound non-linearly over time.
  • Keywords: responsibility, meaning, archetype, sacrifice, trust
  • Core insight: Overcoming personal stagnation requires confronting one's past mistakes, taking complete responsibility, and committing to an unwavering upward moral aim.

Core insights

5
Empirical Resultmedium noveltymoderate evidence

Progress and failure are nonlinear and compounding, so a small avoidable error that is corrected early has an outsized effect, while an uncorrected error accelerates into a downward spiral. Agent reliability work should monitor the trajectory of the error rate and intervene at the first wrong action rather than waiting for aggregate quality metrics to drop.

Why it matters

This changes where failure-detection and correction resources are placed: not on post-hoc rollback or retraining after a task fails, but on continuous early local evaluation and small corrective interventions that prevent an exponential failure trajectory.

Generalization

In any self-improving control system, the expected cost of an error grows nonlinearly with how long it is left uncorrected; therefore early detection must be treated as a first-class architectural requirement.

Progress and failure operate non-linearly; improvement accelerates success, while failure accelerates downward spirals.
Open source video
Catch downward spirals early by making small, consistent correct choices.
Open source video
Architecturemedium noveltymoderate evidence

Betrayal is placed below other failures in Dante's moral hierarchy because it destroys trust, and trust is the substrate that makes community possible. For multi-agent architectures, the equivalent is that an agent corrupting or falsely representing shared state is a deeper, more catastrophic failure than an agent making a normal performance error.

Why it matters

Integrity and non-repudiation between agents should be engineered below the application layer: if any agent can silently alter shared context, commitments, or tool results, then every higher-level coordination mechanism is built on a broken foundation.

Generalization

Collaborative systems should classify failures by whether they undermine the trust substrate; such failures need a different, more severe detection and isolation response than ordinary functional failures.

Betrayal is placed at the bottom of Dante's hell because it undermines trust, the foundation of community.
Open source video
Mechanismmedium noveltymoderate evidence

Behavior is improved when each local action is aligned with a high-level 'upward aim,' not merely optimized locally. The mechanism is a hierarchy: a terminal moral aim transforms daily actions, and small daily corrections accumulate into an upward trajectory.

Why it matters

For agent systems, this argues for separating a stable meta-objective or constitution from a fast step-level policy. The local reward should be insufficient on its own; it must be checked against the higher aim to prevent drift and short-term exploitation.

Generalization

Long-horizon behavior emerges when low-level policies are constrained by an explicit, slow-changing objective that is above and distinct from the immediate reward signal.

Upward aim transforms daily actions by aligning them with a transcendent moral framework.
Open source video
Incremental moral improvements lead to an upward trajectory over time.
Open source video
Practicemedium noveltymoderate evidence

Recovery and growth require a detailed, honest examination of one's own past mistakes, not blaming external circumstances. In an engineered system, this maps to preserving high-fidelity memory of the system's own bad decisions and requiring explicit self-critique before a policy update.

Why it matters

If agent logs preserve only successful traces or summarize failures in terms of environment/tool errors, the system lacks the information needed to find the 'stupid thing' it did that created the failure. Confession-like failure analysis becomes part of the learning loop.

Generalization

A self-correcting system must keep an append-only record of its own faults and be designed to ask not only 'what went wrong in the environment?' but also 'what did I do that increased the probability of this failure?'.

Write down past misdeeds and mistakes in detail to understand where your moral compass went wrong.
Open source video
To move forward, individuals must thoroughly examine and confront their past mistakes and personal inadequacies.
Open source video
Tradeoffmedium noveltymoderate evidence

Cynicism is an intermediate developmental state between naive innocence and wisdom; resting in cynicism prevents trust and growth. For adversarial AI systems, this means a permanent posture of distrust after observing failures is not a mature end-state but a stage to pass through.

Why it matters

Security hardening and adversarial training can make a system globally more suspicious, but if that suspicion becomes the system's fixed disposition, it sacrifices the cooperation and information sharing needed for actual capability.

Generalization

Post-adversarial systems need explicit trust calibration mechanisms, so that trust can be selectively rebuilt where behavior warrants it, rather than globally disabled after one betrayal.

Cynicism is an intermediate step between naive innocence and wisdom, but remaining cynical prevents trust and personal growth.
Open source video

Deep dives

4

Early detection of compounding failure trajectories in agent systems

Research question

What leading indicators reliably identify the onset of a downward spiral before task-level metrics collapse?

Why

Because corrections early in nonlinear failure trajectories are cheaper and more effective; waiting for aggregate quality drops misses the intervention point where errors begin to compound.

Progress and failure operate non-linearly; improvement accelerates success, while failure accelerates downward spirals.
Open source video
Catch downward spirals early by making small, consistent correct choices.
Open source video
Source video

Verifiable state integrity for multi-agent trust substrates

Research question

How can append-only attested shared state bound the blast radius of a compromised agent while preserving coordination latency?

Why

Trust-subversion failure is a distinct and deeper class than ordinary performance error; without a substrate that makes corruption visible and non-repudiable, all higher-level coordination mechanisms are built on a broken foundation.

Betrayal is placed at the bottom of Dante's hell because it undermines trust, the foundation of community.
Open source video
Source video

Hierarchical objective architectures with slow constitutions and fast policies

Research question

Does constraining each local action with an explicit terminal aim reduce long-horizon drift more than reward shaping alone?

Why

Local rewards alone cannot distinguish short-term exploitation from long-horizon trajectory; an upward aim serves as a slow-changing meta-objective that prevents drift and aligns daily actions.

Upward aim transforms daily actions by aligning them with a transcendent moral framework.
Open source video
Incremental moral improvements lead to an upward trajectory over time.
Open source video
Source video

Self-fault attribution in postmortem learning loops

Research question

Does requiring agents to 'confess' their own contributing actions before a policy update improve reliability compared with external-failure-only postmortems?

Why

If agent logs summarize failures only in terms of environment or tool errors, the system lacks the information needed to correct its own contribution; self-attribution may be necessary for durable behavioral change.

Write down past misdeeds and mistakes in detail to understand where your moral compass went wrong.
Open source video
To move forward, individuals must thoroughly examine and confront their past mistakes and personal inadequacies.
Open source video
Source video

Article ideas

4

One Bad Action Away From a Death Spiral: Why Agent Reliability Needs Leading Indicators, Not Final-Score Postmortems

Because failure compounds nonlinearly, monitoring the first derivative of minor-error rates and intervening at the first wrong action is more effective than waiting for aggregate quality metrics to collapse.

Angle

Engineering argument that early local correction is an architectural requirement, using nonlinear dynamics as the core rationale.

Source video

Betrayal Is a Bug Class: Designing Multi-Agent Systems Where Trust Is the Foundation, Not an Assumption

Multi-agent architectures must classify trust-subversion failures separately from functional failures and engineer append-only, attested shared state to make betrayal detectable and non-repudiable.

Angle

Security/architecture argument drawing from moral hierarchy: the deepest failure is the one that destroys the trust substrate.

Source video

Stop Tuning Rewards, Start Aiming Up: The Case for a Constitution Layer Between Agent and Action

Robust long-horizon behavior emerges not from finely shaped local rewards but from a stable terminal aim that every proposed local action must be checked against.

Angle

Alignment architecture argument: separate the slow 'moral' objective from the fast policy to prevent myopic drift.

Source video

AI Confession: Why Self-Correcting Agents Need a Durable Memory of Their Own Faults

A self-correcting system must keep an append-only record of its own faults and be forced to ask 'what did I do to increase failure probability?' before every policy update.

Angle

Learning-loop critique of postmortems that blame the environment, proposing confession-like self-attribution as a practical engineering pattern.

Source video

Project ideas

4

SpiralWatch

beyond-evals

A sustained increase in the rate of minor error-prone actions predicts task failure earlier and with fewer false positives than a threshold on task success rate.

Proof of concept

Instrument a suite of agent trajectories with injected avoidable small errors; compare a derivative-based drift detector against an end-threshold detector.

Measurement

Precision, recall, and mean time-to-intervention on labeled runs that spiral versus recover.

Source video

LedgerAgent

gatehouse

In a multi-agent collaborative task, one compromised agent corrupting shared state causes no downstream task failure when shared state is append-only and attested, but causes measurable task failure with mutable shared state.

Proof of concept

Build two multi-agent task environments (mutable shared state vs append-only attested ledger); inject one agent that writes forged operation results; compare outcomes.

Measurement

Task success rate and number of valid downstream states across 100 runs per condition.

Source video

ConstitutionGuard

movement-lab

Agents whose action selection is checked by a constitution layer will achieve higher cumulative success on a long-horizon task than reward-only agents, despite identical step-level rewards.

Proof of concept

Implement a task where local reward is noisy and step-optimal choices conflict with terminal goals; compare an LLM action proposer with and without a constitution-checking filter.

Measurement

Cumulative success rate over 500 episodes and the amount of short-term reward sacrificed per episode.

Source video

Confession Auditor

new

Policy updates that include structured self-attribution of the model's own actions in failed episodes reduce repeated same-category failures more than updates using only external outcome labels.

Proof of concept

Build an agent benchmark with recurring decision traps; log full traces and, for failures, generate self-critique before fine-tuning; compare against standard supervised learning on successful traces.

Measurement

Rate of same-category failure across held-out evaluation episodes.

Source video

Architectural implications

5

Suffering in the described moral model has levels and eventually becomes self-inflicted pathological repetition.

Before

Failures are classified primarily by observable output error and handled uniformly.

After

Failure classification includes whether the system contributed to its own worsening trajectory and whether the failure is compounding or isolated.

Consequence

Incident response routes early-stage self-inflicted failures to local correction and reserves deep intervention for systemic trust failures.

Source video

Betrayal is the deepest failure mode because it destroys the trust on which community is built.

Before

Multi-agent architectures share state and tool results assuming all components are cooperative, with failures treated as ordinary bugs.

After

Architectures add an integrity/security layer that assumes agents can lie or corrupt shared context; key state is append-only, verifiable, and attributed.

Consequence

A compromised agent can no longer silently poison the whole system because its actions are visible and non-repudiable.

Source video

An upward aim transforms daily actions, suggesting a nested hierarchy of objectives.

Before

Each agent chooses actions to maximize a single local reward or task score.

After

There is a separate high-level aim, and each action is additionally checked against that aim before execution or after selection.

Consequence

Agents sacrifice short-term reward more readily when needed, reducing drift from the intended long-horizon behavior.

Source video

Recovery from a pathological past requires detailed examination and confession of one's own misdeeds.

Before

Postmortems focus on what happened outside the model: bad tools, ambiguous instructions, or environment anomalies.

After

Postmortems include a replayable trace of the model's own actions and a structured step where the system must identify its own contribution to the failure.

Consequence

The same underlying bad decision pattern is less likely to be repeated because the corrective signal is recorded in a machine-reviewable form.

Source video

The Matthew principle means downward spirals are hard to reverse once they gain momentum.

Before

Observability systems alert after a metric such as task success rate has fallen below threshold.

After

Monitoring systems look at the first derivative and consistency of small 'bad choices' to detect spiral initiation before macro metrics collapse.

Consequence

Interventions are cheaper and more surgical because they happen while the deviation is still linear and local.

Source video

Tradeoffs and failure modes

5

Trust vs. adversarial hardening

Benefit

Systems that never trust other agents are resilient to betrayal and deception.

Cost or risk

Permanent distrust blocks the information exchange and collaboration that the system needs to grow and adapt.

A cynic says all hell can break loose so why trust, but this mindset creates a barrier to meaningful relationships.
Open source video
Source video

Early intervention on small errors

Benefit

Catching a small wrong choice early prevents a compounding downward spiral and makes correction cheap.

Cost or risk

If the early signal is noisy, over-sensitive intervention can over-correct healthy exploration and creative behavior.

Catch downward spirals early by making small, consistent correct choices.
Open source video
Source video

Detailed self-examination and confession

Benefit

Identifying one's own contribution to a failure makes the correction specific and durable.

Cost or risk

Detailed review of past mistakes is expensive and slows the loop; if done superficially or performatively, it adds cost without producing a change in behavior.

Write down past misdeeds and mistakes in detail to understand where your moral compass went wrong.
Open source video
Source video

Alignment to a high-level upward aim

Benefit

It gives local actions a long-horizon telos and reduces drift toward short-sighted, locally optimal behavior.

Cost or risk

The aim is abstract and difficult to impossible to formalize exactly; an imperfect or overly rigid formulation can justify destructive sacrifices.

Willingness to make sacrifices
Open source video
Source video

Denial versus owning failure

Benefit

Attributing problems to external circumstances protects the actor from immediate discomfort and blame.

Cost or risk

It multiplies downstream misery because the actual cause is never addressed.

Denying one's faults and insisting on being right multiplies misery.
Open source video
Source video

Open questions

4

How can an 'upward aim' be specified as a stable, non-gameable objective for an AI system?

Why unresolved

The summary treats the aim as transcendent and frames its value precisely in being beyond the local, immediately measurable world.

Research direction

Design experiments in hierarchical value alignment where a slow 'constitution' layer constrains a fast local policy, and measure the effect on long-horizon outcomes.

Source video

What is the earliest observable signal of a downward spiral in agent behavior?

Why unresolved

The summary says failure accelerates nonlinearly, but gives no concrete measurement for detecting the point where small repeated errors begin to compound.

Research direction

Instrument agent traces to measure the first derivative of error-prone action sequences and compare leading indicators against later collapse.

Source video

How should multi-agent architectures detect and survive an insider agent that intentionally corrupts shared context?

Why unresolved

Betrayal is identified as the deepest failure mode, but the summary does not describe an operational detection mechanism.

Research direction

Apply Byzantine fault-tolerance and attestation techniques to agent communication, measuring cost and latency under an adversarial insider.

Source video

Does requiring an agent to periodically 'confess' its own faults improve future reliability, or does it merely overfit it to known failure patterns?

Why unresolved

The summary asserts confession is necessary for recovery, but does not specify whether this is an ongoing process or a one-time unblocking event.

Research direction

Compare policy update processes that include explicit own-fault attribution versus those that only incorporate external failure evidence.

Source video

Key claims

6
causalVerification needed

Progress and failure operate non-linearly; improvement accelerates success, while failure accelerates downward spirals.

Evidence

Economic and psychological observations show that those who have more are given more, and those who fail fall faster.

Question

Can this compounding be empirically measured in agent learning curves so that early intervention thresholds can be derived?

Source video
comparativeVerification needed

Betrayal is placed at the bottom of Dante's hell because it undermines trust, which is the foundation of community.

Evidence

Betrayal is placed at the bottom of Dante's hell because it undermines trust, the foundation of community.

Question

Is the severity ordering in Dante's inferno consistently based on trust-subversion, and can it be mapped to a failure taxonomy for multi-agent systems?

Source video
factualVerification needed

Clinical evidence suggests that exploring past darkness and taking responsibility is necessary for psychological recovery.

Evidence

Clinical evidence suggests that exploring past darkness and taking responsibility is necessary for psychological recovery.

Question

What specific clinical studies or controlled experiments support the causal direction, and do they generalize to AI self-correction?

Source video
opinionVerification not requested

Cynicism is an intermediate step between naive innocence and wisdom, but remaining cynical prevents trust and personal growth.

Evidence

Cynicism is an intermediate step between naive innocence and wisdom, but remaining cynical prevents trust and personal growth.

Source video
factualVerification needed

Sibling rivalry is most likely to emerge between same-sex siblings born close together.

Evidence

Sibling rivalry is most likely to emerge between same-sex siblings born close together.

Question

Does this observation replicate in the developmental psychology literature, and does it depend on the family or culture studied?

Source video
comparativeVerification needed

Jacob represents the schemer and usurper who grabs his brother's heel at birth.

Evidence

Jacob represents the schemer and usurper who grabs his brother's heel at birth.

Question

Which source text and interpretation is being used, and does the original Hebrew support the reading of 'grabs the heel' as usurpation?

Source video

Connections

5