Moral development · beyond the ladderKohlberg is the start of the argument, not the end
Kohlberg's six-stage ladder (preconventional → conventional → postconventional) is real and useful. But three major critiques reshaped the field, and knowing them is the difference between memorizing a ladder and understanding moral psychology.
The dilemmas themselves are worth picturing, because the method shaped the theory. In the most famous, a man named Heinz cannot afford the drug that would save his dying wife, and the interviewer asks whether he should steal it. Kohlberg cared almost nothing about the yes-or-no answer and almost everything about the why: the structure of the justification revealed, in his account, the underlying stage of reasoning (Kohlberg, 1969). A child who says stealing is wrong because Heinz would go to jail reasons from consequences to the self; an adult who says a human life outweighs a property right reasons from an abstract principle. Same act recommended, radically different logic — and it was the logic Kohlberg staged.
This was a deliberately cognitive-developmental theory, built on Piaget's insight that thinking is restructured in qualitative stages rather than simply accumulated (Kohlberg, 1969). The stages were claimed to be universal in sequence, invariant in order, and irreversible: everyone passes through them in the same order, no one skips a rung, and no one slides back down. A twenty-year longitudinal study of American boys offered the strongest support for that sequence claim, finding that reasoning moved upward through the lower and middle stages without skipping or regression, exactly as the model required (Colby et al., 1983). That study is also where the theory's soft spot shows: the highest, principled stage was so rare that it nearly vanished from the data.
Cross-cultural evidence complicates the universality claim in an instructive way. A review of forty-five studies across dozens of societies found that the lower stages appeared everywhere in the predicted order, supporting a genuine developmental core — but that the postconventional stage, as Kohlberg scored it, was largely absent outside urban, educated, Western samples (Snarey, 1985). One reading is that most of humanity is morally underdeveloped, which is absurd. The better reading is that the scoring system mistook a particular culture's individualistic, rights-based rhetoric for the summit of moral maturity, and could not hear communal or relational moral languages as anything but lower-stage. That critique — that the ladder's top is parochial, not universal — is the doorway to the three thinkers who reshaped the field.
Kohlberg scored morality by the sophistication of a person's verbal justifications to hypothetical dilemmas. Three problems: (1) it privileges abstract justice reasoning over other moral concerns; (2) it assumes reasoning drives moral behavior, when much moral judgment is fast and intuitive; and (3) it was built largely on interviews with boys. The corrected view: moral development is multiple systems — care and justice, intuition and reasoning, domain-specific knowledge — not a single ladder everyone climbs at the same rate.
Carol Gilligan (1982) — argued Kohlberg's scale carried a male bias: it treated abstract justice reasoning as the pinnacle and scored women lower because they more often reasoned from an ethic of care (relationships, responsibility, avoiding harm). Not less moral — a different, equally valid orientation. Elliot Turiel (1983) — showed children distinguish moral transgressions (hitting, stealing — wrong even if no rule forbids them) from social-conventional ones (wearing pajamas to school — wrong only because of a rule) far earlier than Kohlberg's stages imply. Even preschoolers get the difference. Jonathan Haidt (2001) — the social-intuitionist model: moral judgment is mostly fast intuition first, with reasoning arriving after to justify the gut call ("moral dumbfounding"). His Moral Foundations work adds concerns beyond harm/fairness — loyalty, authority, sanctity — that a justice-only ladder misses.
Four takes on moral development. Tap one.
The three critiques · in depthCare, domains, and the intuitive dog
Gilligan's charge began with a pattern in the data: women in Kohlberg's samples were often scored at stage three, men at stage four, making women look developmentally behind (Gilligan, 1982). Her reinterpretation was that the scale, not the women, was the problem. Where the justice orientation asks what is fair, what the rules are, and whose rights win, a care orientation asks who will be hurt, what a relationship needs, and how to stay responsible to the people involved. The second voice is not a lower rung on the first ladder; it is a different ladder, and the standardized dilemmas — abstract strangers in hypothetical conflicts — were built to reward the first.
The uncomfortable epilogue is that the empirical version of Gilligan's claim — that men and women actually reason differently — has not held up well. A careful review of the moral-reasoning studies found essentially no sex difference in Kohlberg stage scores once education and the type of dilemma were controlled (Walker, 1984). A later meta-analysis of the care-versus-justice orientation itself found that the gender difference existed but was small and depended heavily on whether the dilemma concerned a relationship or a stranger; the content of the problem predicted the "voice" far better than the sex of the reasoner (Jaffee & Hyde, 2000). The lasting contribution is therefore not that women think in care and men in justice, but the deeper point that care is a genuine moral domain a justice-only scale will systematically miss, in anyone.
Turiel attacked a different assumption — that young children are moral primitives who confuse rules with right and wrong. In his social-domain theory, children as young as three or four already sort transgressions into distinct conceptual domains: moral violations involving harm, fairness, or rights, and conventional violations that break a social agreement (Turiel, 1983). The evidence is that children treat the two differently along several dimensions at once. They judge hitting or stealing as wrong even in a hypothetical school with no rule against it, and wrong in other countries too — harm is treated as authority-independent and universal. But wearing pajamas to class or eating dessert first is judged wrong only because a rule or custom says so, and fine if the rule were changed. That children draw this line long before Kohlberg's conventional stage supposedly arrives means moral understanding is not one slowly-climbed ladder but several knowledge systems developing in parallel.
Haidt's social-intuitionist model targets the most basic assumption of all — that moral judgment is a product of moral reasoning (Haidt, 2001). His claim, drawing on dual-process psychology, is that judgment usually arrives first, fast and affect-laden, and that reasoning is typically a post-hoc lawyer hired to defend a verdict the gut has already returned. The signature evidence is moral dumbfounding: presented with harmless-but-taboo scenarios — consensual sibling incest with contraception, eating a family dog killed by accident — people condemn them instantly and then flounder to explain why, running out of reasons yet refusing to budge (Haidt, Koller, & Dias, 1993; Haidt, 2001). If reasoning drove judgment, losing the reasons should soften the verdict; that it does not suggests the verdict was never built from reasons. Haidt's later Moral Foundations work extends the point by cataloguing intuitions a harm-and-fairness ladder ignores — loyalty, respect for authority, and sanctity or purity — that carry enormous moral weight across cultures and political divides.
None of the three abolished Kohlberg; they bounded him. The current consensus keeps his core finding — that the capacity for moral reasoning genuinely develops and broadens with age — while rejecting three overreaches: that reasoning is the whole of morality (Haidt's intuitions matter), that justice is its summit (Gilligan's care is a peer), and that sophistication arrives late and all at once (Turiel's domains are early and plural). Morality is now studied as several interacting systems rather than a single staircase. The ladder is one true thing about moral development, not the only thing.
Erikson · the engine of middle childhoodIndustry versus inferiority
Moral reasoning does not develop in a vacuum; it develops in a child who is, at the same age, staking out an identity as a competent person. Erikson's psychosocial theory names the central crisis of the school years — roughly ages six to twelve — as industry versus inferiority (Erikson, 1963). The child's world has widened from the family to the classroom, the team, and the peer group, and the developmental task is to become good at the things the culture values: reading, arithmetic, sport, craft, cooperation. Success builds a durable sense of competence — the felt conviction that one can make things and do them well. Repeated failure, or a steady diet of comparison in which the child always comes up short, breeds inferiority: the sense of being fundamentally less capable than others.
What makes this stage matter for the rest of the chapter is that its currency is social comparison, which comes online precisely now. Younger children rate themselves unrealistically highly and rarely measure against peers; by middle childhood they compare relentlessly and accurately, and their self-esteem increasingly reflects where they actually rank (Erikson, 1963). This is the same machinery that makes middle childhood the launch point for peer hierarchies and for bullying, which trades in exactly this coin — public demonstrations of one child's dominance and another's diminished standing. Competence is not a luxury but a foundation: the child who leaves this stage believing "I am capable" carries a buffer into adolescence, while the child who leaves it feeling inferior carries a vulnerability.
Peers · the second world of childhoodFriendship is developmental, not decorative
Adults tend to treat children's friendships as pleasant background. Developmental science treats them as a distinct socializing force, doing work that parents structurally cannot. The classic argument is that the parent-child relationship is vertical — an attachment to someone with more power and knowledge — while peer relationships are horizontal, between equals, and it is only among equals that a child practices negotiation, compromise, and the reciprocity that cooperation and fairness demand (Sullivan, 1953). A parent can love a child unconditionally; only a peer can teach a child what it costs to be a bad partner, because only a peer can walk away.
Two aspects of peer life must be kept separate because they predict different things. Acceptance is a child's standing in the group — how liked or disliked they are by peers at large — while friendship is a specific, mutual, dyadic bond. A child can be widely accepted yet friendless, or have a close friend while being generally rejected, and the two carry different protective value. Both matter: a landmark review concluded that children who are chronically rejected by peers are at genuinely elevated risk for later difficulties, including dropping out of school and externalizing problems, making low peer acceptance one of the better childhood predictors of later maladjustment (Parker & Asher, 1987). Peer standing is not just a barometer of how a child is doing; it feeds forward into how they will do.
Friendship also changes shape across childhood in a way that tracks cognitive development. For young children a friend is largely whoever is nearby and shares toys — friendship is momentary and activity-based. Across middle childhood it deepens into an appreciation of loyalty, trust, and mutual assistance, and by adolescence into intimacy and self-disclosure. This maturation is not incidental to the moral story: the reciprocity and perspective-taking that friendship demands are the same capacities that moral reasoning and the care orientation are built from. Learning to be a good friend and learning to be a moral person are, in middle childhood, nearly the same project — which is exactly why the failure mode of peer life, bullying, is so developmentally costly.
Bullying · what works vs. what doesn'tGood intentions aren't evidence
What matters most here is the intervention science. Some popular responses backfire; others have real evidence behind them.
Start with what bullying actually is, because the definition does real work. The foundational research program defined bullying as aggressive behavior with three defining features: it is intentional, it is repeated over time, and it occurs within a relationship marked by a power imbalance between aggressor and target (Olweus, 1993). Those three criteria are what separate bullying from an ordinary fight between equals or a single cruel remark. The power imbalance — in size, numbers, status, or social savvy — is why the target cannot simply defend themselves and make it stop, and the repetition is why bullying grinds down its victims in a way a one-off conflict does not. Getting the definition right matters for measurement and for policy: a school that treats every playground conflict as bullying and one that recognizes none both fail children.
It is common. A large nationally representative U.S. survey found that roughly three in ten students in grades six through ten reported moderate or frequent involvement in bullying — as a perpetrator, a target, or both — with those involved showing poorer psychosocial adjustment than uninvolved peers (Nansel et al., 2001). And the harm is not confined to childhood. A prospective study that followed children into adulthood found that having been bullied predicted elevated rates of anxiety, depression, and other disorders decades later, with the worst outcomes for the "bully-victims" who were both targeted and aggressive; crucially, these associations survived controls for family hardship and childhood psychiatric problems, arguing that the victimization itself does lasting damage rather than merely marking children who were already struggling (Copeland et al., 2013).
The single most important scientific correction to the popular picture is that bullying is a group process, not a private duel between a bully and a victim. Detailed observation of classrooms shows that most children present are not neutral: they occupy participant roles — assistants who join in, reinforcers who laugh and egg the bully on, outsiders who withdraw and pretend not to see, and defenders who intervene for the victim (Salmivalli et al., 1996). The bully is typically performing for an audience, and the audience's response is the reward. This reframing changes everything about intervention: if aggression is sustained by the reactions of the surrounding group, then the leverage point is not only the bully but the bystanders, whose silence functions as applause and whose defense can extinguish the behavior (Salmivalli, 2010).
What works
Whole-school programs that change the climate, not just punish individuals. KiVa (Finland; Salmivalli et al.; Kärnä et al., 2011) is the flagship: it targets the bystanders, teaching the silent majority to defend rather than reward the bully with an audience. Bullying is a group phenomenon, so shifting the group's response is what moves the needle.
What doesn't
Zero-tolerance policies — reflexive suspension/expulsion — show little effect on bullying and can worsen outcomes (the school-to-exclusion pipeline). One-off assemblies and "just tell them to stop" also underperform. Punishing the individual ignores the audience that sustains the behavior.
This is why the intervention evidence favors what it does. The largest synthesis of the field — a meta-analysis of controlled evaluations of school anti-bullying programs — found that well-implemented whole-school programs reduced bullying perpetration and victimization by roughly a fifth, a real if modest effect, and identified the ingredients that mattered: parent involvement, firm disciplinary methods, improved playground supervision, and program intensity and duration (Ttofi & Farrington, 2011). Programs that were mere awareness campaigns did little; programs that changed the adult-monitored environment and the peer response did more. The flagship of the bystander-focused approach, Finland's KiVa program, embodies exactly this logic by teaching the silent majority to recognize their own role and to defend rather than reward, and its large controlled trials found meaningful reductions in both bullying and victimization (Kärnä et al., 2011; Salmivalli et al., 2011).
The contrast with zero-tolerance discipline is instructive. Reflexive suspension and expulsion feel decisive and satisfy the demand that something be done, but they show little effect on bullying and can worsen outcomes by pushing already-marginal children out of the very structure — school — that socializes them, feeding an exclusionary pipeline rather than repairing the climate (Ttofi & Farrington, 2011). Punishing the individual bully while leaving the reinforcing audience untouched treats a group phenomenon as an individual crime, and the evidence says it underperforms the harder work of changing what the group does.
It has features that make it distinctively corrosive. Anonymity — aggressors can hide, lowering inhibition (the online disinhibition effect). No escape — the phone follows the child home; there's no safe space that ends at the school bell. Permanence and reach — a screenshot lives forever and can spread to thousands instantly, and bystanders multiply silently through likes and shares. These features mean offline anti-bullying tools don't map one-to-one; digital contexts need their own strategies.
The instinct to treat online cruelty as simply bullying with a keyboard is understandable but incomplete. The most comprehensive review and meta-analysis of the research finds that cyberbullying overlaps heavily with traditional bullying — the same children are often involved, and its links to depression and other harms are comparable — yet it carries structural features that offline bullying lacks (Kowalski et al., 2014). The online disinhibition effect captures the core mechanism: the anonymity, invisibility, and asynchrony of screens loosen the social and emotional brakes that operate face to face, so people say and do things online they never would in person, cruelty included (Suler, 2004). You cannot see your target flinch, you may not even use your name, and the reply need not be immediate — every one of those conditions lowers the threshold for aggression.
Three further features make the digital version distinctively corrosive. There is no escape: where the schoolyard bully is left behind at the final bell, the phone follows the child into the bedroom, so victimization becomes continuous rather than bounded. There is permanence and reach: a cruel post or screenshot can be preserved indefinitely and forwarded to thousands in seconds, so a single act of aggression is re-inflicted with every view. And the bystander audience multiplies silently through likes, shares, and comments, each a small reinforcement delivered at scale (Kowalski et al., 2014). Because the mechanisms differ, the tools must too: strategies that work on a contained playground — supervision, immediate adult intervention — map poorly onto a borderless, permanent, anonymous medium, which is why cyberbullying is treated as a related but not identical problem.
Neurodiversity · reframing the labelsDifference, not just deficit
Beyond ADHD and autism as clinical categories, a modern, respectful frame is neurodiversity: these are variations in how brains are wired, carrying real challenges and real strengths, rather than purely deficits to be erased.
The neurodiversity frame emerged from the autistic self-advocacy movement and reframes conditions such as autism and ADHD as natural variation in human neurology rather than as pure pathology to be cured (Kapp et al., 2013). The claim is not that these conditions carry no impairment — they can carry very real challenges — but that a purely deficit-based model misdescribes the phenomenon and harms the people it describes, treating difference as damage and erasing genuine strengths. Research on how autistic people themselves relate to the concept finds that endorsing a neurodiversity view is associated with a more positive personal identity without requiring denial of difficulty; the deficit and difference framings can coexist, and which one dominates has real consequences for wellbeing and self-concept (Kapp et al., 2013). The reframe is thus not mere euphemism — it is a substantive claim about the most accurate and least harmful way to describe atypical development.
Diagnosis, in this light, is not a neutral readout of biology but a judgment made by fallible systems against prototypes — and the prototypes have biases. One well-documented distortion is the relative-age effect: within a single grade, the youngest children, who are up to a year less mature than their oldest classmates, are diagnosed with ADHD substantially more often, because ordinary immaturity is misread as disorder (Morrow et al., 2012). The same behavior is pathologized or normalized depending on an accident of birth month, which should make anyone cautious about treating a diagnosis rate as a simple fact of nature.
The clearest and most consequential bias runs by sex. Both autism and ADHD are diagnosed later and less often in girls and women, and the reasons are now well understood. Diagnostic criteria were historically built around boys' presentations, so the male-to-female ratio in autism diagnoses — long cited as around four to one — appears to overstate the true difference; a meta-analysis suggests a ratio closer to three to one, meaning a substantial number of autistic girls are being missed (Loomes, Hull, & Mandy, 2017). Part of the reason is camouflaging or masking: many autistic girls and women consciously and effortfully suppress their differences and imitate neurotypical social behavior, which hides their difficulties from clinicians and teachers at a real psychological cost (Hull et al., 2017). The result is a generation of women identified only in adolescence or adulthood, after years of being mislabeled as merely anxious, shy, or difficult — a direct consequence of prototypes built on the wrong template.
Diagnosis isn't a neutral readout. There are debates about overdiagnosis in some groups (e.g., the youngest kids in a grade, who are developmentally behind classmates, are diagnosed with ADHD more often — a relative-age effect) and underdiagnosis in others. The clearest case: girls and women are diagnosed with ADHD and autism later and less often, partly because diagnostic prototypes were built on boys and because girls more often "mask" symptoms (internalizing, camouflaging socially). Many are identified only in adolescence or adulthood — after years of being mislabeled as anxious or "just shy."
Two terms students conflate. Mainstreaming = placing a student with disabilities in the general classroom for part of the day, expecting them to keep up with existing instruction. Inclusion (the stronger, current model) = the student is a full member of the general classroom, with the curriculum and supports adapted to them — the "least restrictive environment" principle in action. The shift from mainstreaming to inclusion mirrors the shift from deficit to neurodiversity thinking.
Synthesis · the through-lineDevelopment is plural, and so is judgment
The single idea binding this unit is that the tidy staircase models of an earlier psychology keep giving way to plural, interacting systems. Moral development is not one ladder of reasoning but a braid of intuition and deliberation, care and justice, early-emerging domains of harm and convention. Peer life is not a sideshow to family but a second developmental world where reciprocity and fairness are actually rehearsed. Bullying is not a duel between two children but a group performance sustained by an audience — which is precisely why the interventions that work redirect the group rather than punish the individual, and why the digital version, with its limitless and anonymous audience, is so hard to contain. And atypical development is not simply deficit but difference-with-challenge, misdescribed by prototypes narrow enough to miss half the children who have it. The common lesson across moral psychology, peer relations, bullying, and neurodiversity is the one that separates memorizing a ladder from understanding a mind: look for the several systems, not the single stair.
SourcesCited in APA 7
Colby, A., Kohlberg, L., Gibbs, J., & Lieberman, M. (1983). A longitudinal study of moral judgment. Monographs of the Society for Research in Child Development, 48(1–2), 1–124.
Copeland, W. E., Wolke, D., Angold, A., & Costello, E. J. (2013). Adult psychiatric outcomes of bullying and being bullied by peers in childhood and adolescence. JAMA Psychiatry, 70(4), 419–426.
Erikson, E. H. (1963). Childhood and society (2nd ed.). Norton.
Gilligan, C. (1982). In a different voice: Psychological theory and women's development. Harvard University Press.
Haidt, J. (2001). The emotional dog and its rational tail: A social intuitionist approach to moral judgment. Psychological Review, 108(4), 814–834.
Haidt, J., Koller, S. H., & Dias, M. G. (1993). Affect, culture, and morality, or is it wrong to eat your dog? Journal of Personality and Social Psychology, 65(4), 613–628.
Hull, L., Petrides, K. V., Allison, C., Smith, P., Baron-Cohen, S., Lai, M.-C., & Mandy, W. (2017). "Putting on my best normal": Social camouflaging in adults with autism spectrum conditions. Journal of Autism and Developmental Disorders, 47(8), 2519–2534.
Jaffee, S., & Hyde, J. S. (2000). Gender differences in moral orientation: A meta-analysis. Psychological Bulletin, 126(5), 703–726.
Kapp, S. K., Gillespie-Lynch, K., Sherman, L. E., & Hutman, T. (2013). Deficit, difference, or both? Autism and neurodiversity. Developmental Psychology, 49(1), 59–71.
Kärnä, A., Voeten, M., Little, T. D., Poskiparta, E., Kaljonen, A., & Salmivalli, C. (2011). A large-scale evaluation of the KiVa antibullying program: Grades 4–6. Child Development, 82(1), 311–330.
Kohlberg, L. (1969). Stage and sequence: The cognitive-developmental approach to socialization. In D. A. Goslin (Ed.), Handbook of socialization theory and research (pp. 347–480). Rand McNally.
Kowalski, R. M., Giumetti, G. W., Schroeder, A. N., & Lattanner, M. R. (2014). Bullying in the digital age: A critical review and meta-analysis of cyberbullying research among youth. Psychological Bulletin, 140(4), 1073–1137.
Loomes, R., Hull, L., & Mandy, W. P. L. (2017). What is the male-to-female ratio in autism spectrum disorder? A systematic review and meta-analysis. Journal of the American Academy of Child & Adolescent Psychiatry, 56(6), 466–474.
Morrow, R. L., Garland, E. J., Wright, J. M., Maclure, M., Taylor, S., & Dormuth, C. R. (2012). Influence of relative age on diagnosis and treatment of attention-deficit/hyperactivity disorder in children. Canadian Medical Association Journal, 184(7), 755–762.
Nansel, T. R., Overpeck, M., Pilla, R. S., Ruan, W. J., Simons-Morton, B., & Scheidt, P. (2001). Bullying behaviors among US youth: Prevalence and association with psychosocial adjustment. JAMA, 285(16), 2094–2100.
Olweus, D. (1993). Bullying at school: What we know and what we can do. Blackwell.
Parker, J. G., & Asher, S. R. (1987). Peer relations and later personal adjustment: Are low-accepted children at risk? Psychological Bulletin, 102(3), 357–389.
Salmivalli, C. (2010). Bullying and the peer group: A review. Aggression and Violent Behavior, 15(2), 112–120.
Salmivalli, C., Kärnä, A., & Poskiparta, E. (2011). Counteracting bullying in Finland: The KiVa program and its effects on different forms of being bullied. International Journal of Behavioral Development, 35(5), 405–411.
Salmivalli, C., Lagerspetz, K., Björkqvist, K., Österman, K., & Kaukiainen, A. (1996). Bullying as a group process: Participant roles and their relations to social status within the group. Aggressive Behavior, 22(1), 1–15.
Snarey, J. R. (1985). Cross-cultural universality of social-moral development: A critical review of Kohlbergian research. Psychological Bulletin, 97(2), 202–232.
Suler, J. (2004). The online disinhibition effect. CyberPsychology & Behavior, 7(3), 321–326.
Sullivan, H. S. (1953). The interpersonal theory of psychiatry. Norton.
Ttofi, M. M., & Farrington, D. P. (2011). Effectiveness of school-based programs to reduce bullying: A systematic and meta-analytic review. Journal of Experimental Criminology, 7(1), 27–56.
Turiel, E. (1983). The development of social knowledge: Morality and convention. Cambridge University Press.
Walker, L. J. (1984). Sex differences in the development of moral reasoning: A critical review. Child Development, 55(3), 677–691.