Grades Were Created to Measure Learning. Then They Became the Point of It
Imagine asking a college student how a course is going.
The answer will often begin with a letter.
“I have an A.”
“I’m trying to bring my B up.”
“I can still pass if I do well on the final.”
These answers tell us something about the student’s academic standing. They tell us much less about what the student understands, what the student can do, or how the student’s thinking has changed.
That distinction matters because letter grades have become so deeply embedded in American education that they can seem like a natural and necessary part of learning.
They are neither natural nor particularly old.
Grades began as an educational reform
For much of the history of education, students learned without receiving an A, B, C, D, or F. Early American colleges relied on oral examinations, public demonstrations, written comments, and broad descriptions of performance.
There was no shared grading system. Institutions experimented with numbers, percentages, ranks, divisions, and descriptive labels.
The nineteenth-century education reformer Horace Mann helped introduce several practices that would eventually shape modern grading. At the time, students were frequently examined orally, ranked against their classmates, and sometimes physically moved within the classroom according to their performance.
Mann believed this constant competition could undermine motivation. After studying the Prussian education system, he promoted written examinations and monthly report cards that would document individual progress, inform parents, and help schools maintain consistent records.
His goal was not to make grades the center of education. It was to replace public ranking with a system that could show how students were developing over time.
As Robert Talbert explains in Grading for Growth, Mann’s reforms were legitimate improvements over the systems they replaced. Students and parents previously had few reliable ways to understand what had been learned or what still needed improvement. The problem came later, as a record intended to communicate growth accumulated more purposes and more power.
By the end of the nineteenth century, grades were beginning to take the familiar form of letters, points, and averages. By the early twentieth century, much of the grading structure we now consider traditional was in place.
Grades helped solve real problems.
But eventually, the tool created to document learning began to define it.
A remarkably efficient system
The letter grade is an impressive piece of administrative technology.
A professor can evaluate months of work and reduce the result to a single symbol. That symbol can be entered into a transcript, converted into grade points, averaged with other courses, and used by people who were never present in the classroom.
An A in one course can be combined mathematically with a B in another, even when the courses measure entirely different knowledge and skills.
The resulting grade-point average can then influence:
Academic standing
Scholarships and financial aid
Admission to majors
Graduate and professional school applications
Honors and awards
Internship opportunities
Employment screening
Grades make complex decisions manageable.
That does not necessarily make them precise.
A grade appears objective because it is expressed as a letter, number, or percentage. But every grade is the result of many human decisions.
Faculty determine what work students will complete, how much each assignment will count, what criteria will be used, whether participation matters, how late work will be handled, and what level of performance separates an A from a B.
Two instructors can evaluate similar work differently. Two sections of the same course can have different expectations. A grade of 88 may suggest mathematical precision, but the judgments producing that number are rarely precise to a single percentage point.
Grades provide useful information. The problem begins when we mistake compressed information for complete information.
When the measure becomes the objective
Students are rational participants in the systems we create.
When institutions make grades consequential, students pay attention to grades.
They calculate what assignments are worth. They ask what will be on the test. They determine how much effort a task requires. They prioritize graded work over ungraded learning. They avoid risks that could lower their averages.
None of this means students are lazy or unconcerned about learning.
It means they understand the incentives.
A student who experiments with a difficult idea may produce weaker work than a student who follows the safest path through an assignment. A student who chooses a challenging course may receive a lower grade than one who selects an easier option. A student who struggles, responds to feedback, and improves substantially may still be penalized by an average that preserves every early mistake.
The system may say we value curiosity, intellectual risk, revision, and growth. The gradebook can communicate something different:
Get it right the first time.
Complete what earns points.
Avoid unnecessary risks.
Protect the average.
Grades do not automatically eliminate a student’s interest in learning. Students can care deeply about a subject while also caring about their grades.
But when every learning activity becomes a transaction for points, the transaction can overwhelm the learning.
Feedback and grades are not the same
A grade tells students how their work was judged.
Feedback helps them understand what to do next.
These functions are related, but they are not interchangeable.
Consider the difference between receiving “B-” and receiving this information:
Your argument is clear, but the evidence does not fully support your conclusion. Two sources are summarized rather than analyzed. In your revision, explain why the evidence matters and address the strongest competing interpretation.
The letter efficiently records a result. The feedback creates an opportunity for learning.
In practice, students may focus so intensely on the grade that they barely process the comments accompanying it. Once the grade has been awarded and the class has moved to the next assignment, detailed feedback can feel like an explanation of a closed decision rather than guidance for future improvement.
Why do we provide some of our most detailed feedback after students have lost the opportunity to apply it?
The more useful approach is to give students information while there is still time to revise, practice, and improve. Feedback becomes educational when students are expected to do something with it.
Grades now carry more weight than they can support
A course grade was never designed to provide a complete description of a person’s intelligence, potential, work ethic, creativity, judgment, or professional readiness.
Yet grades are frequently asked to stand in for all of them.
An employer may interpret a GPA as evidence of future performance. A graduate program may use it to estimate academic potential. A scholarship committee may use it as a measure of merit. Students may internalize it as a measure of their own ability.
But the same grade can represent very different realities.
A B might reflect strong work in an unusually demanding course. It might reflect incomplete mastery. It might reflect excellent learning combined with missed deadlines. It might reflect average work under one professor’s standards and exceptional work under another’s.
Even within a single assignment, a grade can combine several dimensions: knowledge, writing ability, organization, compliance with directions, timeliness, participation, and presentation.
Once those dimensions are merged into one symbol, the information cannot easily be separated again.
Grades are effective at ranking and sorting because they reduce complexity.
Learning is complex.
Generative AI is making that tension harder to ignore. When students can use AI to improve the apparent quality of essays, presentations, code, and other products, a grade assigned to the final submission may reveal less about the learner’s independent understanding.
The question is no longer simply whether AI makes academic dishonesty easier.
It is whether our grading systems were already placing too much emphasis on polished products and too little emphasis on the learning process that produced them.
The answer is not necessarily to abolish grades
Grades continue to serve practical purposes. Institutions need to determine whether students are making satisfactory progress. Faculty need a way to certify performance. Other institutions and employers need information they can understand.
Eliminating grades entirely would also create new problems. Narrative evaluations require considerable faculty time. Pass-fail systems can conceal meaningful differences in performance. Alternative approaches can become confusing when students transfer or apply to programs that expect conventional transcripts.
The more productive question is not whether every college should abandon grades tomorrow.
It is whether grades should continue to dominate learning as completely as they do now.
Faculty and institutions can begin with smaller changes:
Use more low-stakes practice before high-stakes evaluation.
Give students opportunities to apply feedback through revision.
Separate academic mastery from behaviors such as punctuality when possible.
Design assessments around authentic demonstrations of ability.
Explain what a grade represents rather than relying on a percentage alone.
Use portfolios, narratives, competencies, and examples of work alongside grades.
Reduce point accumulation that rewards completion without demonstrating learning.
Ask students to reflect on how their work changed and what they learned from the process.
None of these changes eliminates standards.
In many cases, they make standards clearer.
Recovering the purpose of assessment
Assessment should help answer three questions:
What should students learn?
What evidence would demonstrate that they learned it?
What should happen when they have not learned it yet?
The modern grading system often answers a different question:
How many points did the student accumulate before the semester ended?
That question is administratively convenient, but it should not become the organizing principle of education.
Grades developed to help schools record, summarize, compare, and communicate student performance. They still perform those functions remarkably well.
But a tool designed to document learning should not become the reason students learn.
The grade should be evidence of the education.
It should never become the education itself.


