explainx.ainewsletter3.5k
TrendingNewsPathwaysSkills
Pricing
explainx.ai

Upskill in AI — 16 free pathways, live workshops & bootcamps, and 50+ courses from practitioners. Plus the skills, tools, and MCP servers to practice on.

follow us

corporate training

support@explainx.ai

get started

Find your pathTake Free Evaluation

learn

pathways — start freeworkshopsbootcampscoursescertificationsmock testsexplainx universitycorporate traininglearn skills & mcp

discover

skillsmcp serversexplainx mcptoolsagentsllmsdesignsdictionaryagi trackerranks

company

aboutvisionmissionteaminstructorscommunityhackathonscareers

content

daily AI newsstate of AI — live resultsblogreleasespromptsgeneratorsresource libraryfor LLMsexplainx.ai kids

solutions

all solutionsdeveloper upskillingmarketing upskillingproduct manager upskillingleadership upskilling

newsletter · weekly

Get AI news, tools, and insights in your inbox.

supportcontactprivacytermsdata rightshow we create contentsubmission guidelines

© 2026 AISOLO Technologies Pvt Ltd

On this page

  • TL;DR
  • The original critique is useful—but too uncharitable
  • Explanation is a scaffold, not the finish line
  • Dewey's four impulses as an AI-tutor design test
  • How this maps to explainx.ai Learning and Melo
  • A 20-minute learning-by-making loop you can run now
  • What people are asking
  • The better question for every AI learning product
  • Related on explainx.ai
← Back to blog

explainx / blog

AI Tutors Need Learning by Making, Not Just Better Answers

Why explanations and chatbots do not create durable learning by themselves, what Khan Academy gets right, and how Melo turns answers into practice.

Aug 24, 2026·11 min read·Yash Thakker
AI in EducationAI TutorsLearning ScienceMeloKhan Academy
go deep
AI Tutors Need Learning by Making, Not Just Better Answers

Punya Mishra's essay “Why Sal Khan't: On Learning by Making but Teaching by Telling” lands one sharp question: Sal Khan learns by researching, drawing, questioning specialists, making connections, and producing a video—so why should the learner on the other side merely receive the finished explanation?

The question deserved the lively Hacker News debate it received. The best replies also caught the essay's weakest move: Khan Academy has never been only a pile of videos. It includes exercises, quizzes, tests, mastery progression, teacher dashboards, and a flipped-classroom idea intended to free class time for practice and feedback.

Both points can be true. Explanation is valuable scaffolding. It is not, by itself, durable learning. The design problem for any AI tutor—including Melo, explainx.ai's learning copilot—is how to turn a helpful answer into a learner-owned act of inquiry, construction, expression, and revision.

TL;DR

table · 2 cols
QuestionDirect answer
Is Mishra's central criticism fair?Yes: receiving someone else's polished understanding is not the same as constructing your own.
Is “Khan Academy = watch videos” fair?No: the platform also has substantial practice, mastery, assessment, and teacher-support systems.
What did Khanmigo's adoption reveal?A tutor waiting in a sidebar cannot manufacture curiosity, metacognition, or a reason to persist.
What should an AI tutor optimize for?Learner output: questions, attempts, artifacts, explanations, decisions, feedback, and revision.
Where does Melo fit?It can scaffold, quiz, create practice, and ask for an explanation back; the learner still needs a meaningful destination.
What is the practical rule?Use AI to shorten the distance to a good attempt, not to remove the attempt.
Weekly digest3.5k readers

Catch up on AI

Curated AI updates on agents, skills, and MCP — delivered to your inbox. Unsubscribe anytime.

The original critique is useful—but too uncharitable

Mishra contrasts Khan's own learning process with the learner's experience. Khan first builds a scaffold, follows questions textbooks leave unresolved, draws representations, tests analogies against specialists, and finally makes a public artifact. The student receives the residue of that work: a lucid explanation.

That asymmetry is real. A worked explanation gives the learner the destination without reproducing the cognitive journey that made it meaningful to its creator.

But “watch a video” is still an incomplete description of Khan Academy. The HN commenters are right to steelman it:

  • A clear, free explanation can establish the vocabulary and mental model needed before practice begins.
  • Self-paced video can be paused, rewound, accelerated, or revisited without the social pressure of holding up a classroom.
  • Exercises, quizzes, tests, and mastery levels ask learners to do more than watch.
  • The flipped-classroom model was explicitly meant to move lecture out of shared class time so teachers could use that time for problems, projects, and human feedback.
  • A consistent explanation sets a quality floor for learners whose local instruction is unavailable or incomprehensible.

Those are not small wins. Several HN commenters described Khan's videos as the first explanation that made calculus or another difficult subject feel beautiful and possible. Inspiration and scaffolding matter because a learner cannot practice a concept they cannot yet see.

The more precise criticism is this: a platform can provide every ingredient for active learning and still leave the learner with no reason to combine them. Sal Khan described that engagement problem to Chalkbeat: for many students, Khanmigo was a “non-event” because they did not use it much. A capable tutor sitting in the back of the room is still waiting for the student to raise a hand.

Explanation is a scaffold, not the finish line

An explanation can do three valuable jobs:

  1. Compress prior work. The learner does not need to rediscover calculus from first principles.
  2. Provide a representation. A timeline, analogy, diagram, or worked example makes an abstract structure inspectable.
  3. Reduce the cost of beginning. A learner who was stuck can make the first informed attempt.

It cannot prove that the learner can retrieve, transfer, or use the idea. That requires output from the learner.

This is the same distinction visible in recent education research. Our review of homework gains and exam-score losses with unrestricted AI found that producing better work with AI is not the same as being able to reproduce the reasoning without it. By contrast, the Dartmouth Phosphor study centered constructed responses and rubric-based feedback rather than open-ended chat.

The durable loop looks less like “ask → receive” and more like this:

table · 3 cols
StageLearner actionTutor's job
PurposeChoose a destination worth reachingHelp scope it without inventing the stakes
InquiryAsk, compare, notice a gapSupply sources, examples, and better questions
AttemptSolve, build, classify, predict, or decideWithhold the finished answer long enough for effort
FeedbackExpose the work to a rubric or another mindIdentify the smallest consequential gap
RevisionChange the artifact or explanationCheck whether the change fixed the gap
ExpressionDefend or share the resultChallenge assumptions and ask for evidence

If the learner never reaches the attempt, the AI has delivered content but has not yet designed learning.

Dewey's four impulses as an AI-tutor design test

Four dark learning blocks connected by green arcs, representing inquiry, construction, expression, and communication as one active loop

Mishra uses four impulses associated with John Dewey: inquire, construct, express, and communicate. They are more useful as a product checklist than as abstract philosophy.

table · 4 cols
ImpulseWhat it looks like from the learnerFailure mode in a chatbotBetter design requirement
InquireNotices a gap, asks why, compares explanationsLearner types a vague request and receives a complete answerHelp the learner sharpen the question and inspect sources
ConstructMakes a model, solution, diagram, plan, or artifactAI constructs the whole answer while the learner watchesRequire a learner attempt before showing a full solution
ExpressExplains the idea in their own wordsLearner recognizes fluent prose and mistakes recognition for recallAsk for explanation back, prediction, or defense
CommunicateTests the work against another mind or audienceConversation has no real stakes outside the chatCreate a shareable output and invite human feedback

Notice that “communicate” is not identical to “the chatbot replied.” A responsive model can simulate dialogue, but it cannot automatically create a classmate who depends on your explanation, a colleague who will implement your plan, or a user who will reject a confusing result. Audience creates stakes; stakes create reasons to revise.

That is why social learning for AI needs both dialogic tools and human contexts. An AI can make solitary practice more responsive. It cannot make every solitary exercise socially meaningful.

How this maps to explainx.ai Learning and Melo

explainx.ai's vision says learners do not want to be lectured at; they want to build things. Turning that sentence into product behavior means giving each surface a different job instead of pretending one chat box can do everything.

table · 3 cols
SurfaceRole in the learning loopWhat it does not guarantee
/pathwaysProvides sequence and a destination across related conceptsThat the learner will attempt a project or persist through difficulty
/dashboard/learnMelo can scaffold with Teach or ELI5, test retrieval with Quiz, and prompt self-explanation with Explain BackThat a generated explanation or quiz is factually perfect, or that the learner cares
/practiceProvides no-login, hands-on tools for manipulating AI concepts instead of only reading about themTransfer to an unfamiliar production problem
Melo Practice modeGenerates an applied scenario and can evaluate the response against a rubricReal-world consequences, users, teammates, or domain expertise
Melo's interactive visualsLets a learner watch, explore, or recall a concept inside a lessonA fully open-ended artifact built by the learner
Workshops or a real peerAdds accountability, disagreement, and an audienceUnlimited individual pacing or always-on help

The mapping is intentionally not “Dewey proved Melo is good.” A product feature only creates an opportunity for inquiry or construction. Whether the learner takes it depends on the prompt, the task, the surrounding teacher or cohort, and the reason the work matters.

Melo's most important design choice is therefore not that it can explain. Any frontier chatbot can explain. It is that explanation is only one mode among activities that require output: Quiz, Practice, Explain Back, fill-in-the-blank checks, matching, flashcards, interviews, and interactive visual recall. The generative UI system behind those visuals uses structured components so an explanation can become something the learner manipulates and is tested on, not only another paragraph.

The hardest layer remains outside Melo: purposeful making. A learner should leave the loop with something that exists because they understood—a working agent, an evaluation rubric, a prompt experiment, a decision memo, or an explanation another person can use.

A 20-minute learning-by-making loop you can run now

Pick one concept inside an explainx.ai pathway and give the session a destination: “I need to design a retrieval test for my project,” not “teach me RAG.” Then run this loop:

text
/teach Give me only the minimum scaffold I need to design a RAG evaluation.
Stop before giving me the finished design. Ask me one question at a time.

/practice Give me a small RAG failure scenario and a visible rubric.
Wait for my attempt before giving feedback.

/explain-back I will explain why my evaluation catches the failure.
Challenge the weakest assumption in my explanation.

Then leave the chat:

  1. Build the smallest artifact that embodies the idea.
  2. Run it against one case you did not use while designing it.
  3. Show the output to another person, or write a short note for the next person who must use it.
  4. Revise the artifact based on what they misunderstood or what the test exposed.

This preserves the legitimate value of a Khan-style explanation while refusing to stop there. The AI shortens the route to a meaningful attempt; it does not take the attempt away.

What people are asking

“Should an AI tutor refuse to answer directly?”

Not always. Direct answers are appropriate for factual lookup, accessibility, review, and moments when missing context prevents any useful attempt. The design error is defaulting to a complete answer when the learner's actual goal is skill acquisition. A good tutor should know whether this turn is reference, instruction, practice, or assessment.

“Is making automatically better than listening?”

No. Busywork is still busywork, and an artifact without feedback can preserve a misconception. The making must require the target concept, expose a decision, and produce evidence that can change the learner's next attempt. Our review of LLM simulation games for difficult concepts reaches the same conclusion: interaction is useful when it makes a mechanism visible, not merely because something moves on screen.

“Can an AI create motivation?”

It can reduce friction, personalize examples, offer encouragement, and make progress visible. Those can support motivation. It cannot reliably supply the deeper purpose that comes from identity, curiosity, responsibility to others, or a real problem the learner chose to solve. That is why Khan's later emphasis on human systems is not a retreat from technology; it is a more complete account of what the technology cannot originate.

The better question for every AI learning product

The wrong evaluation is “How good was the explanation?” The better evaluation is:

What did the learner have to notice, produce, defend, and revise because this tool existed?

If the answer is “nothing,” the product delivered information. If the answer names a learner-created artifact and the feedback that changed it, the product may have designed a learning experience.

Khan Academy's videos can be excellent scaffolds. Its exercises and mastery system are meaningful attempts to move past passive viewing. Khanmigo can be a useful support inside that system. The lesson from weak voluntary engagement is not that explanation, practice banks, or AI tutors are worthless. It is that availability is not purpose, and assistance is not agency.

That is the standard explainx.ai should be held to as well. Melo succeeds only when its fluent answer becomes the beginning of the learner's work, not the end.

Related on explainx.ai

  • Introducing Melo: the AI learning copilot built into explainx.ai
  • How Melo teaches with generative UI
  • Social learning for AI: why solo chatbots fail
  • AI homework scores rose while exam scores fell
  • Dartmouth's AI tutor study: why quizzes beat optional chat
  • The doer effect and AI-graded interactive textbooks
  • Introducing interactive AI learning pathways
  • What schools should teach in the AI era

Sources: Punya Mishra's original essay · Matt Barnum's Chalkbeat interview with Sal Khan · Hacker News discussion


This analysis reflects the cited product descriptions, reporting, explainx.ai surfaces, and public discussion available on August 24, 2026. AI learning features and education research can change; verify current product behavior and treat generated feedback as a supplement to human judgment.

Spotted something out of date? Let us know.
Yash Thakker

Written by

Yash Thakker

Yash is an AI expert with over 300K learners. Join his workshops →

Related posts

Aug 22, 2026

Harvard’s AI Professor Clones Are Pitch Simulators, Not Replacements

A viral post says Harvard Business School just launched AI clones of professors for pitches, sales calls, and board meetings. The verified story is narrower and more useful: HBS Foundry gives founders repeatable pitch practice with faculty-modeled AI mentors, including video avatars, alongside live experts.

Aug 22, 2026

The /eli5 Claude Code Skill Anthropic Engineers Are Using Daily

A tweet from Anthropic's Thariq about an internal /eli5 Claude Code skill went viral for promising simple, visual explainers on demand. explainx.ai covers what it does, how to install it, the community pushback on AI verbosity, and why Melo has shipped the same idea as a built-in mode since mid-August.

Aug 21, 2026

The Research Behind AI-Graded Quizzes: 20 Studies on Interactive Textbooks

Dartmouth's Phosphor study isn't an outlier — it sits inside a fast-growing 2025-2026 research cluster on LLM-graded interactive textbooks. We mapped 20 papers, from VitalSource's 15.2-million-event doer-effect replication to Google's Learn Your Way RCT to the Bastani-vs-Kestin fight over whether AI helps or harms learning, and pulled out what actually holds up.