In development. No student data yet.
निर्माणाधीन। अभी किसी छात्र का डेटा नहीं।
Navgati · the tutor inside HeyIAS
Learn in Hindi.
Score in English.
हिंदी में सीखिए।
अंग्रेज़ी में स्कोर कीजिए।
An AI that knows exactly where you are weak, and says so plainly. It is the only teacher you get, it works in Hindi as a first language rather than a translation, and it will tell you in year three that you are still not Mains-ready if that is what the evidence says.
एक AI जो ठीक-ठीक जानता है कि आप कहाँ कमज़ोर हैं, और यह बात साफ़ कहता है। यही आपका इकलौता शिक्षक है, यह हिंदी को अनुवाद नहीं बल्कि पहली भाषा मानकर पढ़ाता है, और तीसरे प्रयास में भी अगर आप मेन्स के लिए तैयार नहीं हैं तो वह यही कहेगा।
Four refusals
चार इनकार
Decided now, while they are still cheap to decide. The market's failure mode is a product that flatters. The aspirant's failure mode is believing they know more than they do. Those are the same failure.
ये फ़ैसले अभी लिए गए हैं, जब इन्हें लेना सस्ता है। बाज़ार की सबसे बड़ी ख़ामी है ऐसा उत्पाद जो चापलूसी करे। अभ्यर्थी की सबसे बड़ी ख़ामी है यह मान लेना कि उन्हें जितना आता है उससे ज़्यादा आता है। ये दोनों एक ही ख़ामी हैं।
UPSC marking is judicially described as band-accurate, not point-accurate. Moderation accepts examiner subjectivity and seeks uniformity. So we return a band and our uncertainty, never a number pretending to be precise.
अदालतों ने माना है कि UPSC की मार्किंग बैंड के स्तर पर सही होती है, अंक के स्तर पर नहीं। मॉडरेशन परीक्षक की व्यक्तिगत राय को मानकर एकरूपता लाने की कोशिश करता है। इसलिए हम बैंड और अपनी अनिश्चितता बताते हैं, कोई नक़ली सटीक संख्या नहीं।
If a grader claims higher agreement than the measured human-human ceiling, it has learned an artifact of the data, not the skill. We publish our own ceiling first, then never claim above it.
अगर कोई ग्रेडर यह दावा करे कि वह दो इंसानी परीक्षकों की आपसी सहमति से भी ऊपर है, तो उसने कौशल नहीं, डेटा की एक ख़ामी सीखी है। हम पहले अपनी सीमा प्रकाशित करते हैं, फिर उससे ऊपर का दावा कभी नहीं करते।
Every claim carries a pointer back to the syllabus, a past paper, or the answer itself. A tutor that cannot show its source is guessing in a confident voice.
हर बात के साथ उसका स्रोत जुड़ा होता है: पाठ्यक्रम, कोई पुराना प्रश्नपत्र, या आपका अपना उत्तर। जो शिक्षक स्रोत न दिखा सके, वह भरोसे की आवाज़ में अंदाज़ा लगा रहा है।
No knowledge-tracing model reliably detects what a student does not know. Production evidence puts specificity below 0.5. So a suspected gap is probed with targeted questions before we spend a single hour of your time on it.
कोई भी मॉडल यह भरोसे से नहीं बता सकता कि छात्र को क्या नहीं आता। असली आँकड़ों में इसकी विशिष्टता 0.5 से नीचे रही है। इसलिए किसी संदिग्ध कमी को पहले सीधे सवालों से जाँचा जाता है, उसके बाद ही आपका एक घंटा उस पर लगाया जाता है।
The engine
इंजन
Not a chatbot with a syllabus attached. Content is fuel; the loop is the product. It never stops running, so your plan is never static.
यह पाठ्यक्रम जोड़ा हुआ कोई चैटबॉट नहीं है। सामग्री तो ईंधन है; असली उत्पाद यह चक्र है। यह कभी रुकता नहीं, इसलिए आपकी योजना कभी जड़ नहीं होती।
Every practice session ends in recall, never rereading. Revision arrives on the day you are about to forget something, and you never build a deck to make that happen.
हर अभ्यास याद करने पर ख़त्म होता है, दोबारा पढ़ने पर नहीं। दोहराव ठीक उसी दिन आता है जिस दिन आप कुछ भूलने वाले होते हैं, और इसके लिए आपको कोई डेक नहीं बनाना पड़ता।
The student model
छात्र का मॉडल
Mastery per atomic concept, decaying over time. Not "Polity 62 percent" but "Fundamental Rights solid, Emergency provisions shaky, Fiscal Policy fading".
हर छोटी अवधारणा पर पकड़, जो समय के साथ घटती है। "राजव्यवस्था 62 प्रतिशत" नहीं, बल्कि "मौलिक अधिकार मज़बूत, आपातकालीन उपबंध कमज़ोर, राजकोषीय नीति धुँधली"।
Separate from knowledge, and stage-specific. You can know everything and still fail Mains on answer structure, word economy, or time per answer.
ज्ञान से अलग, और हर चरण का अपना। आपको सब आता हो, फिर भी उत्तर की बनावट, शब्द-संयम या समय के कारण मेन्स निकल सकता है।
Where retention is won or lost. The subjects you silently skip are diagnostic, because people flee their weaknesses. We ask for your predicted score before every test and track the gap.
यहीं टिकना या छूटना तय होता है। जिन विषयों को आप चुपचाप छोड़ देते हैं, वही सबसे बड़ा संकेत हैं, क्योंकि लोग अपनी कमज़ोरी से भागते हैं। हर टेस्ट से पहले हम आपका अनुमान पूछते हैं और अंतर पर नज़र रखते हैं।
What makes advice actionable. Two hours a day after a job is not ten hours. A fourth attempt needs the opposite handling from a first, and the median successful attempt is the fourth.
यही सलाह को व्यावहारिक बनाती है। नौकरी के बाद के दो घंटे, दस घंटे नहीं होते। चौथे प्रयास को पहले प्रयास से उल्टा संभालना पड़ता है, और सफल अभ्यर्थियों का औसत प्रयास चौथा ही है।
Why it is built this way
यह ऐसा क्यों बना है
of CSE 2021 applicants never sat Prelims at all. Retention is the business, not evaluation. A grader wins a customer; the loop keeps them for three years.
CSE 2021 के आवेदक प्रीलिम्स में बैठे ही नहीं। असली काम टिकाए रखना है, मूल्यांकन नहीं। ग्रेडर ग्राहक लाता है; चक्र उसे तीन साल टिकाता है।
chose Hindi for the compulsory qualifying paper, but only that many wrote their scored Mains papers in it. The bet is Hindi-first learning, often with English scoring.
ने अनिवार्य भाषा पत्र हिंदी में चुना, पर मुख्य परीक्षा हिंदी में इतने ही लोगों ने लिखी। इसलिए दाँव है: सीखना हिंदी में, लिखना अक्सर अंग्रेज़ी में।
of recommended candidates came from attempts three and four together. Only 7.6 percent cleared on their first. A first-attempt product is built for the wrong student.
चयनित अभ्यर्थी तीसरे और चौथे प्रयास से आए। पहले प्रयास में सिर्फ़ 7.6 प्रतिशत निकले। पहले प्रयास के लिए बना उत्पाद ग़लत छात्र के लिए बना है।
is when CSAT hardened, and it did not revert in 2024, 2025 or 2026. Courts did not force relief. So CSAT is a skill track from day one, not a last-week qualifier.
से CSAT कठिन हुआ, और 2024, 2025, 2026 में भी आसान नहीं हुआ। अदालतों से राहत नहीं मिली। इसलिए CSAT पहले दिन से एक अलग कौशल है, आख़िरी हफ़्ते की औपचारिकता नहीं।
| Stage | Share choosing Hindi |
|---|---|
| Compulsory qualifying language paper | 69.3% |
| Scored Mains papers | 3.8% |
The hard part
सबसे कठिन हिस्सा
Publishing sample evaluated copies is already table stakes: the strongest competitor publishes six. What nobody publishes is any evidence that their marks correspond to a human examiner's. That is the wedge, and it is defensible because it is expensive, it needs retired examiners, and a rubric-prompted grader architecturally cannot produce a defensible agreement number.
नमूना जाँची हुई कॉपियाँ छापना अब सामान्य बात है: सबसे मज़बूत प्रतिद्वंद्वी छह छापता है। जो कोई नहीं छापता, वह है इसका प्रमाण कि उनके अंक किसी इंसानी परीक्षक से मेल खाते हैं। यही हमारा दाँव है, और यह टिकाऊ है क्योंकि यह महँगा है, इसके लिए सेवानिवृत्त परीक्षक चाहिए, और सिर्फ़ रूब्रिक से चलने वाला ग्रेडर ऐसा आँकड़ा बना ही नहीं सकता।
This is also why we never ask a model for a mark. We ask it which of two answers is better, many times over, and rank from that. The published evidence is unambiguous:
इसीलिए हम मॉडल से कभी अंक नहीं पूछते। हम बार-बार पूछते हैं कि दो उत्तरों में कौन बेहतर है, और उसी से क्रम बनाते हैं। प्रकाशित प्रमाण साफ़ है:
| Method | QWK |
|---|---|
| Rubric prompting (GPT-4) | 0.567 |
| Human vs human ceiling | 0.761 |
| Comparative judgment, fine-grained | 0.776 |
Status
स्थिति
Research is complete for this phase: 24 indexed artifacts, seven decisions that evidence overturned, and six claims that were refuted outright. The engine core is written and tested. Nothing has touched a real student, and this page will say so until that changes.
इस चरण का शोध पूरा है: 24 दस्तावेज़, सात फ़ैसले जो प्रमाण के आगे बदले गए, और छह दावे जो पूरी तरह ग़लत निकले। इंजन का मूल हिस्सा लिखा और जाँचा जा चुका है। अभी तक किसी असली छात्र तक कुछ नहीं पहुँचा है, और जब तक यह नहीं बदलता, यह पन्ना यही कहेगा।
Per-concept Elo mastery, with separate item difficulties for Hindi and English, because translation shifts difficulty.
हर अवधारणा पर Elo आधारित पकड़, जिसमें हिंदी और अंग्रेज़ी के प्रश्नों की कठिनाई अलग-अलग मापी जाती है, क्योंकि अनुवाद से कठिनाई बदल जाती है।
FSRS retention scheduling, gated by mastery so an unlearned concept is never drilled as if it were revision.
FSRS आधारित दोहराव, जो पकड़ से नियंत्रित है, ताकि बिना समझी हुई अवधारणा को दोहराव समझकर न रटाया जाए।
The strategy optimiser, with its anti-burnout, anti-avoidance and anti-collector constraints, and a "why this" on every plan item.
रणनीति अनुकूलक, जिसमें थकान, बचने की आदत और स्रोत जमा करने के विरुद्ध नियम हैं, और हर काम के साथ "यह क्यों" जुड़ा है।
The Phase 0 gates: self-consistency and order-bias probes. If the grader cannot reproduce its own band on the same answer, nothing downstream ships.
चरण 0 की कसौटियाँ: आत्म-संगति और क्रम-पक्षपात की जाँच। अगर ग्रेडर एक ही उत्तर पर अपना ही बैंड दोहरा न सके, तो आगे कुछ भी नहीं जाएगा।
The examiner panel. The long pole of the entire product, and currently unpriced. Calibration is not real until retired examiners have marked against us.
परीक्षकों का पैनल। पूरे उत्पाद की सबसे लंबी कड़ी, और अभी इसकी लागत तय नहीं है। जब तक सेवानिवृत्त परीक्षक हमारे सामने जाँच न करें, कैलिब्रेशन असली नहीं है।
The concept graph and item bank, seeded from all 22 official Prelims papers, 2016 to 2026. Every one is a scan, and the Hindi exists only as pixels.
अवधारणा-ग्राफ़ और प्रश्न-बैंक, जो 2016 से 2026 तक के सभी 22 आधिकारिक प्रीलिम्स पत्रों से बनेगा। हर पत्र स्कैन है, और हिंदी सिर्फ़ पिक्सल के रूप में मौजूद है।