HeyIAS ← Product← उत्पाद

In development. No student data yet.

निर्माणाधीन। अभी किसी छात्र का डेटा नहीं।

नवगति

Navgati · the tutor inside HeyIAS

Learn in Hindi.
Score in English.

हिंदी में सीखिए।
अंग्रेज़ी में स्कोर कीजिए।

An AI that knows exactly where you are weak, and says so plainly. It is the only teacher you get, it works in Hindi as a first language rather than a translation, and it will tell you in year three that you are still not Mains-ready if that is what the evidence says.

एक AI जो ठीक-ठीक जानता है कि आप कहाँ कमज़ोर हैं, और यह बात साफ़ कहता है। यही आपका इकलौता शिक्षक है, यह हिंदी को अनुवाद नहीं बल्कि पहली भाषा मानकर पढ़ाता है, और तीसरे प्रयास में भी अगर आप मेन्स के लिए तैयार नहीं हैं तो वह यही कहेगा।

Four refusals

चार इनकार

Four things Navgati will never do

चार चीज़ें जो नवगति कभी नहीं करेगा

Decided now, while they are still cheap to decide. The market's failure mode is a product that flatters. The aspirant's failure mode is believing they know more than they do. Those are the same failure.

ये फ़ैसले अभी लिए गए हैं, जब इन्हें लेना सस्ता है। बाज़ार की सबसे बड़ी ख़ामी है ऐसा उत्पाद जो चापलूसी करे। अभ्यर्थी की सबसे बड़ी ख़ामी है यह मान लेना कि उन्हें जितना आता है उससे ज़्यादा आता है। ये दोनों एक ही ख़ामी हैं।

01

Give you an exact mark

आपको सटीक अंक देना

UPSC marking is judicially described as band-accurate, not point-accurate. Moderation accepts examiner subjectivity and seeks uniformity. So we return a band and our uncertainty, never a number pretending to be precise.

अदालतों ने माना है कि UPSC की मार्किंग बैंड के स्तर पर सही होती है, अंक के स्तर पर नहीं। मॉडरेशन परीक्षक की व्यक्तिगत राय को मानकर एकरूपता लाने की कोशिश करता है। इसलिए हम बैंड और अपनी अनिश्चितता बताते हैं, कोई नक़ली सटीक संख्या नहीं।

02

Claim it beats two examiners agreeing with each other

यह दावा करना कि वह दो परीक्षकों से भी ज़्यादा सही है

If a grader claims higher agreement than the measured human-human ceiling, it has learned an artifact of the data, not the skill. We publish our own ceiling first, then never claim above it.

अगर कोई ग्रेडर यह दावा करे कि वह दो इंसानी परीक्षकों की आपसी सहमति से भी ऊपर है, तो उसने कौशल नहीं, डेटा की एक ख़ामी सीखी है। हम पहले अपनी सीमा प्रकाशित करते हैं, फिर उससे ऊपर का दावा कभी नहीं करते।

03

State a fact it cannot cite

ऐसा तथ्य कहना जिसका स्रोत वह न दे सके

Every claim carries a pointer back to the syllabus, a past paper, or the answer itself. A tutor that cannot show its source is guessing in a confident voice.

हर बात के साथ उसका स्रोत जुड़ा होता है: पाठ्यक्रम, कोई पुराना प्रश्नपत्र, या आपका अपना उत्तर। जो शिक्षक स्रोत न दिखा सके, वह भरोसे की आवाज़ में अंदाज़ा लगा रहा है।

04

Tell you about a gap it has not tested

ऐसी कमी बताना जिसे उसने जाँचा नहीं

No knowledge-tracing model reliably detects what a student does not know. Production evidence puts specificity below 0.5. So a suspected gap is probed with targeted questions before we spend a single hour of your time on it.

कोई भी मॉडल यह भरोसे से नहीं बता सकता कि छात्र को क्या नहीं आता। असली आँकड़ों में इसकी विशिष्टता 0.5 से नीचे रही है। इसलिए किसी संदिग्ध कमी को पहले सीधे सवालों से जाँचा जाता है, उसके बाद ही आपका एक घंटा उस पर लगाया जाता है।

The engine

इंजन

One loop, run continuously, over one model of one student

एक ही चक्र, लगातार चलता हुआ, एक छात्र के एक मॉडल पर

Not a chatbot with a syllabus attached. Content is fuel; the loop is the product. It never stops running, so your plan is never static.

यह पाठ्यक्रम जोड़ा हुआ कोई चैटबॉट नहीं है। सामग्री तो ईंधन है; असली उत्पाद यह चक्र है। यह कभी रुकता नहीं, इसलिए आपकी योजना कभी जड़ नहीं होती।

The Navgati loop Diagnose, model, optimise, teach, practise by recalling, then re-diagnose. The last step returns to the first, continuously. Diagnose जाँच Model मॉडल Optimise रणनीति Teach पढ़ाई Practise अभ्यास Re-diagnose फिर जाँच IT NEVER STOPS RUNNING यह कभी रुकता नहीं
The dashed return is the part that matters. Most products run this once, at onboarding, and call the result a study plan. Here the last step is the next first step.
बिंदीदार वापसी वाला हिस्सा ही असली है। ज़्यादातर उत्पाद इसे सिर्फ़ शुरुआत में एक बार चलाते हैं और उसी को अध्ययन योजना कह देते हैं। यहाँ आख़िरी क़दम ही अगला पहला क़दम है।

Every practice session ends in recall, never rereading. Revision arrives on the day you are about to forget something, and you never build a deck to make that happen.

हर अभ्यास याद करने पर ख़त्म होता है, दोबारा पढ़ने पर नहीं। दोहराव ठीक उसी दिन आता है जिस दिन आप कुछ भूलने वाले होते हैं, और इसके लिए आपको कोई डेक नहीं बनाना पड़ता।

The student model

छात्र का मॉडल

A student modelled honestly is four layers, not one score

ईमानदारी से देखा जाए तो छात्र चार परतें हैं, एक अंक नहीं

01

Knowledge

ज्ञान

Mastery per atomic concept, decaying over time. Not "Polity 62 percent" but "Fundamental Rights solid, Emergency provisions shaky, Fiscal Policy fading".

हर छोटी अवधारणा पर पकड़, जो समय के साथ घटती है। "राजव्यवस्था 62 प्रतिशत" नहीं, बल्कि "मौलिक अधिकार मज़बूत, आपातकालीन उपबंध कमज़ोर, राजकोषीय नीति धुँधली"।

02

Skill

कौशल

Separate from knowledge, and stage-specific. You can know everything and still fail Mains on answer structure, word economy, or time per answer.

ज्ञान से अलग, और हर चरण का अपना। आपको सब आता हो, फिर भी उत्तर की बनावट, शब्द-संयम या समय के कारण मेन्स निकल सकता है।

03

Psyche

मनोदशा

Where retention is won or lost. The subjects you silently skip are diagnostic, because people flee their weaknesses. We ask for your predicted score before every test and track the gap.

यहीं टिकना या छूटना तय होता है। जिन विषयों को आप चुपचाप छोड़ देते हैं, वही सबसे बड़ा संकेत हैं, क्योंकि लोग अपनी कमज़ोरी से भागते हैं। हर टेस्ट से पहले हम आपका अनुमान पूछते हैं और अंतर पर नज़र रखते हैं।

04

Context

परिस्थिति

What makes advice actionable. Two hours a day after a job is not ten hours. A fourth attempt needs the opposite handling from a first, and the median successful attempt is the fourth.

यही सलाह को व्यावहारिक बनाती है। नौकरी के बाद के दो घंटे, दस घंटे नहीं होते। चौथे प्रयास को पहले प्रयास से उल्टा संभालना पड़ता है, और सफल अभ्यर्थियों का औसत प्रयास चौथा ही है।

Why it is built this way

यह ऐसा क्यों बना है

Four numbers that decided the product

चार आँकड़े जिन्होंने उत्पाद तय किया

53.5%

of CSE 2021 applicants never sat Prelims at all. Retention is the business, not evaluation. A grader wins a customer; the loop keeps them for three years.

CSE 2021 के आवेदक प्रीलिम्स में बैठे ही नहीं। असली काम टिकाए रखना है, मूल्यांकन नहीं। ग्रेडर ग्राहक लाता है; चक्र उसे तीन साल टिकाता है।

69.3% → 3.8%

chose Hindi for the compulsory qualifying paper, but only that many wrote their scored Mains papers in it. The bet is Hindi-first learning, often with English scoring.

ने अनिवार्य भाषा पत्र हिंदी में चुना, पर मुख्य परीक्षा हिंदी में इतने ही लोगों ने लिखी। इसलिए दाँव है: सीखना हिंदी में, लिखना अक्सर अंग्रेज़ी में।

44.7%

of recommended candidates came from attempts three and four together. Only 7.6 percent cleared on their first. A first-attempt product is built for the wrong student.

चयनित अभ्यर्थी तीसरे और चौथे प्रयास से आए। पहले प्रयास में सिर्फ़ 7.6 प्रतिशत निकले। पहले प्रयास के लिए बना उत्पाद ग़लत छात्र के लिए बना है।

2023

is when CSAT hardened, and it did not revert in 2024, 2025 or 2026. Courts did not force relief. So CSAT is a skill track from day one, not a last-week qualifier.

से CSAT कठिन हुआ, और 2024, 2025, 2026 में भी आसान नहीं हुआ। अदालतों से राहत नहीं मिली। इसलिए CSAT पहले दिन से एक अलग कौशल है, आख़िरी हफ़्ते की औपचारिकता नहीं।

The Hindi paradox 69.3 percent chose Hindi for the compulsory qualifying language paper, but only 3.8 percent wrote their scored Mains papers in Hindi. Chose Hindi, qualifying paper अनिवार्य भाषा पत्र में हिंदी 69.3% Wrote scored Mains in Hindi मुख्य परीक्षा हिंदी में लिखी 3.8% SAME ASPIRANTS · SAME EXAM वही अभ्यर्थी, वही परीक्षा
Nearly seven in ten reach for Hindi when it does not affect their rank. Fewer than four in a hundred do it when it does. They are not choosing English because they prefer it; they are choosing it because the Hindi preparation ecosystem is thin and the risk is theirs. That is the gap the product is aimed at.
दस में से लगभग सात लोग हिंदी चुनते हैं, जब उससे रैंक पर असर नहीं पड़ता। सौ में चार से भी कम तब चुनते हैं, जब असर पड़ता है। वे अंग्रेज़ी इसलिए नहीं चुनते कि वह उन्हें पसंद है; वे इसलिए चुनते हैं कि हिंदी की तैयारी का ढाँचा कमज़ोर है और जोखिम उन्हीं का है। उत्पाद इसी खाई को भरने के लिए है।
View as table तालिका में देखें
StageShare choosing Hindi
Compulsory qualifying language paper69.3%
Scored Mains papers3.8%

The hard part

सबसे कठिन हिस्सा

Everyone shows you a marked copy. Nobody shows you calibration.

जाँची हुई कॉपी सब दिखाते हैं। कैलिब्रेशन कोई नहीं दिखाता।

Publishing sample evaluated copies is already table stakes: the strongest competitor publishes six. What nobody publishes is any evidence that their marks correspond to a human examiner's. That is the wedge, and it is defensible because it is expensive, it needs retired examiners, and a rubric-prompted grader architecturally cannot produce a defensible agreement number.

नमूना जाँची हुई कॉपियाँ छापना अब सामान्य बात है: सबसे मज़बूत प्रतिद्वंद्वी छह छापता है। जो कोई नहीं छापता, वह है इसका प्रमाण कि उनके अंक किसी इंसानी परीक्षक से मेल खाते हैं। यही हमारा दाँव है, और यह टिकाऊ है क्योंकि यह महँगा है, इसके लिए सेवानिवृत्त परीक्षक चाहिए, और सिर्फ़ रूब्रिक से चलने वाला ग्रेडर ऐसा आँकड़ा बना ही नहीं सकता।

This is also why we never ask a model for a mark. We ask it which of two answers is better, many times over, and rank from that. The published evidence is unambiguous:

इसीलिए हम मॉडल से कभी अंक नहीं पूछते। हम बार-बार पूछते हैं कि दो उत्तरों में कौन बेहतर है, और उसी से क्रम बनाते हैं। प्रकाशित प्रमाण साफ़ है:

Agreement with human examiners, by grading method Quadratic weighted kappa. Rubric prompting 0.567. Comparative judgment 0.776. Two human examiners agreeing with each other reach 0.761. Rubric prompting रूब्रिक से 0.567 Comparative judgment तुलनात्मक निर्णय 0.776 TWO HUMANS AGREEING · 0.761 दो इंसानी परीक्षकों की आपसी सहमति · 0.761
Rubric prompting sits well below what two human examiners manage between themselves. Comparative judgment clears it. That gap is the entire reason the grader is never asked for a mark, only for which of two answers is better.
रूब्रिक से चलने वाला ग्रेडर दो इंसानी परीक्षकों की आपसी सहमति से काफ़ी नीचे रहता है। तुलनात्मक निर्णय उससे ऊपर निकलता है। यही फ़र्क़ पूरी वजह है कि हम ग्रेडर से कभी अंक नहीं पूछते, सिर्फ़ यह पूछते हैं कि दो उत्तरों में कौन बेहतर है।
View as table तालिका में देखें
MethodQWK
Rubric prompting (GPT-4)0.567
Human vs human ceiling0.761
Comparative judgment, fine-grained0.776

Status

स्थिति

What actually exists today

आज असल में क्या मौजूद है

Research is complete for this phase: 24 indexed artifacts, seven decisions that evidence overturned, and six claims that were refuted outright. The engine core is written and tested. Nothing has touched a real student, and this page will say so until that changes.

इस चरण का शोध पूरा है: 24 दस्तावेज़, सात फ़ैसले जो प्रमाण के आगे बदले गए, और छह दावे जो पूरी तरह ग़लत निकले। इंजन का मूल हिस्सा लिखा और जाँचा जा चुका है। अभी तक किसी असली छात्र तक कुछ नहीं पहुँचा है, और जब तक यह नहीं बदलता, यह पन्ना यही कहेगा।

built

Per-concept Elo mastery, with separate item difficulties for Hindi and English, because translation shifts difficulty.

हर अवधारणा पर Elo आधारित पकड़, जिसमें हिंदी और अंग्रेज़ी के प्रश्नों की कठिनाई अलग-अलग मापी जाती है, क्योंकि अनुवाद से कठिनाई बदल जाती है।

built

FSRS retention scheduling, gated by mastery so an unlearned concept is never drilled as if it were revision.

FSRS आधारित दोहराव, जो पकड़ से नियंत्रित है, ताकि बिना समझी हुई अवधारणा को दोहराव समझकर न रटाया जाए।

built

The strategy optimiser, with its anti-burnout, anti-avoidance and anti-collector constraints, and a "why this" on every plan item.

रणनीति अनुकूलक, जिसमें थकान, बचने की आदत और स्रोत जमा करने के विरुद्ध नियम हैं, और हर काम के साथ "यह क्यों" जुड़ा है।

built

The Phase 0 gates: self-consistency and order-bias probes. If the grader cannot reproduce its own band on the same answer, nothing downstream ships.

चरण 0 की कसौटियाँ: आत्म-संगति और क्रम-पक्षपात की जाँच। अगर ग्रेडर एक ही उत्तर पर अपना ही बैंड दोहरा न सके, तो आगे कुछ भी नहीं जाएगा।

next

The examiner panel. The long pole of the entire product, and currently unpriced. Calibration is not real until retired examiners have marked against us.

परीक्षकों का पैनल। पूरे उत्पाद की सबसे लंबी कड़ी, और अभी इसकी लागत तय नहीं है। जब तक सेवानिवृत्त परीक्षक हमारे सामने जाँच न करें, कैलिब्रेशन असली नहीं है।

next

The concept graph and item bank, seeded from all 22 official Prelims papers, 2016 to 2026. Every one is a scan, and the Hindi exists only as pixels.

अवधारणा-ग्राफ़ और प्रश्न-बैंक, जो 2016 से 2026 तक के सभी 22 आधिकारिक प्रीलिम्स पत्रों से बनेगा। हर पत्र स्कैन है, और हिंदी सिर्फ़ पिक्सल के रूप में मौजूद है।