·The Hindu

Lower margin for error in second phase of Census: Registrar-General

In this note
  1. At a Glance
  2. Why in the News
  3. Background & Evolution
  4. Core Static Facts
  5. Multi-Dimensional Analysis
  6. Recent Developments (last 12-18 months)
  7. Prelims Hooks
  8. Why an Open-Text Caste Field Is the Actual Failure Point
  9. What Rides on Phase II That Did Not Ride on Phase I
  10. The Case That Self-Enumeration Improves Caste Accuracy
  11. Fixes with a Named Owner
  12. Anchors for Answers
  13. Mains Relevance
  14. Related Topics to Study Next
  15. Common Errors / Trap Areas

1. At a Glance

  • Census 2027 is India's first fully digital census, conducted via a mobile app and self-enumeration portal instead of paper schedules [S4].
  • The Registrar-General and Census Commissioner of India (RG&CCI), Mritunjay Kumar Narayan, has flagged that Phase II (Population Enumeration) carries a "narrower" margin for error than Phase I, since its questions are more numerous, sensitive, and consequential [3].
  • Involves training of nearly 34 lakh enumerators and supervisors, mostly government school teachers/staff, ahead of the caste-inclusive population count [3].
  • High-yield for Prelims (numbers, dates, phases) and Mains GS-II (governance/federalism) and GS-I (social issues around caste data).

2. Why in the News

  • RG&CCI wrote to all States on September 9, 2026, cautioning that the second phase's margin for error is "narrower" than the first, given the sensitivity of questions (including caste) [3].
  • The letter asked for gender sensitisation to be integrated into enumerator training, stressing respectful interaction, accurate recording, and avoidance of bias/assumptions [3].
  • The circular confirmed self-enumeration will continue into Phase II, with enumerators trained to validate, edit, and accept self-submitted responses [3].

3. Background & Evolution

  • Cabinet approved the "Conduct of Census of India 2027" scheme [1].
  • Cabinet Committee on Political Affairs decision of 30 April 2025 mandated inclusion of caste enumeration in Census 2027 — a first since 1931 [2].
  • Phase I — House Listing Operations (HLO): April 1, 2026 to September 30, 2026, covering housing/household characteristics [1][2].
  • Phase II — Population Enumeration (PE): scheduled for February 2027, covering demographic, socio-economic, and caste data [1].
  • Self-Enumeration Portal (se.census.gov.in) introduced for the first time, offering a 15-day pre-enumeration digital entry window; over 82 lakh households used it during HLO [2].

4. Core Static Facts

Item Detail
Nodal authority Office of the Registrar-General & Census Commissioner of India (RG&CCI), Ministry of Home Affairs
Current RG&CCI Mritunjay Kumar Narayan (in office since Nov 1, 2022) [1]
Phase I House Listing Operations (HLO), April 1 – September 30, 2026 [1]
Phase II Population Enumeration (PE), February 2027 [1]
Enumerators/supervisors ~34 lakh, mostly government school teachers and staff [3]
Digital tools Mobile app, Self-Enumeration Portal, Census Management & Monitoring System (CMMS) [2]
New feature Caste enumeration added to PE phase per Cabinet Committee on Political Affairs decision, 30 April 2025 [2]
Key letter RG&CCI circular to States, dated September 9, 2026, on Phase II training and error margins [3]

5. Multi-Dimensional Analysis

Administrative

  • Reliance on teachers/school staff as enumerators raises capacity and academic-calendar-disruption concerns.
  • Self-enumeration plus enumerator-validation is a hybrid model requiring dual-layer quality control.

Social

  • Caste enumeration is administratively and politically sensitive; enumerator bias or assumption-driven recording risks under/over-counting specific caste categories.
  • Gender-sensitisation mandate reflects concern over respectful data collection from women and third-gender respondents.

Governance/Ethical

  • Emphasis on "accurate recording" and "avoidance of bias" signals accountability concerns given the consequential nature of caste data for reservation policy.
  • Digital-first design (CMMS, geo-referencing) aims at real-time monitoring and error reduction, a governance-technology dimension.

Legal/Constitutional

  • Conducted under the Census Act, 1948, and Census Rules, 1990, though caste enumeration is a Cabinet policy addition, not a fresh legislative mandate.

6. Recent Developments (last 12-18 months)

  • 30 April 2025: Cabinet Committee on Political Affairs approves inclusion of caste enumeration in Census 2027 [2].
  • April 1, 2026: Phase I (HLO) commences nationwide [2].
  • 2 August 2026: Self-Enumeration for Census 2027 begins in Assam [2].
  • September 9, 2026: RG&CCI issues letter to States on Phase II training, error margins, and gender sensitisation [3].
  • September 30, 2026: Phase I (HLO) scheduled to conclude [3].
  • February 2027: Phase II (Population Enumeration, with caste data) to be conducted [1].

7. Prelims Hooks

  • Census 2027 is India's first digital census [1].
  • Current RG&CCI: Mritunjay Kumar Narayan, in office since November 1, 2022 [1].
  • Phase I (HLO) window: April 1 – September 30, 2026.
  • Phase II (Population Enumeration): scheduled for February 2027.
  • Caste enumeration decision taken by the Cabinet Committee on Political Affairs on 30 April 2025.
  • Nearly 34 lakh enumerators and supervisors to be deployed for Census 2027, mostly government school teachers.
  • Self-Enumeration Portal URL: se.census.gov.in.
  • Over 82 lakh households used self-enumeration during Phase I.
  • RG&CCI's cautionary letter to States on Phase II was dated September 9, 2026.
  • Digital monitoring backbone: Census Management & Monitoring System (CMMS).
  • Self-enumeration continues into Phase II, with enumerators validating/editing/accepting self-submitted data.
  • Last caste-inclusive census in India was 1931 (context fact; Census 2027 breaks this 96-year gap for caste data).

8. Why an Open-Text Caste Field Is the Actual Failure Point

  • The error is generated at the keyboard, not at the door — caste in Phase II is recorded as declared by the respondent, not matched against a pre-defined state list [6]. A 34-lakh-strong enumerator force typing free-text strings produces spelling, language, surname, gotra and sub-caste variants that no downstream algorithm reliably collapses.
  • SECC 2011 is the precedent that failed on exactly this fault line — the same open-declaration method yielded roughly 46 lakh distinct caste/sub-caste names, and the Centre flagged 8.19 crore caste-related errors to States, of which about 6.74 crore were corrected and 1.45 crore were not [4].
  • Ex-post classification did not rescue it — an Expert Group under then NITI Aayog Vice-Chairperson Arvind Panagariya was constituted precisely to classify SECC caste data [5]; its classification was never released, and the SECC caste tables remain unpublished to this day [4].
  • Digital capture does not solve a taxonomy problem — CMMS and the mobile app catch format errors (blank fields, invalid ages, geo-mismatch) [2]; they cannot adjudicate whether "Yadav", "Ahir" and "Gwala" are one entry or three. The RG&CCI's "narrower margin" warning is therefore about a class of error the digital stack is structurally unable to validate.
  • No published standardisation dictionary exists ahead of February 2027 — unlike SC/ST enumeration, which is bounded by the notified Presidential Lists under Articles 341/342, the general caste column has no legal universe to validate against; the enumeration is caste-inclusive by Cabinet decision of 30 April 2025, not by a statutory schedule [2].

9. What Rides on Phase II That Did Not Ride on Phase I

  • Phase I errors are correctable; Phase II errors are consequential — HLO produces housing-amenity tables. PE produces the population figure that feeds delimitation of Lok Sabha and Assembly seats, the constitutional trigger being the first census after 2026 (Article 82) [7].
  • Caste counts become reservation arithmetic — OBC sub-categorisation, State quota ceilings and creamy-layer debates will be argued off these numbers; an undercount of a specific sub-caste is not a statistical footnote but a distributional loss with no appellate remedy under the Census Act, 1948.
  • Asymmetric political scrutiny — a household missed in house-listing is invisible; a caste undercounted in PE will be contested publicly by that community's organisations. This, not questionnaire length alone, is why the tolerance band genuinely narrows.
  • Sequencing risk — the non-synchronous belt (Ladakh, snow-bound J&K, parts of Uttarakhand and Himachal) ran PE in September 2026, months ahead of the February 2027 mainland count [2], so reference-date comparability across States rests on the standard-reference-moment adjustment rather than on simultaneity.

10. The Case That Self-Enumeration Improves Caste Accuracy

  • The strongest opposing argument — enumerator-mediated caste recording is precisely where assumption and bias enter (the risk the September 9 circular itself names). Self-enumeration removes the intermediary: the household types its own caste string, so the RG&CCI's gender- and bias-sensitisation worry shrinks as portal uptake rises.
  • Concede what is right — for a sensitive, self-ascribed attribute, respondent self-declaration is methodologically superior to a third party's inference, and the hybrid design (self-entry, then enumerator validate/edit/accept) preserves a coverage check the portal alone would lack.
  • But the scale answer is unflattering — about 82 lakh households used self-enumeration during HLO [2], against an all-India base of roughly 30 crore households; on that share the portal is a supplement, not a mitigation of enumerator-side error.
  • And it shifts the error rather than removing it — free self-declaration widens, not narrows, the string space: SECC's 46-lakh-name problem arose from respondent-declared entries in the first place [4]. Self-enumeration fixes bias; it worsens standardisation.
  • Residual validation gap — the enumerator may "edit" a self-submitted caste entry [3], which reintroduces third-party discretion at the exact field the sensitisation drive is meant to protect, without any recorded audit trail disclosed so far.

11. Fixes with a Named Owner

  • RG&CCI: publish the caste-name standardisation directory before February 2027 — SECC's mistake was ex-post classification via an expert group [5] after 46 lakh raw strings were already in the database [4]. A pre-loaded, State-specific auto-suggest list in the app, with free-text retained as a fallback, converts a post-facto cleaning problem into a point-of-capture one.
  • RG&CCI: log the enumerator edit — where an enumerator changes a self-submitted caste entry, the app should store both the original and edited string, making the accuracy claim auditable rather than asserted [3].
  • States: complete correction at source during the enumeration window — SECC pushed error correction to States after the field phase and 1.45 crore flagged corrections were never made [4]; CMMS real-time monitoring [2] allows the same validation to run while the enumerator is still in the ward.
  • MHA: commit to publishing the caste tables with a stated timeline — the governance failure of SECC was not collection but non-release; the Ministry of Social Justice and Empowerment's position remained that there was no proposal to release the SECC caste data [4]. A declared release calendar is what distinguishes 2027 from 2011.
  • RG&CCI: pre-empt the trainer-cascade dilution — training runs through master trainers down to ~30 lakh-plus field functionaries [2]; a standardised national video/module plus a mandatory in-app competency test for the caste and gender modules bounds the loss at each cascade layer better than State-level discretion does.

12. Anchors for Answers

  • Data: SECC 2011 generated ~46 lakh caste/sub-caste/surname/gotra names; 8.19 crore caste-related errors flagged to States, ~6.74 crore corrected, ~1.45 crore pending — and the caste tables were never published [4]
  • Data: ~82 lakh households used the Self-Enumeration Portal in Phase I (HLO) [2]
  • Report/Committee: Expert Group to classify SECC 2011 Caste/Tribe data, chaired by then NITI Aayog Vice-Chairperson Arvind Panagariya (Cabinet-approved, 2015) — classification never released [5][4]
  • Law/Case: Census Act, 1948 and Census Rules, 1990 (statutory basis); Article 82 — delimitation on the first census after 2026, which Census 2027 supplies [7]; Articles 341/342 — Presidential Lists bound SC/ST enumeration, but no comparable notified universe bounds the general caste column
  • Scheme: SECC 2011 — the closest Indian precedent, useful as a failure case on caste-name standardisation and on non-release of collected data [4]
  • Scheme: Census Management & Monitoring System (CMMS) — real-time supervisory layer intended to catch field errors during, not after, enumeration [2]

13. Mains Relevance

14. Related Topics to Study Next

  • Census Act, 1948 & Census Rules, 1990 — legal basis for census operations.
  • Socio-Economic Caste Census (SECC) 2011 — earlier attempt at caste data collection, useful comparison.
  • Reservation policy and OBC sub-categorisation — downstream policy use of caste data.
  • Delimitation exercise post-2026 Census — Census population figures feed into delimitation, a live UPSC theme.
  • Digital India / e-Governance initiatives — CMMS and app-based data collection fit this broader theme.
  • NPR (National Population Register) — related but distinct exercise, often confused with Census.
  • Right to Privacy (Justice K.S. Puttaswamy judgment) — relevant to sensitive personal/caste data collection.

15. Common Errors / Trap Areas

  • Confusing Census (under Census Act 1948, MHA/RGI) with NPR (under Citizenship Rules, also MHA) — distinct legal bases and purposes.
  • Assuming caste enumeration is legislated by Parliament — it is a Cabinet Committee decision (30 April 2025), not a fresh Act.
  • Mixing up phase names: Phase I = House Listing Operations; Phase II = Population Enumeration (caste data collected in PE, not HLO).
  • Assuming the RG&CCI's September 9, 2026 letter changed the Census law — it is an administrative/training circular, not a legal amendment.
  • Overlooking that self-enumeration is optional and supplementary, not a replacement for enumerator-led door-to-door visits.

Sources

  1. 1Registrar General and Census Commissioner of India addresses Press Conference on Census-2027pib.gov.in · tier 1
  2. 2Census 2027: India's First Digital Enumeration Exercisepib.gov.in · tier 1
  3. 3Lower margin for error in second phase of Census: Registrar-General, The Hindu (Vijaita Singh), September 16, 2026thehindu.com · tier 4
  4. 4Explained: What is caste census, when was it last held and why is it back? — Business Standardbusiness-standard.com · tier 4
  5. 5Cabinet approves setting up of an expert group to classify the Caste/Tribe data of the Socio Economic and Caste Census (SECC), 2011 — PIBpib.gov.in · tier 1
  6. 6Census 2027 pre-test: What India is testing before counting its citizens — Business Standardbusiness-standard.com · tier 4
  7. 7Explained: India's 2027 census to include caste count, trigger delimitation — Business Standardbusiness-standard.com · tier 4

Mains Q&A on this note

Also on 16 September

All 16 September articles →