Data sources: the 20 public sources behind the database
Published 22 Aug 2026 · IndianSchoolsDatabase.com
Every record in the database is built from publicly published institutional contact information: the details schools, colleges and institutes publish so that they can be reached. 179,472 institutions are compiled from 20 source extracts in three categories, then deduplicated into one record per institution. Each record keeps the names of its sources and a source count, so on any row you can see how many independent extracts agree. The database was last updated on 21 Aug 2026.
The three source categories
Board affiliation registries (4 extracts)
Official affiliation records published by the Central Board of Secondary Education (three extracts: the affiliation repository, school detail pages and a legacy list) and the CISCE list of ICSE/ISC schools.
Trust: Highest: the board itself publishes the record.
Business and education directory listings (12 extracts)
Eleven extracts of publicly listed education businesses from a national business directory, covering schools, colleges and universities, coaching and training centres, research institutes and one state-level listing.
Trust: Medium: self-listed by the institution; cross-checked against other extracts where possible.
Curated school and college lists (5 extracts)
Public lists of schools and colleges, including city-level lists for Mumbai and a list of medical colleges.
Trust: Medium to low: used to fill gaps and corroborate, never as the sole source for a contact field.
How records are deduplicated
Records from every extract are normalised (name, location, phone, email) and merged into one record per institution; each merged record keeps the list of source keys and a source count so agreement between sources is visible on every row.
Where two extracts disagree on a field, the value from the higher trust source wins and the other is kept as evidence; where only one extract has a field, that value is used and the source count shows it stands alone. Board affiliation is recorded only where an official registry publishes it: 39,071 records carry CBSE and 4,599 ICSE/ISC.
What enrichment adds, and what it never does
After compilation, an enrichment pipeline fills gaps only from verifiable evidence: for example, the state of an institution is derived from the STD code embedded in its landline number, and an institution's type is classified from unambiguous vocabulary in its name. Every such change is recorded with its evidence and can be reversed. Nothing is generated from a pattern: no guessed email addresses, no inferred phone numbers.
Coverage today
- 77,760 institutions (43%) have an email address; 166,495 (93%) a phone number, 110,935 of them a mobile number.
- 105,459 name the principal or a primary contact; 22,275 have a website.
- 123,977 have state and district or city; 36 states and union territories and 668 districts are represented.
Refresh cadence
The headline numbers on this page refresh every hour from the live database. The dataset itself is re-enriched in batches, and every run is listed on the updates page. Each data refresh is posted on the updates page with its date and counts, and full access includes every refresh at no extra cost.
Corrections and removals
Institutions can ask for their record to be corrected or removed by emailing hello@indianschoolsdatabase.com; the data policy explains the process and the responsibilities of buyers who contact institutions from the list.
Try the database free
100 institutions with full contact details, every filter and export. ₹2,499 once for the schools database.
Create a free account