Indian datasets · Demographics
These are published records assigned the Demographics category label. Check each title, description and source to confirm the subject; licence and previews are shown when available.
Showing 25–48 datasets in Demographics.
708 extractive question-and-answer pairs built from Dh 2011 0401 Part a Dchb Chandigarh, published by censusindia.gov.in. The source is a 136-page publication. Source-evidenced topics include chandigarh, district census handbook, primary census abstract, 2011 census, census data, non-census data, chandigarh administration, capitol complex. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 282 factual, 147 table, 96 list, 79 definition, 64 summary, 26 relationship, 14 comparison. 279 of the 708 pairs (39.4%) ask explanatory questions; the remaining 429 are factual or table-based questions. 96.33% of rows pass the corpus quality gate.
275 extractive question-and-answer pairs built from Village & Townwise Primary Census Abstract, Karnal, Parts XIII A & B, Series-6, Haryana, published by censusindia.gov.in. The source is a 266-page publication. Source-evidenced topics include series-6 haryana, district census handbook, primary census abstract, karnal district, panipat, handloom industry, scheduled castes, scheduled tribes. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 133 factual, 68 list, 34 summary, 24 definition, 14 relationship, 2 comparison. 142 of the 275 pairs (51.6%) ask explanatory questions; the remaining 133 are factual or table-based questions. 97.45% of rows pass the corpus quality gate.
1422 extractive question-and-answer pairs built from Census of India 1891, A General Report, published by censusindia.gov.in. The source is a 326-page publication. Source-evidenced topics include general report, madras, bombay, central provinces, mysore, baroda, rajputana, urban population. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 740 factual, 257 relationship, 159 summary, 137 list, 68 comparison, 61 definition. 682 of the 1422 pairs (48.0%) ask explanatory questions; the remaining 740 are factual or table-based questions. 97.96% of rows pass the corpus quality gate.
600 extractive question-and-answer pairs built from Dh 2011 1024 Part a Dchb Munger, published by censusindia.gov.in. The source is a 636-page publication. Source-evidenced topics include district census handbook, munger, bihar, village directory, town directory, primary census abstract, c.d. block, population. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 359 factual, 82 definition, 82 list, 57 summary, 16 relationship, 4 comparison. 241 of the 600 pairs (40.2%) ask explanatory questions; the remaining 359 are factual or table-based questions. 97.33% of rows pass the corpus quality gate.
205 extractive question-and-answer pairs built from Dh 2011 1028 Part B Dchb Patna, published by censusindia.gov.in. The source is a 462-page publication. Source-evidenced topics include nagar parishad, persons males females, males females, persons males, females persons, census abstract, dinapur nizamat, primary census abstract. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 134 factual, 26 list, 25 summary, 12 definition, 6 relationship, 2 comparison. 71 of the 205 pairs (34.6%) ask explanatory questions; the remaining 134 are factual or table-based questions. 92.68% of rows pass the corpus quality gate.
65 extractive question-and-answer pairs built from Estimated Population by Castes, Madras, published by censusindia.gov.in. The source is a 34-page publication. Source-evidenced topics include 1951 census, madras, scheduled castes, scheduled tribes, backward classes, backward classes commission, hindus, muslims. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 39 factual, 14 relationship, 8 list, 2 comparison, 1 definition, 1 summary. 26 of the 65 pairs (40.0%) ask explanatory questions; the remaining 39 are factual or table-based questions. 92.31% of rows pass the corpus quality gate.
8 extractive question-and-answer pairs built from Estimated Population by Castes, 24, 1951, published by censusindia.gov.in. The source is a 13-page publication. Source-evidenced topics include castes, 1951 census, delhi, registrar general, india, scheduled castes, scheduled tribes, hindus, muslims. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 6 factual, 2 list. 2 of the 8 pairs (25.0%) ask explanatory questions; the remaining 6 are factual or table-based questions. 100.00% of rows pass the corpus quality gate.
12 extractive question-and-answer pairs built from Estimated Population By Castes, 12 Bombay, published by censusindia.gov.in. The source is a 27-page publication. Source-evidenced topics include castes, 1951 census, bombay, registrar general, india, scheduled castes, scheduled tribes, backward classes. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 8 factual, 2 summary, 1 list, 1 relationship. 4 of the 12 pairs (33.3%) ask explanatory questions; the remaining 8 are factual or table-based questions. 100.00% of rows pass the corpus quality gate.
325 extractive question-and-answer pairs built from Final Population Totals,, published by censusindia.gov.in. The source is a 579-page publication. Source-evidenced topics include 1961 census, final population totals, india, states, union territories, districts, population increase, administrative states. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 149 factual, 50 list, 46 summary, 44 relationship, 21 definition, 15 comparison. 176 of the 325 pairs (54.2%) ask explanatory questions; the remaining 149 are factual or table-based questions. 98.77% of rows pass the corpus quality gate.
8 extractive question-and-answer pairs built from Final Population of Tamil Nadu, published by censusindia.gov.in. The source is a 175-page publication. Source-evidenced topics include tamil nadu, 1991 census, registrar general, india, provisional population totals, final population totals, madras district, chengalpattu-mgr district, primary census abstract. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 4 factual, 3 list, 1 summary. 4 of the 8 pairs (50.0%) ask explanatory questions; the remaining 4 are factual or table-based questions. 100.00% of rows pass the corpus quality gate.
413 extractive question-and-answer pairs built from Bombay Presidency General Report, Part I, Vol-VIII, published by censusindia.gov.in. The source is a 666-page publication. Source-evidenced topics include bombay presidency, a. h. dbacu, h. t. borley, indian civil service, government central press, 1933, natural divisions, urban population. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 245 factual, 49 relationship, 43 summary, 40 list, 26 definition, 10 comparison. 168 of the 413 pairs (40.7%) ask explanatory questions; the remaining 245 are factual or table-based questions. 96.37% of rows pass the corpus quality gate.
475 extractive question-and-answer pairs built from 23902 1951 LA, published by censusindia.gov.in. The source is a 740-page publication. Source-evidenced topics include 1951 census, languages, tribal languages, indian languages, non-indian languages, mother-tongue, bi-lingualism, scheduled tribes. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 288 factual, 54 relationship, 53 list, 41 summary, 29 definition, 10 comparison. 187 of the 475 pairs (39.4%) ask explanatory questions; the remaining 288 are factual or table-based questions. 93.26% of rows pass the corpus quality gate.
73 extractive question-and-answer pairs built from Population According to Religion, Tables-6, Pakistan, published by censusindia.gov.in. The source is a 40-page publication. Source-evidenced topics include religion, population, muslim, hindu, christian, east bengal, west pakistan, baluchistan. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 33 factual, 18 summary, 14 relationship, 4 definition, 2 comparison, 2 list. 40 of the 73 pairs (54.8%) ask explanatory questions; the remaining 33 are factual or table-based questions. 98.63% of rows pass the corpus quality gate.
171 extractive question-and-answer pairs built from Religion, Series-1, published by censusindia.gov.in. The source is a 129-page publication. Source-evidenced topics include india, religion, population, hindus, muslims, christians, sikhs, buddhists. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 125 factual, 21 list, 10 summary, 9 relationship, 4 definition, 2 comparison. 46 of the 171 pairs (26.9%) ask explanatory questions; the remaining 125 are factual or table-based questions. 94.15% of rows pass the corpus quality gate.
24 extractive question-and-answer pairs built from Punjab, Vol-VI, published by censusindia.gov.in. The source is a 76-page publication. Source-evidenced topics include census enumeration, british india, community, province, state, district, tehsil, town. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 16 factual, 4 summary, 2 definition, 2 relationship. 8 of the 24 pairs (33.3%) ask explanatory questions; the remaining 16 are factual or table-based questions. 100.00% of rows pass the corpus quality gate.
263 extractive question-and-answer pairs built from Report, Part I, Vol-V, Bengal, published by censusindia.gov.in. The source is a 617-page publication. Source-evidenced topics include 1911, bengal, bihar, orissa, sikkim, royal statistical society, population, statistics. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 192 factual, 27 relationship, 15 list, 15 summary, 13 definition, 1 comparison. 71 of the 263 pairs (27.0%) ask explanatory questions; the remaining 192 are factual or table-based questions. 98.10% of rows pass the corpus quality gate.
732 extractive question-and-answer pairs built from Report, Part -I, Vol-I, India, published by censusindia.gov.in. The source is a 543-page publication. Source-evidenced topics include 1931, j. h. hutton, l. s. vaidyanathan, population, provinces, states, migration, age distribution. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 507 factual, 97 relationship, 46 summary, 42 list, 21 comparison, 19 definition. 225 of the 732 pairs (30.7%) ask explanatory questions; the remaining 507 are factual or table-based questions. 95.90% of rows pass the corpus quality gate.
1 extractive question-and-answer pairs built from Report on the census of Assam For 1881, published by censusindia.gov.in. The source is a 432-page publication. Source-evidenced topics include have been, per cent, has been, garo hills, deputy commissioner, those who, hill tribes, assam valley. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 1 factual. 100.00% of rows pass the corpus quality gate.
497 extractive question-and-answer pairs built from Report of the Census of Bengal, Vol-I, published by censusindia.gov.in. The source is a 317-page publication. Source-evidenced topics include bengal civil service, calcutta, enumeration, village lists, district officers, survey mouzahs, thannahs, circles. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 326 factual, 52 relationship, 49 summary, 43 list, 16 definition, 11 comparison. 171 of the 497 pairs (34.4%) ask explanatory questions; the remaining 326 are factual or table-based questions. 94.16% of rows pass the corpus quality gate.
99 extractive question-and-answer pairs built from State Census Handbook, Vol-I, Manipur, published by censusindia.gov.in. The source is a 89-page publication. Source-evidenced topics include manipur, state census handbook, naga hills, cachar, lushai hills, burma, part c state, deputy commissioner. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 56 factual, 17 summary, 15 list, 6 definition, 4 relationship, 1 comparison. 43 of the 99 pairs (43.4%) ask explanatory questions; the remaining 56 are factual or table-based questions. 91.92% of rows pass the corpus quality gate.
2 extractive question-and-answer pairs built from Census Of India 1941, Vol-IV, Bengal, published by censusindia.gov.in. The source is a 18-page publication. Source-evidenced topics include appendix bengal caste tables, r. a. dutch, bengal, scheduled castes, agarwallas, rajputs, baruis, tribes. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 1 definition, 1 factual. 1 of the 2 pairs (50.0%) ask explanatory questions; the remaining 1 are factual or table-based questions. 100.00% of rows pass the corpus quality gate.
1573 extractive question-and-answer pairs built from Census of India 1951, Appendices to the Census Report, 1951, Volume I, Part !-B, India, published by censusindia.gov.in. The source is a 424-page publication. Source-evidenced topics include 1951, population, land use, cultivation, rainfall belts, irrigation, mineral production, foodgrains. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 916 factual, 206 relationship, 181 list, 141 summary, 77 definition, 52 comparison. 657 of the 1573 pairs (41.8%) ask explanatory questions; the remaining 916 are factual or table-based questions. 94.41% of rows pass the corpus quality gate.
160 extractive question-and-answer pairs built from Dh 2011 1034 Part B Dchb Gaya, published by censusindia.gov.in. The source is a 504-page publication. Source-evidenced topics include bihar, gaya district, district census handbook, mahabodhi temple, bodh gaya, unesco world heritage site, siddhartha gautama, bodhi tree. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 105 factual, 21 summary, 19 list, 6 definition, 6 relationship, 3 comparison. 55 of the 160 pairs (34.4%) ask explanatory questions; the remaining 105 are factual or table-based questions. 98.12% of rows pass the corpus quality gate.
931 extractive question-and-answer pairs built from Dh 35 2001 and, published by censusindia.gov.in. The source is a 370-page publication. Source-evidenced topics include nicobars, village directory, town directory, primary census abstract, district census handbook, andamans district, nicobars district, community development block. Every answer is a verbatim span of text the source prints, and each row carries the passage it sits in, its offset in that passage, the source quote, the page and the location in the document, so any row can be checked against the original. Question types in the export: 551 factual, 146 list, 99 definition, 87 summary, 34 relationship, 14 comparison. 380 of the 931 pairs (40.8%) ask explanatory questions; the remaining 551 are factual or table-based questions. 97.85% of rows pass the corpus quality gate.