Cross-cultural encounters in the Survey of India in the mid-nineteenth century

Collection Details
Now showing 1 - 4 of 4
  • Thumbnail Image
    Some of the metrics are blocked by your 
    Item type:Dataset,
    Survey of India Topographical Survey Report 1865 Comparison of Extracted Entities
    (2026-06-10)
    Rowlands, Huw
    This dataset contains a compilation of the tables extracted by ChatGPT, in parallel to the Python extraction workflow. In this file, the tables have been analysed to identify the longest-serving survey staff, the proportions of Native to to other survey staff, and a comparison of the areas of land they surveyed. The longest-serving staff identified became targets for the biographical research undertaken towards the end of the project. The project as a whole covers the years 1864 to 1873. It is an Excel file.
      3  1
  • Thumbnail Image
    Some of the metrics are blocked by your 
    Item type:Dataset,
    Survey of India Topographical Survey Reports 1864-1873 Images, TXT, and XML:
    (2026-06-10)
    Hossen-Mamode, Effie
    This dataset contains the source data from the Coleridge Fellowship 2025-26. Digital images of printed reports from the Survey of India's Topographical Survey from 1864 to 1873 were uploaded to Optical Character Recognition platform Transkribus. The images were transcribed there and two files exported for each report, one as txt and one as xml. The dataset comprises one image and xml file per page of each report, and one txt file per report, identified by its year of publication. There are ten subfolders, one for each report from 1864 to 1873.
      6  2
  • Thumbnail Image
    Some of the metrics are blocked by your 
    Item type:Dataset,
    Survey of India Topographical Survey Reports Complete Table-Only Extractions
    (2026-06-10)
    Rowlands, Huw
    This dataset contains a comparison of output data from The Coleridge Fellowship 2025-26. Python code was compiled to extract entities relating to people and places from xml files exported from the Transkribus Optical Character Recognition platform. The project as a whole covers the years 1864 to 1873. For one year, 1865, a subset of data was also extracted manually and using several online AI services. This dataset brings together the different outputs of people's names and of places from the Python, manual and AI extractions to facilitate comparison of the output data. It is an Excel file.
      3  1
  • Thumbnail Image
    Some of the metrics are blocked by your 
    Item type:Dataset,
    Survey of India Topographical Survey Reports 1864-1873 Extracted Entities
    (2026-06-10)
    Lloyd, Harry
    This dataset contains the output data from The Coleridge Fellowship 2025-26. Python code was compiled to extract entities relating to people and places from xml files exported from the Transkribus Optical Character Recognition platform. The data covers the years 1864 to 1873. It is in two formats - csv and gexf (for use with Gephi) - and includes the names, titles, years of service and survey parties of Survey of India surveyors, together with placenames relevant to their surveying work. A data dictionary is also provided.
      2  1