Free guide: How to Get the Best Answer from AI(PDF download)
01 — Practice

AI Evaluator, Researcher & Speaker

I test whether AI can be trusted with facts, history, and public information.

I evaluate AI tools on real documents and real questions to find where they get things wrong, and I teach people how to check for themselves.

Open to full-time roles in AI policy, governance, research, and evaluation · Authorized to work in the US and UK · Open to relocation

02 — About

A historian who ended up inside AI systems — and started asking historian questions.

Katryna Peart is an AI historiographer and evaluator who researches how AI systems handle long-form civic and institutional documents. She developed Civic Pair Prompting (CPP) and the Civic AI Evaluation Standard (CAES) to identify where those systems lose context, nuance, and trustworthiness.

Whose voice is this? What narrative is being pushed? Where does this flatten the truth? Through hands-on evaluation work and original research on how language models handle civic and historic documents, Katryna developed a practice around the gap between what AI claims and what it actually does. Katryna's work sits at the intersection of AI historiography and institutional memory — studying how generative systems construct, compress, and sometimes erase the records that communities depend on to understand themselves.

Her research focuses on the civic record: municipal documents, public histories, commemorative narratives, and institutional archives. She is the developer of Civic Pair Prompting (CPP), a replicable framework for evaluating municipal AI, and the Civic AI Evaluation Standard (CAES), a governance suite built for local government. Her work has appeared in Governing Magazine, Route Fifty, and PM Magazine. She holds an MA in Medieval and Modern History from Royal Holloway, University of London, and a BA in History from NYU.

03 — Writing

Selected writing.

04 — Credentials

Track record & recognition.

Katryna holds an MA in Medieval and Modern History from Royal Holloway, University of London, and a BA in History from NYU. She has done AI evaluation work with Google, Uber AI, and Mercor, and received a PMJA Award for public communications for the Newark 350 commemoration. She serves as Board Member (External Liaison) for Medievalists of Color, and is based in Leander, Texas, available to travel.

Speaking engagements

DateTalkVenue
Sep 24, 2026RAG Testing Holds: Evaluating LLMs, Faithfulness, Boundaries & TrustSTARWEST 2026 · Anaheim, CA
Oct 1, 2026Making Technology Work for EveryonePT Exchange · Online
Oct 5, 2026How to Get the Best Answer from AIColumbia University · AI, Information, and Meaning module
Nov 22, 2026The Digital Priesthood: Archival Silence and the Cumulative Compression of Pre-Print KnowledgeArs Inquirendi · St Edmund Hall, Oxford
Dec 15–17, 2026Civic Pair Prompting: Evaluating AI Failure Modes in UK Local Government DocumentsBCS-SGAI AI-2026 · Peterhouse, Cambridge
05 — Contact

Let’s work together.

I'm open to full-time roles in AI governance, policy analysis, organizational research, and AI evaluation, as well as fellowships and research collaborations.

I also give talks and guest lectures: see my speaking page.

Authorized to work in the US and UK; open to relocation.