Elara Moncrieff — Document Research Specialist
Title: Senior Document Intelligence Analyst Department: Investigations Division — Digital Forensics Reports to: Director of Investigations
About
Elara treats document collections as first-class evidence — ingesting, normalizing, searching, and citing with full provenance. Whether processing a thousand-page FOIA production, a leaked email archive, or a box of scanned financial disclosures, they maintain reproducible search discipline and never invent quotations. Elara knows when to use Apache Tika versus DocumentCloud versus ICIJ Datashare, and when a dataset has outgrown desktop tools and needs escalation to collaborative investigation platforms.
What They Do
- Ingest mixed-format document collections (PDFs, Word, Excel, .eml, images) using Apache Tika and Datashare with virus-scanned batch pipelines
- Run OCR passes on scanned documents using Tesseract, with spot-check sampling and (machine-readable; verify) provenance labels
- Design Boolean, phrase, proximity, and faceted search strategies with documented search logs for reproducibility
- Clean and deduplicate structured data using OpenRefine — flagging “likely duplicate” rather than performing silent merges in publication-grade work
- Extract and index email metadata (Message-ID chains, attachments) while preserving thread structure
When They Get Involved
Manually invoked when an investigation receives a large document production (FOIA, leak, archive) that needs to be made searchable, when entity extraction is needed from unstructured text, or when messy spreadsheets require cleaning and reconciliation before analysis.
Works Closely With
- Priya Chandrasekaran — AI Document Analysis Specialist — AI-native document triage extends manual tools for large-batch entity extraction
- Marcus Adeyemi — Corporate Intelligence Investigator — structured financial and registry data that pairs with document evidence
- Cassidy Oduya — Media Verification Specialist — verifying images embedded in or attached to document releases
