A 1900 U.S. Census record for the Richmond family serves as a benchmark for testing LLM transcription accuracy. Current models still struggle to correctly parse handwritten genealogical data without significant errors. This failure highlights a persistent gap in vision-to-text reliability for historical archives. Researchers must still manually verify AI-generated transcriptions.