I am looking for an AI-driven approach to digitizing photos of old handwritten journals and church documents from the 1600s and 1700s.
I have found models that can extract and read individual images, but the process quickly becomes a very long chat, especially when I keep adding more images. Extracting information from an image is one thing, but adding that information to individual profiles, creating profiles for people, mapping relationships, and storing the relevant details from the text becomes much more complicated. I would like to avoid copying information from one chat into another just to format it and manually create these records.
How would I approach a scenario like this? What tools could be useful? So far, I have used OpenRouter and a standard OpenWebUI setup, but this workflow does not seem very practical—at least not when simply copying images into the chat window and asking the model to extract the information.
I currently copy the extracted information into a Markdown file and try to organize everything that way.
I need help improving this entire workflow and automating as much of it as possible.
https://i.redd.it/exg13s67boqh1.png
Source: r/learnAIAgents · by /u/Ai_MOON_SHOT
