Budget: 15000 UAH Deadline: 10 days
Good day.
For the task of three-column layout and extracting blocks of text along with images, a custom coordinate parser can be written, but as a more reliable alternative, I suggest considering specialized APIs like AWS Textract or Google Document AI. They natively recognize complex multi-column layouts and provide a ready structure, which will significantly reduce the number of errors before sending the text for verification.
I will implement all server-side logic with routing, validation through Claude API in three attempts, and saving results on Node.js with Typescript. The admin interface for managing the queue of books, displaying statistics, and viewing logs for problematic sections will be built on Next.js.
In private messages, I will show examples of scripts for extracting data from documents with complex structures and integration with LLM API. I would be happy to review the extended technical assignment.