发布于 2026年8月31日 · 我们于 2026年8月31日 确认该职位仍然有效
₹ 12.500 – ₹ 37.500 (每个项目)
A batch of multi-page PDF files needs to become a clean, analysis-ready Excel workbook. Each text element—whether it appears as straight digital text or inside a scanned image—has to land in the correct cell so the spreadsheet mirrors the logical flow of the original documents. The job is straightforward but accuracy is critical: no dropped characters, no shifted columns, no merged cells where they don’t belong. I’m happy for you to use any reliable method—Python (tabula-py, camelot, pdfplumber), Power Query, Acrobat automation, ABBYY FineReader OCR, or a manual approach—so long as the end result is an .xlsx file I can filter and pivot without cleanup. Deliverables • One Excel workbook containing every record from the supplied PDFs, organised consistently and ready for immediate use. • A brief note (or script) describing the extraction method so the process is reproducible if new PDFs arrive later. Final file must be double-checked for completeness; spot checks should prove 100 % coverage and less than 1 % transcription error. If that sounds routine to you, let’s get started.