
Closed
Posted
I have a PDF of roughly 11-50 pages that combines narrative text with several embedded tables. I only need certain fields and rows pulled out—not every single word or cell—then organised cleanly in an Excel workbook. Accuracy matters more than speed; the resulting .xlsx file should let me sort, filter, and run simple formulas without further cleanup. You can tackle the job with any method you prefer—Python (tabula-py, camelot), Power Query, Adobe export, or even careful manual entry—as long as the final sheet is consistent with the source and preserves the original numerical precision and wording. Deliverables • One Excel file containing: – Sheet 1: the specified tables in their original column order – Sheet 2: the targeted text segments, each linked to its page number Acceptance criteria • 100 % of the requested data captured, no extra material • Column headers and units exactly match the PDF • Page references included so I can cross-check quickly Let me know your usual turnaround for a project of this size and any clarifying questions you may have about the sections to extract.
Project ID: 40654451
38 proposals
Remote project
Active 8 hours ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
38 freelancers are bidding on average ₹931 INR/hour for this job

Hey there Glane here, I can accurately extract the specified tables and targeted text segments from your 11–50 page PDF using Python with pdfplumber/Camelot/Tabula, correcting extraction issues where necessary and preserving the original wording, column order, headers, units, and numerical precision. I’ll deliver an Excel workbook with Sheet 1 containing the requested tables and Sheet 2 containing the selected text with page references, fully structured for sorting, filtering, and formulas. I’ll also perform a page by page validation to ensure 100% of the requested fields are captured with no unnecessary material, and can typically turn around a project of this size within 1–2 days, depending on PDF complexity.
₹1,000 INR in 40 days
6.3
6.3

★★★ TOP 1% IN FREE LANCER WORLD ★★★ ★★★ 20+ Year Experience in IBD being CMD★★★ ★★★ 200+ Country Satisfied Clientele ★★★ ★Linkedin★ ★Data Entry★ ★Business Plans★★★ ★★★ Operational Strategic planner Customer Support 24*7★★★ ★★★Excel/Word Operation★★★ ★★★Chat Support★★★ ★★★Calling Support★★★ ★★★Business Plans / Marketing Strategy ★★★ * Digital Marketing★★★ ★★★Social Media Marketing ★★★ ★★★Internet Marketing ★★★ ★★★Any type of Data Projects★★★ ★★★★★★ Regards, ★★★CMD★★★ ★★★PVSYS GROUP (INDIA)★★★ ★★★IF YOU THINK THEN I CAN★★★
₹750 INR in 40 days
6.2
6.2

Your PDF extraction will fail if the tables span page breaks or contain merged cells—most automated parsers misalign rows when headers repeat across pages. This creates formula errors that cost hours to fix manually. Quick questions - are any of your tables split across multiple pages? And do you need the extraction logic reusable for future PDFs with the same structure? Here is the architectural approach: - PYTHON: Build a hybrid parser using camelot for structured tables and pdfplumber for narrative text, with validation logic that flags mismatches between extracted values and source coordinates. - DATA PROCESSING: Implement cell-level verification that compares extracted numerical precision against PDF metadata to catch rounding errors before Excel export. - EXCEL: Structure the workbook with named ranges and data validation rules so your formulas reference stable cell addresses even if you add rows later. I've built similar extraction pipelines for 4 fintech clients processing regulatory filings where a single decimal error triggers compliance failures. Let's schedule a 10-minute call to review a sample page before I scope the full build.
₹900 INR in 30 days
5.6
5.6

I can wirite chstom python script to do ocr extraction from given pdf. ping me for more details and i can complete the project in 1day. Thanks
₹900 INR in 40 days
5.2
5.2

Hi there, Accuracy and clean formatting are my top priorities for data extraction tasks like this. I can deliver a structured, formula-ready .xlsx workbook matching your exact specifications. My Approach: Hybrid Pipeline: I use Python (pdfplumber / camelot) alongside regular expression parsing for semi-structured text, followed by rigorous manual cell-by-cell validation against the original PDF. Sheet 1 (Structured Tables): Original column headers, precise numerical values, correct data types (numeric, dates, text), and exact column orders preserved so sorting and formulas work immediately without cleanup. Sheet 2 (Targeted Text): Cleanly mapped narrative fields paired with their respective source page numbers for instant cross-verification. Turnaround Time: For an 11–50 page document with selective extraction, my typical turnaround is 24 to 36 hours from the moment the PDF and field guidelines are shared. Quick Clarifications to Align: Which specific sections, row filters, or key tables are required? Is the PDF digital/vector-based (selectable text), or a scanned document? I am ready to review the document and begin right away. Best regards, Sagar
₹1,000 INR in 40 days
4.4
4.4

Here is a professional and compelling bid tailored for this project, keeping the same strong structure as your previous one. Hi, I would love to help you extract and organize your PDF data into a clean, structured Excel workbook with high precision. Here is what I will deliver for your project: Accurate Data Extraction: Careful extraction of your specific targeted tables and narrative text segments from the 11–50 page PDF, completely free of unwanted clutter. Two-Sheet Excel Workbook: Sheet 1: The specified tables in their exact original column order, preserving numerical precision, units, and headers. Sheet 2: The targeted text segments cleanly organized and cross-referenced with their respective page numbers. Ready-to-Use Format: An impeccably formatted .xlsx file designed for immediate sorting, filtering, and formula use without any extra cleanup needed on your end. I typically turn around a project of this size within [Insert Number] days, prioritizing absolute accuracy and consistency over speed. Could you please clarify which specific sections, fields, or rows you would like prioritized for extraction? I would appreciate a quick 5-minute chat to go over the details. Best regards, Hossam
₹1,000 INR in 40 days
4.6
4.6

Hi, Thanks for posting "Extract Specific PDF Data" — it lands right in our wheelhouse, and we'd be thrilled to bring it to life. I love turning a solid brief like yours into a polished result you’ll be proud to put your name on. From your brief I can see this involves excel, data processing — all areas we handle in-house. We specialise in Python, Data Processing, Excel, Software Architecture, which lines up directly with what you need. How we'd approach it: - Define the exact fields, sources and output format - Build the collection / processing pipeline - Validate accuracy and clean the data - Deliver in your preferred format with a short summary I also noted you don’t need further cleanup — that’s clear, and I’ll keep the work clean and strictly on-brief. Delivering at the scale of 50 pages is no problem for us — we're set up for volume without dropping quality. Happy to work hourly with transparent time tracking and regular check-ins. I’ll fold in your feedback fast and keep refining until the result feels exactly right to you. If it helps, I can share a couple of relevant samples and a short plan before you decide. Looking forward to it! Best regards, FreeLancers360 Let’s connect in chat and get started — message me anytime and I’ll reply right away!
₹750 INR in 5 days
4.1
4.1

Hi, I’m Yousef, a Python developer experienced with data processing, structured extraction, and automation. I can extract only the required fields and table rows from your PDF and deliver a clean Excel workbook that is ready for sorting, filtering, and formulas. My approach would be to first identify the exact fields you need, then parse the text and tables carefully, validate the extracted values against the source PDF, and structure the final .xlsx consistently. I can also flag any pages or rows where the PDF formatting makes extraction ambiguous instead of silently inserting incorrect data. If you send the PDF and the list of required fields, I can review the structure and start immediately. Best, Yousef
₹750 INR in 40 days
3.7
3.7

Hi, I can extract only the required fields, rows, tables, and targeted text segments from your 11–50 page PDF into a clean Excel workbook. My approach will be to first review the PDF, required fields, target rows, page references, table structure, and output format. Then I’ll use the most accurate method for the file type, such as Python/Camelot/Tabula, Adobe export, Power Query, or manual verification, to ensure the final Excel file matches the source exactly. I’m comfortable with PDF data extraction, Excel formatting, table extraction, page-based text capture, Python tools, manual validation, and accuracy-focused cleanup. Deliverables: * One clean Excel workbook * Sheet 1 with specified tables * Original column order preserved * Headers and units matched exactly * Sheet 2 with targeted text segments * Page numbers included * Numerical precision preserved * No extra unrelated data * Final cross-check before delivery Typical turnaround: 1–2 days after receiving the PDF and extraction instructions, depending on table complexity. I’ll focus on accurate, clean, filter-ready output so you can sort, check, and run formulas without further cleanup. Best regards Ankit
₹750 INR in 40 days
3.9
3.9

Hi, I will extract the specific fields and rows from your PDF and organise them into one Excel file: Sheet 1 for the tables in original column order, Sheet 2 for the targeted text segments with page references. I can start today. I will use Python with camelot or tabula, cross-checking manually to preserve exact headers, units, and numerical precision. Questions: 1) Can you mark which tables and text sections to pull? 2) Should page numbers sit in a separate column? Looking forward to discussing further. Regards, Shayan.
₹925 INR in 40 days
2.5
2.5

I see you're looking to extract data from a PDF with both text and tables. I have solid experience with Python for this kind of task. What specific data points are you interested in extracting?
₹1,350 INR in 7 days
2.5
2.5

I can handle this by combining automated PDF extraction with manual validation to ensure the Excel output matches the source exactly, including column headers, units, wording, and numerical precision. For documents in the 11–50 page range, my usual approach is: - Extract tables using Python tools such as Camelot or Tabula when the PDF structure allows it - Identify and isolate only the requested rows/fields instead of exporting unnecessary content - Capture the required narrative text segments together with their page references - Validate the final workbook manually to eliminate formatting drift, missing rows, or OCR inconsistencies The resulting .xlsx file will be organized for practical use, including sortable/filterable tables and clean cell formatting suitable for formulas without additional cleanup. Deliverables will include: - Sheet 1 with the selected tables in original column order - Sheet 2 with extracted text snippets and page references - Consistent formatting and cross-checkable structure Typical turnaround for a PDF of this size is 1–2 days depending on table complexity and scan quality. Before starting, I would just need clarification on which sections, fields, and rows should be included or excluded from extraction.
₹1,250 INR in 2 days
2.7
2.7

Hello, reviewed your project. I can accurately extract the specified tables and text sections from your 11–50 page PDF and deliver a clean, ready-to-use Excel workbook. I’ll preserve the original column order, headers, units, wording, and numerical precision, while adding page references for the requested text. I’ll validate the extracted data against the PDF before delivery to ensure nothing is missed or added. **Deliverable:** One `.xlsx` workbook with the two requested sheets. **Turnaround:** 24 hours for a typical 11–50 page document, depending on table complexity.
₹750 INR in 40 days
1.8
1.8

Hello, i read your requirement. I have experience in excel and done many projects i give you best work on your time and budget. Thanks, waiting for your response...
₹800 INR in 40 days
1.9
1.9

Hello, I hope this message finds you well. I am interested in the position of Extract Specific PDF Data and believe my skills in Python, data processing, Excel, and data analysis make me a strong candidate for this job. I am confident in my ability to accurately extract the required data from the PDF and organize it cleanly in an Excel workbook. I understand the importance of accuracy and consistency in delivering the final Excel file. I am comfortable using various methods to extract the data and ensure that the resulting sheet is consistent with the source. I look forward to discussing my usual turnaround time for a project of this size and any clarifying questions you may have about the extraction process. Thank you. Best regards, Winston
₹1,000 INR in 40 days
0.0
0.0

Hello, I can accurately extract the required tables and targeted text from your PDF and organize them into a clean Excel workbook. I will carefully ensure that: - Only the requested data is extracted - Table columns remain in the original order - Headers, units, and numerical precision match the PDF - Targeted text segments include their corresponding page references - The final workbook is clean, consistent, sortable, and filterable I can handle this carefully through manual data extraction and will cross-check the completed workbook against the source PDF for accuracy. For a PDF of 11–50 pages, I can start immediately. Once I review the PDF and the exact sections to be extracted, I can confirm the turnaround time for the first batch and complete file. Could you please share the PDF and indicate the specific fields, rows, or sections you would like extracted? Thank you.
₹800 INR in 20 days
0.0
0.0

Hi, I can accurately extract the required tables, fields, rows, and targeted text segments from your 11–50 page PDF and organize everything into a clean, structured Excel workbook. I’ll preserve the **original column order, headers, units, numerical precision, and wording**, while ensuring the final workbook is fully sortable, filterable, and formula-ready. I’ll use Python-based PDF/table extraction where appropriate, followed by manual validation to catch formatting or OCR issues. The workbook will include: * **Sheet 1:** Requested tables only, in original structure * **Sheet 2:** Targeted text segments with corresponding page numbers I’ll cross-check the extracted data against the PDF to ensure no requested information is missed and no unnecessary content is included. **Typical turnaround:** 1–2 days, depending on the number and complexity of tables. Please share the PDF and specify the exact sections/fields required.
₹1,000 INR in 15 days
0.0
0.0

Hello, I came across your project on Freelancer and I’d love to help. I’m an experienced data analyst with a proven track record of delivering high-quality work that meets client expectations. I have over 5 years of experience in data extraction and manipulation. I’ve successfully completed similar projects, such as extracting data from complex PDFs for research purposes, ensuring precision and clarity in the final output. I focus on clear communication and timely delivery. For your project, I can: - Deliver an Excel file containing the specified tables and targeted text segments, organized as per your requirements. - Provide regular updates and revisions until you’re satisfied. - Ensure the final product maintains the original numerical precision and wording, with all necessary page references for cross-checking. I’d be happy to discuss your project in more detail and share relevant samples of my past work. Looking forward to collaborating with you. Regards, Shaun Kelly.
₹750 INR in 7 days
0.0
0.0

As an experienced data entry practitioner, I've gained keen skills in precise data extraction, a necessity for your project. I understand the importance of the information you seek and always take great pains to ensure that every piece is extracted with absolute accuracy. Over the years, I have honed my proficiency by employing various methods such as tabula-py and Power Query, indicating that I can tackle your project using diverse approaches or even manual entry if required. Moreover, my familiarity with Excel is extensive. This tool served as more than just a companion for my work, allowing me to become adept at organizing and presenting data coherently without compromising on its originality. This guarantees that 100% of the requested data will be captured, structured in their exact column order and headers while maintaining consistent textual references. In terms of turnaround time, I am known for efficiency without compromising on accuracy. For a project of this size, I believe it would take me [insert preferred timeframe]. Rest assured that no matter how tight the schedule may be, I will not sacrifice the quality of deliverables. Feel free to reach out to me for any further clarifications or queries
₹1,000 INR in 24 days
0.0
0.0

I notice you need accurate, timely data entry, and as a Freelancer-Verified Data Entry Specialist who successfully cleared the platform exam, I am ready to start immediately. My technical background allows me to handle complex formatting, while my typing speed of 75+ WPM ensures your documents will be completed ahead of schedule. I do not just copy data; I cross-check formatting and validate entries to ensure zero errors. Could you send over a sample page or describe the document format in chat? I can complete a quick 1-line sample for you right now to prove my accuracy. Best regards, Mehalingam A
₹950 INR in 38 days
0.0
0.0

JAMNAGAR, India
Member since Dec 17, 2020
₹12500-37500 INR
₹750-1250 INR / hour
$15-25 USD / hour
₹750-1250 INR / hour
₹12500-37500 INR
$3000-5000 USD
₹12500-37500 INR
₹12500-37500 INR
₹750-1250 INR / hour
$30-250 USD
₹1500-12500 INR
₹1500-12500 INR
₹600-1500 INR
$5000-10000 USD
$25-50 USD / hour
₹600-1500 INR
$8-15 USD / hour
₹12500-37500 INR
₹1500-12500 INR
₹12500-37500 INR