
Closed
Posted
Paid on delivery
I want to push our resume-parsing engine to its limits by feeding it deliberately tricky files. I will send you a clean résumé in plain text; your job is to engineer multiple adversarial versions of it—both DOCX and PDF right from the start—using every document-level sleight of hand you know. Hidden text, layered images covering text, font or glyph substitution, OOXML rewrites that mis-align the visual and logical order, even extraction quirks that only show up when the file is flattened to PDF: anything that could fool an ATS should appear somewhere in the series. I am flexible on tooling—Python, C#, or a mix—so choose whichever lets you script the generation process cleanly and document it for repeatability. The end goal is twofold: realistic files that open without warning in Word/Reader, and a transparent methodology I can re-run as we harden our software. Please provide: • A stepped set of adversarial resumes (easy, moderate, hard, extreme) in both DOCX and PDF • Well-commented source code and any helper scripts that create or post-process the files • A concise manifest mapping each file to the exact manipulation techniques used • Notes on test methodology so my team can extend the suite later I’ll review portfolios that show prior work with DOCX/PDF programmatic generation, OOXML hacking, PDF object editing or similar parser-focused QA. When you reply, include a short summary of relevant projects, a rough timeline for producing the first batch, and your fixed cost or milestone breakdown. Let’s make our parser bullet-proof.
Project ID: 40648241
36 proposals
Remote project
Active 3 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
36 freelancers are bidding on average ₹3,321 INR for this job

Hi Sir, I am from Banglore.I have 7 years of experience in Python software engineering.I will build solution using Python based solution and build similar earlier for other client project. I have worked here with 130+ clients. Let’s connect
₹5,000 INR in 2 days
6.4
6.4

Hi, Your project "DOCX/PDF Resume Manipulation Test Suite" is a good fit -- Python backend and automation work is my main line of work. How I would run it: 1. Confirm the inputs, outputs and edge cases in writing first, so there is no ambiguity about what the script or service has to handle. 2. Build it in small reviewable pieces with tests around the parts that touch real data, rather than one large drop at the end. 3. Deliver clean, documented code with a requirements file and setup notes, so you or another developer can run and extend it without me. Matching your listed skills: C# Programming, Software Testing, Automation, Scripting, Python, C++ Programming, Software Architecture, OCR. My bid is ₹3400, within your ₹2000-4000 range. Let's connect to discuss this further -- happy to walk you through how I would structure it and answer anything you want covered first. Thanks for your time. Best regards, Ashish & Team
₹3,400 INR in 7 days
5.4
5.4

With extensive experience in OCR and Python, I am confident in my ability to create a comprehensive adversarial test suite for your resume-parsing engine. My portfolio showcases previous work with DOCX/PDF programmatic generation, as well as an understanding of how to manipulate OOXML and PDF files to challenge parsers like yours. The toolset flexibility you've mentioned aligns perfectly with my skill set; I will leverage Python script automation to precisely engineer adversarial versions of your clean résumé. Rather than just delivering a bunch of manipulated resumes, I provide coded transparency and methodology that you can not only use for this project but also extend later as you harden your software. My milestones include providing a concise manifest mapping each file manipulation technique used, notes on the test methodology, and well-commented source codes with any helper scripts necessary for post-processing. Ultimately, my aim is to help bullet-proof your ATS' parsing capabilities using realistic files that actually open seamlessly in Word/Reader. With my dedication for high-quality software development and a passion for building solutions that scale, I'm committed to producing a top-tier test suite that truly challenges your resume-parsing engine capability developed in precise timelines and pricing to ensure the success of this project. Let’s discuss further how we can make your parser resilient together!
₹3,000 INR in 7 days
5.0
5.0

Hello, I can build a reproducible document-generation test suite that creates controlled adversarial DOCX and PDF resume variants for parser/ATS testing. I would organize the suite into progressive difficulty levels: - Easy - Moderate - Hard - Extreme Each generated document would have a manifest describing the exact manipulation applied, allowing your team to identify which parser behavior is being tested. The implementation can use Python-based document generation and PDF processing tools, with OOXML-level manipulation where necessary. I would also separate generation, transformation and validation into independent components so additional test cases can easily be added later. Deliverables would include: - DOCX/PDF adversarial test files - Source code and helper scripts - Technique manifest - Reproduction instructions - Testing methodology - Notes for extending the suite Estimated delivery: 10 days for the first complete batch. I would prioritize deterministic test cases so the same input produces reproducible outputs, which is important when this is used as a regression suite for your parser. Best regards, Albert.
₹2,000 INR in 10 days
4.8
4.8

Hi, I enjoy this kind of parser-stress work. I can script adversarial resume generation in Python (python-docx / reportlab or LaTeX-to-PDF, plus OOXML-level rewrites) to produce the DOCX and PDF variants with hidden text, glyph substitution, layered images and visual/logical misalignment, documented so your team can re-run and extend the suite. I will deliver the stepped set (easy/moderate/hard/extreme), commented source, a manifest mapping each technique, and methodology notes. Could you share the parser's current extraction approach (which ATS or libraries it uses) so I can target the weaknesses that matter most for your stack?
₹2,500 INR in 3 days
4.0
4.0

Hi, I can create the DOCX/PDF adversarial resume test suite for your resume parser using scripted, repeatable document-generation and manipulation methods. My approach will be to first review your clean plain-text resume, parser behavior, target edge cases, and required difficulty levels. Then I’ll generate easy, moderate, hard, and extreme versions in both DOCX and PDF with a clear manifest explaining each manipulation. I’m comfortable with Python/C# scripting, DOCX generation, OOXML structure edits, PDF generation/editing, parser QA, hidden text tests, layout manipulation, extraction-order issues, font/glyph edge cases, layered content, and automation. Deliverables: * Easy/moderate/hard/extreme DOCX files * Matching PDF versions * Scripted generation workflow * Well-commented source code * Helper post-processing scripts * Manifest of techniques per file * Test methodology notes * Extension guide for future cases I’ll focus on realistic files that open normally in Word/Reader while giving your parsing engine strong controlled test cases for visual-vs-logical text, layout, extraction, and document-structure robustness. Timeline: first batch in 2–3 days after receiving the clean resume. Best regards Ankit
₹4,000 INR in 1 day
4.2
4.2

Hi, I can build a reproducible adversarial resume test suite covering both DOCX/OOXML and PDF formats, with progressively harder cases from easy to extreme. I’ll start from your clean resume and generate controlled variations targeting hidden content, visual/logical text-order mismatches, layered objects, font/glyph edge cases, OOXML structure variations, and PDF extraction behavior. Each file will remain valid and open normally in Word/Reader. I’ll provide the generation/post-processing scripts, a technique-by-technique manifest, expected extraction behavior, and clear instructions for reproducing and extending the test suite. I’ll also keep the original resume content as ground truth so your team can directly compare parser output against expected results. First batch: 2–3 days, with the final timeline depending on the number of adversarial cases required. I’m comfortable working in Python with DOCX/OOXML and PDF processing tools and can structure the implementation for repeatable QA rather than one-off files.
₹3,000 INR in 7 days
4.2
4.2

Hi, I can help you build a repeatable adversarial test suite for your resume-parsing engine, covering DOCX and PDF manipulation from basic extraction edge cases through highly deceptive document structures. I have experience with Python, document automation, PDF/DOCX generation, parser-focused QA, API testing, and building reproducible test utilities. For this project, I’d structure the suite into **Easy → Moderate → Hard → Extreme**, with each level introducing progressively more challenging techniques such as hidden/overlaid content, font/glyph variations, reading-order inconsistencies, OOXML-level changes, and PDF object/extraction quirks. Deliverables will include: • Adversarial resumes in both DOCX and PDF • Clean, well-commented Python-based generation/post-processing scripts • A manifest mapping every file to its exact manipulation techniques • Validation checks to ensure files open normally in Word/Reader • Testing methodology and extension guidelines for your QA team I’ll keep the entire process deterministic and repeatable so you can regenerate the suite whenever the parser changes. I’m ready to start with the clean résumé and build the test framework around your parser’s current weaknesses and target extraction behavior.
₹3,000 INR in 7 days
3.8
3.8

I understand you need a controlled adversarial resume dataset to stress-test a resume parser with difficult DOCX and PDF structures while keeping the generation process repeatable for QA. I build document automation and testing tools using Python, scripting, PDF/DOCX generation, and parser-focused workflows. I can create a structured test suite with multiple difficulty levels, generate files programmatically, document each manipulation method, and provide reproducible scripts for your team. I have experience with backend automation, data processing, and production systems where reliability and validation are critical. For this project, I can focus on safe parser testing scenarios such as document structure variations, extraction order issues, formatting differences, metadata handling, and OCR-related edge cases. I will provide the generated files, source scripts, manifest documentation, and testing notes so your team can expand the suite later.
₹5,000 INR in 2 days
3.5
3.5

Hi there, let's have short meeting if you wanna discuss the test scope quickly. I can build a repeatable DOCX/PDF adversarial resume test suite using Python and OOXML/PDF tooling. I’ll create easy, moderate, hard, and extreme variants covering hidden text, image overlays, font/glyph tricks, XML order issues, extraction edge cases, and PDF flattening behavior. I’ll also provide clean commented scripts, a manifest for every manipulation, and short methodology notes so your team can easily add new cases later. I’ll make sure the generated files open normally in Word/Reader and are useful for parser/OCR QA. Relevant work: document automation, structured file generation, parser testing, and scripting for repeatable QA workflows. Budget: $150 Timeline: 5 days
₹15,000 INR in 5 days
2.7
2.7

Understood! Crafting adversarial DOCX and PDF versions of a resume to challenge your parsing engine is right up my alley. With expertise in Python and C# and experience in manipulating document formats for parser testing, I am equipped to meet your project needs effectively. I'll create a series with varying complexity levels, thoroughly document the generation processes, and provide a detailed manifest and notes for future extensions. Excited to help bullet-proof your system! Anticipating two weeks for the first batch, I can provide a detailed cost breakdown upon further discussion.
₹3,000 INR in 7 days
2.5
2.5

Hi, This is an interesting match for my Python and QA experience. I can turn your clean résumé into a structured adversarial test suite designed to expose differences between what a document visually shows and what an ATS/parser actually extracts. I’ll build the files in progressive levels—easy, moderate, hard and extreme—covering techniques such as hidden text, image overlays, font/glyph variations, document structure/order manipulation, and PDF extraction edge cases. Each generated DOCX/PDF will remain openable in normal Word/Reader workflows. I’ll also provide the reusable generation/post-processing scripts, a manifest mapping every file to its exact manipulation techniques, and concise methodology notes so your team can reproduce and extend the suite. My approach will be systematic: generate → verify the document visually → extract/inspect the underlying content → document the expected adversarial behavior. This will make the test set useful for regression testing as you harden the parser. I can provide the first batch within 1–2 days and complete the full suite in approximately 3–4 days. Fixed price: ₹4,000 Available to start immediately.
₹2,500 INR in 2 days
2.2
2.2

Hi, I will build a stepped adversarial resume suite in DOCX and PDF, using hidden text, layered images, glyph substitution, and OOXML rewrites to stress your parser. Python scripts will generate each file with a manifest mapping techniques per level. I can start today. For the PDF layer, I will edit objects directly so extraction quirks appear only after flattening. Questions: 1) Python only, or a Python/C# mix? 2) Should the manifest be Markdown or CSV? Looking forward to discussing further. Regards, Shayan.
₹2,700 INR in 3 days
2.2
2.2

Your goal is hardening the parser, so the value here is not "weird files" but a corpus where every single file maps to a named technique your team can re-run after each change. What I would build: a tier ladder (easy, moderate, hard, extreme), each tier in both DOCX and PDF, techniques stacking as you go up: white and zero-size text, off-canvas and behind-image text, OOXML rewrites where drawing order and logical run order disagree, text split across runs and rPr boundaries so keyword matching breaks, glyph and CMap substitution so extraction yields different characters than the render, image-only pages whose text layer contradicts the OCR, and flatten-only artifacts that surface only after the DOCX is printed to PDF. Generator: Python, python-docx plus direct OOXML editing for what the library will not express, pikepdf and reportlab on the PDF side. Fully scripted, one command regenerates the whole suite from your clean plain-text resume, so you can re-target it at a new resume or extend a tier later without me. Manifest: a CSV/JSON table mapping each file to the techniques used, the expected failure mode, and what a correct parse should return. That last column is what turns this into a regression suite instead of a pile of samples. Every file opens clean in Word and Reader with no repair prompt, and I verify that before hand-off, it is easy to break by accident when editing OOXML by hand. Relevant background: I work in Python document and data pipelines daily, and I have 12 merged pull requests into third-party open-source projects, mostly a 182-star Go security tool, each reviewed and accepted by the maintainers. That is the closest match to this job, adversarial thinking aimed at a parser plus code somebody else has to read and extend. This account also has one completed project rated 5 out of 5, delivered on time and on budget. Timeline and cost: first batch in 2 days (easy and moderate tiers, both formats, generator skeleton and manifest), full ladder with methodology notes in 5 days. Fixed INR 3500 for the whole scope, split 40% at first batch and 60% at final delivery. One question so I aim the extreme tier where it pays off: does your engine read PDFs through an OCR path, a text-layer extraction path, or both? Petro Pankov, BotCraft Group
₹3,500 INR in 5 days
1.5
1.5

The core requirement is a repeatable DOCX/PDF test suite that exposes parser weaknesses while keeping every generated file valid and openable. I’d begin with the clean résumé and establish a controlled baseline, then build the manipulations incrementally so each file has a known, isolated purpose. Using Python and Software Testing, I’d script the generation of easy, moderate, hard, and extreme variants, covering document structure, text extraction, visual/logical ordering, fonts, overlays, and PDF conversion behavior. Each output would be paired with a manifest describing exactly what was changed. I’d also keep the source modular and commented so your team can add new cases without rebuilding the workflow. The first batch can be produced within 3 days, followed by the remaining variants, documentation, and methodology notes. The next step is to provide the clean résumé and confirm the parser environments you want the suite tested against.
₹3,000 INR in 2 days
0.0
0.0

? High-Impact Action Plan for "DOCX/PDF Resume Manipulation Test Suite" I reviewed your requirements and I am ready to deliver outstanding results for you. Experienced Full-Stack Developer skilled in web applications, mobile apps, React, Node.js, and Python. ✅ Deliverables & Execution: 1. Project Setup & Architecture Review (Days 1-2) 2. Core Development & Integration (Days 3-4) 3. Quality Testing & Final Launch (Day 5) Let's connect via chat to discuss your specific goals and get started! Best regards, todayintech
₹3,700 INR in 5 days
0.0
0.0

Hi, I’m a Software Engineer with experience in software testing, debugging, automation, and building reproducible test workflows. This project is a strong match for my background because it focuses on finding edge cases, reproducing parser failures, and turning them into structured test cases. I can create a progressive adversarial resume suite in both DOCX and PDF: Easy: formatting, tables, columns, unusual fonts Moderate: hidden text, overlapping objects, text boxes, extraction-order tricks Hard: OOXML modifications, run/glyph manipulation, visual vs. logical ordering Extreme: combined techniques and PDF-specific extraction inconsistencies You’ll receive the DOCX/PDF samples, well-commented Python scripts, a manifest describing every manipulation, and testing notes so your team can extend the suite later. I can deliver the first Easy/Moderate batch within 1–2 days and complete the full suite in 4 days. Fixed bid: ₹3,500 INR Delivery: 4 days My focus will be making every test reproducible, realistic, and useful for identifying weaknesses in your parser.
₹3,000 INR in 7 days
0.0
0.0

Hi Client, I haven't built document forensics tools specifically, but I'm highly proficient at using AI coding assistants to generate clean, production-ready test generator Python scripts. For this project, I'll script the generation process using python-docx, direct OOXML ZIP rewriting, and PyPDF with custom cmap/ToUnicode manipulation to produce 8 adversarial resumes (easy/moderate/hard/extreme × DOCX + PDF). Techniques will include zero-width characters, invisible text, font glyph redirection, OOXML logical-vs-visual misalignment, and PDF object-level ToUnicode garble—all files will open without warnings in Word/Reader. Deliverables: 8 files + commented source code + manifest mapping techniques + methodology notes. Cheers,
₹2,000 INR in 1 day
0.0
0.0

I've built several document parsing test suites and have hands-on experience with OOXML structure manipulation and PDF object-level editing. Here's what I'll deliver: a Python-based generation framework using python-docx and reportlab for controlled DOCX/PDF creation, plus targeted post-processing scripts for advanced tricks like hidden text layers, glyph substitution, and logical/visual misalignment. You'll get four progression levels (easy through extreme) in both formats, fully commented source code organized for repeatability, a detailed manifest CSV mapping each file to its exact techniques, and a methodology guide so your team can extend the suite with new adversarial patterns. I'll make sure every file opens cleanly in Word and Reader while containing the edge cases your parser needs to handle. Fixed cost 3500 INR for the complete suite with full documentation.
₹2,020 INR in 4 days
0.0
0.0

I would treat this as a repeatable parser QA suite, not a collection of one-off files. I can generate paired DOCX/PDF cases in Python, preserve files that open cleanly in Word and Reader, and provide a manifest that maps each sample to its expected extraction failure. The first batch can cover hidden text, overlays, reading-order changes, font/glyph substitution, and PDF flattening quirks within 48 hours. I would then use your parser results to refine the moderate, hard, and extreme cases. Full scripted suite, source, manifest, and test notes: 5 days at INR 2,000. Which extraction engines or ATS parsers should the baseline target?
₹2,000 INR in 5 days
0.0
0.0

Lakhimpur kheri, India
Member since May 14, 2025
₹1500-12500 INR
₹1500-12500 INR
₹1500-12500 INR
$8-15 USD / hour
$15-25 USD / hour
₹37500-75000 INR
₹750-1250 INR / hour
$30-250 USD
$25-50 USD / hour
$15-25 USD / hour
$500 USD
₹12500-37500 INR
$30-250 USD
$30-250 NZD
$250-750 USD
$30-250 USD
$30-250 USD
₹12500-37500 INR
$30-250 USD
₹600-1500 INR
$25-50 USD / hour
$250-750 USD
₹600-1500 INR