
Senior Python Data Scraping Engineer – Freelance
Posted 1 day ago

Posted 1 day ago
This is a fully remote position, open to applicants in Argentina.
• Manage comprehensive data extraction processes across intricate websites.
• Ensure thorough coverage, precision, and dependable delivery of organized datasets.
• Utilize tools and customized workflows to expedite data gathering, validation, and task execution.
• Extract information from dynamic and interactive web sources, including content rendered by JavaScript.
• Adapt scraping techniques in response to changing website behavior.
• Uphold data quality standards through validation checks and cross-source consistency measures.
• Adhere to formatting guidelines and systematically verify data prior to delivery.
• Scale scraping operations for extensive datasets using effective batching or parallel processing.
• Monitor for failures and maintain stability against minor alterations in site structure.
• Employ tools like Apify, OpenRouter, and other technologies along with technical expertise.
• A minimum of 5+ years of pertinent experience in data engineering, web scraping, automation, or software development.
• A Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical disciplines is preferred.
• Strong technical foundation and hands-on experience with scripting, automation, and data extraction processes.
• Proven ability to solve complex problems and work with contemporary development tools and technologies.
• Capability to systematically collect, organize, and validate data from various sources.
• Methodical and detail-oriented mindset.
• Ability to work autonomously.
• Extensive experience in Python web scraping utilizing BeautifulSoup, Selenium, or similar tools.
• Proficiency in handling dynamic content, including JavaScript, AJAX, and infinite scroll.
• Experience working with APIs through proxies.
• Competence in extracting data from intricate structures such as hierarchies, archived pages, and inconsistent HTML.
• Experience in data cleaning, normalization, and validation processes.
• Ability to provide structured datasets in formats like CSV, JSON, and Google Sheets.
• Proven experience managing anti-bot mechanisms and dynamic site architectures at scale.
• Familiarity with cloud infrastructure such as AWS or equivalent.
• Experience with containerization technologies like Docker.
• Practical experience with LLM frameworks such as LangChain, OpenRouter, or similar applied to automation tasks.
• Strong attention to detail and a commitment to data accuracy.
• Self-motivated work ethic with the ability to troubleshoot independently.
• Proficiency in English at an Upper-intermediate (B2) level or higher.
• A GitHub link is an advantage.
• Opportunity for freelance work.
• Part-time remote working arrangement.
• Estimated commitment of 10–20 hours per week during active project phases.
• Compensation of up to $25 per hour equivalent, based on experience and contribution pace.
RR Donnelley
plotdesk
CmdScale GmbH
Colsubsidio
Get handpicked remote jobs straight to your inbox weekly.