Experienced Web Scraping Engineer - Python

Oxylabs Io — Poland · Posted ~1 hour ago

Senior Full-time Remote

Skills

Python web scraping web data parsing infrastructure maintenance Ceph Kafka

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join an experienced engineering team building and maintaining large-scale web scraping and data infrastructure. You will tackle complex Python-based scraping and parsing challenges, work with high-volume distributed systems, and help maintain reliable services operating at significant scale. The role offers remote work, technically demanding problems, and collaboration with an international group of specialists.

Highlights

Work remotely on technically challenging, large-scale web data infrastructure with substantial traffic and high request volumes. Collaborate with an international team and solve complex engineering problems in a globally impactful environment.

Description

We’re a team of 500+ professionals who develop cutting-edge proxy and web data scraping solutions for thousands of the world’s best known businesses, including Fortune 500 companies. What’s in store for you: You’ll be solving complex challenges and maintaining our own infrastructure with 60PB+ monthly data traffic. Here are its scale and maturity in numbers: 6PB+ Ceph storage 60PB+ monthly data traffic through our systems 300k+ service requests/sec processed 500k+ Kafka messages/sec streamed A word from the team: We run one of the most advanced and largest scraping and parsing products in the world. We serve thousands of requests per second with a very high success rate. Our scrapers and parsers are used by leading e-commerce, market intelligence, and AI industry players making the work challenging and truly global. The team is a blend of different interesting personalities from different walks of life and nationalities. Here you can find people who are experts in gaming, playing guitar, riding bicycles, and other areas. We, as a team, will support you in learning how to build your own scrapers and will share all the tips, tricks and hacks we know to ensure that you are onboard in no time. In this role, you’ll: Develop scalable scrapersDefine resilient scraping strategies, unblock websites for scrapingImprove observability in the systemDevelop back-end solutions for scraping & parsing problems of various magnitudesMaintain the current system and develop new features related to scraping & parsing Your skills & experience: Experience working with PythonUnderstanding of computer science, including data structures, algorithms, computability and complexityVersion Control skills using GitKnowledge on how to unblock websites for scrapingIs able to use different scraping techniques & open-source tools to build scrapersIs comfortable with using Dev ToolsNetwork (TLS/SSL) knowledgeWorked with browser automationsKnows their way around asynchronous programming Nice to have: Web development knowledgeKnows how to use CSS Selectors / XPaths for parsingExperience working with Go & C++Worked on browser source codeKnowledge of any front-end frameworkExperience working with Pydantic, FastAPI, SQLAlchemyHas experience working with Redis, MySQL, Docker, Kubernetes, Elasticsearch, Kibana and monitoring tools like Grafana, PrometheusExperience with machine learning that is scraping domain-specificHas experience in building scalable systems Salary: Gross salary: from 23 000 PLN/month. Keep in mind that we are open to discussing a different salary based on your skills and experience Up for the challenge? Let’s talk! We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.