Experienced Web Scraping Engineer - Python (Remote)
Oxylabs Warsaw, Masovian Voivodeship, Poland · PLN 276K/yr
IT Services and IT Consulting · 201-500 employees
About the role
You will develop and maintain scalable web scraping and parsing solutions while improving system observability. The role involves defining resilient scraping strategies and managing high-traffic infrastructure.
What they look for
Requirements
Candidates must have strong experience with Python, computer science fundamentals, and web scraping techniques. Proficiency in network protocols, browser automation, and asynchronous programming is required.
Full description
We’re a team of 500+ professionals who develop cutting-edge proxy and web data scraping solutions for thousands of the world’s best known businesses, including Fortune 500 companies.
What’s in store for you:
You’ll be solving complex challenges and maintaining our own infrastructure with 60PB+ monthly data traffic. Here are its scale and maturity in numbers:
- 6PB+ Ceph storage
- 60PB+ monthly data traffic through our systems
- 300k+ service requests/sec processed
- 500k+ Kafka messages/sec streamed
A word from the team:
We run one of the most advanced and largest scraping and parsing products in the world. We serve thousands of requests per second with a very high success rate. Our scrapers and parsers are used by leading e-commerce, market intelligence, and AI industry players making the work challenging and truly global. The team is a blend of different interesting personalities from different walks of life and nationalities. Here you can find people who are experts in gaming, playing guitar, riding bicycles, and other areas. We, as a team, will support you in learning how to build your own scrapers and will share all the tips, tricks and hacks we know to ensure that you are onboard in no time.
\n
In this role, you’ll:
- Develop scalable scrapers.
- Define resilient scraping strategies, unblock websites for scrapingImprove observability in the system.
- Develop back-end solutions for scraping & parsing problems of various magnitudes.
- Maintain the current system and develop new features related to scraping & parsing.
Your skills & experience:
- Experience working with Python.
- Understanding of computer science, including data structures, algorithms, computability and complexity.
- Version Control skills using Git.
- Knowledge on how to unblock websites for scraping.
- Is able to use different scraping techniques & open-source tools to build scrapers.
- Is comfortable with using Dev Tools.
- Network (TLS/SSL) knowledge.
- Worked with browser automations.
- Knows their way around asynchronous programming.
Nice to have:
- Web development knowledge.
- Knows how to use CSS Selectors / XPaths for parsing.
- Experience working with Go & C++.
- Worked on browser source code.
- Knowledge of any front-end framework.
- Experience working with Pydantic, FastAPI, SQLAlchemy.
- Has experience working with Redis, MySQL, Docker, Kubernetes, Elasticsearch, Kibana and monitoring tools like Grafana, Prometheus.
- Experience with machine learning that is scraping domain-specific.
- Has experience in building scalable systems.
Salary:
- Gross salary: from 23 000 PLN/month. Keep in mind that we are open to discussing a different salary based on your skills and experience.
\nUp for the challenge? Let’s talk!
Similar roles
-
Python Full Stack Engineer - Professional
HEXAWARE United States
-
Experienced Software Engineer Java / Python (Full Stack or Back End)
JPMorgan Chase & Co. Plano, Texas, United States
-
Senior Python Engineer – Generative AI
Avenga Yerevan, Armenia
-
Senior Software Entwickler (m/w/d) – KI-gestützte Automatisierung | Python & .NET/C#
IUNworld GmbH Ismaning, Bavaria, Germany
-
Solution Architect (m/w/d) – KI-gestützte Automatisierung | Python & .NET/C#
IUNworld GmbH Ismaning, Bavaria, Germany
-
AI Engineer / Python Developer
Logicalis Spain Madrid, Community of Madrid, Spain · €45K–€50K/yr