We are looking for a specialist to refine the automation processes in the OLX parser.
Requirements for the specialist
Confident knowledge of Python
Knowledge of Python 3.10 (or similar versions)
Experience in developing parsers (web scraping) in Python, including using the Selenium library
Experience with Selenium and managing browsers in containers
Ability to configure Selenium WebDriver (Chrome/Firefox) in Docker containers
Understanding the specifics of running and configuring browsers without a graphical interface (headless mode)
Knowledge of REST API and client-server interaction principles
Ability to understand existing REST API and modify/optimize it if necessary
Experience with tokens, cookies, authorization, etc.
Practical experience in debugging and bypassing anti-spam/anti-bot systems
Understanding of rate limits, timeouts, proxies, and other tools for stable parser operation
Experience with version control systems and containerization
Git (GitHub, GitLab)
Docker (basic skills: building images, setting up environments, docker-compose, etc.)
Will be a plus: CI/CD (GitLab CI, GitHub Actions, etc.)
Additional qualities
Ability to work with someone else's code: refactoring, documentation
Willingness to closely interact with the team and the client to clarify requirements
Self-organization and responsibility for results
Analysis of the current solution:
Study the existing OLX parser on Python 3.10, determine the reasons for the non-working authorization process.
Refactoring authorization:
Fix/redesign the login process, ensure correct operation with cookies/sessions.
Transition to Selenium:
Transfer the main parsing process from REST API to Selenium using a managed browser in Docker (headless mode).
Set up authorization and data collection through Selenium.
The cost of completing the task is negotiable.
Python Developer for improving the parser
We are looking for a specialist to improve the automation processes in the OLX parser. Task: fix authorization, set up work with cookies/sessions, migrate parsing to Selenium (Docker, headless mode).
Requirements:
✅ Proficient in Python 3.10
✅ Experience in web scraping (Selenium, REST API)
✅ Setting up Selenium WebDriver in Docker
✅ Working with tokens, cookies, authorization
✅ Understanding of rate limits, proxies, anti-spam systems
✅ Git, Docker (building images, docker-compose)
Will be a plus:
🔹 CI/CD (GitLab CI, GitHub Actions)
🔹 Experience in code refactoring and documentation
📩 If you have the necessary experience – write to us! 🚀
Log in
or
register,
to view the original