Position: Web Scraper / Python Developer
Hours: Monday–Friday, up to 45 hours per week
Work Hours: EST Shift – 6:00 PM to 03:00 AM (Fixed Shift)
Location: Office 801, 8th Floor, Cerebrum IT Park – B3, Kalyani Nagar, Pune – 411014
Length: Permanent position with a three-month probation period
Valasys Media is a globally acclaimed company based in Pune, specializing in Lead Generation, Lead Nurturing, Content Syndication, and Data Intelligence services.
This is a full-time on-site role for a Web Scraper / Python Developer based in Pune. The candidate will be responsible for developing and maintaining web scraping solutions, automation scripts, API integrations, data processing workflows, and database operations to support large-scale data collection and business requirements.
-
Strong Python programming and OOP concepts.
-
Web scraping using Selenium, Playwright, BeautifulSoup, Requests, and/or Scrapy.
-
Knowledge of HTML DOM, XPath, CSS selectors, JavaScript/AJAX, sessions, and cookies.
-
Experience with REST APIs, JSON, API authentication, headers, and parameters.
-
Strong knowledge of Pandas and data processing.
-
Data cleaning, validation, transformation, and deduplication.
-
Knowledge of PostgreSQL/MySQL/MongoDB.
-
Experience with FastAPI.
-
Linux and AWS EC2.
-
Git/GitHub and version control.
-
Understanding of automation, scheduling, retries, and background jobs.
-
Good debugging and problem-solving skills.
-
Experience with data enrichment and lead intelligence workflows is highly preferred.
-
Experience using AI/LLMs for classification, research, enrichment, entity extraction, or automation is an advantage.
-
Experience with approved email/outreach platform integrations is an advantage.
-
Qualification – Graduation / Post Graduation in Computer Science, IT, Engineering, or a related field.
-
Python development, web scraping, automation, or a similar role.
-
Strong hands-on experience with Python and web scraping tools.
-
Experience handling dynamic websites, pagination, sessions, cookies, and JavaScript-based content.
-
Good understanding of APIs and data processing.
-
Basic to good knowledge of databases and SQL/NoSQL databases.
-
Knowledge of Linux, AWS EC2, and Git/GitHub.
-
Ability to work with large-volume datasets.
-
Strong analytical and problem-solving skills.
-
Ability to work independently as well as with a team.
-
Willingness to work in a fixed EST shift from the Pune office.
-
Develop and maintain Python-based web scraping and automation scripts.
-
Build scraping solutions using Selenium, Playwright, BeautifulSoup, Requests, and/or Scrapy.
-
Handle dynamic websites, pagination, AJAX, sessions, cookies, XPath, and CSS selectors.
-
Integrate REST and third-party APIs.
-
Develop basic backend APIs using FastAPI.
-
Perform data cleaning, transformation, validation, and deduplication using Pandas.
-
Work with CSV, Excel, JSON, and database data.
-
Manage data using PostgreSQL, MySQL, or MongoDB.
-
Implement error handling, logging, retries, and monitoring.
-
Automate scheduled scraping and data-processing jobs using Cron/Celery or task queues.
-
Debug and optimize scraping processes for performance and reliability.
-
Deploy and manage Python applications on AWS EC2/Linux servers.
-
Maintain code using Git/GitHub.
-
Support internal data and automation requirements.