{"schemaVersion":"jobsearcher.job.v1","id":"f0218cf892fef465fb2c6a4d","url":"https://jobsearcher.com/jobs/f0218cf892fef465fb2c6a4d","canonicalUrl":"https://jobsearcher.com/jobs/f0218cf892fef465fb2c6a4d","title":"Data Engineer - Web Scraping","description":"This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Data Engineer - Web Scraping based in Greece.\r\nThis role is an exciting opportunity to build scalable web scraping solutions and data pipelines that power business-critical insights. You will work closely with analysts, engineers, and cross-functional teams to develop reliable datasets that support strategic decision-making. With significant ownership over your projects, you'll design, automate, and optimize data collection workflows while ensuring data quality and reliability. The environment encourages innovation, collaboration, and continuous learning, giving you the freedom to experiment with new technologies and improve existing processes. If you enjoy solving complex data challenges and building automation at scale, this role offers an excellent platform to make a meaningful impact.\r\nAccountabilities: As a Data Engineer specializing in web scraping, you will design, develop, and maintain automated data collection systems while ensuring the quality, accuracy, and availability of large-scale datasets. You will collaborate across teams to build efficient, reliable, and scalable data solutions.\r\nCollaborate with analysts and stakeholders to understand data requirements and deliver tailored data solutions.\r\nDesign, develop, and maintain web scrapers for a wide range of structured and unstructured data sources.\r\nClean, transform, validate, and manipulate large datasets using Python and Pandas.\r\nBuild and maintain data ingestion pipelines into databases or data warehouses.\r\nSchedule, monitor, and optimize scraping workflows using orchestration tools such as Apache Airflow.\r\nDevelop quality control checks to ensure data integrity, consistency, and availability.\r\nInvestigate and resolve data pipeline issues and time-sensitive production incidents.\r\nDesign and enhance internal tools, automation frameworks, and platform capabilities to improve operational efficiency.\r\nWork closely with cross-functional engineering teams to implement scalable and maintainable data processing workflows.\r\nRequirements: The ideal candidate combines strong software engineering skills with hands-on experience in web scraping, data processing, and automation. Success in this role requires both technical expertise and a proactive, problem-solving mindset.\r\nBachelor's or Master's degree in Computer Science or a related technical discipline.\r\n2–4 years of professional software development experience.\r\nStrong programming skills in Python and solid SQL/database knowledge.\r\nAdvanced experience using the Pandas library for data cleaning, transformation, and analysis.\r\nExperience working with web technologies, including HTML, JavaScript, APIs, and related protocols.\r\nProven experience processing, cleaning, and transforming large datasets.\r\nFamiliarity with web scraping frameworks and tools such as Selenium, Scrapy, XPath, Fiddler, or Postman.\r\nExperience with workflow orchestration tools such as Apache Airflow or similar platforms.\r\nKnowledge of Docker containerization; Kubernetes experience is an advantage.\r\nFamiliarity with CI/CD tools such as Jenkins or GitLab CI/CD.\r\nExperience working with cloud services, particularly AWS technologies such as S3, RDS, Lambda, SNS, or SQS, is preferred.\r\nStrong analytical thinking, attention to detail, communication skills, and a passion for automation and continuous improvement.\r\nBenefits: Opportunity to work on challenging projects supporting a leading global asset management environment.\r\nHigh level of ownership and autonomy in a collaborative, team-oriented culture.\r\nExposure to modern data engineering, web scraping, cloud, and automation technologies.\r\nCollaborative environment with experienced engineering, product, and data professionals.\r\nOpportunities for continuous learning, professional growth, and technical skill development.\r\nMerit-driven culture that values innovation, initiative, and individual contributions.\r\nFlexible, technology-focused environment with opportunities to work on impactful data products.\r\nWe appreciate your interest and wish you the best!\r\nLI-CL1\r\nJ-18808-Ljbffr","company":"JobGether","rawCompany":"jobgether","city":"Greece","state":"NY","isRemote":false,"isActive":false,"createdAt":"2026-08-08T01:18:23.124Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Engineer - Web Scraping","description":"This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Data Engineer - Web Scraping based in Greece.\r\nThis role is an exciting opportunity to build scalable web scraping solutions and data pipelines that power business-critical insights. You will work closely with analysts, engineers, and cross-functional teams to develop reliable datasets that support strategic decision-making. With significant ownership over your projects, you'll design, automate, and optimize data collection workflows while ensuring data quality and reliability. The environment encourages innovation, collaboration, and continuous learning, giving you the freedom to experiment with new technologies and improve existing processes. If you enjoy solving complex data challenges and building automation at scale, this role offers an excellent platform to make a meaningful impact.\r\nAccountabilities: As a Data Engineer specializing in web scraping, you will design, develop, and maintain automated data collection systems while ensuring the quality, accuracy, and availability of large-scale datasets. You will collaborate across teams to build efficient, reliable, and scalable data solutions.\r\nCollaborate with analysts and stakeholders to understand data requirements and deliver tailored data solutions.\r\nDesign, develop, and maintain web scrapers for a wide range of structured and unstructured data sources.\r\nClean, transform, validate, and manipulate large datasets using Python and Pandas.\r\nBuild and maintain data ingestion pipelines into databases or data warehouses.\r\nSchedule, monitor, and optimize scraping workflows using orchestration tools such as Apache Airflow.\r\nDevelop quality control checks to ensure data integrity, consistency, and availability.\r\nInvestigate and resolve data pipeline issues and time-sensitive production incidents.\r\nDesign and enhance internal tools, automation frameworks, and platform capabilities to improve operational efficiency.\r\nWork closely with cross-functional engineering teams to implement scalable and maintainable data processing workflows.\r\nRequirements: The ideal candidate combines strong software engineering skills with hands-on experience in web scraping, data processing, and automation. Success in this role requires both technical expertise and a proactive, problem-solving mindset.\r\nBachelor's or Master's degree in Computer Science or a related technical discipline.\r\n2–4 years of professional software development experience.\r\nStrong programming skills in Python and solid SQL/database knowledge.\r\nAdvanced experience using the Pandas library for data cleaning, transformation, and analysis.\r\nExperience working with web technologies, including HTML, JavaScript, APIs, and related protocols.\r\nProven experience processing, cleaning, and transforming large datasets.\r\nFamiliarity with web scraping frameworks and tools such as Selenium, Scrapy, XPath, Fiddler, or Postman.\r\nExperience with workflow orchestration tools such as Apache Airflow or similar platforms.\r\nKnowledge of Docker containerization; Kubernetes experience is an advantage.\r\nFamiliarity with CI/CD tools such as Jenkins or GitLab CI/CD.\r\nExperience working with cloud services, particularly AWS technologies such as S3, RDS, Lambda, SNS, or SQS, is preferred.\r\nStrong analytical thinking, attention to detail, communication skills, and a passion for automation and continuous improvement.\r\nBenefits: Opportunity to work on challenging projects supporting a leading global asset management environment.\r\nHigh level of ownership and autonomy in a collaborative, team-oriented culture.\r\nExposure to modern data engineering, web scraping, cloud, and automation technologies.\r\nCollaborative environment with experienced engineering, product, and data professionals.\r\nOpportunities for continuous learning, professional growth, and technical skill development.\r\nMerit-driven culture that values innovation, initiative, and individual contributions.\r\nFlexible, technology-focused environment with opportunities to work on impactful data products.\r\nWe appreciate your interest and wish you the best!\r\nLI-CL1\r\nJ-18808-Ljbffr","datePosted":"2026-08-08T01:18:23.124Z","dateModified":"2026-08-08T01:18:23.124Z","hiringOrganization":{"@type":"Organization","name":"JobGether","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Greece","addressRegion":"NY","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"f0218cf892fef465fb2c6a4d"},"url":"https://jobsearcher.com/jobs/f0218cf892fef465fb2c6a4d"}}