{"schemaVersion":"jobsearcher.job.v1","id":"f3b81f0688fbddbc059e9fca","url":"https://jobsearcher.com/jobs/f3b81f0688fbddbc059e9fca","canonicalUrl":"https://jobsearcher.com/jobs/f3b81f0688fbddbc059e9fca","title":"Data Engineer Intern at Liba Space","description":"About Jobnova\r\nJobnova.ai is building the AI-powered job and people discovery infrastructure of the future. Our platform connects job seekers and companies through intelligent matching, AI agents, and large-scale data aggregation across job boards, social platforms, and talent networks.\r\nWe're looking for a Data Engineer Intern who is excited about working with real-world data at scale — web scraping, cleaning, structuring, and building robust data pipelines that power AI models and recommendations.\r\nThis role is perfect for someone who loves building systems, automating data collection, and turning messy data into reliable signals.\r\nResponsibilities\r\nBuild and maintain web scrapers to extract job postings, candidate profiles, company data, and social signals from multiple platforms.\r\nDesign and implement ETL / ELT pipelines to clean, normalize, and transform unstructured data into usable formats.\r\nWork with vector databases and structured stores to support retrieval, ranking, and AI matching.\r\nDevelop automated workflows for recurring data collection and enrichment.\r\nCollaborate with AI engineers to support model training, evaluation, and production inference.\r\nMonitor data quality, reliability, and scraper performance, and implement optimizations.\r\nExplore new data sources and propose scalable approaches for long-term data infrastructure.\r\nQualifications\r\nStrong programming skills in Python.\r\nExperience with web scraping tools / frameworks (Playwright, Selenium, Scrapy, BeautifulSoup, Apify, etc.).\r\nUnderstanding of data cleaning, parsing, and normalization techniques.\r\nKnowledge of databases (SQL, NoSQL) and familiarity with data pipeline tools.\r\nAbility to design and automate ETL workflows.\r\nCuriosity, fast learning ability, strong problem-solving skills.\r\nComfortable working in a fast-paced startup environment with ambiguity.\r\nJ-18808-Ljbffr","company":"Wayne State","rawCompany":"wayne state","city":"Madison","state":"WI","isRemote":false,"isActive":true,"createdAt":"2026-08-07T00:51:15.209Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"519290","title":"Web Search Portals and All Other Information Services","slug":"web-search-portals-and-all-other-information-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Engineer Intern at Liba Space","description":"About Jobnova\r\nJobnova.ai is building the AI-powered job and people discovery infrastructure of the future. Our platform connects job seekers and companies through intelligent matching, AI agents, and large-scale data aggregation across job boards, social platforms, and talent networks.\r\nWe're looking for a Data Engineer Intern who is excited about working with real-world data at scale — web scraping, cleaning, structuring, and building robust data pipelines that power AI models and recommendations.\r\nThis role is perfect for someone who loves building systems, automating data collection, and turning messy data into reliable signals.\r\nResponsibilities\r\nBuild and maintain web scrapers to extract job postings, candidate profiles, company data, and social signals from multiple platforms.\r\nDesign and implement ETL / ELT pipelines to clean, normalize, and transform unstructured data into usable formats.\r\nWork with vector databases and structured stores to support retrieval, ranking, and AI matching.\r\nDevelop automated workflows for recurring data collection and enrichment.\r\nCollaborate with AI engineers to support model training, evaluation, and production inference.\r\nMonitor data quality, reliability, and scraper performance, and implement optimizations.\r\nExplore new data sources and propose scalable approaches for long-term data infrastructure.\r\nQualifications\r\nStrong programming skills in Python.\r\nExperience with web scraping tools / frameworks (Playwright, Selenium, Scrapy, BeautifulSoup, Apify, etc.).\r\nUnderstanding of data cleaning, parsing, and normalization techniques.\r\nKnowledge of databases (SQL, NoSQL) and familiarity with data pipeline tools.\r\nAbility to design and automate ETL workflows.\r\nCuriosity, fast learning ability, strong problem-solving skills.\r\nComfortable working in a fast-paced startup environment with ambiguity.\r\nJ-18808-Ljbffr","datePosted":"2026-08-07T00:51:15.209Z","dateModified":"2026-08-07T00:51:15.209Z","hiringOrganization":{"@type":"Organization","name":"Wayne State","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Madison","addressRegion":"WI","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"f3b81f0688fbddbc059e9fca"},"url":"https://jobsearcher.com/jobs/f3b81f0688fbddbc059e9fca"}}