{"schemaVersion":"jobsearcher.job.v1","id":"cee811c73e14f59bacc34033","url":"https://jobsearcher.com/jobs/cee811c73e14f59bacc34033","canonicalUrl":"https://jobsearcher.com/jobs/cee811c73e14f59bacc34033","title":"Data Engineer","description":"PURPOSE\n\nThe Data Engineer is responsible for designing, building, and maintaining the data pipelines and infrastructure that power Baker Group's Microsoft Fabric data warehouse, serving as the organization's single source of truth. This role owns the ingestion, transformation, and orchestration of data from disparate internal systems (ERP, HRIS, MRP and other structured data sources) into governed, reliable data products used by Data Analysts and developers to deliver insights to executive and operational teams and ensures that data is structured to support both traditional reporting and emerging AI and machine learning use cases. The Data Engineer curates and maintains core datasets spanning employees, finance, construction and manufacturing projects, and service, and partners with the Data Scientist, Data Analyst, and Software Development roles to ensure data is trustworthy, well-structured, and fit for downstream use.\n\nESSENTIAL FUNCTIONS AND RESPONSIBILITIES\n\nThe following duties are typical for this job. These are not to be constructed as exclusive or all inclusive. Other duties may be required and assigned.\n\nDesigns, builds, and maintains ETL/ELT pipelines that ingest data from enterprise systems into Microsoft Fabric.\nArchitects and maintains the Fabric medallion Lakehouse structure (bronze, silver, gold layers) as Baker Group's single source of truth.\nDevelop and implement best practices for the data infrastructure and environment (e.g. Development/Test/Production environments, Git for version control).\nOwns pipeline orchestration, scheduling, and monitoring to ensure reliable, timely, and accurate data availability.\nCurates and maintains core datasets across employee, finance, project, service, and manufacturing domains.\nEstablishes and enforces data quality, validation, and reconciliation processes across all pipelines.\nDesigns and manages data models, schemas, and semantic layers that support Data Analyst reporting and Data Scientist modeling work.\nDefines and maintains data ontologies and canonical business definitions (for example, what constitutes a \"project,\" \"employee,\" or \"cost code\") to ensure consistent meaning across systems and consumers.\nPrepares and structures data to support AI and machine learning use cases, including feature-ready datasets, retrieval-augmented generation (RAG) pipelines, and vector embedding storage.\nManages Fabric capacity planning, workspace organization, and performance optimization.\nImplements data governance practices, including access controls, lineage tracking, and metadata management, consistent with Baker Group's data classification standards.\nPartners with business system owners (ERP, HRIS, MRP, etc.) to understand upstream data structures and manage change impacts.\nCollaborates with the Data Scientist to ensure pipeline outputs support analytical and machine learning use cases.\nCollaborates with Data Analysts to ensure data products support paginated reporting, dashboards, and self-service BI needs.\nCollaborates with Software Development and DevOps Teams to ensure data products support application development needs.\nCoordinates with 3rd party consultants when necessary to deliver data engineering projects and augment capacity for demanding business needs.\nDevelops and maintains documentation for pipelines, schemas, and integration logic.\nTroubleshoots and resolves pipeline failures, latency issues, and data quality incidents.\nMonitors and maintains data-specific infrastructure, including Fabric capacity, pipeline orchestration tools, and monitoring/alerting systems.\nEvaluates and recommends new data engineering tools, patterns, and best practices.\nStays current on emerging trends in data engineering, cloud data platforms, and integration techniques.\n\nMINIMUM EDUCATION and EXPERIENCE REQUIRED TO PERFORM ESSENTIAL FUNCTIONS\n\nBachelor's degree in Computer Science, Data Engineering, Information Systems, or other relevant quantitative field\nThree to five years of experience in data engineering, ETL/ELT development, or a related field\nProficiency with SQL and database technologies for data extraction, transformation, and loading\nExperience with Microsoft Fabric, Azure Data Factory, or similar cloud ETL/orchestration tools\nExperience with medallion architecture and modern data warehousing patterns\nExperience with a programming language such as Python, PySpark, or T-SQL for data transformation\nFamiliarity with data modeling techniques (dimensional modeling, star schema)\nUnderstanding of data governance, data quality, and metadata management practices\nExperience preparing data for AI/ML consumption (e.g., vector embeddings, RAG architectures) is a plus\nBusiness acumen and understanding of construction or related industries is a plus\n\nCERTIFICATES, LICENSES, REGISTRATIONS\n\nNo specific requirements; however, relevant certifications such as Microsoft Certified: Fabric Data Engineer Associate, Azure Data Engineer Associate, or similar cloud platform certifications are a plus\n\nMENTAL AND PHYSICAL COMPETENCIES REQUIRED TO PERFORM ESSENTIAL FUNCTIONS\n\nStrong analytical and troubleshooting skills with the ability to diagnose and resolve complex pipeline and data quality issues\nExcellent time and project management skills with the ability to prioritize across multiple pipeline and infrastructure projects\nCurrent with industry trends in data engineering, cloud platforms, and integration best practices\nStrong communication skills with the ability to translate technical data structures for non-technical stakeholders\nTeam player with strong collaboration skills, particularly with the Data Scientist, Data Analysts, and business system owners\nMust be able to focus on complex technical problems and work independently with minimal supervision\nAbility to work in a fast-paced environment and adapt to changing business priorities\nMeticulous attention to detail and commitment to producing reliable, well-documented data infrastructure\n\nENVIRONMENTAL ADAPTABILITY\n\nProlonged periods of sitting at a desk and working on a computer\nMust be able to lift 10 pounds occasionally\nMay have occasional visits to a job site which would require periods of standing, walking and/or climbing stairs\n\nEQUIPMENT/TOOLS\n\nUse a computer for 8 hours a day\n\nBaker Group is an Equal Opportunity Employer. In compliance with the Americans with Disabilities Act, Baker Group will consider reasonable accommodations for qualified individuals with disabilities and encourage prospective employees and incumbents to discuss potential accommodations with the Employer.","company":"Baker Group","rawCompany":"baker group","city":"Ankeny","state":"IA","isRemote":false,"isActive":false,"createdAt":"2026-09-06T13:29:01.023Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Engineer","description":"PURPOSE\n\nThe Data Engineer is responsible for designing, building, and maintaining the data pipelines and infrastructure that power Baker Group's Microsoft Fabric data warehouse, serving as the organization's single source of truth. This role owns the ingestion, transformation, and orchestration of data from disparate internal systems (ERP, HRIS, MRP and other structured data sources) into governed, reliable data products used by Data Analysts and developers to deliver insights to executive and operational teams and ensures that data is structured to support both traditional reporting and emerging AI and machine learning use cases. The Data Engineer curates and maintains core datasets spanning employees, finance, construction and manufacturing projects, and service, and partners with the Data Scientist, Data Analyst, and Software Development roles to ensure data is trustworthy, well-structured, and fit for downstream use.\n\nESSENTIAL FUNCTIONS AND RESPONSIBILITIES\n\nThe following duties are typical for this job. These are not to be constructed as exclusive or all inclusive. Other duties may be required and assigned.\n\nDesigns, builds, and maintains ETL/ELT pipelines that ingest data from enterprise systems into Microsoft Fabric.\nArchitects and maintains the Fabric medallion Lakehouse structure (bronze, silver, gold layers) as Baker Group's single source of truth.\nDevelop and implement best practices for the data infrastructure and environment (e.g. Development/Test/Production environments, Git for version control).\nOwns pipeline orchestration, scheduling, and monitoring to ensure reliable, timely, and accurate data availability.\nCurates and maintains core datasets across employee, finance, project, service, and manufacturing domains.\nEstablishes and enforces data quality, validation, and reconciliation processes across all pipelines.\nDesigns and manages data models, schemas, and semantic layers that support Data Analyst reporting and Data Scientist modeling work.\nDefines and maintains data ontologies and canonical business definitions (for example, what constitutes a \"project,\" \"employee,\" or \"cost code\") to ensure consistent meaning across systems and consumers.\nPrepares and structures data to support AI and machine learning use cases, including feature-ready datasets, retrieval-augmented generation (RAG) pipelines, and vector embedding storage.\nManages Fabric capacity planning, workspace organization, and performance optimization.\nImplements data governance practices, including access controls, lineage tracking, and metadata management, consistent with Baker Group's data classification standards.\nPartners with business system owners (ERP, HRIS, MRP, etc.) to understand upstream data structures and manage change impacts.\nCollaborates with the Data Scientist to ensure pipeline outputs support analytical and machine learning use cases.\nCollaborates with Data Analysts to ensure data products support paginated reporting, dashboards, and self-service BI needs.\nCollaborates with Software Development and DevOps Teams to ensure data products support application development needs.\nCoordinates with 3rd party consultants when necessary to deliver data engineering projects and augment capacity for demanding business needs.\nDevelops and maintains documentation for pipelines, schemas, and integration logic.\nTroubleshoots and resolves pipeline failures, latency issues, and data quality incidents.\nMonitors and maintains data-specific infrastructure, including Fabric capacity, pipeline orchestration tools, and monitoring/alerting systems.\nEvaluates and recommends new data engineering tools, patterns, and best practices.\nStays current on emerging trends in data engineering, cloud data platforms, and integration techniques.\n\nMINIMUM EDUCATION and EXPERIENCE REQUIRED TO PERFORM ESSENTIAL FUNCTIONS\n\nBachelor's degree in Computer Science, Data Engineering, Information Systems, or other relevant quantitative field\nThree to five years of experience in data engineering, ETL/ELT development, or a related field\nProficiency with SQL and database technologies for data extraction, transformation, and loading\nExperience with Microsoft Fabric, Azure Data Factory, or similar cloud ETL/orchestration tools\nExperience with medallion architecture and modern data warehousing patterns\nExperience with a programming language such as Python, PySpark, or T-SQL for data transformation\nFamiliarity with data modeling techniques (dimensional modeling, star schema)\nUnderstanding of data governance, data quality, and metadata management practices\nExperience preparing data for AI/ML consumption (e.g., vector embeddings, RAG architectures) is a plus\nBusiness acumen and understanding of construction or related industries is a plus\n\nCERTIFICATES, LICENSES, REGISTRATIONS\n\nNo specific requirements; however, relevant certifications such as Microsoft Certified: Fabric Data Engineer Associate, Azure Data Engineer Associate, or similar cloud platform certifications are a plus\n\nMENTAL AND PHYSICAL COMPETENCIES REQUIRED TO PERFORM ESSENTIAL FUNCTIONS\n\nStrong analytical and troubleshooting skills with the ability to diagnose and resolve complex pipeline and data quality issues\nExcellent time and project management skills with the ability to prioritize across multiple pipeline and infrastructure projects\nCurrent with industry trends in data engineering, cloud platforms, and integration best practices\nStrong communication skills with the ability to translate technical data structures for non-technical stakeholders\nTeam player with strong collaboration skills, particularly with the Data Scientist, Data Analysts, and business system owners\nMust be able to focus on complex technical problems and work independently with minimal supervision\nAbility to work in a fast-paced environment and adapt to changing business priorities\nMeticulous attention to detail and commitment to producing reliable, well-documented data infrastructure\n\nENVIRONMENTAL ADAPTABILITY\n\nProlonged periods of sitting at a desk and working on a computer\nMust be able to lift 10 pounds occasionally\nMay have occasional visits to a job site which would require periods of standing, walking and/or climbing stairs\n\nEQUIPMENT/TOOLS\n\nUse a computer for 8 hours a day\n\nBaker Group is an Equal Opportunity Employer. In compliance with the Americans with Disabilities Act, Baker Group will consider reasonable accommodations for qualified individuals with disabilities and encourage prospective employees and incumbents to discuss potential accommodations with the Employer.","datePosted":"2026-09-06T13:29:01.023Z","dateModified":"2026-09-06T13:29:01.023Z","hiringOrganization":{"@type":"Organization","name":"Baker Group","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Ankeny","addressRegion":"IA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"cee811c73e14f59bacc34033"},"url":"https://jobsearcher.com/jobs/cee811c73e14f59bacc34033"}}