{"schemaVersion":"jobsearcher.job.v1","id":"cd31b00111f8cc3562bbfc12","url":"https://jobsearcher.com/jobs/cd31b00111f8cc3562bbfc12","canonicalUrl":"https://jobsearcher.com/jobs/cd31b00111f8cc3562bbfc12","title":"Data Engineer (Databricks/AWS)","description":"Job Information\nDate Opened 08/07/2026\nIndustry Pharma\nJob Type Full time\nYears of Experience 5\nCity Indianapolis\nState/Province Indiana\nCountry United States\nZip/Postal Code 46201\nAbout Us\nFounded in 2015, RADcube is a leading technology consulting and software development firm headquartered in Carmel, Indiana. The company specializes in transforming enterprise ideas into real-world innovations by leveraging emerging technologies such as Artificial Intelligence, Blockchain, and Cloud Computing.\n\nWith nearly a decade of industry experience, RADcube serves diverse sectors, including healthcare, finance, government, and manufacturing. Their core service portfolio includes:\n\nDigital Transformation and strategy consulting.\nCustom Software Development tailored to specific business needs.\nAdvanced Data Analytics and AI-driven platforms.\nCybersecurity and risk management.\n\nRecognized for its innovation-led culture, RADcube operates RADlabs, an R&D hub focused on high-impact solutions like Responsible AI and Intelligent Automation. The firm is committed to a human-centric approach, ensuring cutting-edge technology delivers measurable business outcomes and long-term success for global clients.\n\nThe company’s commitment to innovation has earned significant industry honors:\n\n2026 TechPoint Mira Awards Finalist: Named a finalist for Tech Company of the Year, recognizing high-growth pioneers that demonstrate extraordinary leadership.\n\nPublic Sector Excellence: Awarded the Utah NASPO Cloud & Software Solutions Contract, solidifying their role as a trusted partner for large-scale government digital initiatives and more.\n\nJob Description\nAWS Data Engineer - Databricks\n\nHybrid – Indianapolis, IN\n\nAbout the Role\n\nWe are seeking a Data Engineer with 3–5 years of experience working specifically within the pharma industry to join a pharma-focused data team. This is a senior-flavored engineering role that combines hands-on pipeline and platform work with significant business-facing responsibility — including translating business needs into technical specs, presenting to executive-level stakeholders, and helping stand up new data domains from the ground up. You will design, build, and govern the data infrastructure that powers analytics and reporting across the business, while also acting as a trusted technical partner to non-technical stakeholders.\n\nKey Responsibilities\n\nDesign, build, and maintain scalable ETL/ELT pipelines (batch and streaming) using Databricks, AWS, and related orchestration tools\n\nWrite and optimize advanced SQL, and build data transformations in Python or Scala\n\nIntegrate external data sources via APIs and manage pipeline orchestration (Airflow, Databricks Workflows, AWS Glue)\n\nApply data quality, governance, cataloging, and lineage practices aligned with regulated-industry standards\n\nWork within GxP-regulated data environments and apply awareness of data privacy/compliance considerations (e.g., 21 CFR Part 11, GDPR where applicable)\n\nPartner with business stakeholders across the pharma value chain (R&D, Manufacturing & Quality, Commercial, Drug Development) to gather and translate requirements into technical specifications\n\nPresent technical work and data strategy to executive-level audiences\n\nPrioritize high-impact data initiatives and proactively identify and avoid duplicated data efforts\n\nSupport change management and adoption of new data solutions across business teams\n\nHelp stand up new data domains from scratch (green-field build), not just maintain existing ones\n\nRequirements\nRequired Qualifications\n\nData Engineering & Pipelines\n\nETL/ELT development (batch and streaming)\n\nAdvanced SQL (joins, window functions, query optimization)\n\nPython or Scala for data transformation\n\nData pipeline orchestration (Airflow, Databricks Workflows, AWS Glue)\n\nAPI integration for external data source ingestion\n\nPlatforms & Tools\n\nDatabricks (Delta Lake, Unity Catalog, Genie)\n\nCloud platforms — AWS (S3, Glue, Athena) and/or Azure/GCP equivalents\n\nData warehousing concepts (dimensional modeling, star schema)\n\nBI/visualization tools (Tableau, Power BI, or similar) to understand downstream consumption\n\nData Quality & Governance\n\nData profiling and cleansing techniques\n\nMetadata management and data cataloging\n\nMaster data management (MDM) principles\n\nData lineage tracking\n\nData governance frameworks (especially regulated-industry standards)\n\nPharma / Life Sciences Domain Knowledge\n\nFamiliarity with GxP-regulated data environments\n\nUnderstanding of the pharma value chain (R&D, Manufacturing & Quality, Commercial, Drug Development)\n\nAwareness of data privacy/compliance considerations (21 CFR Part 11, GDPR where applicable)\n\nKnowledge of common pharma data domains (clinical, manufacturing, quality, commercial)\n\nStakeholder Management\n\nRequirements gathering and translation (business need technical spec)\n\nCross-functional communication (Business IT)\n\nExecutive-level presentation skills (given EC visibility)\n\nChange management / adoption support\n\nAnalytical & Strategic Thinking\n\nPrioritization frameworks (identifying high-impact vs. low-value data asks)\n\nCost-avoidance mindset (spotting duplication before it happens)\n\nAbility to work with ambiguity and evolving priorities\n\nProject & Program Skills\n\nAgile/Scrum familiarity\n\nDocumentation discipline (data dictionaries, source-to-target mappings)\n\nVendor/partner coordination (if external data sources are involved)\n\nNice-to-Have Differentiators\n\nPrior consulting or client-facing delivery experience\n\nExperience standing up new data domains from scratch (green-field vs. maintenance)\n\nFamiliarity with AI/GenAI-enabled analytics tools","company":"Radcube","rawCompany":"radcube","city":"Indianapolis","state":"IN","isRemote":false,"isActive":false,"createdAt":"2026-08-08T11:31:53.676Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Engineer (Databricks/AWS)","description":"Job Information\nDate Opened 08/07/2026\nIndustry Pharma\nJob Type Full time\nYears of Experience 5\nCity Indianapolis\nState/Province Indiana\nCountry United States\nZip/Postal Code 46201\nAbout Us\nFounded in 2015, RADcube is a leading technology consulting and software development firm headquartered in Carmel, Indiana. The company specializes in transforming enterprise ideas into real-world innovations by leveraging emerging technologies such as Artificial Intelligence, Blockchain, and Cloud Computing.\n\nWith nearly a decade of industry experience, RADcube serves diverse sectors, including healthcare, finance, government, and manufacturing. Their core service portfolio includes:\n\nDigital Transformation and strategy consulting.\nCustom Software Development tailored to specific business needs.\nAdvanced Data Analytics and AI-driven platforms.\nCybersecurity and risk management.\n\nRecognized for its innovation-led culture, RADcube operates RADlabs, an R&D hub focused on high-impact solutions like Responsible AI and Intelligent Automation. The firm is committed to a human-centric approach, ensuring cutting-edge technology delivers measurable business outcomes and long-term success for global clients.\n\nThe company’s commitment to innovation has earned significant industry honors:\n\n2026 TechPoint Mira Awards Finalist: Named a finalist for Tech Company of the Year, recognizing high-growth pioneers that demonstrate extraordinary leadership.\n\nPublic Sector Excellence: Awarded the Utah NASPO Cloud & Software Solutions Contract, solidifying their role as a trusted partner for large-scale government digital initiatives and more.\n\nJob Description\nAWS Data Engineer - Databricks\n\nHybrid – Indianapolis, IN\n\nAbout the Role\n\nWe are seeking a Data Engineer with 3–5 years of experience working specifically within the pharma industry to join a pharma-focused data team. This is a senior-flavored engineering role that combines hands-on pipeline and platform work with significant business-facing responsibility — including translating business needs into technical specs, presenting to executive-level stakeholders, and helping stand up new data domains from the ground up. You will design, build, and govern the data infrastructure that powers analytics and reporting across the business, while also acting as a trusted technical partner to non-technical stakeholders.\n\nKey Responsibilities\n\nDesign, build, and maintain scalable ETL/ELT pipelines (batch and streaming) using Databricks, AWS, and related orchestration tools\n\nWrite and optimize advanced SQL, and build data transformations in Python or Scala\n\nIntegrate external data sources via APIs and manage pipeline orchestration (Airflow, Databricks Workflows, AWS Glue)\n\nApply data quality, governance, cataloging, and lineage practices aligned with regulated-industry standards\n\nWork within GxP-regulated data environments and apply awareness of data privacy/compliance considerations (e.g., 21 CFR Part 11, GDPR where applicable)\n\nPartner with business stakeholders across the pharma value chain (R&D, Manufacturing & Quality, Commercial, Drug Development) to gather and translate requirements into technical specifications\n\nPresent technical work and data strategy to executive-level audiences\n\nPrioritize high-impact data initiatives and proactively identify and avoid duplicated data efforts\n\nSupport change management and adoption of new data solutions across business teams\n\nHelp stand up new data domains from scratch (green-field build), not just maintain existing ones\n\nRequirements\nRequired Qualifications\n\nData Engineering & Pipelines\n\nETL/ELT development (batch and streaming)\n\nAdvanced SQL (joins, window functions, query optimization)\n\nPython or Scala for data transformation\n\nData pipeline orchestration (Airflow, Databricks Workflows, AWS Glue)\n\nAPI integration for external data source ingestion\n\nPlatforms & Tools\n\nDatabricks (Delta Lake, Unity Catalog, Genie)\n\nCloud platforms — AWS (S3, Glue, Athena) and/or Azure/GCP equivalents\n\nData warehousing concepts (dimensional modeling, star schema)\n\nBI/visualization tools (Tableau, Power BI, or similar) to understand downstream consumption\n\nData Quality & Governance\n\nData profiling and cleansing techniques\n\nMetadata management and data cataloging\n\nMaster data management (MDM) principles\n\nData lineage tracking\n\nData governance frameworks (especially regulated-industry standards)\n\nPharma / Life Sciences Domain Knowledge\n\nFamiliarity with GxP-regulated data environments\n\nUnderstanding of the pharma value chain (R&D, Manufacturing & Quality, Commercial, Drug Development)\n\nAwareness of data privacy/compliance considerations (21 CFR Part 11, GDPR where applicable)\n\nKnowledge of common pharma data domains (clinical, manufacturing, quality, commercial)\n\nStakeholder Management\n\nRequirements gathering and translation (business need technical spec)\n\nCross-functional communication (Business IT)\n\nExecutive-level presentation skills (given EC visibility)\n\nChange management / adoption support\n\nAnalytical & Strategic Thinking\n\nPrioritization frameworks (identifying high-impact vs. low-value data asks)\n\nCost-avoidance mindset (spotting duplication before it happens)\n\nAbility to work with ambiguity and evolving priorities\n\nProject & Program Skills\n\nAgile/Scrum familiarity\n\nDocumentation discipline (data dictionaries, source-to-target mappings)\n\nVendor/partner coordination (if external data sources are involved)\n\nNice-to-Have Differentiators\n\nPrior consulting or client-facing delivery experience\n\nExperience standing up new data domains from scratch (green-field vs. maintenance)\n\nFamiliarity with AI/GenAI-enabled analytics tools","datePosted":"2026-08-08T11:31:53.676Z","dateModified":"2026-08-08T11:31:53.676Z","hiringOrganization":{"@type":"Organization","name":"Radcube","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Indianapolis","addressRegion":"IN","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"cd31b00111f8cc3562bbfc12"},"url":"https://jobsearcher.com/jobs/cd31b00111f8cc3562bbfc12"}}