{"schemaVersion":"jobsearcher.job.v1","id":"2729d9c90dafe1748d9986c9","url":"https://jobsearcher.com/jobs/2729d9c90dafe1748d9986c9","canonicalUrl":"https://jobsearcher.com/jobs/2729d9c90dafe1748d9986c9","title":"Remote Data Engineer (Databricks + Python + Azure)","description":"Allata is a global consulting and technology services firm with offices in the US, India, and Argentina. We help organizations accelerate growth, drive innovation, and solve complex challenges by combining strategy, design, and advanced technology. Our expertise covers defining business vision, optimizing processes, and creating engaging digital experiences. We architect and modernize secure, scalable solutions using cloud platforms and top engineering practices.\r\nAllata also empowers clients to unlock data value through analytics and visualization and leverages artificial intelligence to automate processes and enhance decision-making. Our agile, cross-functional teams work closely with clients, either integrating with their teams or providing independent guidance—to deliver measurable results and build lasting partnerships.\r\nWe are seeking a skilled Data Engineer to join our team and contribute to data-driven initiatives within the healthcare industry. This role focuses on designing, building, and optimizing scalable data solutions that support analytics, reporting, and advanced data use cases in regulated environments.\r\nRole & Responsibilities:\r\nDesign, develop, and maintain scalable data pipelines using Databricks (PySpark) and Python .\r\nBuild and optimize ETL/ELT processes within Azure cloud environments .\r\nImplement data models following modern Data Lakehouse principles (e.g., Medallion architecture).\r\nEnsure data quality, consistency, and performance across ingestion, staging, and curated layers.\r\nCollaborate with data architects, analysts, and business stakeholders to translate healthcare data requirements into technical solutions.\r\nDevelop reusable data transformation logic and modular processing components.\r\nSupport deployment processes following CI/CD and DevOps best practices.\r\nMonitor and optimize data workflows for performance, scalability, and reliability.\r\nContribute to data governance, security, and compliance practices relevant to healthcare environments.\r\nHard Skills - Must have:\r\nCurrent knowledge of an using modern data tools like (Databricks,FiveTran, Data Fabric and others); Core experience with data architecture, data integrations, data warehousing, and ETL/ELT processes\r\nApplied experience with developing and deploying custom whl and or in session notebook scripts for custom execution across parallel executor and worker nodes\r\nApplied experience in SQL, Stored Procedures, and Pysparkbased on area of data platform specialization.\r\nStrong knowledge of cloud and hybrid relational database systems, such as MS SQL Server, PostgresSQL, Oracle, Azure SQL, AWS RDS, Auroraor a comparable engine.\r\nStrong experience with batch and streaming data processing techniques and file compactization strategies.\r\nHard Skills - Nice to have/It's a plus:\r\nStrong hands-on experience with Databricks in Azure environments.\r\nAdvanced proficiency in Python and PySpark for distributed data processing.\r\nExperience building and optimizing data pipelines in Azure (Azure Data Factory, Azure SQL, Data Lake Storage, etc.) .\r\nSolid understanding of data warehousing, data lakehouse concepts, and ETL/ELT frameworks.\r\nExperience working with relational databases such as SQL Server, PostgreSQL, Oracle, or similar.\r\nKnowledge of batch and streaming data processing patterns.\r\nExperience working with large, complex datasets in cloud-based distributed environments.\r\nSoft Skills / Business Specific Skills:\r\nStrong analytical and problem-solving skills.\r\nAbility to work effectively in cross-functional and distributed teams.\r\nClear communication skills, with the ability to explain technical concepts to non-technical stakeholders.\r\nProactive mindset with a strong sense of ownership.\r\nCommitment to delivering high-quality, reliable data solutions.\r\nAt Allata, we value differences.\r\nAllata is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.\r\nAllata makes employment decisions without regard to race, color, creed, religion, age, ancestry, national origin, veteran status, sex, sexual orientation, gender, gender identity, gender expression, marital status, disability or any other legally protected category.\r\nThis policy applies to all terms and conditions of employment, including but not limited to, recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation, and training. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.","company":"GrabJobs","rawCompany":"grabjobs","city":"San Jose","state":"CA","isRemote":true,"isActive":false,"createdAt":"2026-08-08T01:53:07.004Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541690","title":"Other Scientific and Technical Consulting Services","slug":"other-scientific-and-technical-consulting-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Remote Data Engineer (Databricks + Python + Azure)","description":"Allata is a global consulting and technology services firm with offices in the US, India, and Argentina. We help organizations accelerate growth, drive innovation, and solve complex challenges by combining strategy, design, and advanced technology. Our expertise covers defining business vision, optimizing processes, and creating engaging digital experiences. We architect and modernize secure, scalable solutions using cloud platforms and top engineering practices.\r\nAllata also empowers clients to unlock data value through analytics and visualization and leverages artificial intelligence to automate processes and enhance decision-making. Our agile, cross-functional teams work closely with clients, either integrating with their teams or providing independent guidance—to deliver measurable results and build lasting partnerships.\r\nWe are seeking a skilled Data Engineer to join our team and contribute to data-driven initiatives within the healthcare industry. This role focuses on designing, building, and optimizing scalable data solutions that support analytics, reporting, and advanced data use cases in regulated environments.\r\nRole & Responsibilities:\r\nDesign, develop, and maintain scalable data pipelines using Databricks (PySpark) and Python .\r\nBuild and optimize ETL/ELT processes within Azure cloud environments .\r\nImplement data models following modern Data Lakehouse principles (e.g., Medallion architecture).\r\nEnsure data quality, consistency, and performance across ingestion, staging, and curated layers.\r\nCollaborate with data architects, analysts, and business stakeholders to translate healthcare data requirements into technical solutions.\r\nDevelop reusable data transformation logic and modular processing components.\r\nSupport deployment processes following CI/CD and DevOps best practices.\r\nMonitor and optimize data workflows for performance, scalability, and reliability.\r\nContribute to data governance, security, and compliance practices relevant to healthcare environments.\r\nHard Skills - Must have:\r\nCurrent knowledge of an using modern data tools like (Databricks,FiveTran, Data Fabric and others); Core experience with data architecture, data integrations, data warehousing, and ETL/ELT processes\r\nApplied experience with developing and deploying custom whl and or in session notebook scripts for custom execution across parallel executor and worker nodes\r\nApplied experience in SQL, Stored Procedures, and Pysparkbased on area of data platform specialization.\r\nStrong knowledge of cloud and hybrid relational database systems, such as MS SQL Server, PostgresSQL, Oracle, Azure SQL, AWS RDS, Auroraor a comparable engine.\r\nStrong experience with batch and streaming data processing techniques and file compactization strategies.\r\nHard Skills - Nice to have/It's a plus:\r\nStrong hands-on experience with Databricks in Azure environments.\r\nAdvanced proficiency in Python and PySpark for distributed data processing.\r\nExperience building and optimizing data pipelines in Azure (Azure Data Factory, Azure SQL, Data Lake Storage, etc.) .\r\nSolid understanding of data warehousing, data lakehouse concepts, and ETL/ELT frameworks.\r\nExperience working with relational databases such as SQL Server, PostgreSQL, Oracle, or similar.\r\nKnowledge of batch and streaming data processing patterns.\r\nExperience working with large, complex datasets in cloud-based distributed environments.\r\nSoft Skills / Business Specific Skills:\r\nStrong analytical and problem-solving skills.\r\nAbility to work effectively in cross-functional and distributed teams.\r\nClear communication skills, with the ability to explain technical concepts to non-technical stakeholders.\r\nProactive mindset with a strong sense of ownership.\r\nCommitment to delivering high-quality, reliable data solutions.\r\nAt Allata, we value differences.\r\nAllata is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.\r\nAllata makes employment decisions without regard to race, color, creed, religion, age, ancestry, national origin, veteran status, sex, sexual orientation, gender, gender identity, gender expression, marital status, disability or any other legally protected category.\r\nThis policy applies to all terms and conditions of employment, including but not limited to, recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation, and training. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.","datePosted":"2026-08-08T01:53:07.004Z","dateModified":"2026-08-08T01:53:07.004Z","hiringOrganization":{"@type":"Organization","name":"GrabJobs","sameAs":"https://jobsearcher.com"},"jobLocationType":"TELECOMMUTE","applicantLocationRequirements":{"@type":"Country","name":"US"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"San Jose","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"2729d9c90dafe1748d9986c9"},"url":"https://jobsearcher.com/jobs/2729d9c90dafe1748d9986c9"}}