{"schemaVersion":"jobsearcher.job.v1","id":"fc3abd949e9de02cb077d8ae","url":"https://jobsearcher.com/jobs/fc3abd949e9de02cb077d8ae","canonicalUrl":"https://jobsearcher.com/jobs/fc3abd949e9de02cb077d8ae","title":"Databricks Engineer","description":"JPG Infotech LLC is seeking a *Hands-On Senior Databricks Engineer & Architect* to lead data modernization and engineering initiatives on a mission-critical Federal contract. In this role, you will own the end-to-end Databricks ecosystem—from infrastructure deployment and security configuration to designing scalable Lakehouse architectures and executing complex data migrations from legacy environments.\n\nYou will bridge the gap between high-level systems architecture and hands-on PySpark/SQL development, working closely with federal stakeholders to transform complex data requirements into secure, high-performance analytical pipelines.\n\nKey Responsibilities1. Databricks Installation, Configuration & Governance\n\n* *Platform Setup:* Architect, install, and configure enterprise Databricks workspaces within a secure Federal Cloud environment (AWS GovCloud, Azure Government, or GCP).\n* *Cluster & Compute Management:* Configure compute clusters, auto-scaling policies, and job orchestration environments to optimize performance and cloud consumption.\n* *Governance & Security:* Implement *Unity Catalog* for centralized data governance, metadata management, fine-grained access control (RBAC/ABAC), and data lineage tracking.\n* *Compliance:* Ensure all Databricks deployments align with federal security frameworks (e.g., FedRAMP, FISMA, NIST SP 800-53).\n\n2. Data Architecture & Lakehouse Design\n\n* *Lakehouse Architecture:* Design and deploy modern *Medallion Architectures* (Bronze, Silver, Gold layers) using Delta Lake.\n* *Data Modeling:* Create conceptual, logical, and physical data models for transactional, dimensional, and analytical data stores.\n* *System Integration:* Establish seamless API and pipeline integrations across multi-cloud services, enterprise data warehouses, and downstream reporting tools.\n\n3. Data Migration & Pipeline Engineering\n\n* *Legacy Modernization:* Lead the migration of on-premises data warehouses, relational databases, and legacy ETL workloads into Databricks Delta Lake.\n* *Pipeline Development:* Build, test, and deploy robust ETL/ELT pipelines using *PySpark, Python, and SQL* for both batch and real-time streaming data.\n* *Orchestration:* Configure automated scheduling, dependency management, and monitoring using Lakeflow / Databricks Workflows or Apache Airflow.\n\n4. Data Analysis & Performance Tuning\n\n* *Query & Cluster Optimization:* Diagnose and troubleshoot bottlenecks, optimize Spark query execution plans, and implement caching and partitioning strategies.\n* *Data Quality & Validation:* Implement automated data-reconciliation, completeness, and schema-enforcement checks to ensure auditability and compliance.\n* *Stakeholder Collaboration:* Partner with federal program managers, analysts, and engineering teams to translate operational requirements into actionable technical solutions.\n\nRequired QualificationsSecurity & Location\n\n* *Citizenship:* *Must be a U.S. Citizen* (dual citizenship or non-citizen status cannot be accommodated due to federal contract mandate).\n* *Clearance:* Ability to pass a federal background investigation and obtain/maintain a *Public Trust* or *Secret clearance*.\n* *Location:* Preference for candidates residing in the *Washington, D.C. Metropolitan Area (DC/MD/VA)* with the flexibility to work on-site a few days per week as required by project milestones.\n\nExperience & Technical Skills\n\n* *Experience:* *5+ years* of hands-on data engineering, data architecture, or big data platform experience.\n* *Databricks Expertise:* Demonstrated hands-on experience installing, configuring, and maintaining production *Databricks*environments.\n* *Programming Languages:* Deep proficiency writing production-grade code in *Python (PySpark)* and *SQL* (Scala is a plus).\n* *Core Technologies:*\n* Strong expertise with *Apache Spark™* distributed processing and runtime internals.\n* Extensive experience building Lakehouses with *Delta Lake* and *Unity Catalog*.\n* Familiarity with major cloud infrastructure platforms (*AWS, Azure, or GCP*).\n* *Engineering Practices:* Experience with Git-based CI/CD workflows, automated testing, and Infrastructure as Code (e.g., Terraform) for data platform deployment.\n\nPreferred Qualifications\n\n* *Certifications:* _Databricks Certified Data Engineer Professional_ or _Databricks Certified Data Engineer Associate_.\n* *Public Sector Experience:* Prior experience working on federal, state, or defense IT contracts.\n* *Advanced Databricks Tools:* Experience with Delta Live Tables (DLT) / Lakeflow Declarative Pipelines, Databricks SQL Warehouses, or MLOps integrations.\n* *Data Integration:* Familiarity with data contract standards, source-to-target mapping, and open-data government reporting requirements.\n\nWhy Join JPG Infotech LLC?\n\nAt *JPG Infotech LLC*, we deliver cutting-edge AI, cloud, and digital transformation solutions to federal and enterprise clients. You will be part of an agile, high-impact technical team working on meaningful government missions with access to the latest data and AI technologies.\n\nPay: $117,766.11 - $142,557.03 per year\n\nBenefits:\n* 401(k)\n* 401(k) matching\n* Health insurance\n* Paid time off\n* Referral program\n* Retirement plan\n\nWork Location: Hybrid remote in Ballston, VA","company":"Jpginfotechllc","rawCompany":"jpginfotechllc","city":"Arlington","state":"VA","isRemote":false,"isActive":false,"createdAt":"2026-08-22T15:17:13.156Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Databricks Engineer","description":"JPG Infotech LLC is seeking a *Hands-On Senior Databricks Engineer & Architect* to lead data modernization and engineering initiatives on a mission-critical Federal contract. In this role, you will own the end-to-end Databricks ecosystem—from infrastructure deployment and security configuration to designing scalable Lakehouse architectures and executing complex data migrations from legacy environments.\n\nYou will bridge the gap between high-level systems architecture and hands-on PySpark/SQL development, working closely with federal stakeholders to transform complex data requirements into secure, high-performance analytical pipelines.\n\nKey Responsibilities1. Databricks Installation, Configuration & Governance\n\n* *Platform Setup:* Architect, install, and configure enterprise Databricks workspaces within a secure Federal Cloud environment (AWS GovCloud, Azure Government, or GCP).\n* *Cluster & Compute Management:* Configure compute clusters, auto-scaling policies, and job orchestration environments to optimize performance and cloud consumption.\n* *Governance & Security:* Implement *Unity Catalog* for centralized data governance, metadata management, fine-grained access control (RBAC/ABAC), and data lineage tracking.\n* *Compliance:* Ensure all Databricks deployments align with federal security frameworks (e.g., FedRAMP, FISMA, NIST SP 800-53).\n\n2. Data Architecture & Lakehouse Design\n\n* *Lakehouse Architecture:* Design and deploy modern *Medallion Architectures* (Bronze, Silver, Gold layers) using Delta Lake.\n* *Data Modeling:* Create conceptual, logical, and physical data models for transactional, dimensional, and analytical data stores.\n* *System Integration:* Establish seamless API and pipeline integrations across multi-cloud services, enterprise data warehouses, and downstream reporting tools.\n\n3. Data Migration & Pipeline Engineering\n\n* *Legacy Modernization:* Lead the migration of on-premises data warehouses, relational databases, and legacy ETL workloads into Databricks Delta Lake.\n* *Pipeline Development:* Build, test, and deploy robust ETL/ELT pipelines using *PySpark, Python, and SQL* for both batch and real-time streaming data.\n* *Orchestration:* Configure automated scheduling, dependency management, and monitoring using Lakeflow / Databricks Workflows or Apache Airflow.\n\n4. Data Analysis & Performance Tuning\n\n* *Query & Cluster Optimization:* Diagnose and troubleshoot bottlenecks, optimize Spark query execution plans, and implement caching and partitioning strategies.\n* *Data Quality & Validation:* Implement automated data-reconciliation, completeness, and schema-enforcement checks to ensure auditability and compliance.\n* *Stakeholder Collaboration:* Partner with federal program managers, analysts, and engineering teams to translate operational requirements into actionable technical solutions.\n\nRequired QualificationsSecurity & Location\n\n* *Citizenship:* *Must be a U.S. Citizen* (dual citizenship or non-citizen status cannot be accommodated due to federal contract mandate).\n* *Clearance:* Ability to pass a federal background investigation and obtain/maintain a *Public Trust* or *Secret clearance*.\n* *Location:* Preference for candidates residing in the *Washington, D.C. Metropolitan Area (DC/MD/VA)* with the flexibility to work on-site a few days per week as required by project milestones.\n\nExperience & Technical Skills\n\n* *Experience:* *5+ years* of hands-on data engineering, data architecture, or big data platform experience.\n* *Databricks Expertise:* Demonstrated hands-on experience installing, configuring, and maintaining production *Databricks*environments.\n* *Programming Languages:* Deep proficiency writing production-grade code in *Python (PySpark)* and *SQL* (Scala is a plus).\n* *Core Technologies:*\n* Strong expertise with *Apache Spark™* distributed processing and runtime internals.\n* Extensive experience building Lakehouses with *Delta Lake* and *Unity Catalog*.\n* Familiarity with major cloud infrastructure platforms (*AWS, Azure, or GCP*).\n* *Engineering Practices:* Experience with Git-based CI/CD workflows, automated testing, and Infrastructure as Code (e.g., Terraform) for data platform deployment.\n\nPreferred Qualifications\n\n* *Certifications:* _Databricks Certified Data Engineer Professional_ or _Databricks Certified Data Engineer Associate_.\n* *Public Sector Experience:* Prior experience working on federal, state, or defense IT contracts.\n* *Advanced Databricks Tools:* Experience with Delta Live Tables (DLT) / Lakeflow Declarative Pipelines, Databricks SQL Warehouses, or MLOps integrations.\n* *Data Integration:* Familiarity with data contract standards, source-to-target mapping, and open-data government reporting requirements.\n\nWhy Join JPG Infotech LLC?\n\nAt *JPG Infotech LLC*, we deliver cutting-edge AI, cloud, and digital transformation solutions to federal and enterprise clients. You will be part of an agile, high-impact technical team working on meaningful government missions with access to the latest data and AI technologies.\n\nPay: $117,766.11 - $142,557.03 per year\n\nBenefits:\n* 401(k)\n* 401(k) matching\n* Health insurance\n* Paid time off\n* Referral program\n* Retirement plan\n\nWork Location: Hybrid remote in Ballston, VA","datePosted":"2026-08-22T15:17:13.156Z","dateModified":"2026-08-22T15:17:13.156Z","hiringOrganization":{"@type":"Organization","name":"Jpginfotechllc","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Arlington","addressRegion":"VA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"fc3abd949e9de02cb077d8ae"},"url":"https://jobsearcher.com/jobs/fc3abd949e9de02cb077d8ae"}}