{"schemaVersion":"jobsearcher.job.v1","id":"4b2efa6c43db402fb2bcb217","url":"https://jobsearcher.com/jobs/4b2efa6c43db402fb2bcb217","canonicalUrl":"https://jobsearcher.com/jobs/4b2efa6c43db402fb2bcb217","title":"Data Engineer","description":"Data Engineer\nOverview\nWe're looking for a Data Engineer to design, build, and operate scalable data architectures and pipelines that transform diverse structured and unstructured sources into high-quality data repositories. You'll develop robust ETL/ELT processes in AWS, create parsers and extraction logic for complex formats (e.g., PDFs, contracts, procurement documents, budgetary reports), and deliver curated datasets that serve as reliable sources for analytics and AI/ML applications.\nOur CompanyN2IA Technologies is a consulting company specializing in acquisition/contracting support, cost/FinOps, and technology optimization for federal clients. We deliver tailored strategies, robust software solutions, and streamlined operations to help organizations achieve their goals.As a growing, remote-first organization, N2IA relies on secure, reliable, and scalable IT operations to support both internal teams and federal mission delivery.\nKey Responsibilities\nDesign cloud data architectures for ingestion, storage, transformation, and consumption (batch and, where needed, near real-time).\nBuild and maintain ETL/ELT pipelines that are reliable, testable, observable, and cost efficient.\nIngest and integrate data from diverse sources including APIs, relational databases, file drops, event streams, SaaS platforms, and external data providers.\nWork extensively with structured and unstructured data, including normalization, enrichment, and metadata management.\nDevelop data parsers and extraction logic for complex unstructured sources such as PDFs, contracts, procurement documents, and budgetary reports; implement validation and error handling for imperfect inputs.\nImplement and optimize data storage patterns (e.g., lake/lakehouse/warehouse), indexing/partitioning strategies, and query performance tuning.\nBuild and manage data repositories designed to support AI (feature-ready datasets, document corpora, embeddings-ready stores, retrieval-oriented schemas, lineage and provenance).\nApply data quality practices (automated checks, anomaly detection, reconciliation, SLAs) and implement governance-friendly patterns (cataloging, RBAC, encryption).\nPartner with stakeholders (product, analytics, data science, engineering) to translate requirements into scalable datasets and interfaces.\nCreate and maintain documentation: data models, interfaces, lineage, runbooks, and operational playbooks.\nRequired Qualifications\nBachelor's Degree A Bachelor's degree in a quantitative or business field (e.g., Statistics, Mathematics, Engineering, Computer Science). (Required)\n8+ years of experience in data engineering (or 3–5 years with demonstrable senior-level impact), building production-grade pipelines and data systems.\nStrong proficiency in SQL and at least one general-purpose language (Python strongly preferred).\nProven experience designing data architectures (e.g., data lake/lakehouse/warehouse patterns) and selecting fit-for-purpose storage/compute.\nHands-on experience with AWS data engineering, including several of the following:\nS3, IAM, KMS, VPC, CloudWatch\nGlue, Athena, EMR, Lambda, Step Functions\nRedshift (or alternative warehouse)\nKinesis/MSK (streaming) and/or EventBridge (eventing)\nPractical understanding of data reliability practices: testing, CI/CD, monitoring/alerting, backfills, and cost/performance optimization.\nStrong communication skills-able to explain technical tradeoffs to both technical and non-technical audiences.\nPreferred Qualifications\nExperience supporting AI/ML data products, such as building curated corpora, document stores, vector/embedding pipelines, and retrieval-optimized datasets.\nFamiliarity with search and indexing concepts (e.g., OpenSearch/Elasticsearch) and/or graph/metadata systems.\nExposure to Infrastructure-as-Code (Terraform/CDK/CloudFormation) and containerization (Docker/Kubernetes).\nCertifications (Relevant / Preferred)\nCandidates may have one or more of the following (or equivalent):\nAWS Certified Data Engineer – Associate\nAWS Certified Solutions Architect – Associate or Professional\nAWS Certified Developer – Associate\nAWS Certified Database – Specialty\nDatabricks Certified Data Engineer (Associate/Professional)\nEqual Employment OpportunityN2IA is committed to fostering a diverse and inclusive work environment. We are an Equal Employment Opportunity Employer and encourage applications from all qualified individuals, regardless of gender, race, ethnicity, sexual orientation, disability, or veteran status.","company":"N2ia Technologies","rawCompany":"n2ia technologies","city":"Washington","state":"DC","isRemote":false,"isActive":false,"createdAt":"2026-04-14T10:46:35.951Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Engineer","description":"Data Engineer\nOverview\nWe're looking for a Data Engineer to design, build, and operate scalable data architectures and pipelines that transform diverse structured and unstructured sources into high-quality data repositories. You'll develop robust ETL/ELT processes in AWS, create parsers and extraction logic for complex formats (e.g., PDFs, contracts, procurement documents, budgetary reports), and deliver curated datasets that serve as reliable sources for analytics and AI/ML applications.\nOur CompanyN2IA Technologies is a consulting company specializing in acquisition/contracting support, cost/FinOps, and technology optimization for federal clients. We deliver tailored strategies, robust software solutions, and streamlined operations to help organizations achieve their goals.As a growing, remote-first organization, N2IA relies on secure, reliable, and scalable IT operations to support both internal teams and federal mission delivery.\nKey Responsibilities\nDesign cloud data architectures for ingestion, storage, transformation, and consumption (batch and, where needed, near real-time).\nBuild and maintain ETL/ELT pipelines that are reliable, testable, observable, and cost efficient.\nIngest and integrate data from diverse sources including APIs, relational databases, file drops, event streams, SaaS platforms, and external data providers.\nWork extensively with structured and unstructured data, including normalization, enrichment, and metadata management.\nDevelop data parsers and extraction logic for complex unstructured sources such as PDFs, contracts, procurement documents, and budgetary reports; implement validation and error handling for imperfect inputs.\nImplement and optimize data storage patterns (e.g., lake/lakehouse/warehouse), indexing/partitioning strategies, and query performance tuning.\nBuild and manage data repositories designed to support AI (feature-ready datasets, document corpora, embeddings-ready stores, retrieval-oriented schemas, lineage and provenance).\nApply data quality practices (automated checks, anomaly detection, reconciliation, SLAs) and implement governance-friendly patterns (cataloging, RBAC, encryption).\nPartner with stakeholders (product, analytics, data science, engineering) to translate requirements into scalable datasets and interfaces.\nCreate and maintain documentation: data models, interfaces, lineage, runbooks, and operational playbooks.\nRequired Qualifications\nBachelor's Degree A Bachelor's degree in a quantitative or business field (e.g., Statistics, Mathematics, Engineering, Computer Science). (Required)\n8+ years of experience in data engineering (or 3–5 years with demonstrable senior-level impact), building production-grade pipelines and data systems.\nStrong proficiency in SQL and at least one general-purpose language (Python strongly preferred).\nProven experience designing data architectures (e.g., data lake/lakehouse/warehouse patterns) and selecting fit-for-purpose storage/compute.\nHands-on experience with AWS data engineering, including several of the following:\nS3, IAM, KMS, VPC, CloudWatch\nGlue, Athena, EMR, Lambda, Step Functions\nRedshift (or alternative warehouse)\nKinesis/MSK (streaming) and/or EventBridge (eventing)\nPractical understanding of data reliability practices: testing, CI/CD, monitoring/alerting, backfills, and cost/performance optimization.\nStrong communication skills-able to explain technical tradeoffs to both technical and non-technical audiences.\nPreferred Qualifications\nExperience supporting AI/ML data products, such as building curated corpora, document stores, vector/embedding pipelines, and retrieval-optimized datasets.\nFamiliarity with search and indexing concepts (e.g., OpenSearch/Elasticsearch) and/or graph/metadata systems.\nExposure to Infrastructure-as-Code (Terraform/CDK/CloudFormation) and containerization (Docker/Kubernetes).\nCertifications (Relevant / Preferred)\nCandidates may have one or more of the following (or equivalent):\nAWS Certified Data Engineer – Associate\nAWS Certified Solutions Architect – Associate or Professional\nAWS Certified Developer – Associate\nAWS Certified Database – Specialty\nDatabricks Certified Data Engineer (Associate/Professional)\nEqual Employment OpportunityN2IA is committed to fostering a diverse and inclusive work environment. We are an Equal Employment Opportunity Employer and encourage applications from all qualified individuals, regardless of gender, race, ethnicity, sexual orientation, disability, or veteran status.","datePosted":"2026-04-14T10:46:35.951Z","dateModified":"2026-04-14T10:46:35.951Z","hiringOrganization":{"@type":"Organization","name":"N2ia Technologies","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Washington","addressRegion":"DC","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"4b2efa6c43db402fb2bcb217"},"url":"https://jobsearcher.com/jobs/4b2efa6c43db402fb2bcb217"}}