{"schemaVersion":"jobsearcher.job.v1","id":"ba97ee84c92baed4a77376be","url":"https://jobsearcher.com/jobs/ba97ee84c92baed4a77376be","canonicalUrl":"https://jobsearcher.com/jobs/ba97ee84c92baed4a77376be","title":"Data Migration & Batch Engineer","description":"Data Migration & Batch EngineerLocation: Charlotte, NC/ Austin, TX / San Diego , CA and NYC/ NYType: ContractExperience: 8+ years in relational data migration and/or ETL engineering, with cutovers delivered to production.You own the deterministic core of the database and batch workstreams. Two systems have minted identifiers independently for many years across about a million accounts; merging them incorrectly mis-maps an investor's assets. This seat is where correctness is defended.ResponsibilitiesOwn schema-drift detection and reconciliation (Red Gate / SQL comparison) - the deterministic ground truth on which every downstream decision rests.Build the deterministic rule ladder for table and job dispositions: staging / truncate-and-reload detection, dead-table detection, target-equivalent matching, straight-through repoint, duplicate-job merge.Own the ID-collision engine: a full (never sampled) overlap scan across every shared ID space; classification into true collision, phantom and historical overlap; analysis of the resolution strategy; and a non-bypassable validation harness on every script that touches an ID column.Own data reconciliation: row counts, column checksums, referential integrity, stored-procedure output parity, business-key parity; staging promotion gates and validated rollback scripts.Batch: build the dependency graph, blast-radius (butterfly) scoring, Simple/Medium/Complex classification, and the migration approach per job.Own the SLA validation harness - proving every migrated job and every SLA chain runs within its window under combined nightly volume.Emit the downstream artifacts: migration scripts, job modifications, rollback scripts, reconciliation specifications.QualificationsSQL Server (2024) - deep: DDL, stored procedures, indexing, referential integrity, query performance.Relational data migration at scale - Red Gate, AWS DMS or equivalent; cutover, reconciliation and rollback delivered in production.Identity and key-collision resolution - surrogate versus natural keys, reconciliation and historical-mapping tables, preserving referential integrity under a merge. Rare, and highly valued for this role.Informatica PowerCenter; ETL dependency analysis; SLA and batch-window modelling.Python and strong SQL; data-quality and validation frameworks.The discipline to keep correctness-critical steps deterministic while working alongside an LLM reasoning layer.Working knowledgeInformatica IDMC and/or Apache Airflow; PySpark.Graph queries; handling of regulated / PII data.","company":"VBeyond","rawCompany":"vbeyond","city":"Charlotte","state":"NC","isRemote":false,"isActive":false,"createdAt":"2026-09-11T13:00:13.833Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-1242.00","title":"Database Administrators","slug":"database-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541519","title":"Other Computer Related Services","slug":"other-computer-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Migration & Batch Engineer","description":"Data Migration & Batch EngineerLocation: Charlotte, NC/ Austin, TX / San Diego , CA and NYC/ NYType: ContractExperience: 8+ years in relational data migration and/or ETL engineering, with cutovers delivered to production.You own the deterministic core of the database and batch workstreams. Two systems have minted identifiers independently for many years across about a million accounts; merging them incorrectly mis-maps an investor's assets. This seat is where correctness is defended.ResponsibilitiesOwn schema-drift detection and reconciliation (Red Gate / SQL comparison) - the deterministic ground truth on which every downstream decision rests.Build the deterministic rule ladder for table and job dispositions: staging / truncate-and-reload detection, dead-table detection, target-equivalent matching, straight-through repoint, duplicate-job merge.Own the ID-collision engine: a full (never sampled) overlap scan across every shared ID space; classification into true collision, phantom and historical overlap; analysis of the resolution strategy; and a non-bypassable validation harness on every script that touches an ID column.Own data reconciliation: row counts, column checksums, referential integrity, stored-procedure output parity, business-key parity; staging promotion gates and validated rollback scripts.Batch: build the dependency graph, blast-radius (butterfly) scoring, Simple/Medium/Complex classification, and the migration approach per job.Own the SLA validation harness - proving every migrated job and every SLA chain runs within its window under combined nightly volume.Emit the downstream artifacts: migration scripts, job modifications, rollback scripts, reconciliation specifications.QualificationsSQL Server (2024) - deep: DDL, stored procedures, indexing, referential integrity, query performance.Relational data migration at scale - Red Gate, AWS DMS or equivalent; cutover, reconciliation and rollback delivered in production.Identity and key-collision resolution - surrogate versus natural keys, reconciliation and historical-mapping tables, preserving referential integrity under a merge. Rare, and highly valued for this role.Informatica PowerCenter; ETL dependency analysis; SLA and batch-window modelling.Python and strong SQL; data-quality and validation frameworks.The discipline to keep correctness-critical steps deterministic while working alongside an LLM reasoning layer.Working knowledgeInformatica IDMC and/or Apache Airflow; PySpark.Graph queries; handling of regulated / PII data.","datePosted":"2026-09-11T13:00:13.833Z","dateModified":"2026-09-11T13:00:13.833Z","hiringOrganization":{"@type":"Organization","name":"VBeyond","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Charlotte","addressRegion":"NC","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"ba97ee84c92baed4a77376be"},"url":"https://jobsearcher.com/jobs/ba97ee84c92baed4a77376be"}}