{"schemaVersion":"jobsearcher.job.v1","id":"356552d2175abeadbd47b081","url":"https://jobsearcher.com/jobs/356552d2175abeadbd47b081","canonicalUrl":"https://jobsearcher.com/jobs/356552d2175abeadbd47b081","title":"Data Analyst","description":"Responsibilities:\nBuild, optimize, and operate production Apache Spark pipelines that generate attributed conversion and optimization feeds from large-scale impression and conversion data.\nDesign and evolve data processing and models across modern data lake and warehouse technologies, orchestrated with a production workflow scheduler on AWS.\nBuild and operate partner optimization feeds and integrations with external partners, including schemas, identity fields and unique IDs, file delivery, reconciliation, and SLAs.\nBuild, change, and operate production REST APIs, including PHP APIs, delivering secure, backward-compatible changes end to end (implementation, validation, authorization, data access, documentation, automated testing, and production diagnostics).\nTrace and modify behavior across controllers, database queries, templates, JavaScript, and automated tests.\nDiagnose production failures, reconcile partner-facing data, execute backfills, and manage staged releases and rollbacks.\nWrite and maintain unit and integration tests across pipelines, APIs, and feeds.\nQualifications and Education Requirements:\n7+ years of professional software or data engineering experience.\n4+ years building, optimizing, and maintaining production Apache Spark pipelines, with strong Java and Python skills across JVM Spark and PySpark.\nAdvanced SQL, relational data modeling, and large-scale data processing experience, including production work with Snowflake, MySQL, and Iceberg.\n3+ years operating AWS data workloads using S3 and Airflow (or equivalent production workflow orchestration); container orchestration and data-catalog experience a plus.\n3+ years building, changing, and operating production REST APIs, including PHP APIs built with Symfony or a closely equivalent PHP MVC framework.\nDemonstrated proficiency using AI development tools to deliver production-quality work efficiently — including AI-assisted coding, code review, documentation, and investigation, working with agentic code harnesses, and spec-based (spec-driven) AI development — with the judgment to know when to rely on AI output and when not to.\nAbility to independently deliver secure, backward-compatible API changes, including implementation, validation, authorization, data access, documentation, automated testing, and production diagnostics.\nExperience implementing external provider or client integrations involving schemas, APIs, file delivery, identity fields, reconciliation, privacy-sensitive data, and SLAs.\nProduction full-stack experience with PHP, Symfony (or a comparable framework), Doctrine, Twig (or another server-rendered template system), and JavaScript.\nAbility to trace and modify behavior across controllers, database queries, templates, JavaScript, and automated tests.\nExperience writing unit and integration tests for data pipelines, APIs, and web applications.\nAbility to diagnose production failures, reconcile data, execute backfills, manage staged releases and rollbacks, and support delivery commitments.\nHands-on experience building and running big data systems, with a focus on performance, reliability, and data quality.\nExperience in ad tech or advertising measurement, such as attribution, conversions, audience and impression data, identity matching, or partner optimization feeds.\nAbility to ramp quickly and deliver with minimal onboarding in an existing, complex codebase.\nPreferred Skills:\nDirect experience building partner optimization or advertising feeds with external platforms (DSPs, publishers, or measurement partners).\nExperience with deterministic and probabilistic identity matching — unique IDs, device IDs, IP-based matching, and identity graphs.\nFamiliarity with edge/log delivery infrastructure and pixel/impression tracking.\nExperience with data lake table formats and query engines at large scale.\nExperience with privacy and compliance-driven data workflows (deletion, opt-out, suppression, data retention).\nComfort operating in uncharted territory — turning ambiguity into production systems without a detailed guide.","company":"Thoughtfocus","rawCompany":"thoughtfocus","city":"Denver","state":"CO","isRemote":false,"isActive":false,"createdAt":"2026-09-24T11:38:59.907Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Analyst","description":"Responsibilities:\nBuild, optimize, and operate production Apache Spark pipelines that generate attributed conversion and optimization feeds from large-scale impression and conversion data.\nDesign and evolve data processing and models across modern data lake and warehouse technologies, orchestrated with a production workflow scheduler on AWS.\nBuild and operate partner optimization feeds and integrations with external partners, including schemas, identity fields and unique IDs, file delivery, reconciliation, and SLAs.\nBuild, change, and operate production REST APIs, including PHP APIs, delivering secure, backward-compatible changes end to end (implementation, validation, authorization, data access, documentation, automated testing, and production diagnostics).\nTrace and modify behavior across controllers, database queries, templates, JavaScript, and automated tests.\nDiagnose production failures, reconcile partner-facing data, execute backfills, and manage staged releases and rollbacks.\nWrite and maintain unit and integration tests across pipelines, APIs, and feeds.\nQualifications and Education Requirements:\n7+ years of professional software or data engineering experience.\n4+ years building, optimizing, and maintaining production Apache Spark pipelines, with strong Java and Python skills across JVM Spark and PySpark.\nAdvanced SQL, relational data modeling, and large-scale data processing experience, including production work with Snowflake, MySQL, and Iceberg.\n3+ years operating AWS data workloads using S3 and Airflow (or equivalent production workflow orchestration); container orchestration and data-catalog experience a plus.\n3+ years building, changing, and operating production REST APIs, including PHP APIs built with Symfony or a closely equivalent PHP MVC framework.\nDemonstrated proficiency using AI development tools to deliver production-quality work efficiently — including AI-assisted coding, code review, documentation, and investigation, working with agentic code harnesses, and spec-based (spec-driven) AI development — with the judgment to know when to rely on AI output and when not to.\nAbility to independently deliver secure, backward-compatible API changes, including implementation, validation, authorization, data access, documentation, automated testing, and production diagnostics.\nExperience implementing external provider or client integrations involving schemas, APIs, file delivery, identity fields, reconciliation, privacy-sensitive data, and SLAs.\nProduction full-stack experience with PHP, Symfony (or a comparable framework), Doctrine, Twig (or another server-rendered template system), and JavaScript.\nAbility to trace and modify behavior across controllers, database queries, templates, JavaScript, and automated tests.\nExperience writing unit and integration tests for data pipelines, APIs, and web applications.\nAbility to diagnose production failures, reconcile data, execute backfills, manage staged releases and rollbacks, and support delivery commitments.\nHands-on experience building and running big data systems, with a focus on performance, reliability, and data quality.\nExperience in ad tech or advertising measurement, such as attribution, conversions, audience and impression data, identity matching, or partner optimization feeds.\nAbility to ramp quickly and deliver with minimal onboarding in an existing, complex codebase.\nPreferred Skills:\nDirect experience building partner optimization or advertising feeds with external platforms (DSPs, publishers, or measurement partners).\nExperience with deterministic and probabilistic identity matching — unique IDs, device IDs, IP-based matching, and identity graphs.\nFamiliarity with edge/log delivery infrastructure and pixel/impression tracking.\nExperience with data lake table formats and query engines at large scale.\nExperience with privacy and compliance-driven data workflows (deletion, opt-out, suppression, data retention).\nComfort operating in uncharted territory — turning ambiguity into production systems without a detailed guide.","datePosted":"2026-09-24T11:38:59.907Z","dateModified":"2026-09-24T11:38:59.907Z","hiringOrganization":{"@type":"Organization","name":"Thoughtfocus","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Denver","addressRegion":"CO","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"356552d2175abeadbd47b081"},"url":"https://jobsearcher.com/jobs/356552d2175abeadbd47b081"}}