{"schemaVersion":"jobsearcher.job.v1","id":"fae39cdfda12748cc7fca783","url":"https://jobsearcher.com/jobs/fae39cdfda12748cc7fca783","canonicalUrl":"https://jobsearcher.com/jobs/fae39cdfda12748cc7fca783","title":"Site Reliability Architect","description":"Position: Senior Consultant - Site Reliability ArchitectLocation: Austin, TXType: Full-TimeAbout the company:Incedo is a global AI and data transformation specialist empowering companies to realize sustainable business impact from their digital investments by delivering ROI from AI@Scale.As a long-term partner for strategy to execution, we operate at the intersection of business and technology. Our integrated services and platforms are built on the foundation of AI & Data, digital engineering, and operations transformation, bringing deep domain expertise and full stack capabilities together.With over 4,000 people in the US, Canada, Latin America and India and a large, diverse portfolio of Fortune 500 enterprises and fast-growing clients worldwide, we work across banking & payments, wealth management, telecom, hi-tech and life sciences.Job OverviewWe are seeking a highly experienced Senior Consultant / SRE Architect to lead the strategy, design, and implementation of enterprise-wide observability and reliability frameworks supporting business-critical transaction flows across distributed systems.In this role, you will act as a thought leader and architect, driving end-to-end design, architecture, and implementation of scalable, resilient, and secure cloud-native platforms on AWS.. You will partner with engineering, architecture, and business stakeholders to define standards, influence technical direction, and implement scalable observability solutions.This is a high-impact role focused on transforming SRE maturity, improving advisor experience, and enabling proactive, data-driven operations through modern observability practices. The ideal candidate is passionate about SRE, observability, and system design, with a proven ability to drive large-scale transformation initiatives.Required Qualifications10+ years of experience in Site Reliability, Observability, Production Support, Cloud Architecture or related roles, with a strong focus on architecture and strategyDeep hands-on expertise with observability platforms such as Dynatrace, ELK, Datadog, Splunk, OpenTelemetry, JaegerStrong understanding of microservices architecture, APIs, and distributed systemsProficiency in programming/scripting (e.g., Python, Go, Java) for automation and integrationStrong hands-on experience with AWS services, including:Compute & Networking: VPC, EC2, ECS/EKS, LambdaDatabases: RDS, Aurora, DynamoDBStorage & CDN: S3, CloudFrontSecurity: IAM, KMS, Security Groups, NACLsProven experience designing multi-account, multi-region AWS architecturesDeep understanding of:Cloud networking and distributed systemsSecurity and compliance best practicesScalability, resiliency, and fault-tolerant design patternsHands-on expertise with Terraform (or similar IaC tools)Experience with monitoring and observability tools (CloudWatch, Prometheus, Grafana, etc.)Strong experience with DevSecOps principles and CI/CD pipelinesExcellent problem-solving and analytical skillsDemonstrated ability to lead cross-functional initiatives and influence technical directionPreferred QualificationsAWS Certifications (e.g., Solutions Architect - Associate or Professional)Experience working in financial services, banking, or regulated environmentsBackground in Site Reliability Engineering (SRE) practices and production support modelsKey ResponsibilitiesDesign, architect, and build cloud-native infrastructure and application services on AWSLead end-to-end infrastructure design for application platforms, microservices, and shared servicesImplement and manage Infrastructure as Code (IaC) using TerraformDesign and maintain highly available, scalable, secure, and cost-optimized AWS architecturesTroubleshoot and resolve complex infrastructure and application service issuesProvide architectural guidance and technical leadership across engineering teamsDrive adoption of DevSecOps best practices across the SDLCEstablish and enhance monitoring, observability, and alerting frameworks","company":"Value Maximizer","rawCompany":"value maximizer","city":"Austin","state":"TX","isRemote":false,"isActive":false,"createdAt":"2026-04-22T18:50:30.579Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Site Reliability Architect","description":"Position: Senior Consultant - Site Reliability ArchitectLocation: Austin, TXType: Full-TimeAbout the company:Incedo is a global AI and data transformation specialist empowering companies to realize sustainable business impact from their digital investments by delivering ROI from AI@Scale.As a long-term partner for strategy to execution, we operate at the intersection of business and technology. Our integrated services and platforms are built on the foundation of AI & Data, digital engineering, and operations transformation, bringing deep domain expertise and full stack capabilities together.With over 4,000 people in the US, Canada, Latin America and India and a large, diverse portfolio of Fortune 500 enterprises and fast-growing clients worldwide, we work across banking & payments, wealth management, telecom, hi-tech and life sciences.Job OverviewWe are seeking a highly experienced Senior Consultant / SRE Architect to lead the strategy, design, and implementation of enterprise-wide observability and reliability frameworks supporting business-critical transaction flows across distributed systems.In this role, you will act as a thought leader and architect, driving end-to-end design, architecture, and implementation of scalable, resilient, and secure cloud-native platforms on AWS.. You will partner with engineering, architecture, and business stakeholders to define standards, influence technical direction, and implement scalable observability solutions.This is a high-impact role focused on transforming SRE maturity, improving advisor experience, and enabling proactive, data-driven operations through modern observability practices. The ideal candidate is passionate about SRE, observability, and system design, with a proven ability to drive large-scale transformation initiatives.Required Qualifications10+ years of experience in Site Reliability, Observability, Production Support, Cloud Architecture or related roles, with a strong focus on architecture and strategyDeep hands-on expertise with observability platforms such as Dynatrace, ELK, Datadog, Splunk, OpenTelemetry, JaegerStrong understanding of microservices architecture, APIs, and distributed systemsProficiency in programming/scripting (e.g., Python, Go, Java) for automation and integrationStrong hands-on experience with AWS services, including:Compute & Networking: VPC, EC2, ECS/EKS, LambdaDatabases: RDS, Aurora, DynamoDBStorage & CDN: S3, CloudFrontSecurity: IAM, KMS, Security Groups, NACLsProven experience designing multi-account, multi-region AWS architecturesDeep understanding of:Cloud networking and distributed systemsSecurity and compliance best practicesScalability, resiliency, and fault-tolerant design patternsHands-on expertise with Terraform (or similar IaC tools)Experience with monitoring and observability tools (CloudWatch, Prometheus, Grafana, etc.)Strong experience with DevSecOps principles and CI/CD pipelinesExcellent problem-solving and analytical skillsDemonstrated ability to lead cross-functional initiatives and influence technical directionPreferred QualificationsAWS Certifications (e.g., Solutions Architect - Associate or Professional)Experience working in financial services, banking, or regulated environmentsBackground in Site Reliability Engineering (SRE) practices and production support modelsKey ResponsibilitiesDesign, architect, and build cloud-native infrastructure and application services on AWSLead end-to-end infrastructure design for application platforms, microservices, and shared servicesImplement and manage Infrastructure as Code (IaC) using TerraformDesign and maintain highly available, scalable, secure, and cost-optimized AWS architecturesTroubleshoot and resolve complex infrastructure and application service issuesProvide architectural guidance and technical leadership across engineering teamsDrive adoption of DevSecOps best practices across the SDLCEstablish and enhance monitoring, observability, and alerting frameworks","datePosted":"2026-04-22T18:50:30.579Z","dateModified":"2026-04-22T18:50:30.579Z","hiringOrganization":{"@type":"Organization","name":"Value Maximizer","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Austin","addressRegion":"TX","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"fae39cdfda12748cc7fca783"},"url":"https://jobsearcher.com/jobs/fae39cdfda12748cc7fca783"}}