JOBSEARCHER

Hiring for Devops And SRE Engineer - Remote

ARCHIVED
Tekone It ServicesRemoteL6 LeadApril 12th, 2026

We can't find an active application page for this role right now. It may reopen or be listed elsewhere. Use Next Steps to search for an active apply link and similar live jobs.

HelloHope you are doing great.This is Satya Chittidi from Intellectt INC; we have an immediate opportunity. Please find the below job description and if you are interested, please forward your resume to satya.c@intellectt.com or call me on 7329976982.Position: DevOps EngineerLocation: Remote (Must be able to travel to Dallas on needed basis)Candidate needs to travel to Dallas every quarter for PI planning workshops; allexpenses will be reimbursed after the travelThe Site Reliability &DevOps Engineer is accountable for the availability, reliability, andperformance of the services and platforms in a highly transactional 24x7 environment.When error budget is below the threshold/within tolerance limits, SRE works onapplication development and bug fixes activities as part of DevOps responsibilities.Role & ResponsibilitiesHelp build a Site Reliability Engineering culture by sharing best practices approaches, documentation, and code with other engineering teamsDefine and setup KPIs to monitor Error BudgetsImplement strategies to ensure Error Budgets stay above the defined-acceptance levelsDefine and implement response mechanisms when Error Budget thresholds are breachedApply automation and software to any tasks or parts of the system that would benefit from it or are performed manually;Able to troubleshoot complicated issues handling OS, Networking, Database in a cloud-based SaaS environment and handle live production incidents, debug/troubleshoot infrastructure and application issues, including development and testingMonitor application performance, take steps to improve overall application performance and stability and follow through with implementation (design, develop and test);Conduct system analysis, configuration management and develops improvements for system software performance, availability and reliability;Design, write, ship, and motivate the creation of software and systems to increase observability, product reliability and organizational efficiency;Work closely with software engineers and QAs to ensure the system is responding properly to no-functional requirements such as performance, security, and availability;Document your system knowledge as you acquire it over time, create runbooks, and ensure critical system information is readily available to those who need it;Maintain and monitoring deployment, orchestration, of the servers, dockercontainers, databases, and general backend infrastructure;Keep up-to-date with security and proactively identify, diagnose, and solvecomplex security issues.Design, Develop & Test Java, SpringBoot, GraphQL based REST/JSON WebServices deployed on AWS ECS Fargate.Design, Develop & Test Typescript, NodeJS based REST/JSON Web Services deployed on AWS Lambda.Design, Develop & Test AWS AppSync based GraphQL services.Design, Develop & Test Terraform based Infrastructure as Code scripts to automate AWS infrastructure setup