Sr. Site Reliability Engineer
Comtech is a woman-owned small business founded in 1998 and headquartered in Reston, VA. We offer IT solutions across the disciplines of program/project management, applications development, infrastructure, Cyber security, and enterprise content/data management services. We have developed our methodologies and processes based on the IT Infrastructure Library (ITIL) v.3 Framework across enterprise infrastructure operations. These methodologies and processes are reinforced through our organization’s externally accredited certifications, which include ISO 9001:2008 Quality Management System (QMS), ISO/IEC 20000-1:2011 IT Service Management Systems (SMS, corporate ITIL certification), ISO 27001:2005 Information Security Management System (ISMS), and CMMI-DEV Level 3.
Job Description
Sr. Site Reliability Engineer
Location – Seattle, WA
Duration – 12 months
Interview – in-person if local or Phone + Skype
Minimum Requirements
30% C# / Java coding project experience involving API and code/bug/error handling.
Platform experience with Chef, Puppet, Azure, Ansible; hands‑on implementation of at least two; language skills required in Python, PowerShell, Ruby, or Perl (at least two).
Infrastructure experience – Windows/Linux/Unix server management, network and security protocols, load balancing and system engineering support.
Agile and project methodology knowledge.
Roles & Responsibilities
As a Sr. Site Reliability Engineer, you will be responsible for the day‑to‑day maintenance and administration of Internet‑based enterprise systems. You will identify root causes of operational issues, resolve them, and develop tools and scripts to facilitate maintenance and administration.
This position works closely with other teams to document enterprise infrastructure and monitoring systems, and plans and executes small to large‑scale projects under the manager’s direction.
We are looking for a technical expert with deep proficiency in enterprise‑scale systems and next‑gen cloud‑native applications. If you believe a cup of coffee can change a life and change our world, join us to deliver that experience to customers worldwide.
Must Haves / Nice to Haves
Experience in high‑capacity, highly scalable mission‑critical web serving environments.
Proven ability to participate with other functional teams in system integration and design, including writing operational specifications, test plans, and requirements management.
UNIX/Linux and Windows server experience: system installation, configuration, administration, troubleshooting, performance tuning, preventative maintenance, capacity planning, monitoring, and security procedures.
Web (IIS, Apache), .NET & Java application (Tomcat, JBoss, etc.) server expertise: installation, administration, configuration, troubleshooting, performance tuning, preventative maintenance, capacity planning, monitoring, and security procedures.
Experience with at least two scripting/programming languages (Ruby, Perl, Python, Shell, PowerShell, etc.).
Experience with configuration management platforms (Chef, Ansible, CFEngine, Puppet, etc.).
Database administration: setup, configuration, and basic DB troubleshooting skills.
Understanding of internet standards such as HTTP, DNS, FTP, SSH, HTML, XML, JDBC, ODBC, SNMP and other protocols.
Knowledge of high‑availability hardware and database systems design, implementation, cluster management, redundancy, and failover testing.
Knowledge of storage systems (SAN, NAS, RAID arrays, etc.).
Experience in hardening and maintaining secure systems (Safe Harbor or PCI experience a plus).
Network hardware architecture experience with load balancing equipment, switches, routers, and network troubleshooting.
Ability to produce system documentation: requirements, operational specifications, system architecture, test plans, and as‑built documentation.
Experience with ITIL and Service Management best practices is a plus.
Ability to build strong relationships and influence others across the organization.
Demonstrated knowledge of agile project methodologies.
5+ years experience designing, supporting, and deploying Internet‑based products or services.
4+ years operating complex, large‑scale enterprise, guest‑facing applications or websites.
#J-18808-Ljbffr