Server Infrastructure Engineer
Introduction We are seeking an experienced professional to join our global IT team. This role requires strong technical expertise in server hardware, out-of-band management, firmware operations, and infrastructure design. The ideal candidate will operate with a high degree of independence, take ownership of complex technical initiatives, provide guidance to peers, and contribute meaningfully to architecture-level discussions and decisions. Required Skills & Qualifications10 years of experience in server infrastructure, systems engineering, or related IT roles, with a minimum of 5 years working at an enterprise levelAssociate's degree or equivalent work experienceIn-depth knowledge of x86 server hardware, specifically CPU architecture, bus architecture, and expansion cardsExpert-level knowledge of out-of-band (OOB) server management systems (e.g., iLO, DRAC) and hardware management systems (e.g., HPE OneView, HPE COM, OpenManage Enterprise)Extensive experience with Windows and Linux server operating systems; strong knowledge of TCP/IP, including DNS, DHCP, broadcast domains, and routing protocolsProficient scripting/automation skills (PowerShell, Python, Ansible, or similar); experience with CI/CD development, automated server builds, and REST API integrationsSolid experience with FMEA methodology, firmware lifecycle management, and hardware testing/certificationWorking knowledge of virtualization technologies, distributed storage systems (SAN, HDFS, SSD), and containerizationHands-on experience with enterprise monitoring systems such as Dynatrace and NetcoolDemonstrated ability to lead technical projects independently and mentor team membersStrong analytical, documentation, and problem-solving skillsExcellent communication skills, with the ability to collaborate across technical and business teamsPrior work experience at client or in client's industryApplicants must be able to work directly for Artech on W2 Preferred Skills & QualificationsKnowledge of MS Active Directory, GCP, MS Azure, Kubernetes, and other cloud-like solutionsFamiliarity with storage systems, networking, data center facilities, and overall systems architectureBasic knowledge of cloud computing Day-to-Day ResponsibilitiesProvide Global L3 hardware support and conduct in-depth root cause forensic analysis for complex server issuesServe as a subject matter resource for out-of-band (OOB) management, such as HPE iLO, HPE OneView, HPE COM, Dell iDRAC, and Dell OpenManage EnterpriseIndependently resolve complex, cross-system technical issues with minimal oversightLead scripting initiatives and develop automated server configurations to streamline operationsDesign and maintain scripts using the Redfish API to automate server discovery, configuration profiles, BIOS updates, and firmware lifecyclesGuide application owners in automating application testing and certification processesOwn firmware update methodologies and firmware version management for assigned segments of the global server fleetIdentify opportunities to improve automation coverage and firmware management practices across the teamDeliver robust server designs and lead FMEA (Failure Mode and Effects Analysis) testing effortsManage server and component testing, including planning, execution, and results analysisLead hardware-related projects such as trusted SSL certificate deployment across the fleet and MFA initiativesIdentify and troubleshoot systemic hardware issues, recommending long-term solutionsCreate and maintain comprehensive standards documentation and user documentation for OOB managementEstablish best practices and contribute to team-wide technical standardsReview and improve existing documentation for accuracy and completenessDesign, implement, and maintain enterprise servers and general-purpose hybrid/cloud systems for optimal business operationsResearch, design, and implement infrastructure solutions for internal customer needs, such as containerization, Kubernetes orchestration, and on-demand scalingDesign and configure secure virtualized environments aligned with enterprise architecture blueprints, including virtual machines, virtual networks, and distributed storage systems (SAN, HDFS, SSD, etc.)Contribute technical input into enterprise architecture blueprints and roadmaps based on hands-on engineering experienceEngineer and refine monitoring solutions, leveraging Dynatrace and Netcool for observability, monitoring, and alerting to ensure systems are configured correctly, run efficiently, and remain secure against potential threatsMonitor and analyze resource utilization (CPU, memory, storage, network traffic) using Dynatrace, Netcool, and other tools to ensure effective utilization and prevent runaway costsConfigure and maintain Dynatrace dashboards to monitor and alert on hardware firmware versions and drift from current supported versionsProactively identify and resolve bottlenecks; lead capacity planning efforts based on usage trends and Dynatrace analyticsMentor and guide the team, fostering a culture of technical excellenceLead knowledge-sharing sessions and contribute to the team's learning For immediate consideration please click APPLY to begin the screening process with Alex