Submit Resume

Senior DevOps Engineer

  • Texas, El Paso

  • 08/10/2026

  • Contract

  • Active

Job Description:

  • Job Summary
    Client is seeking a Senior DevOps Engineer for a 3-month contingent engagement supporting the High Performance Computing (HPC) and Electronic Design Automation (EDA) infrastructure team. This role operates within the IT Datacenter (ITDC) organization and requires independent senior-level execution with minimal ramp-up time. The ideal candidate brings strong hands-on experience with Linux HPC environments, infrastructure automation, SLURM workload management, datacenter migration, and enterprise identity integration.

    Key Responsibilities
    • HPC / EDA Platform Operations
    • Support SLURM HPC environments including partition configuration and migration planning
    • Plan and execute datacenter migrations for compute and storage infrastructure
    • Develop migration strategies and author MOPs/runbooks for infrastructure changes
    • Coordinate cross-functionally with EDA, storage, and IDAM teams
    • Verify service continuity following migrations
    • Define HPC storage tiers and gather workload requirements
    • Linux Systems Engineering & OS Deployment
    • Administer SLES systems in production HPC environments
    • Build custom OS images using Kiwi NG
    • Enable bare-metal provisioning via RackN/Digital Rebar Provision
    • Provision VMware vSphere VMs for HPC workloads
    • Troubleshoot Linux HPC services and system daemons
    • Automation & Infrastructure as Code
    • Develop Ansible playbooks for Linux system setup and configuration
    • Ensure multi-version compatibility across SLES versions
    • Manage Git repositories and migrate artifacts to Artifactory
    • Contribute to GitHub repositories and conduct reviews
    • Drive changes via ServiceNow workflows
    • Identity & Access Management
    • Integrate enterprise identity systems for Linux/HPC environments
    • Audit UID/GID data across domains
    • Validate authentication methods and extend SSSD authentication
    • Monitoring, Logging & Operational Readiness
    • Implement log management strategies including Splunk integration
    • Investigate production issues in Linux services
    • Produce technical documentation in Confluence

    Required Qualifications
    • 5+ years of experience in DevOps, Platform Engineering, or Linux Systems Engineering
    • Hands-on HPC cluster administration experience with SLURM or equivalent workload managers
    • Experience supporting EDA or scientific computing environments
    • Strong Ansible automation skills with production-grade playbook development
    • Experience with bare-metal provisioning tools (RackN, Cobbler, or equivalent)
    • Proven ability to plan and execute datacenter migrations with minimal disruption
    • Familiarity with enterprise Linux identity/authentication stacks (SSSD, LDAP, AD, NIS, Okta)
    • Experience with NetApp or comparable enterprise storage platforms
    • Ability to author technical documentation (MOPs, runbooks, diagrams)
    • Strong written and verbal communication skills

    Preferred Qualifications
    • Experience with SUSE Linux Enterprise Server (SLES) 12/15 in enterprise environments
    • Familiarity with RackN/Digital Rebar Provision for OS deployment
    • Hands-on experience with Kiwi NG for OS image creation
    • Knowledge of VMware vSphere for HPC VM provisioning
    • Experience migrating artifacts to Artifactory
    • Background in semiconductor, storage, or high-tech manufacturing IT environments

    Certifications
    • Relevant Linux, DevOps, or HPC certifications preferred (e.g., RHCSA, RHCE, VMware, Ansible)

.

.

.