3+ years of experience in Linux systems engineering, SRE, DevOps, or infrastructure support.
Strong Linux administration skills, including systemd, package management, permissions, filesystems, log
analysis, and performance troubleshooting.
Good understanding of networking fundamentals such as DNS, NTP/PTP, routing, and general host
Experience with automation, scripting, and operational tooling.
Familiarity with Kubernetes, virtualization, or clustered platforms.
Experience with configuration management or infrastructure-as-code tools such as Ansible, Salt, or
Ability to troubleshoot production issues methodically and communicate clearly during incidents.
Experience with Git-based workflows and maintainable documentation.
Hands-on, practical problem solver with a strong ownership mindset.
Comfortable working close to production and balancing support with continuous improvement.
Collaborative communicator who works well across compute, storage, networking, and application teams.
Experience with Slurm, HPC-style environments, GPU infrastructure, or researcher-facing Linux platforms.
Working familiarity with shared storage clients such as NFS, autofs, or GPFS / IBM Storage Scale from a
host and application perspective.
Experience with observability tools such as Prometheus, Grafana, or equivalent platforms.
Exposure to identity and access services such as LDAP, Kerberos, SSSD, or PAM.
Exposure to on-premises datacenter operations, hardware lifecycle support, or vendor escalations.
Interest in using AI/ML techniques for infrastructure optimization, anomaly detection, or predictive