Job Description
We are seeking a skilled IT Systems Engineer with deep expertise in Linux Administration, Disaster Recovery (DR), and Business Continuity Planning (BCP). This role is responsible for managing and optimizing our Linux-based infrastructure, ensuring robust data protection, and supporting business continuity initiatives across our enterprise environment.
The Senior Linux Administrator lead automation efforts and serve as a key escalation resource for complex system and application issues. This is a high-impact role with significant ownership and cross-team visibility.
What you'd be doing:
Installing, configuring, and maintaining their RHEL environment with a focus on performance, security, and availability. Building and maintaining automation with Bash, Python, and Ansible or Terraform. Acting as the expert escalation resource for kernel panics, virtualization failures, and complex networking issues. There's also a backup and disaster recovery component, including runbook documentation and recovery testing.
What they're looking for:
7+ years of hands-on Linux administration, expert level RHEL including kernel tuning, storage configuration, and network stack management, proficiency in Bash and/or Python, and production experience with Ansible, Terraform, or similar automation tooling. Solid networking fundamentals in a data center context. Rubrik, Zerto, or other backup and DR tool experience is a plus.
CORE RESPONSIBILITIES
• Linux Administration: Install, configure, and maintain Linux systems (e.g., RHEL, CentOS), ensuring optimal performance, security, and availability across all environments.
• System Monitoring: Monitor system health proactively, troubleshoot performance degradations, and implement tuning optimizations to consistently meet SLA targets.
• Security Management: Apply security best practices, manage user accounts and access controls, and maintain a rigorous software patching and vulnerability remediation cadence.
• Automation: Develop and maintain automation scripts and workflows using Bash, Python, and Ansible/Terraform to drive consistency, reduce operational toil, and accelerate deployments.
• Advanced Support: Provide expert-level troubleshooting for application issues, kernel panics, virtualization failures, and complex networking problems across Linux environments.
• Vendor Management: Interact, coordinate and at times lead efforts that involve one or more vendors
• Backup & Disaster Recovery: Develop, maintain, and regularly test backup and disaster recovery strategies — ensuring documented runbooks and validated recovery time objectives (RTO/RPO).
• Business Continuity Planning: Collaborate with stakeholders to build and refine business continuity plans, minimizing downtime risk and ensuring rapid, coordinated recovery from disruptions.
Pay Rate: $60-70/hour
We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/.
Required Skills & Experience
• Experience: 7+ years of in-depth hands-on Linux Administration with a strong understanding of industry best practices, automation, backups, disaster recovery, and business continuity.
• Work Style: It’s crucial that the person in this role is a "self starter" and "motivated" individual who is a team player and communicates well
• Linux Expertise: Expert-level proficiency with RHEL and CentOS, including kernel tuning, storage configuration, and network stack management.
• Scripting: Proficiency in Bash and/or Python for system automation and operational tooling development.
• Automation Tools: Proven experience authoring and maintaining Ansible playbooks, Terraform, or other automation tools in production environments.
• Backup Tools: Demonstrated experience with on-premise and cloud-based backup services, cloud-to-on-premises recovery scenarios, and determining best fit RTOs and RPOs
• DR / BCP Ownership: Experience writing DR documentation, conducting tabletop exercises, and executing recovery validation tests.
• Networking: Solid understanding of TCP/IP, DNS, routing, firewalls, and VPN concepts in a data center context.
• Monitoring: Familiarity with monitoring and observability platforms such as Dynatrace, Nagios, Prometheus, or Zabbix.
Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.