Hi
Role: Site Reliability Engineer
REMOTE
Description:
•Need in-depth knowledge of Linux, Windows Administration, setting up racks, servers
• Infrastructure as code automations in Terraform, Puppet to automate aspects migrations
•Hybrid cloud environment, do have AWS and building a private cloud.
• Different data centers are currently running on different infrastructure, need to bring under 1 central stack: consisting of HPE Performance Cluster, Windows (older and newer versions), Linux (older and newer versions), JBOSS, Tomcat servers
• Terraform IaC, Octopus, Jenkins, Puppet
• Day to day: incident management, looking at alerts (Dynatrace, Datadog),certificate management, setting up the core systems--CPU, memory, storage, adjustments/configurations. Where can we automate?
• 2 profiles: 1 more focused on building the data centers; 1 more on the SRE side, supporting current systems that are already out there. More of a Systems Engineer with Data Center focus vs SRE.
Required Skills:
• Linux
• Windows administration
• Terraform
• Puppet or similar
• Scripting: Python, Shell, BASH
• Public or Private Cloud experience (Private preferred)
Nice to Have:
• Data center consolidation experience
Required Skills :
• Linux
• Windows administration
• Terraform
• Puppet
• Ansible Scripting: Python, Shell, BASH
• Public or Private Cloud experience (Private preferred)
• Able to take lead
Thank you,
Naresh (Nick)
Email:
ni...@sapphiresoftwaresolutions.com