Dear friends,
Please respond with suitable profiles with contact details, visa copy, DL, Passport number and lilnkedin for the below urgent requirement.
Position: cloud operations engineer
Location: phoenix, AZ(Hybrid/remote)
Type: long term contract
Rate: open
Position summary:
Experience with administrating or setting up observability/monitoring in tools such as Prometheus, elastic, Dynatrace, and some level of programming ability (python and or Java as an example) and experience with CI/CD.
1. Prometheus administration
2. Elastic/OpenSearch administration
3. CI/CD experience
4. Ability to code in Python or Java
5. Ansible/Jenkins experience (possibly already accounted for in the original request/requirements)
Required Skills: Dynatrace AppMon, DevOps, SRE
• Dynatrace AppMon DevOps SRE Administrating/ setting up observability/monitoring tools such as prometheus, elastic, dynatrace, and programming ability (python and or Java) and experience with CI/CD
• Accountable for creating application and infrastructure performance & Chaos test plans/models for a highly scalable, low-latency, highly-available and high-throughput payment processing system.
• Perform chaos testing on a distributed system in order to build
Roles & Responsibilities
• Work with the architects and development team to ensure proper metrics instrumentation done in software components, to help facilitate real time and remote troubleshooting/performance monitoring.
• Evaluate, develop and execute load test tools to stress the limits of AMEXs most critical payment applications.
• Defining Performance Strategy and reporting performance baselines required to certify Go-Lives. Should have strong experience in handling multiple tasks and stakeholders.
• Drive end to end performance & Chaos test activities.
• Help optimizing system components such as CPU/Memory/Disk/Network & OS/Application software for maximizing the computer resources utilization.
• Develop automated chaos testing in pre-production systems
• Support in triaging and troubleshooting of issues related to performance degradation incidents in production environment
• Monitor application performance, optimize performance bottlenecks and usage to create an application capacity model.
• Analyze complex problems in the application space relating to resilience
Regards,
Radha Venkatraman
Recruitment Lead
ReqRoute,Inc
Desk: 408-600-2008; Fax: 888-400-2698
Email: ra...@reqroute.com