Pay Rate Range: $ 40.71 - 41.91/hr.
Site Reliability Engineer
Job Description:
Relevant Experience
(in Yrs)
• 5+ years of experience as SRE and knowledge of Platform ( AWS/Kubernetes )
• 3+ years of Telemetry experience, Obsessive elimination of Single points of failure, Application config standards.
Technical/Functional Skills
• Deep knowledge of platform (AWS/ Kubernetes etc) as platform engineer
• Bridge between Platform and app engineering/ partners with application SRE
• Telemetry, Obsessive elimination of single points of failure, application config standards
• Works with enterprise platform, network, storage, etc. and external vendor teams to ensure Upgrade planning, platform migrations, app config standards, and high HA
• Skill set is high in monitoring tools such as Grafana, Data Dog, EAPM, Splunk
Roles & Responsibilities
• Primary roles have been to build telemetry for business and to support engineering in the same.
• One to one dotted line relationship with Eng Leader, And Service Director over a suite of applications/ services.
• High attention to reliability engineering, single points of failure, infrastructure capacity and tuning, and related telemetry/ trend analytics.
• Defines and measures SLA/SLO/SLI.
Generic Managerial Skills
• Ability to interact comfortably with all levels of team/organization, vendors and third parties
• Architecture or other technical experience (DevOps)
• Demonstrated ability to achieve goals in a matrix environment
• Demonstrated ability to work collaboratively and influence others
• Experience with integration testing/automation practices
• Understanding of LeSS Framework and product focused mindset
Experience Required: 6-8 years
Skills:
Category
Name
Importance
Experience
SkillCategoryTest1_MN
Business Analysis
Yes
1
>7 years
SkillCategoryTest1_MN
Digital : Site Reliability Engineering (SRE)
Yes
1
>7 years