Site Reliability Engineer 2
@ WEX IncSite Reliability Engineer 2
About the job
WEX's SRE team focuses on observability, incident response, reliability, and performance. Support internal stakeholders and solve complex challenges to improve service quality and operational excellence.
Requirements
- 2+ years SRE or similar experience
- Development experience with major programming language
- Experience with cloud platforms (AWS, Azure, GCP)
- Knowledge of observability and logging tools
- Strong communication skills
Qualifications
- BA/BS in Computer Science or related field
- Experience with containers like Docker or Kubernetes
Full job description
About the Team & Role
The WEX Site Reliability Engineering (SRE) team is looking for individuals passionate about developing software and solutions focused on observability, incident response, reliability and performance, operational excellence, and compliance. The team will be part of the Platform Reliability organization which supports our internal stakeholders included in our Corporate Payments line of business. As part of the Platform Reliability organization you’ll have the opportunity to solve complex challenges and improve the quality of life of our engineering teams as well as our ability to service our customers.
The successful candidate should have a strong aptitude for learning new technologies and the ability to drive complex and meaningful projects to a conclusion. Tight-knit collaboration with the engineering teams and an ability to thrive under pressure are key skills required to succeed in this role.
How you'll make an impact
Willingness to dig deep into code, networking, operating systems, and/or storage solutions to solve complex issues
Develop automation and utilize monitoring tools to ensure system reliability
Participate in incident response and troubleshooting
Participate in 24x7 Site Reliability rotations and escalation workflows
Identify and address performance bottlenecks. This will include code optimization, configuration changes, or infrastructure upgrade recommendations.
Collaborate with development teams to ensure software design meets operational requirements
Continuously improve processes and procedures to increase system reliability and efficiency
Stay up-to-date with the latest industry trends and technologies
Experience you'll bring
2+ years of hands-on experience as a Site Reliability Engineer or equivalent role
2+ years of development experience with at least one major programming language
Experience with Cloud Computing platforms (AWS, Azure, GCP)
Ability to thrive in a fast paced, development and operations world
Strong communication and collaboration skills
Experience with observability and logging technologies
Experience with at least one major RDBMS and NoSQL data store
Experience with containerization technologies such as Docker or Kubernetes
BA/BS degree in Computer Science or related technical field, or equivalent job experience
Nice to haves
Experience with one or more of the following languages: C#, Java, GoLang, Python
Experience with infrastructure as code, preferably Terraform
Working knowledge in building and designing RESTful APIs.
Experience with Grafana and Splunk
Familiarity with Agile methodologies and practices
Experience with GitOps
Similar jobs in Portland, ME
- W
Software Development Engineer 2- Data Engineering
WEX Inc · Portland, ME
Posted 3 weeks ago - C
Engineering Analyst, NERC Compliance
Constellation Energy · Limington, ME
Posted 4 weeks ago - 2
Manufacturing Automation Engineer II
2010 Abbott Diagnostics Scarborough, Inc. · Scarborough, ME
Posted 1 week ago - 2
Principal Manufacturing Molding Engineer
2001 Alere Inc. · Scarborough, ME
Posted 3 weeks ago - E
Assistant Site Manager
Express Wash Operations LLC · Scarborough, ME
Posted 1 day ago - W
Staff Software Engineer - Semantic Foundation
WEX Inc · Portland, ME
Posted 2 days ago