Staff Engineer, Site Reliability

LearnUpon

Job Description

Work Mode: Flex 1+ days per week in our Dublin office Department: Engineering   ## What you'll do be a principal technical leader and a key catalyst for our infrastructure's evolution. In this role, you will take ownership of our core platform resilience, driving the strategy to build out an advanced, cost-effective observability function spanning metrics, logs, and transaction tracking. This opportunity requires a strategic thinker who can design cross-team ### Slo/Sli frameworks, navigate complex distributed system ## What you'll bring , and mentor talent to ensure LearnUpon scales efficiently to support our ambitious global goals. In addition, you’ll be responsible for: - Infrastructure Optimization: Identify opportunities to improve and scale our infrastructure for performance, observability, maintainability, and cost, by creating innovative solutions. - Observability Function Strategy: Lead our efforts to build an observability function that incorporates application metrics, application transaction tracking, and event log management. - Resilience & Scaling: Drive the processes to maintain resilient, scalable, and cost-effective infrastructure while working with other Engineering teams to provide solutions that meet their ongoing requirements. - Tooling & Self-Service: Build tools focused on measuring, monitoring, and alerting, with an eye towards self-service in order to promote Engineers’ ownership of observability. - Operational Agility & Support: React quickly to changing customer and business needs and actively participate in the team's on-call rota.  Team Up-Leveling: Mentor junior talent and effectively communicate complex technical ideas to both technical and non-technical peers.   Skills & Experience  Must-Haves                                                                          - 7+ years of experience in a software or Ops role. - 5+ years of cloud engineering experience, with at least 2 years of experience with AWS. - Experience deploying Microservice environments using containerisation technologies such as Kubernetes and Docker. - Experience designing and implementing Observability tech stacks, championing its …

Requirements

See the listing for full requirements.

Hybrid • Dublin
full time
Posted 5 days ago

Related jobs

More roles you might like

View all

Work Mode: Hybrid 3+ days per week in our Dublin office Department: Partnerships   ## What you'll do be the single-threaded owner of the entire funnel for partner-generated and partner-influenced revenue. In this role, you will transition the partnership function into a…

Hybrid • Dublin
full time
5 days ago

Work Mode: Hybrid 3+ days per week in our Dublin office Department: Customer Success   ## What you'll do partner with our enterprise customers to drive adoption, deliver measurable success, and secure long-term value. Acting as a trusted advisor, you will combine consul…

Hybrid • Dublin
full time
5 days ago

Work Mode: Hybrid 3+ days per week in our Dublin office Department: Finance   ## What you'll bring                                …

Hybrid • Dublin
full time
5 days ago