Site Reliability Engineer, Observability
职位要求 / 描述
<div class="content-intro"><p><span style="font-weight: 400;">At Ripple, we’re building a world where value moves like information does today. It’s big, it’s bold, and we’re already doing it. Through our crypto solutions for financial institutions, businesses, governments and developers, we are improving the global financial system and creating greater economic fairness and opportunity for more people, in more places around the world. And we get to do the best work of our career and grow our skills surrounded by colleagues who have our backs. </span></p> <p><span style="font-weight: 400;">If you’re ready to see your impact and unlock incredible career growth opportunities, join us, and build real world value.</span></p></div><div> <p>At Ripple, we’re building a world where value moves like information does today. Through our crypto solutions for financial institutions, businesses, governments, and developers, we are improving the global financial system and creating greater economic fairness and opportunity for more people, in more places around the world.</p> <p>Ripple Treasury, now a Ripple solution acquired in 2025, marks a significant expansion into the multi-trillion-dollar corporate finance arena. With more than 40 years of experience supporting some of the world’s largest and most sophisticated companies, Ripple Treasury integrates a treasury command center into Ripple’s technology stack—giving corporates the ability to move, manage, and optimize liquidity in real-time, across traditional and digital assets, under one expanded umbrella.</p> <p><strong>THE WORK:</strong></p> <p>This is an engineering-first role with a coaching dimension—not the other way around. You will spend the majority of your time doing hands-on observability and reliability engineering work: building instrumentation, designing alert configurations, authoring Terraform, and troubleshooting production systems. Alongside that, you will coach and consult with stream-aligned product teams, helping them build operational maturity over time.</p> <p>You will join Ripple’s Technical Operations team and work across Azure (80%) and AWS (20%) environments supporting infrastructure that is predominantly Windows-based (80%), handling significant payment volume for enterprise treasury customers. The incident management program you will help build is early-stage—you will be establishing practices, not inheriting a mature playbook.</p> <p><br><strong>WHAT YOU’LL DO:</strong></p> <ul> <li> <h2><span style="font-size: 12pt;"><strong>Observability Engineering</strong></span></h2> <ul> <li>Design and implement monitoring, alerting, and dashboards in New Relic (APM, Infrastructure, Logs, Synthetics) across Azure and AWS; write NRQL queries for troubleshooting, analysis, and reporting.</li> <li>Define and implement SLOs/SLIs and error budgets; coach teams on using them to balance feature velocity with reliability and communicate system health to stakeholders.</li> <li>Lead alert noise reduction and signal quality engineering—tune thresholds, eliminate false positives, and ensure every alert is actionable.</li> <li>Optimize observability costs through log ingestion management, pipeline rules, and New Relic configuration governance.</li> <li>Partner with engineering teams to improve observability maturity: structured logging, metrics instrumentation (RED/USE methods), distributed tracing, and effective dashboard patterns.</li> </ul> <h2><span style="font-size: 12pt;"><strong>Infrastructure & IaC</strong></span></h2> <ul> <li>Develop and maintain Terraform infrastructure as code for provisioning and managing monitoring resources, alert configurations, and observability infrastructure—this is a primary engineering responsibility, not an occasional task.</li> <li>Establish and enforce IaC governance standards for observability infrastructure across teams, providing a repeatable, auditable model for how monitoring resources are managed.</li> <li>Author and troubleshoot Azu
技能关键字
职责方向
相似岗位
Cloud Infrastructure Engineer AlchemyStaff Infrastructure/DevOps EngineerSuperstateSenior Engineering Manager, Cloud InfrastructureMeshSenior Site Reliability EngineerChainalysis数据来自公开渠道整理,薪资为公开 JD 或聚合估算,仅供参考,以面试谈薪为准。 ← 返回链聘 ChainHire 职位看板