Staff Site Reliability Engineer
Full Job Description
About AlphaSense
AlphaSense removes uncertainty from decision-making with sophisticated AI-driven market intelligence. Trusted by the S&P 500 and global enterprises, our platform empowers professionals to make smarter decisions using trusted data.
About The Role
We are seeking a highly experienced Staff Site Reliability Engineer (SRE) to shape reliability, scalability, and performance at AlphaSense. You will architect core platforms, lead incident response, and drive the adoption of SRE best practices across our global engineering organization.
Your Mission
Engineer systems for mission-critical standards targeting 99.99% uptime while pioneering new tools and cultures that enable scaling. Act as a force multiplier through mentorship, architectural influence, and technical leadership in reliability.
Who You Are
- 8+ years of experience in SRE/DevOps with 3+ years at a Senior level.
- Strong background running production SaaS systems at scale.
- Proficient in Python, Go, or similar programming/scripting languages.
- Expertise with cloud platforms (AWS/GCP/Azure) and Kubernetes.
- Deep understanding of networking fundamentals (TCP/IP, DNS, HTTP/S).
- Familiarity with monitoring & alerting tools like Prometheus, Grafana, Datadog, ELK.
- Experience in advanced observability using OTEL and continuous profiling.
- Proven track record managing high-severity incidents and postmortems.
Company
AlphaSense
AlphaSense powers critical business decisions for over 6,500 leading companies globally, including major banks and pharmaceutical firms.We deliver AI-driven search and market intelligence that helps i...