Site Reliability Engineer (4ECF57B)
RemoteLondon, England, United KingdomseniorFull-time
- Posted
- today
- Source
- LinkedIn (remote, Europe)
- Field
- Engineering
Skills
TypeScriptTerraformSecurityPythonGoReactAWS
Description
Referment is working with a healthcare technology company whose systems support clinicians delivering ADHD and autism assessments. The platform processes sensitive health information and sits directly in important clinical workflows, making availability, security and data integrity central to the value it provides.
This Site Reliability Engineer will take end-to-end responsibility for business-critical backend systems, from technical design and deployment through monitoring and incident response. It is a hands-on role in a small engineering environment, suited to someone who can improve cloud infrastructure without relying on a separate platform team and who understands that dependable systems have a visible effect on clinical care.
The Role
- Design, build, deploy and operate production backend systems that clinicians can rely on in day-to-day assessment work.
- Develop the AWS cloud infrastructure supporting the product and improve its reliability, scalability and operational clarity.
- Establish advanced observability, monitoring and incident-response practices, including the implementation of tracing and logging pipelines.
- Design robust APIs, databases and schemas for structured information at scale, ensuring performance under high-traffic conditions.
- Apply security and data-protection principles throughout design and delivery, including encryption, access control and automated compliance monitoring.
- Work with engineering, operational and clinical stakeholders to explain trade-offs and solve problems in an evolving startup environment.
What We're Looking For
- At least five years of relevant professional experience in a production environment.
- A track record of owning business-critical backend systems through design, deployment, monitoring and on-call support.
- Extensive experience with AWS infrastructure and services, including ECS, Fargate, Lambda, Aurora and DynamoDB.
- Proficiency in Infrastructure as Code using Terraform to manage and evolve cloud environments.
- Advanced observability skills, with experience using tools such as OpenTelemetry, Prometheus or Honeycomb to manage high-volume telemetry data.
- Experience handling high-scale data or high-traffic services (e.g. 1,000+ TPS).
- Practical experience handling sensitive personal or health information and applying GDPR and core data-protection controls.
- Security-first engineering judgement and the ability to preserve reliability and data integrity while working quickly through ambiguity.
Desirable Experience
- Go (Golang) or Python.
- React, TypeScript or Auth0.
- Security and compliance automation, including experience with frameworks like SOC 2 or ISO 27001.
- Asynchronous processing pipelines, structured-data extraction or healthcare-system integrations.
- Corporate IT fundamentals such as device management, identity and networking.
This could suit a senior backend, platform or reliability engineer who wants broad production ownership in a regulated healthcare setting. Applicants must already have the right to work in the UK; visa sponsorship is not currently available, and a background check is required before starting.
#Referment
JobMatch aggregates public listings. Always apply through the original posting.