Akoya External
Senior SRE Engineer
Boston, MA · Senior · Contract
Sponsorship not specified$55k-$75kDetected 5 days ago
PythonGoSwiftBashRedisDynamoDBAWSCloud PlatformsKubernetesTerraformCI/CDNginxDatadogSite Reliability EngineeringPlatform EngineeringIncident ResponseFirewallVPNLeadershipCommunicationCollaborationProblem SolvingMentoringAdaptability
About the role
- driving initiatives, mentoring engineers, owning SLA/SLO adherence, leading incident management, and automating toil out of day-to-day operations.
- If you thrive in fast-paced environments and are passionate about system reliability, security, and operational excellence, we invite you to join our Platform Engineering team.
- This is a 6-month contract role, with a potential to convert to full-time on the Platform Engineering team, working closely with global development, operations, and delivery teams.
Responsibilities
- Champion robust system design and operational excellence across engineering and product teams.
- Design, implement, and maintain scalable systems for uptime, resilience, and performance, while defining and enforcing SLOs, SLIs, and SLAs with product teams.
- Own incident detection, escalation, and resolution, developing playbooks for failure scenarios and leading post-mortem analyses with corrective actions.
- Build monitoring, logging, and alerting systems, and develop automation for deployments, maintenance, and toil reduction.
- Drive CI/CD improvements and maintain operational documentation, runbooks, and best practices.
- Partner with developers on reliability by design, fostering shared responsibility for observability and fault tolerance.
- System Reliability Advocacy: Champion robust system design and operational excellence across engineering and product teams.
Requirements
- Knowledge of certificate and PKI lifecycle management, including HSM-backed key protection and mTLS.
- Proficiency in Python, Go, and Bash for automation and infrastructure-as-code.
- Hands-on experience with Datadog for metrics, distributed tracing, log management, and APM.
Nice to have
- Preferred Skills & Attributes
Skills
- Deep working knowledge across compute, networking, security, storage, and data services in multi-account, multi-region environments.
- Skilled at diagnosing and resolving complex, multi-layer issues in distributed, cloud-native environments.
- Client API integration via VPN to on-premises or cloud environments with HSM-backed mTLS.
- Route53, ELB (NLB/ALB), WAF, VPC, NAT, IGW, VGW, VPN, Lambda
- Analytical Problem-Solving: Skilled at diagnosing and resolving complex, multi-layer issues in distributed, cloud-native environments.
Compensation
- This role is a 6-month contract role, with a potential to convert to full-time, with an expected hourly rate of $55 - $75.
- The actual hourly rate offered considers the candidate's work location, relevant education, job-related knowledge, skills, and experience, among other factors.
- The actual base pay offered may take into account the candidate's work location, relevant education, job-related knowledge, skills, and experience, among other factors.
Company info
- Fosters accountability and a blameless culture focused on reliability and continuous improvement.
Apply directly at Akoya External →Create a free account for alerts like thisView Akoya External immigration profile
This listing is sourced directly from Akoya External's careers page and normalized into a canonical job model.