Site Reliability Engineer at Writer | Hybrid Hired

About the role

Site Reliability Engineer enhancing platform reliability for AI workflows at WRITER. Overseeing automated solutions and cloud infrastructure supporting high-trafficked AI systems.

Responsibilities

Automate operational tasks and infrastructure management by developing robust tools and platforms using Python, Go, or similar languages, significantly reducing manual toil across our production environment
Design and implement scalable, fault-tolerant infrastructure solutions on public cloud providers (AWS, GCP, Azure) to support WRITER's rapidly expanding, high-traffic AI platform
Own the reliability, performance, and efficiency of WRITER’s core services, defining and upholding stringent Service Level Objectives (SLOs) and Error Budgets
Own the observability stack for monitoring, logging, and alerting systems to ensure rapid detection of issues across our complex distributed systems
Lead incident response, post-mortems, and root cause analyses, applying learnings to proactively prevent future outages and build a more resilient system architecture
Collaborate closely with product and engineering teams, providing expert guidance on system design for reliability, performance, and scalability from conception through launch

Requirements

A solid 7+ years of experience in site reliability engineering, DevOps, or a similar role focused on building and operating large-scale, high-availability production systems
Deep expertise with cloud platforms (AWS strongly preferred), containerization technologies like Docker and Kubernetes, and Infrastructure-as-Code tools such as Terraform
Strong proficiency in programming languages such as Python, Java, Go for automation and monitoring
Knowledge of monitoring and logging tools (e.g., Prometheus, Grafana, ELK Stack) to maintain system health and performance
Demonstrated ability to Challenge the status quo, proactively identify systemic weaknesses, and propose innovative solutions to complex reliability problems
Excellent communication, collaboration, and problem-solving skills, with a talent for building strong relationships and Connecting with cross-functional teams
A strong sense of ownership and accountability, eager to Own mission-critical systems and drive them toward peak performance and unparalleled reliability

Benefits

Generous PTO, plus company holidays
Medical, dental, and vision coverage for you and your family
Paid parental leave for all parents (12 weeks)
Fertility and family planning support
Early-detection cancer testing through Galleri
Flexible spending account and dependent FSA options
Health savings account for eligible plans with company contribution
Annual work-life stipends for:
Wellness stipend for gym, massage/chiropractor, personal training, etc.
Learning and development stipend
Company-wide off-sites and team off-sites
Competitive compensation, company stock options and 401k

Similar roles

Browse all Devops Engineer jobs

1 hour ago

PA

Senior Software Engineer – Cloud Infrastructure, DevOps

PayPal

Senior Software Engineer at PayPal managing cloud infrastructure and DevOps solutions. Delivering complete SDLC solutions and guiding engineering teams for scalable and reliable services.

Hybrid Role

San Jose United States Devops Engineer

$143,500 - $212,850 per year

11 hours ago

VS

Senior Site Reliability Engineer

VALCE Talent Solutions

Senior Site Reliability Engineer at Diligent leading reliability, automation, and observability across cloud infrastructure. Build tools for incident response and enhance performance in fast - paced environments.

Hybrid Role

Guadalajara Mexico Devops Engineer

13 hours ago

CI

Perception Deployment Engineer

Caterpillar Inc.

Perception Deployment Engineer deploying deep learning models on embedded systems at Caterpillar. Collaborating with cross - functional teams for integration and optimization of perception modules in vehicles.

Onsite Role

Wuxi China Devops Engineer

14 hours ago

AT

Principal Site Reliability Engineer, SRE

AT&T

Principal Site Reliability Engineer at AT&T required to design scalable solutions for critical operations with minimal downtime. Collaborating with teams to monitor and improve system performance in cloud environments.

Onsite Role

Plano United States Devops Engineer

$174,100 - $261,100 per year

14 hours ago

CO

DevOps Engineer, AI SaaS

Coach4expats

DevOps Engineer managing AI SaaS infrastructure at a high - growth European company. Supporting AI model deployment and ensuring platform security and compliance with multiple systems integration.

Hybrid Role

Europe Devops Engineer

16 hours ago

LE

Observability & DevOps Tools Engineering Manager

LexisNexis

Engineering Manager leading teams for observability platforms at LexisNexis. Owns operational excellence across software delivery lifecycle in Raleigh, NC.

Hybrid Role

Raleigh United States Devops Engineer

$118,300 - $219,800 per year

16 hours ago

RO

Reliability Engineer

Roche

Reliability Engineer optimizing site facility infrastructure and utility systems at Roche. Conducting root cause analyses and developing maintenance plans to enhance reliability and efficiency.

Onsite Role

Santa Clara United States Devops Engineer

$76,900 - $142,700 per year

16 hours ago

TL

DevOps Subject Matter Expert

The Missing Link

DevOps SME designing, implementing, and operating multi - cloud platforms for The Missing Link. Collaborating with engineering, security, and operations teams while embedding DevOps best practices.

Hybrid Role

Melbourne Australia Devops Engineer

17 hours ago

SE

Site Reliability Engineer, SRE

SESTEK

Site Reliability Engineer improving reliability of cloud infrastructure for an AI - specialized company. Taking ownership of monitoring and incident response processes in hybrid - working style.

Hybrid Role

Ankara Turkey Devops Engineer

18 hours ago

SE

Staff DevOps Engineer

Securonix

DevOps Engineer leading automation for sophisticated release/deployment pipelines at Securonix. Focused on Python, Ansible, and cloud services to enhance security operations.

Hybrid Role

Pune India Devops Engineer