Senior Site Reliability Engineer supporting mission-critical environments and ensuring automation for financial technology clients. Collaborating closely with engineers and stakeholders in Azure-based environments.
Responsibilities
Support the operation and enhancement of mission-critical environments for both new and existing clients
Align client requirements to platform capabilities; contribute to platform evolution
Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues
Schedule, optimize, and transition batch jobs to event-driven patterns using enterprise schedulers
Configure and support third-party software in client environments
Contribute to incident, problem, and change management processes
Execute disaster recovery, configuration management, and infrastructure readiness tasks
Provide weekend or on-call support as needed
Collaborate in Agile teams and take part in design discussions with clients, vendors, and stakeholders
Contribute to knowledge sharing within the Product Area
Leverage a solid foundation in ITIL practices, including problem, change, and incident management
Requirements
Bachelor’s degree in Computer Science or related field (Master’s is a plus)
3-5+ years in Site Reliability, DevOps, or Cloud Engineering roles
Expertise with Microsoft Azure; AWS exposure is helpful
Proficiency in Infrastructure as Code (IaC) using Terraform, Bicep, ARM, Ansible
Practical experience in monitoring and logging tools (Azure Monitor, Application Insights, DataDog, Log Analytics)
Experience in IdP Onboarding and configuring IdP solutions like Azure Entra, Okta, KeyCloak or PingFederate
Experience in centralizing authentication, managing user identities, and implementing secure access protocols (SAML, OAuth, OIDC)
Familiarity with SimCorp Dimension is a strong plus
Experience managing both onboarding projects and live production operations
Understanding of networking, virtualization, containerization (Kubernetes, Docker)
Comfort with Linux and Windows systems, APIs, scripting (PowerShell, Bash), and SQL
Collaborative mindset and ability to work in cross-functional teams
Interest in continuous learning and growth within your Product Area
Lead DevOps Engineer focused on AWS and Azure data platform solutions. Collaborating with teams to deliver scalable, secure, and highly available solutions.
DevOps Engineer working at GRÜN Software Group to automate and maintain stable infrastructures. Collaborating with teams to improve deployments and processes for better performance.
Linux System Administrator managing IT infrastructures for educational institutions and research. Collaborating on DevOps and HPC projects while ensuring system security and performance.
Azure SRE Engineer responsible for designing and maintaining secure, scalable Azure cloud infrastructure. Driving automation and operational excellence for leading organizations in technology transformation.
Senior Manager of Site Reliability Engineering overseeing Workday Kubernetes based platform. Leading teams while ensuring high availability and collaborating with federal agencies.
Site Reliability Engineer focusing on AWS cloud environments, SRE practices, and system reliability within GFT's team. Collaborating on cloud migrations and observability initiatives.
Senior DevOps Analyst enhancing infrastructure automation in a transformative technology firm. Collaborating on innovative projects in sectors like healthcare, finance, and utilities in Brazil.
Consultant at Minsait supporting technical decisions in infrastructure automation and developing solutions. Collaborating with teams for maintaining and evolving automation platforms.
Practical Trainee focusing on hardware reliability engineering at Sonova. Support reliability improvement initiatives and work closely with experienced engineers on real - life product challenges.