Description
A high-performance fintech environment in Estonia is looking for a Senior Fintech Reliability Engineer to strengthen the availability, resilience, and observability of systems supporting critical financial services.
This role is highly technical and focuses on engineering reliability into the platform rather than treating production stability as an afterthought. You will work alongside software, infrastructure, security, and product teams to make complex systems more resilient at scale.
Engineering Responsibilities
- Design and implement reliability improvements across critical fintech platforms.
- Define and maintain service-level objectives, indicators, and error budgets.
- Develop sophisticated monitoring, alerting, logging, and observability solutions.
- Investigate complex production incidents and identify systemic causes.
- Build automation that reduces manual operational work and improves service consistency.
- Improve application performance, scalability, fault tolerance, and recovery capabilities.
- Work with software engineers to incorporate reliability principles into application architecture.
- Conduct capacity planning and performance analysis for high-volume services.
- Design and test disaster-recovery and failure-management strategies.
- Develop tooling for automated health checks, deployment validation, and operational response.
- Participate in architecture reviews with a focus on resilience and operational risk.
- Share reliability practices across engineering teams and mentor other engineers.
Candidate Profile
- 6+ years of experience in SRE, platform engineering, DevOps, infrastructure engineering, or distributed-systems engineering.
- Strong programming skills in Go, Python, Java, Rust, or another production-grade language.
- Deep understanding of Linux, networking, distributed systems, and cloud infrastructure.
- Experience with Kubernetes, containers, CI/CD, and infrastructure-as-code.
- Strong knowledge of observability platforms and production monitoring.
- Experience managing incidents in high-availability environments.
- Strong understanding of scalability, resilience, performance engineering, and automation.
Reliability Environment
Kubernetes, cloud infrastructure, distributed services, infrastructure-as-code, CI/CD, metrics, tracing, structured logging, automated remediation, performance testing, and high-availability financial platforms.
Why This Opportunity
You will help engineer the reliability foundations that allow fintech platforms to operate consistently as transaction volumes, integrations, and product complexity grow.
Ideal Candidate: A hands-on reliability engineer who enjoys solving difficult production problems and building systems that remain dependable under demanding conditions.
Are you interested in this position?
Apply by clicking on the “Apply Now” Button below!
#JobsHubEstonia #GlobalRecrument
#CareerOpportunities #HiringNow
#JobSeekersNetwork #EstoniaJobs
#RecruitmentServices #EmploymentPortal.