Confirmed Site Reliability Engineer
swan.io Paris, Ile-de-France, France · €56K–€66K/yr
Banking · 201-500 employees
About the role
You will ensure the reliability, scalability, and security of financial service platforms by managing infrastructure and responding to production incidents. Responsibilities include automating operational tasks, maintaining CI/CD pipelines, and collaborating with cross-functional teams to improve system resilience.
What they look for
Requirements
Candidates should have 2 to 4 years of experience in SRE, DevOps, or platform engineering with hands-on cloud infrastructure and automation skills. Proficiency in tools like Terraform, scripting languages, and a strong understanding of distributed systems and observability are required.
Benefits
Full description
About
Swan is Europe’s embedded banking specialist. We empower software companies to embed banking features like accounts, cards, and payments directly into their products, under their own brand.Swan processes over €2.5 billion in monthly transactions for more than 150 companies - like Pennylane, Indy, Agicap, Libeo, and Lucca.
Founded in 2019, the company has received growth capital from leading investors such as Lakestar, Accel, Creandum, Bpifrance and Eight Roads. Swan is a principal member of Mastercard and a licensed financial institution, regulated by the French banking authority (ACPR).
Our mission
Banking belongs in business software
Many software companies already serve small businesses incredibly well: helping them send invoices, run payroll, manage inventory, and more. They’re on a mission to become the central hub for managing every aspect of business life.
But when it comes to financial workflows, there’s still a gap. Too many critical tasks like managing cash flow, tracking payments, or reconciling accounts happen outside the software, across spreadsheets, email threads, banking portals.
It’s a missed opportunity. Business software shouldn’t just record financial activity — it should run it.
To learn more about us: About Swan, Our story.
Job description
Working within Swan’s Core Infrastructure team, you will help ensure the reliability, scalability, security, and performance of the platforms that support our financial services. You will take ownership of well-scoped services and operational incidents, improve observability, automate repetitive tasks, and collaborate closely with development, product, and security teams.
This is an independent engineering role for someone who has developed solid operational foundations and is ready to take greater ownership of production systems. You will contribute to incident response, infrastructure improvements, service design reviews, and the continuous improvement of our reliability practices.
Main responsibilities
On a daily basis, you will:
- Act as a primary responder for well-understood production incidents and participate independently in the on-call rotation.
- Assess the impact of incidents, including transaction volume affected, potential revenue impact, and implications for data integrity.
- Investigate operational issues using logs, metrics, dashboards, and distributed tracing, then document clear incident updates for stakeholders.
- Contribute to postmortems, update runbooks, identify recurring incident patterns, and suggest preventive measures.
- Create and maintain dashboards, alerts, and basic service-level indicators for the services you support.
- Tune alert thresholds to reduce noise and improve the quality of operational signals.
- Participate in system design reviews, with a particular focus on reliability, operability, failure modes, and production readiness.
- Implement reliability improvements such as health checks, retries with exponential backoff, circuit breakers, and appropriate monitoring.
- Manage cloud resources and contribute to Infrastructure as Code using tools such as Terraform.
- Write automation scripts and small internal tools in Bash, Python, or Go to reduce manual toil and improve operational efficiency.
- Contribute to CI/CD pipelines and automate routine maintenance tasks such as backup verification, certificate renewal, and log management.
- Participate in infrastructure code reviews and help maintain high standards for safe, repeatable changes.
- Support security and compliance activities, including PCI DSS controls, ISO 27001 initiatives, security remediation, data classification, and encryption requirements.
- Monitor resource utilisation, provide basic capacity forecasts, and implement practical cost optimisation measures such as rightsizing resources and removing unused infrastructure.
- Collaborate with development, product, and security teams to improve the resilience and operability of services.
- Provide clear handovers, maintain high-quality documentation, and communicate technical topics effectively to both technical and non-technical stakeholders.
- Use approved AI tools responsibly to support tasks such as code generation, documentation, and log analysis, while validating outputs and protecting sensitive information.
Your team
Core Infrastructure is responsible for building and operating the foundations that enable Swan’s products to remain reliable as the business grows. We work closely with development and other technical teams to improve system resilience, operational efficiency, and customer experience.
We value ownership, pragmatism, knowledge sharing, and open communication. Engineers are encouraged to challenge ideas constructively, document what they learn, and continuously improve the way we build and operate services. You will work in a supportive environment where reliability is a shared responsibility and where operational excellence is developed through collaboration.
Together alongside Engineering Productivity, our squad constitutes the broader Platform Engineering team.
Preferred experience
✨ You’re a great match if:
- You have typically 2 to 4 years of experience in Site Reliability Engineering, DevOps, platform engineering, infrastructure engineering, software engineering, or a related field.
- You have hands-on experience supporting production services and participating in an on-call rotation.
- You can independently respond to well-understood incidents, follow escalation procedures, and contribute to postmortems and runbook improvements.
- You are comfortable working with logs, metrics, dashboards, alerting, and basic distributed tracing.
- You understand the Four Golden Signals: latency, traffic, errors, and saturation.
- You have experience creating dashboards and meaningful alerts, and understand the fundamentals of SLIs and SLOs.
- You have practical experience with cloud infrastructure, ideally AWS, including compute, storage, networking, and managed services.
- You have experience with Infrastructure as Code, particularly Terraform, CloudFormation, or equivalent tools.
- You can write automation scripts in one or more languages such as Bash, Python, or Go.
- You understand CI/CD practices and have contributed to deployment or infrastructure automation.
- You have a working understanding of high availability, fault tolerance, redundancy, health checks, retries, circuit breakers, and failure recovery.
- You are familiar with event-driven architectures and message queuing technologies such as Kafka or AWS SQS.
- You understand the operational implications of distributed systems, including consistency, availability, and partition tolerance.
- You have an awareness of security and compliance requirements relevant to financial services, such as PCI DSS, ISO 27001, encryption, least privilege, and data classification.
- You are comfortable monitoring resource usage, thinking about capacity, and identifying basic cloud cost optimisation opportunities.
- You communicate clearly, document your work thoroughly, and collaborate effectively across technical and non-technical teams.
Nice to have
- AWS certification, such as AWS Certified Solutions Architect, Associate level.
- Experience with Kubernetes and container orchestration.
- Experience with observability platforms such as Grafana, Datadog, or similar.
- Experience operating services that process financial transactions or other highly sensitive data.
- Experience improving service-level objectives, performance, or capacity for production systems.
- Familiarity with Go or another programming language used for internal tooling.
- Our ideal teammate: Empathetic. Skilled. Frank. We love to challenge each other, and we leave our egos at the door.
It’s okay if you don’t tick all the boxes - don’t let imposter syndrome prevent you from applying! 🙌
Swan is committed to providing a caring work environment for all employees, regardless of age, sex, disability, sexual orientation, race, religion, or belief.
When it comes to recruitment, we’re interested in your work experience, skills, and overall personality. Because diversity makes the workplace stronger and is necessary for Swan’s success, we are intensifying efforts to incorporate concrete actions to help us improve in this area.
Our ideal teammate
You are empathetic, skilled, pragmatic, and direct. You take ownership without ego, communicate clearly during both calm and high-pressure situations, and are comfortable asking questions when context is missing.
You enjoy understanding how systems work, learning from incidents, and turning operational lessons into better automation, documentation, and engineering practices. You challenge ideas constructively, support your teammates, and leave systems more reliable than you found them.
About Swan
✨ Perks of being a Swanee:
- Meal Vouchers: We provide a meal voucher card to cover your meals on work days. 🥗
- Transport: Monthly mobility package for remote employees. In accordance with the company agreement for sustainable mobilities, you can now use your mobility package to pay for alternative commuting modes. 🚇
- Holidays : 25 days + RTT 🏝️.
- Health insurance: Alan. This is Swan's health and welfare insurance. 🚑
- Sports: Thanks to our partnership with Classpass, you can enjoy advantageous discounts on subscriptions. They offer a wide range of sports activities as well as wellness activities. The offers are valid in several European cities, as well as online. 🏋
- Well-being support: access to Moka Care for mental health and wellness. 🧘
- Hybrid Remote: Our hybrid remote policy offers the best of both worlds: a great office setting and the flexibility to work remotely with at least 3 days each month in our Parisian office. 🏡
- Offsite: Once a year we gather to reconnect, deep-dive into big topics, and relax. 🤝
- This isn’t a perk, it should be the rule, but diversity and inclusion are important at Swan. We’re working hard to get better every day.
✨ Our values:
Swan’s core values guide our actions daily. Individually, they may seem obvious, but together, they form a unique culture.
Simplicity: Leonardo Da Vinci said: “simplicity is the ultimate sophistication.” If something's convoluted or confusing, we work extra hard to break it down - Making complex things simple is what we do.
Long Term: We always play the long game, whether it's to support our partners in their growth journey, or make tangible commitments to climate action.
Excellence: We are a team of experts who consistently go all out to create pixel-perfect banking services and exceed our partners' expectations - whatever it takes.
Be Human: We believe in the power of kindness and the importance of acting with integrity. But embracing our humanity extends beyond interpersonal interactions, it means caring about greater issues that affect our planet.
You can find out more about our culture.
Recruitment process
- A 30-min call with our Talent Acquisition Manager, to get to know you, understand your career expectations and answer your questions
- An interview with our Platform Director
- A tech test & peer interview
- An interview with our CTO
Similar roles
-
Senior Site Reliability Engineer I
Axon Seattle, Washington, United States · $134K–$215K/yr
-
Senior Site Reliability Engineer
Akamai Cambridge, Massachusetts, United States · $186K–$219K/yr
-
Site Reliability Engineer II
Akamai United States · $138K–$171K/yr
-
Analista de SRE Pleno - Vaga Afirmativa para Mulheres
Experian Sao Carlos, Southeast, Brazil
-
Site Reliability Engineer (SRE) Manager
Plume Ljubljana, Slovenia
-
Senior Site Reliability Engineer (SRE)
Tubi - Canada Toronto, Ontario, Canada · CA$116K–CA$235K/yr