At Swissquote, we’re all in.All in to shake things up. All in to build the bank people actually want to use. All in to make finance less boring — and a lot more powerful. We’re Switzerland’s leading digital bank — 1,400+ people across Europe, the Middle East and Asia, building real financial solutions for over a million clients worldwide. From trading and investing to everyday banking, we cover the full picture. We move fast, but we build things we’re genuinely proud of. Have a look behind the scenes by checkingHumans of Swissquote on Instagram. Growing fast creates room. Room to try things, own things, and grow at a pace most places can’t offer. Whether you like the spotlight or prefer to just put your head down and do great work, there’s space for both here. We’ve ditched the dress code — but never the chance to celebrate. Big win or small, we make it count. The kind of place where the atmosphere takes care of itself. As an equal opportunity employer, we welcome candidates from all backgrounds, experiences and perspectives to join our team and contribute to our shared success. Job Description We are looking for a hands-onSite Reliability Engineer (SRE)to help ensure the reliability, availability, scalability, and performance of our production eFX/Crypto platforms. The role has a strong focus onKubernetes, production operations, automation, observability, and infrastructure reliability. You will work closely with Infrastructure and Development engineers to operate and continuously improve our production environment across on-premise and private cloud infrastructure. Key Responsibilities Monitor production systems, respond to incidents, and ensure optimal uptime and performance of applications and infrastructure. Operate, maintain, and troubleshootKubernetes clustersand containerized applications in production,and on-premise infrastructure. Contribute to the design and implementation of scalable and reliable infrastructure solutions. Automate IT and operational processes to reduce manual work and improve efficiency. Implement and maintain CI/CD pipelines and deployment automation. Monitor and improve system observability using tools such asPrometheus, Grafana, and Elastic/Kibana. Analyse faults, perform root-cause analysis, and implement corrective and preventive actions. Perform stress, resilience, disaster recovery, and BCP testing to validate platform stability. Work closely with Infrastructure and Development teams to improve production readiness and system reliability. Participate in theon-call rotationand provide operational support for production systems. Continuously improve operational processes, automation, monitoring, and reliability practices. Qualifications 3+ years of experiencein SRE, DevOps, Systems Engineering, Platform Engineering, or a similar production engineering role. Strong hands-on Kubernetes experience in production environments — mandatory. Solid understanding of Kubernetes architecture and day-to-day operations, including: Services mesh and Istio Users management Health checks and probes Kubernetes networking and DNS Troubleshooting and performance issues Linux administration and troubleshooting skills. Experience with Docker and container technologies. Hands-on experience withon-premise infrastructure and/or private cloud environments. Experience with Infrastructure as Code and automation tools such asTerraform and Ansible. Experience with CI/CD pipelines. Experience with observability and monitoring tools such asPrometheus, Grafana, Elastic/Kibana. Good programming and scripting skills inPython and Shell. Strong understanding of networking fundamentals, includingTCP/IP, DNS, HTTPS, HTTP, and load balancing. Good understanding ofSRE principles, including SLIs, SLOs, SLAs, availability, reliability, and service performance. Good understanding of Java web applications and web servers such asSpring Boot and Apache Tomcat. Fluent English, both written and spoken, with willingness to learn French. Nice to Have CKA or CKAD certification. Experience withHelm, ArgoCD, Github Actions, or Kubernetes operators. Knowledge ofApache Kafka or RabbitMQ. Knowledge ofMT4/MT5. Experience withWindows Server 2016/2022. Experience ineFX, trading, cryptocurrency, banking, or other financial services. Experience working with high-availability or low-latency production systems. Who You Are: A team player who loves helping colleagues and bridging gaps between departments, Detail-oriented and meticulous in your work, Fluency in English (spoken and written) and willing to learn french. Passionate about technology and eager to stay up-to-date with AI innovations. Additional Information Please note that Swissquote never requests sensitive personal information or payment of any kind during the recruitment process. Any such request is fraudulent. SQ2 By clicking the link above or any third-party link within this posting, you are leaving this site and going to a third-party website where the third-party website's terms and privacy policy apply #J-18808-Ljbffr
Site Reliability Engineer - eFX/ Crypto
Site Reliability Engineer - eFX/ Crypto
Lichtenfels Vollzeit Kein Homeoffice möglich