Scan your resume against ATS criteria for this Senior Site Reliability Engineer - PSRE role at Arcesium.
Arcesium is a global financial technology firm that solves complex data-driven challenges faced by some of the world’s most sophisticated financial institutions. We constantly innovate our platform and capabilities to meet tomorrow’s challenges, anticipate the risks our clients encounter, and design advanced solutions to help our clients achieve transformational business outcomes.
Financial technology is a high-growth industry as change and innovation continue to disrupt the status-quo and prompt major transformation. Arcesium is at a particularly interesting time in our own growth as we look to leverage our successfully established market position and expand operations in pursuit of strategic new business opportunities. We value intellectual curiosity, proactive ownership, and collaboration with colleagues, and we empower you to meaningfully contribute from day and accelerate your professional development.
We are looking for an intelligent, resourceful, and highly skilled Senior Site Reliability Engineer (SRE) to join our Platform Site Reliability Engineering (PSRE) team . This team plays a critical role in ensuring the stability, reliability, and availability of mission-critical production applications on the Arcesium platform.
The
is responsible for:
Observability, monitoring, logging, and tracing to proactively detect and prevent issues.
Building tools and infrastructure that enhance system stability and resilience.
Troubleshooting live production issues with a deep focus on rapid incident resolution.
Governing, declaring, managing, and recovering from platform-wide incidents to minimize downtime and business impact .
As an SRE in this high-impact team , you will work under tight timelines in a high-pressure environment , where every second counts in resolving critical production incidents. This means you must be quick-thinking, highly analytical, and proactive in preventing and resolving disruptions.
, with a deep understanding of SRE principles and best practices.
Incident management expertise , including triaging, escalation, and resolution of high-severity outages .
Proficiency in at least coding language (Python or Java) for automation and debugging.
Hands-on experience in Kubernetes (K8s) for managing and orchestrating containerized applications.
C loud experience (AWS preferred) with exposure to key services like EC2, S3, Lambda, and CloudWatch.
Excellent communication skills to articulate technical challenges and solutions effectively .
Strong troubleshooting and problem-solving skills , with experience diagnosing complex production issues.
Ability to stay calm under pressure , multitask, and prioritize effectively in fast-moving environments .
Fluency in English (spoken and written) is required.
Must have the legal right to work in the country.
Experience with monitoring tools (e.g., Datadog, Prometheus, Grafana)
Familiarity with web application architectures and best practices.
Exposure to CI/CD pipelines and DevOps workflows.
At Arcesium, we offer:
Flexible work arrangements (hybrid model) and a casual dress code
Opportunity to work on challenging projects in a dynamic, global environment
Continuous learning and development opportunities
Collaborative and innovative work culture
Competitive compensation and benefits package
Modern and comfortable office located at Avenida da Liberdade (Lisbon)
Join our team and play a crucial role in shaping Arcesium's future!
Arcesium's Personal Data Privacy Notice for Candidates is linked here .
Emails from genuine Arcesium recruiters who are employees of the company will always come from the @arcesium.com domain. In some cases, you may also be contacted by independent search firms engaged to recruit on our behalf; emails from their employees should always come from their firm's applicable domain. We'll never ask for your banking information or any payment as part of the recruiting process. If something seems off or you're contacted by an unexpected third party, please reach out to us at [email protected] (US/UK), [email protected] (India) or [email protected] (Portugal/Sweden) .
Arcesium is an equal opportunity employer.
Incident Management:
Proactive Monitoring & Analysis:
Troubleshooting & Problem Solving:
Collaboration & Communication:
Automation & Optimization:
Continuous Improvement:
What we’re looking for:
Site Reliability Engineering (SRE), DevOps, or Production Engineering role
Incident management expertise , including triaging, escalation, and resolution of high-severity outages .
Proficiency in at least coding language (Python or Java) for automation and debugging.
Hands-on experience in Kubernetes (K8s) for managing and orchestrating containerized applications.
C loud experience (AWS preferred) with exposure to key services like EC2, S3, Lambda, and CloudWatch.
Excellent communication skills to articulate technical challenges and solutions effectively .
Strong troubleshooting and problem-solving skills , with experience diagnosing complex production issues.
Ability to stay calm under pressure , multitask, and prioritize effectively in fast-moving environments .
Fluency in English (spoken and written) is required.
Must have the legal right to work in the country.
Nice-to-Have Skills:
Experience with monitoring tools (e.g., Datadog, Prometheus, Grafana)
Familiarity with web application architectures and best practices.
Exposure to CI/CD pipelines and DevOps workflows.
At Arcesium, we offer:
Flexible work arrangements (hybrid model) and a casual dress code
Opportunity to work on challenging projects in a dynamic, global environment
Continuous learning and development opportunities
Collaborative and innovative work culture
Competitive compensation and benefits package
Modern and comfortable office located at Avenida da Liberdade (Lisbon)
Join our team and play a crucial role in shaping Arcesium's future!
Arcesium's Personal Data Privacy Notice for Candidates is linked here .
Recruiting Security
**Emails from genuine Arcesium recruiters who are employees of the company will always come from the @arcesium.com domain. In some cases, you may also be contacted by independent search firms engaged to recruit on our behalf; emails from their employees should always come from their firm's applicable domain. We'll never ask for your banking information or any payment as part of the recruiting process. If something seems off or you're contacted by an unexpected third party, please reach out to us at [email protected] (US/UK), [email protected] (India) or [email protected] (Portugal/Sweden) .
Arcesium is an equal opportunity employer.
How to apply for Senior Site Reliability Engineer - PSRE at Arcesium?
Click the "Apply on Company Website" button on this page to submit your application directly on the employer's official portal.
What is the salary for this role?
Salary details will be discussed during the interview.
What experience is required?
5+ yrs of experience is required.
Is this position still open?
Yes, currently active and accepting applications.
Explore related active roles in Software Engineering
Senior Site Reliability Engineer - PSRE
Arcesium · Lisbon