Senior SRE Engineer (m/f/d)
About the job
At SciSure, we're building software that helps laboratories and research organizations do their best work. As our platform grows, so does the importance of the infrastructure behind it.
We're looking for a Senior Site Reliability Engineer (m/f/d) who enjoys solving complex infrastructure challenges, automating everything that should never be done twice, and making systems more reliable every day. You'll help shape how we run our cloud infrastructure while also improving and modernizing parts of our existing environment. We are strongly data- and automation-oriented and value continuous delivery practices. We make data-driven decisions, treat manual toil as an engineering problem, and value engineers who simplify complex workflows by understanding the underlying requirements and aligning improvements with all stakeholders.
Responsibilities
Build and evolve our platform
Design, build, and maintain reliable, scalable infrastructure on AWS.
Manage and continuously improve our Kubernetes platform (primarily Amazon EKS).
Develop and maintain Infrastructure as Code using Terraform to ensure repeatable, consistent deployments.
Improve deployment pipelines and help engineering teams ship software safely and efficiently.
Automate and simplify
Identify repetitive or manual processes and replace them with scalable automation.
Challenge existing ways of working by understanding the real problem, proposing simpler solutions, and bringing stakeholders along with the change.
Continuously improve operational workflows to reduce complexity and increase reliability.
Keep our systems healthy
Maintain and modernize Windows Server environments, including Active Directory, LDAP, patching, and user management.
Strengthen our observability by improving monitoring, alerting, and incident response capabilities.
Use data, metrics, and service levels to drive operational improvements instead of relying on assumptions.
Own reliability
Take part in the on-call rotation and help resolve production incidents when they happen.
Balance project work with operational support, ensuring issues are followed through to resolution.
Work closely with Engineering and IT to build resilient systems that enable our customers to rely on SciSure every day.
About you
Experience & Expertise
You have 8+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
You have built and operated production environments on AWS.
You have experience balancing project work with operational support and on-call responsibilities.
You have professional proficiency in English, both written and spoken.
You hold a Bachelor's degree in Computer Science, Information Technology, Software Engineering, or a related field (or equivalent professional experience).
You are located within the European Union.
You have the legal right to work in the European Union.
Technical Skills
Strong AWS knowledge (EC2, RDS, S3, IAM, networking, VPNs, multi-account environments).
Hands-on experience with Terraform and Infrastructure as Code.
Kubernetes experience, preferably Amazon EKS, including Helm.
Comfortable working with Windows Server, Active Directory, LDAP, patching, and user management.
Experience with observability tools such as Grafana, Prometheus, Zabbix, Alloy, New Relic, Datadog, or similar.
Mindset
You take an automation-first approach and enjoy eliminating manual work.
You simplify complex processes by understanding the root problem and proposing better solutions.
You use data and service levels to drive decisions and improvements.
You take ownership and see problems through to resolution.
You communicate well, collaborate across teams, and enjoy improving how things work.
Nice to Have
Experience modernizing on-premise or legacy infrastructure.
AWS certification (Solutions Architect – Associate or higher) and/or CKA .
Experience with release orchestration tools such as Octopus Deploy.
Experience administering Windows/IIS application pools (.NET Framework and .NET Core).
Experience administering MariaDB/Galera clusters, including backup, restore, and recovery.
Experience working in highly regulated or SaaS environments.
Experience improving developer experience and CI/CD pipelines.
What we offer
Meaningful ownership and direct impact on company growth.
Competitive, market-aligned compensation based on experience and track record.
Fully remote position.
Flexible, trust-based work environment with the ability to work abroad up to 30 working days per year.
30 days of paid vacation leave per year.
Support for courses, certifications, and conferences.
Yearly on-site events in Europe to spend time as a team.
Apply for the job
Do you want to join our team as our Senior SRE Engineer (m/f/d)?
We'd love to hear from you.
If you don't meet every single requirement but believe you'd be a great fit, we still encourage you to apply.
