Site Reliability Engineer
The Cardano Foundation is the independent, non-profit organization responsible for stewarding the advancement of the public, permissionless blockchain...
The Cardano Foundation is the independent, non-profit organization responsible for stewarding the advancement of the public, permissionless blockchain platform Cardano. Our mission is to establish the Cardano blockchain as the future financial and social system of the world for generations to come, driving adoption and facilitating development of the protocol. We aim to de-risk decentralization for regulators and organizations, while also giving the Cardano community the necessary tools and support to leverage the Cardano protocol to solve real world problems.
Based in Switzerland, the Foundation works to facilitate the use of Cardano in mission critical applications across a wide range of industries and markets, anchoring use cases in the off-line world and encouraging active on-chain participation and governance.
The Site Reliability Engineer is part of the Cardano Foundation’s Infrastructure team and will report to the Director of Infrastructure. In this position, you will manage infrastructure, maintain systems, handle operational issues and advise our team on best practices that will make our operations more reliable, secure and scalable. You will also advance innovation with the support of our entire engineering team.
Implement GitOps-based operations across various environments.
Implement automation and effective monitoring with infrastructure as code.
Automate and improve the Foundations productivity tools and processes.
Maintain and monitor network infrastructure and related security.
Deploy applications to monitor service stability and performance.
Implement tools and dashboards for logging, alerting, monitoring.
Deploy and maintain applications on cloud-native microservices architectures.
Ensuring that systems are safe and secure against cybersecurity threats.
Provide guidance and educate team members on development and operations.
Document and design processes and update existing ones.
Automate and maintain development and test environments.
Deploy and monitor blockchain environments for benchmarking and integration testing.
Identifying technical issues and deploying software updates.
Analyse systems performance and resource usage.
Perform root cause analysis for production errors.
Investigate and resolve technical issues.
Be on-call when required for production services.
Degree in computer science, computer engineering or related fields (Masters preferred).
Experience with AWS, Google Compute, or other cloud providers.
Experience with build and release engineering.
Linux operational excellence and automation experience.
Familiarity with Security Ops such as intrusion and vulnerability scanning, security best practices.
Experience with continuous integration tools.
Scripting and programming skills, bash, python.
Sysadmin experience administering application servers, containers, and web servers.
Familiarity with GitOps and git based workflows.
Experience with Docker and container orchestration platforms.
Experience with Prometheus/Grafana, ELK logging.
Experience with Kubernetes.
Experience with configuration management tools (e.g. Ansible, Chef, Puppet etc).
Experience with infrastructure-as-code (e.g. Terraform, Cloudformation)
Proficient English language with strong communication skills.
You have strong analytical and problem-solving skills.
You are focused on achieving goals in a fast-paced environment.
Ability to work autonomously with minimal supervision.
Ability to resolve complex issues in creative and effective ways.
Excellent interpersonal and communication skills.
Ability to understand user requirements and develop solutions.
Nix experience using the tools within the nix ecosystem, specifically using nix as a configuration language.
Interest or knowledge of, statically-typed functional programming languages such as Haskell, Scala, Purescript, or Rust.
Below are some other jobs we think you might be interested in.
-
Engineer - Site Reliability
- Cboe Digital
- London, United Kingdom
Jul 04 -
Site Reliability Engineer
- LayerZero
- Vancouver, BC
Jun 20 -
Site Reliability Engineer
- Strike
- Anywhere
- Remote
Jun 26 -
Senior Site Reliability Engineer
- SSV Network
- Anywhere
- Remote
Jul 10 -
Director, Site Reliability Engineering
- Stellar
- New York
Jul 01 -
Site Reliability Engineer - Algorithmic Trading
- DRW
- Chicago
Jun 25 -
Site Reliability Engineer - Algorithmic Trading
- DRW
- Tel Aviv
Jun 15 -
Senior Site Reliability Engineer, Observability
- Ripple
- Chicago, Illinois, United States
Jul 28 -
Software Engineer - Data Engineering
- Akuna Capital
- Chicago, IL
Jun 23 -
Growth Engineer / Integration Engineer
- Injective Labs
- Anywhere
- Remote
Jul 18 -
Senior Engineer, Trading Product Engineering
- CoinDesk
- London
Jul 13 -
Principal Engineer, CoinDesk Data Engineering
- CoinDesk
- London
Jul 24 -
Lead Engineer, Trading Platform Engineering
- CoinDesk
- London
Jul 18 -
Lead Software Engineer, Clearing Engineering
- CoinDesk
- New York
Jun 29 -
Software Engineer, Cumberland/FICCO Tools Engineering
- DRW
- New York City
Jul 01 -
Growth Engineer
- Goldsky
- Anywhere
- Remote
Jun 28 -
Software Engineer
- Renegade
- San Francisco
Jul 02 -
Sr. Information Security Engineer (Systems Engineer)
- Cboe Digital
- Kansas City, MO
Jun 30 -
Senior Software Engineer, Backend | Product Engineering
- TRM Labs
- South America
- Remote
Jul 12 -
DevOps Engineer
- Sardine
- Germany
- Remote
Jul 25

