About the job
Location Details
Remote (EU)
About Us
At GreenCompute, we are building a sovereign and sustainable cloud for Europe. Our platform spans infrastructure, edge computing, development tooling, and consulting, all designed to give customers more control over where data lives, how workloads are powered, and how cloud services are governed. We combine high-performance cloud infrastructure with a strong commitment to digital sovereignty and environmental responsibility, using approaches such as renewable energy, zero-water cooling, and heat reuse to reduce the footprint of modern computing. As we grow, we are focused on turning a powerful infrastructure foundation into an intelligent, resilient platform that can operate with greater autonomy and efficiency.
Role Introduction
As a Senior Infrastructure Engineer, you will play a central role in shaping the next generation of GreenCompute’s platform. This is a high-impact engineering role focused on transforming a reliable bare-metal cloud foundation into a self-healing, distributed system that can detect failure, recover capacity, and reintegrate nodes with minimal human intervention. You will work deeply across the control plane, automation systems, and core infrastructure architecture, helping define how resilient cloud systems should be built for a more sustainable and sovereign European cloud.
What You'll Do
Lead the evolution of our infrastructure toward autonomous, self-healing operation across distributed cloud units and micro data centers.
Design and implement failure detection, health signaling, consensus-based recovery, and automated node provisioning and reintegration workflows.
Own and strengthen the core control plane, including DNS, proxying, authentication, access control, and portal-related infrastructure components.
Advance our infrastructure model into a declarative, versioned, and observable infrastructure-as-code system with strong operational clarity.
Improve platform resilience and performance across ZFS replication, LXC lifecycle orchestration, secure secrets distribution, and network design.
Integrate AI-driven tooling into infrastructure workflows and the control plane to reduce repetitive work and increase engineering leverage.
Mentor other engineers in distributed systems thinking, resilience engineering, and practical approaches to building intelligent infrastructure.
What You Bring
Deep expertise with Linux systems at scale, along with strong hands-on experience in ZFS-based environments.
Strong distributed systems knowledge, including failure modes, replication, eventual consistency, reconciliation loops, anti-entropy patterns, and recovery design.
A proven track record of evolving infrastructure platforms toward higher levels of autonomy, resilience, and operational efficiency.
Expert-level Python and Bash skills, with a software engineering mindset that treats infrastructure as code and systems design rather than one-off scripting.
Hands-on experience building or operating control loops, agents, operators, or similar automation mechanisms in infrastructure environments.
A strong security mindset and practical experience with authentication, access control, hardening, and protective tooling such as Authelia, Fail2ban, ModSecurity, or comparable solutions.
A pragmatic approach to engineering, with a preference for simple, understandable systems and clear technical decision-making.
Nice to Have
Experience working on private cloud or bare-metal cloud platforms.
Familiarity with hermetic or isolated computing environments.
Knowledge of Knot DNS, reverse proxy architectures, and SSO or identity systems.
Practical experience using AI tooling to accelerate infrastructure engineering workflows and output.
Ability to write clear architecture, design, and strategy documents for technical and cross-functional audiences.
If you’re excited by resilient distributed systems, infrastructure autonomy, and the chance to build foundational technology with real long-term impact, we’d love to hear from you.