Domits is looking for a DevOps / SRE (Site Reliability Engineer) for our digital infrastructure for luxury short term vacation rental managers

Interns can fill in this form.

DevOps / SRE (Site Reliability Engineer)

Location: Netherlands / Hybrid / Remote
Job Type: Internship, Freelance, Parttime, Fulltime
Industry: Hospitality / Property Management / PropTech
Responsibilities: In consultation

What You’re Going to Do:

You’ll design, operate, and continuously improve the infrastructure that powers our platform.

  • Build and maintain scalable, reliable cloud infrastructure
  • Automate deployments, monitoring, and incident response
  • Ensure high availability, performance, and security
  • Collaborate with engineering to design resilient systems
  • Proactively identify risks and improve system reliability
  • Enable fast, safe product iteration through strong DevOps practices

Your Key Responsibilities

  • Design and maintain CI/CD pipelines
  • Manage cloud infrastructure (IaC, scaling, failover)
  • Implement monitoring, logging, and alerting systems
  • Lead incident response, root-cause analysis, and postmortems
  • Improve system performance, reliability, and cost efficiency
  • Partner with engineering on architecture and deployment strategies
  • Champion security, observability, and operational excellence

Bonus If You Have Experience With

  • Hospitality management systems and property management platforms
  • AWS, GCP, or Azure
  • Kubernetes, Docker, Terraform, Helm
  • Monitoring tools (Prometheus, Grafana, Datadog, New Relic)
  • Infrastructure as Code (Terraform, Pulumi, CloudFormation)
  • High-traffic SaaS or marketplace platforms
  • On-call rotations and incident management
  • Security best practices and compliance

What We Ask

  • Experience in DevOps, SRE, or infrastructure engineering
  • Strong understanding of distributed systems and cloud architecture
  • Automation-first mindset
  • Comfort working in fast-moving startup environments
  • Strong collaboration and communication skills
  • Ownership mentality—you treat reliability as a product
  • Self-motivated, independent, and great at remote communication
  • Fluent in English (written and spoken)

Extra Information:

  • Compensation: In consultation + equity/revenue share options
  • Tools & Stack: AWS, JavaScript/TypeScript, React, Node.js, GitHub, Discord, Notion
  • Work Style: Remote-first, async-friendly, outcome-driven
  • Growth Path: Opportunity to become Lead or Head of DevOps/SRE as we scale

Future Career Path 

1. Intern / Junior DevOps – SRE

Foundation stage: learning infrastructure, reliability fundamentals, and professional habits

Technical Skills

  • Linux fundamentals and networking basics (DNS, TCP/IP, HTTP)
  • Scripting basics (Bash, Python)
  • Version control with Git
  • Understanding systems, containers, and application lifecycles

AWS

  • Introductory knowledge of AWS core services (EC2, S3, IAM, VPC)
  • Basic understanding of cloud networking and security
  • Exposure to Infrastructure as Code concepts

DevOps / SRE Skills

  • CI/CD pipeline basics and deployment workflows
  • Monitoring, logging, and alerting fundamentals
  • Understanding uptime, availability, and incident response basics
  • Supporting deployments and reliability improvements

Mindset / Soft Skills

  • Strong learning mindset and curiosity
  • Attention to detail and documentation discipline
  • Clear communication and collaboration
  • Ownership mentality and reliability-first thinking
2. Mid-Level DevOps – SRE / Associate DevOps – SRE

Ownership stage: operating systems, improving reliability, and scaling practices

Technical Skills

  • Strong Linux and networking expertise
  • Automation and scripting proficiency
  • Understanding distributed systems and failure modes

AWS

  • Hands-on experience with AWS services (EKS/ECS, RDS, ALB, Route53)
  • Infrastructure as Code (Terraform, CloudFormation, Pulumi)
  • Managing environments, secrets, and access controls

DevOps / SRE Skills

  • Designing and maintaining CI/CD pipelines
  • Implementing SLOs, SLIs, and error budgets
  • Incident response ownership, root-cause analysis, postmortems
  • Observability stacks (metrics, logs, tracing)
  • Performance optimization and capacity planning

Professional Skills

  • Mentoring junior engineers
  • Leading reliability initiatives or automation projects
  • Writing runbooks, documentation, and standards
  • Collaborating closely with engineering and product teams
3. Senior DevOps – SRE / Associate DevOps – SRE Lead / Chief DevOps – SRE

Leadership stage: defining reliability strategy and organizational impact

Technical Skills

  • Expert-level system design for scalability, reliability, and security
  • Deep understanding of distributed systems and cloud-native architectures

AWS

  • Ownership of cloud platform strategy and tooling
  • Advanced infrastructure governance and cost optimization
  • Designing organization-wide deployment and reliability standards

DevOps / SRE Skills

  • Defining long-term reliability and availability strategy
  • Organization-wide SLO frameworks and incident management processes
  • Driving automation-first culture and platform resilience

Leadership & Impact

  • Leading DevOps / SRE teams and mentoring engineers
  • Influencing executive decisions on platform strategy
  • Ensuring reliability enables product velocity and business growth
  • Embedding reliability as a core company value