Principal DevOps Engineer

Reperio Human Capital · Recruitment agency
Full-time•DevOps•€80k-100k/year (EUR)•Ireland · Remote within Ireland
Apply Now

TechJobs.ie Job Insights

At a glance

Employment type
Full-time
Workplace
Remote
Salary
€80k-100k/year (EUR)
Category
DevOps

Technologies & skills

Primary technologies

Cloud & infrastructure

Other technical skills

What you'll be doing

Principal DevOps Engineer at a fast-growing AI-driven data analytics company. Own platform infrastructure from design to production, working remotely with occasional Galway office visits. Individual contributor role reporting to CTO, influencing technical direction and supporting cloud, on-premise, and air-gapped deployments.

  • Build and run all environments as infrastructure-as-code, from CI through to production
  • Design and operate Kubernetes clusters, both in the cloud and on-premise
  • Build and maintain an offline installer for air-gapped customer deployments
  • Own the CI/CD pipelines and keep local developer environments consistent with production
  • Run secure private networks with controlled egress
  • Manage self-hosted observability: monitoring, logging, tracing and alerting
  • Deploy and operate AI model endpoints and inference serving, including GPU capacity planning
  • Support customer IT teams during installations and upgrades

Key requirements

Must-have

  • Extensive experience running production Kubernetes, both in the cloud and on-premise
  • Track record of owning infrastructure end to end
  • Strong Linux administration, networking and troubleshooting skills
  • Infrastructure-as-code and CI/CD experience
  • Scripting ability in Bash, Python or similar
  • Experience with self-hosted observability tools such as Prometheus, Loki or similar
  • Clear technical writing and excellent spoken and written English
  • Experience working directly with customer IT teams on on-prem installs

Nice-to-have

  • Experience delivering software into air-gapped environments
  • Experience with RKE2/K3s
  • Experience with Postgres, ClickHouse, OpenSearch, Temporal and S3-compatible storage
  • Experience with LLM and model serving, such as vLLM, Ray Serve or similar

Role signals

Technical focus
DevOps/Infrastructure
Architecture / system design
Indicated in the listing
Hands-on vs management
Hands-on

Full job description

Principal DevOps Engineer

A fast-growing, venture-backed software company building an AI-driven data analytics platform. Its clients are government, security and financial-services organisations. The platform runs in the cloud, on-premise and in fully air-gapped environments.

The role
You'll be a senior, hands-on engineer who owns the platform's infrastructure from design through to production. You'll report directly to the CTO and work alongside a small DevOps team. This is an individual contributor role with real influence over technical direction. Most of your time will be remote, with 1-2 days a month in the Galway office.

What you'll do

  • Build and run all environments as infrastructure-as-code, from CI through to production

  • Design and operate Kubernetes clusters, both in the cloud and on-premise

  • Build and maintain an offline installer for air-gapped customer deployments

  • Own the CI/CD pipelines and keep local developer environments consistent with production

  • Run secure private networks with controlled egress

  • Look after self-hosted observability: monitoring, logging, tracing and alerting

  • Deploy and operate AI model endpoints and inference serving, including GPU capacity planning

  • Support customer IT teams during installations and upgrades

  • Keep architecture docs, a decision log and runbooks up to date
    What you'll bring

  • Extensive experience running production Kubernetes, both in the cloud and on-premise

  • A track record of owning infrastructure end to end

  • Strong Linux administration, networking and troubleshooting skills

  • Infrastructure-as-code and CI/CD experience

  • Scripting ability in Bash, Python or similar

  • Experience with self-hosted observability tools such as Prometheus, Loki or similar

  • Clear technical writing and excellent spoken and written English

  • Experience working directly with customer IT teams on on-prem installs

Nice to have

  • Experience delivering software into air-gapped environments
  • RKE2/K3s
  • Postgres, ClickHouse, OpenSearch, Temporal and S3-compatible storage
  • LLM and model serving, such as vLLM, Ray Serve or similar
    Interested? Apply now or get in touch for a confidential conversation.

Reperio Human Capital acts as an Employment Agency and an Employment Business.

Interview prep pack

Grounded in this listing. Use it to prepare examples before you apply.

Your interview focus

Based on this listing, the Principal DevOps Engineer role is a senior, hands-on position responsible for designing, building, and operating infrastructure across cloud, on-premise, and air-gapped environments for an AI-driven data analytics platform.

  • Kubernetes (cloud and on-premise)·Medium
  • Infrastructure-as-Code & CI/CD·High
  • Linux administration & networking·High
  • Self-hosted observability·High

Only have 30 minutes?

Follow a focused preparation plan based on this job.

Start 30-minute prep

Your 30-minute plan

  1. Review Kubernetes and On-Premise Deployments

    0–8 min

    List and reflect on your most relevant Kubernetes projects, especially those involving both cloud and on-premise clusters.

  2. Summarize IaC and CI/CD Experience

    8–15 min

    Prepare concise stories about automating infrastructure and maintaining CI/CD pipelines, highlighting tools and outcomes.

  3. Document Observability Implementations

    15–20 min

    Recall specific examples of deploying and operating self-hosted monitoring and logging solutions.

  4. Prepare for Air-Gapped and Secure Deployment Questions

    20–25 min

    Gather details on any experience with offline installers, secure networking, or restricted environments.

  5. Draft Questions for the Interviewer

    25–30 min

    Select and tailor 2-3 questions from the provided list to ask during your interview.

Likely questions

, 7 items

Priority reflects how strongly this topic is emphasised in the job listing, not whether it will be asked.

Talking points

, 6 items
  • End-to-End Infrastructure Ownership

    Prepare examples where you have designed, implemented, and maintained infrastructure, demonstrating your ability to take full responsibility for reliability and scalability.

  • Kubernetes Cluster Design and Operations

    Showcase your experience managing Kubernetes in both cloud and on-premise settings, including challenges faced and solutions implemented.

  • Infrastructure-as-Code and CI/CD Pipelines

    Be ready to discuss how you have automated environment provisioning and maintained consistency between development and production using IaC and CI/CD tools.

  • Self-Hosted Observability Solutions

    Provide examples of implementing and maintaining monitoring, logging, and alerting systems (e.g., Prometheus, Loki) to ensure platform reliability.

  • Supporting Air-Gapped and Secure Environments

    If applicable, describe your experience with offline installers, secure networking, and deployments in restricted or air-gapped environments.

  • Technical Communication and Documentation

    Demonstrate your ability to produce clear technical documentation and communicate complex concepts to both technical and non-technical stakeholders.

What to research

, 4 items
  • Production Kubernetes Operations

    Review your experience managing Kubernetes clusters in both cloud and on-premise environments, focusing on architecture, upgrades, and troubleshooting.

  • Infrastructure-as-Code and CI/CD Tools

    Refresh your knowledge of IaC tools (e.g., Terraform, Ansible) and CI/CD pipelines, including strategies for environment consistency and automation.

  • Self-Hosted Observability Stack

    Prepare to discuss your setup and maintenance of monitoring and logging tools like Prometheus and Loki, including alerting and incident response.

  • Supporting Air-Gapped and Secure Deployments

    Gather examples of working with offline installers, secure networking, and deployments in restricted environments.

Questions to ask

, 6 items
  1. What are the main challenges your team faces with on-premise and air-gapped deployments?

    Why ask this? To understand the technical and operational pain points you would help address.

  2. How is technical direction and architectural decision-making handled within the DevOps team?

    Why ask this? To clarify your influence and collaboration with the CTO and peers.

  3. What observability and monitoring stack is currently in use, and are there plans for future changes?

    Why ask this? To assess the maturity of the monitoring setup and potential areas for improvement.

  4. How do you support customer IT teams during installations and upgrades, especially in secure environments?

    Why ask this? To gauge expectations for customer interaction and support processes.

  5. What is the roadmap for AI model serving and GPU infrastructure within the platform?

    Why ask this? To understand the direction and scale of AI/ML infrastructure work.

  6. How is documentation and knowledge sharing managed across the DevOps and engineering teams?

    Why ask this? To learn about documentation standards and collaboration practices.

Register now to upload your CV

Create a free account, save a PDF or Word CV, and quick apply on roles that take applications here.

Apply NowApply before: 26 Oct 2026