Principal DevOps Engineer
On this page
TechJobs.ie Job Insights
At a glance
- Employment type
- Full-time
- Location
- Ireland · Remote within Ireland
- Workplace
- Remote
- Salary
- €80k-100k/year (EUR)
- Category
- DevOps
Technologies & skills
Primary technologies
Cloud & infrastructure
Other technical skills
What you'll be doing
Principal DevOps Engineer at a fast-growing AI-driven data analytics company. Own platform infrastructure from design to production, working remotely with occasional Galway office visits. Individual contributor role reporting to CTO, influencing technical direction and supporting cloud, on-premise, and air-gapped deployments.
- Build and run all environments as infrastructure-as-code, from CI through to production
- Design and operate Kubernetes clusters, both in the cloud and on-premise
- Build and maintain an offline installer for air-gapped customer deployments
- Own the CI/CD pipelines and keep local developer environments consistent with production
- Run secure private networks with controlled egress
- Manage self-hosted observability: monitoring, logging, tracing and alerting
- Deploy and operate AI model endpoints and inference serving, including GPU capacity planning
- Support customer IT teams during installations and upgrades
Key requirements
Must-have
- Extensive experience running production Kubernetes, both in the cloud and on-premise
- Track record of owning infrastructure end to end
- Strong Linux administration, networking and troubleshooting skills
- Infrastructure-as-code and CI/CD experience
- Scripting ability in Bash, Python or similar
- Experience with self-hosted observability tools such as Prometheus, Loki or similar
- Clear technical writing and excellent spoken and written English
- Experience working directly with customer IT teams on on-prem installs
Nice-to-have
- Experience delivering software into air-gapped environments
- Experience with RKE2/K3s
- Experience with Postgres, ClickHouse, OpenSearch, Temporal and S3-compatible storage
- Experience with LLM and model serving, such as vLLM, Ray Serve or similar
Role signals
- Technical focus
- DevOps/Infrastructure
- Architecture / system design
- Indicated in the listing
- Hands-on vs management
- Hands-on
Similar jobs
Full job description
Principal DevOps Engineer
A fast-growing, venture-backed software company building an AI-driven data analytics platform. Its clients are government, security and financial-services organisations. The platform runs in the cloud, on-premise and in fully air-gapped environments.
The role
You'll be a senior, hands-on engineer who owns the platform's infrastructure from design through to production. You'll report directly to the CTO and work alongside a small DevOps team. This is an individual contributor role with real influence over technical direction. Most of your time will be remote, with 1-2 days a month in the Galway office.
What you'll do
-
Build and run all environments as infrastructure-as-code, from CI through to production
-
Design and operate Kubernetes clusters, both in the cloud and on-premise
-
Build and maintain an offline installer for air-gapped customer deployments
-
Own the CI/CD pipelines and keep local developer environments consistent with production
-
Run secure private networks with controlled egress
-
Look after self-hosted observability: monitoring, logging, tracing and alerting
-
Deploy and operate AI model endpoints and inference serving, including GPU capacity planning
-
Support customer IT teams during installations and upgrades
-
Keep architecture docs, a decision log and runbooks up to date
What you'll bring -
Extensive experience running production Kubernetes, both in the cloud and on-premise
-
A track record of owning infrastructure end to end
-
Strong Linux administration, networking and troubleshooting skills
-
Infrastructure-as-code and CI/CD experience
-
Scripting ability in Bash, Python or similar
-
Experience with self-hosted observability tools such as Prometheus, Loki or similar
-
Clear technical writing and excellent spoken and written English
-
Experience working directly with customer IT teams on on-prem installs
Nice to have
- Experience delivering software into air-gapped environments
- RKE2/K3s
- Postgres, ClickHouse, OpenSearch, Temporal and S3-compatible storage
- LLM and model serving, such as vLLM, Ray Serve or similar
Interested? Apply now or get in touch for a confidential conversation.
Reperio Human Capital acts as an Employment Agency and an Employment Business.
Interview prep pack
Grounded in this listing. Use it to prepare examples before you apply.
Your interview focus
Based on this listing, the Principal DevOps Engineer role is a senior, hands-on position responsible for designing, building, and operating infrastructure across cloud, on-premise, and air-gapped environments for an AI-driven data analytics platform.
- Kubernetes (cloud and on-premise)·Medium
- Infrastructure-as-Code & CI/CD·High
- Linux administration & networking·High
- Self-hosted observability·High
Only have 30 minutes?
Follow a focused preparation plan based on this job.
Start 30-minute prep
Your 30-minute plan
Review Kubernetes and On-Premise Deployments
0–8 minList and reflect on your most relevant Kubernetes projects, especially those involving both cloud and on-premise clusters.
Summarize IaC and CI/CD Experience
8–15 minPrepare concise stories about automating infrastructure and maintaining CI/CD pipelines, highlighting tools and outcomes.
Document Observability Implementations
15–20 minRecall specific examples of deploying and operating self-hosted monitoring and logging solutions.
Prepare for Air-Gapped and Secure Deployment Questions
20–25 minGather details on any experience with offline installers, secure networking, or restricted environments.
Draft Questions for the Interviewer
25–30 minSelect and tailor 2-3 questions from the provided list to ask during your interview.
Talking points
, 6 itemsEnd-to-End Infrastructure Ownership
Prepare examples where you have designed, implemented, and maintained infrastructure, demonstrating your ability to take full responsibility for reliability and scalability.
Kubernetes Cluster Design and Operations
Showcase your experience managing Kubernetes in both cloud and on-premise settings, including challenges faced and solutions implemented.
Infrastructure-as-Code and CI/CD Pipelines
Be ready to discuss how you have automated environment provisioning and maintained consistency between development and production using IaC and CI/CD tools.
Self-Hosted Observability Solutions
Provide examples of implementing and maintaining monitoring, logging, and alerting systems (e.g., Prometheus, Loki) to ensure platform reliability.
Supporting Air-Gapped and Secure Environments
If applicable, describe your experience with offline installers, secure networking, and deployments in restricted or air-gapped environments.
Technical Communication and Documentation
Demonstrate your ability to produce clear technical documentation and communicate complex concepts to both technical and non-technical stakeholders.
What to research
, 4 itemsProduction Kubernetes Operations
Review your experience managing Kubernetes clusters in both cloud and on-premise environments, focusing on architecture, upgrades, and troubleshooting.
Infrastructure-as-Code and CI/CD Tools
Refresh your knowledge of IaC tools (e.g., Terraform, Ansible) and CI/CD pipelines, including strategies for environment consistency and automation.
Self-Hosted Observability Stack
Prepare to discuss your setup and maintenance of monitoring and logging tools like Prometheus and Loki, including alerting and incident response.
Supporting Air-Gapped and Secure Deployments
Gather examples of working with offline installers, secure networking, and deployments in restricted environments.
Questions to ask
, 6 itemsWhat are the main challenges your team faces with on-premise and air-gapped deployments?
Why ask this? To understand the technical and operational pain points you would help address.
How is technical direction and architectural decision-making handled within the DevOps team?
Why ask this? To clarify your influence and collaboration with the CTO and peers.
What observability and monitoring stack is currently in use, and are there plans for future changes?
Why ask this? To assess the maturity of the monitoring setup and potential areas for improvement.
How do you support customer IT teams during installations and upgrades, especially in secure environments?
Why ask this? To gauge expectations for customer interaction and support processes.
What is the roadmap for AI model serving and GPU infrastructure within the platform?
Why ask this? To understand the direction and scale of AI/ML infrastructure work.
How is documentation and knowledge sharing managed across the DevOps and engineering teams?
Why ask this? To learn about documentation standards and collaboration practices.
Register now to upload your CV
Create a free account, save a PDF or Word CV, and quick apply on roles that take applications here.