On this page
TechJobs.ie Job Insights
At a glance
- Employment type
- Full-time
- Location
- Ireland · Remote within Ireland
- Workplace
- Remote
- Category
- DevOps
Technologies & skills
Cloud & infrastructure
Other technical skills
What you'll be doing
Join Twilio's platform engineering observability team to lead the architecture and delivery of a unified, OpenTelemetry-first observability stack. Drive technical execution, mentor engineers, and collaborate across teams to build scalable, developer-friendly observability solutions in a remote-first environment.
- Lead end-to-end architecture and delivery of observability platform components
- Drive consistency and quality across observability signals (logs, metrics, traces, profiling)
- Serve as technical advisor and mentor across the platform organization
- Focus on problem areas such as high-cardinality telemetry and distributed tracing correlation
- Collaborate with product teams, SREs, and developer experience groups
- Design and build developer-friendly tooling and APIs for incident response and debugging
- Leverage and optionally contribute to open-source standards like OpenTelemetry
Key requirements
Must-have
- Expertise in building and scaling observability systems
- Technical leadership in observability platform components
- Proficiency in at least one modern programming language (Go, Python, Java)
- Familiarity with high-cardinality data challenges and telemetry correlation
- Experience designing high-scale telemetry systems
- Understanding of distributed systems and microservice-based environments
- Experience with AWS, Kubernetes, and infrastructure-as-code tools
- Ability to provide architectural guidance and thought leadership
Role signals
- Technical focus
- DevOps/Observability
- Leadership
- Mentoring
- Architecture / system design
- Indicated in the listing
- Hands-on vs management
- Hands-on
Similar jobs
Full job description
Who we are
At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences.
Our dedication to remote-first work, and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands.
.
- Hiring and how we work :
We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions!
Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings.
.
See yourself at Twilio
Join the team as our next Software Engineer on Twilio’s platform engineering observability team.
About the job
This position is needed to help our platform engineering observability team. Twilio is undergoing a large-scale observability transformation—and you can help shape the foundation. Observability is a strategic pillar and a key enabler for faster incident response, deeper customer-centric insights, and more cost-effective platform operations.
As a Software Engineer on the Platform Observability team, you’ll play a critical role in re-architecting how telemetry flows and is utilized through Twilio—making it structured, accessible, affordable, and actionable. Over the next 3 years, Twilio is rebuilding nearly every component of our observability platform, from data collection to real-time analytics. You will drive core initiatives that shift Twilio from fragmented tooling and wasteful data sprawl to a unified, OpenTelemetry-first observability stack built for scale.
You’ll lead technically and strategically—designing platform components, influencing org-wide architectural decisions, mentoring engineers, and engaging directly with teams across Platform Engineering and R&D
Responsibilities
In this role, you’ll:
-
Lead the end-to-end architecture and delivery of key observability platform components, with a focus on reliability, scalability, and usability.
-
Drive consistency and quality across all observability signals—logs, metrics, traces, and continuous profiling—building intuitive workflows for engineers.
-
Serve as a technical advisor and mentor across the platform org, guiding design decisions and aligning cross-team efforts with long-term architectural goals.
-
Go deep in one or more problem areas (e.g., high-cardinality telemetry, distributed tracing correlation, compute cost insights), while ensuring the platform scales horizontally.
-
Collaborate with product teams, SREs, and developer experience groups to deeply understand telemetry needs and integrate observability into core engineering workflows.
-
Design and build developer-friendly tooling and APIs to support incident response, performance analysis, and platform debugging at scale.
-
Leverage (and optionally contribute to) open-source standards like OpenTelemetry to ensure interoperability and extensibility.
Qualifications
Twilio values diverse experiences from all kinds of industries, and we encourage everyone who meets the required qualifications to apply. If your career is just starting or hasn't followed a traditional path, don't let that stop you from considering Twilio. We are always looking for people who will bring something new to the table!
*Required:
-
Proven expertise in building and scaling observability systems (e.g., logging platforms, metrics pipelines, tracing infrastructure, or profiling tools).
-
Lead technical execution for major components of Twilio’s observability overhaul, including our shift to centralized S3-based data lakes, OpenTelemetry instrumentation, and ClickHouse-backed query engines.
-
Proficiency in at least one modern programming language (e.g., Go, Python, Java).
-
Familiarity with high-cardinality data challenges and telemetry correlation techniques.
-
Experience designing high-scale telemetry systems (e.g., Prometheus, ClickHouse, OpenTelemetry, Kafka, or equivalent).
-
Solid understanding of distributed systems and the challenges of observability in complex, microservice-based environments.
-
Experience with AWS, Kubernetes, and infrastructure-as-code tools.
-
Provide architectural guidance and thought leadership across teams, helping to establish clear telemetry standards, efficient usage patterns, and scalable platform abstractions.
-
Ability to make forward-looking technical decisions and lead others through ambiguity and chan
-
Desired:
- Familiarity with ClickHouse, Grafana Mimir, Athena, or equivalent systems for log and metrics querying.
-
Contributions to open-source observability tools or communities.
Location
This role will be remote from Ireland.
Travel
We prioritize connection and opportunities to build relationships with our customers and each other. For this role, you may be required to travel occasionally to participate in project or team in-person meetings.
What We Offer
Working at Twilio offers many benefits, including competitive pay, generous time off, ample parental and wellness leave, healthcare, a retirement savings program, and much more. Offerings vary by location.
Twilio thinks big. Do you?
We like to solve problems, take initiative, pitch in when needed, and are always up for trying new things. That's why we seek out colleagues who embody our values — something we call Twilio Magic. Additionally, we empower employees to build positive change in their communities by supporting their volunteering and donation efforts.
So, if you're ready to unleash your full potential, do your best work, and be the best version of yourself, apply now! If this role isn't what you're looking for, please consider other open positions.
.
Stay alert to recruitment fraud
We care about your safety. Scammers sometimes impersonate Twilio recruiters through fake job postings, emails, websites, or messages. Please ensure you are engaging with an official @twilio.com email address. We will never ask for payment, gift cards, cryptocurrency, or banking information during the recruiting process. We do not make job offers without a formal interview process or conduct interviews exclusively through text-based messaging apps.
.
Twilio is proud to be an equal opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or other applicable legally protected characteristics. We also consider qualified applicants with criminal histories, consistent with applicable federal, state and local law. Qualified applicants with arrest or conviction records will be considered for employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Additionally, Twilio participates in the E-Verify program in certain locations, as required by law.
Interview prep pack
Grounded in this listing. Use it to prepare examples before you apply.
Your interview focus
Based on this listing, the DevOps Engineer (Observability) role at Twilio focuses on leading the architecture and delivery of scalable observability platform components, driving technical strategy, and mentoring engineers in a remote-first environment.
- Observability system design·Medium
- Technical leadership·High
- Distributed systems·High
- Cloud and Kubernetes·High
Only have 30 minutes?
Follow a focused preparation plan based on this job.
Start 30-minute prep
Your 30-minute plan
Review Observability Projects and Leadership Examples
0–8 minList and reflect on your most relevant observability projects, focusing on your leadership, architectural decisions, and outcomes.
Deep Dive into OpenTelemetry and High-Scale Telemetry Systems
8–15 minStudy OpenTelemetry, Prometheus, ClickHouse, and related tools, emphasizing integration, scaling, and high-cardinality data handling.
Refresh AWS, Kubernetes, and Infrastructure-as-Code Skills
15–21 minReview your experience with AWS, Kubernetes, and automation tools, preparing concrete examples of observability pipeline deployments.
Prepare STAR Stories for Mentorship and Cross-Team Collaboration
21–26 minCraft concise stories that showcase your mentorship, technical guidance, and collaboration with other teams.
Research Twilio’s Observability Strategy and Prepare Questions
26–30 minRead Twilio’s public materials on observability and prepare thoughtful questions to ask during the interview.
Talking points
, 6 itemsBuilding and Scaling Observability Systems
You should be ready to discuss your experience designing, implementing, and scaling logging, metrics, tracing, or profiling platforms, as this is central to the role.
Technical Leadership and Mentorship
Prepare examples of how you have guided teams, influenced architectural decisions, or mentored engineers, since the role emphasizes leadership and cross-team impact.
Handling High-Cardinality Telemetry and Data Correlation
Expect to explain your approach to managing high-cardinality data and correlating telemetry signals, as these are highlighted problem areas.
Experience with Modern Programming Languages
Be ready to demonstrate proficiency in Go, Python, or Java, and how you have used these languages in observability or platform engineering contexts.
Distributed Systems and Microservices Observability
You should be able to articulate your understanding of observability challenges in distributed, microservice-based environments.
AWS, Kubernetes, and Infrastructure-as-Code
Prepare to discuss your hands-on experience with cloud infrastructure, container orchestration, and automation tools, as these are required skills.
What to research
, 4 itemsTwilio’s Observability Transformation
Review Twilio’s public materials and blog posts about their observability initiatives and platform engineering goals.
OpenTelemetry and Related Tooling
Deepen your understanding of OpenTelemetry, Prometheus, Grafana, and ClickHouse, focusing on their integration and scaling in cloud environments.
High-Cardinality Data Management
Prepare to discuss strategies for handling high-cardinality telemetry and correlating signals across distributed systems.
AWS, Kubernetes, and Infrastructure Automation
Refresh your knowledge of AWS services, Kubernetes operations, and infrastructure-as-code practices relevant to observability pipelines.
Questions to ask
, 6 itemsWhat are the biggest challenges Twilio is currently facing in its observability transformation?
Why ask this? Shows your interest in the company’s current priorities and readiness to address real-world problems.
How does the observability team collaborate with product teams, SREs, and developer experience groups?
Why ask this? Clarifies cross-team dynamics and your potential role in broader organizational initiatives.
What is the team’s approach to balancing reliability, scalability, and cost in observability platform design?
Why ask this? Demonstrates your architectural thinking and concern for operational trade-offs.
How are decisions made regarding adoption or contribution to open-source observability standards like OpenTelemetry?
Why ask this? Shows your interest in open-source engagement and technical influence.
What opportunities exist for technical mentorship and leadership within the platform engineering organization?
Why ask this? Highlights your interest in leadership and professional growth.
How does Twilio measure the success and impact of its observability initiatives?
Why ask this? Helps you understand how your work will be evaluated and its business impact.
Register now to upload your CV
Create a free account, save a PDF or Word CV, and quick apply on roles that take applications here.
