Live opening · Posted 7 hours ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Senior Observability Engineer-IT
Be a part of a team that’s ensuring Dell Technologies' product integrity and customer satisfaction. Our IT Software Engineer team turns business requirements into technology solutions by designing, coding and testing/debugging applications, as well as documenting procedures for use and constantly seeking quality improvements.
What you’ll achieve:
You will develop and implement Observability solutions for the enterprise-wide technical requirements of customers. Integrates hardware, processes, methodologies, and software within the Observability environment. Under the guidance of senior engineers and architects, participates in hands-on implementation and deployment of comprehensive observability solutions that enable end-to-end visibility across distributed systems, applications, and infrastructure. Completes documentation and procedures for installation and maintenance.
Join us to do the best work of your career and make a profound social impact as a Senior Observability Engineer on our Software Engineer-IT team.
Take the first step towards your dream career
Every Dell Technologies team member brings something unique to the table. Here’s what we are looking for with this role:
Essential Requirements
5 years of experience in DevOps, platform engineering, or SRE roles with exposure to telemetry and observability concepts ·
Working knowledge of modern observability stacks, such as: Dynatrace, Splunk, Grafana, DataDog, Elastic/ELK and data-pipelines such as: Fluent Bit, Vector, Cribl ·
Familiarity with tools like Terraform, Ansible, or Helm for automated provisioning and configuration management ·
Basic knowledge of OpenTelemetry concepts, collectors, and instrumentation principles
Strong analytical skills with a systematic approach to debugging distributed systems and telemetry pipelines under guidance
Desirable Requirements
Interest or prior exposure to AI Native Observability strategies for production AI systems, including leading platforms (Langfuse, Langsmith, Arize) that support this visibility
You will:
Assist with deploying, configuring, and maintaining observability platforms (covering metrics, logs, distributed traces, and events) in production environments under senior engineer guidance, including troubleshooting basic performance issues and monitoring platform health
Implement industry standards such as OpenTelemetry under senior engineer guidance, including collector configuration, SDK integration, and supporting telemetry data pipelines for reliable data collection and routing ·
Monitor telemetry data ingestion rates and storage; assist with implementing sampling and retention strategies to balance system visibility with operational costs, and support automation of observability instrumentation and configuration using Infrastructure as Code (IaC) and CI/CD pipelines ·
Respond to alerts and participate in supporting observability platform health, assist with troubleshooting production issues related to observability platforms and telemetry pipelines, and support integration of observability platforms with enterprise systems ·
Support implementation of observability strategies for LLM applications, including model performance monitoring, prompt evaluation metrics, token usage tracking, and inference latency measurement under senior engineer guidance, and follow observability standards and best practices
Work arrangement
No
More openings worth a look
Recently tracked roles with full details and direct application links.