Description
We are looking for a Lead DevOps / Observability Engineer to lead the migration of the client's observability stack from New Relic and Datadog to Grafana, using technologies such as Prometheus, Loki, and Tempo.
This is a hands-on role combining:
- DevOps & platform engineering - deploying and configuring the new observability stack using AWS, Kubernetes, and Terraform.
- Application instrumentation - modifying application code to add or adapt metrics, logs, and traces.
- Migration & collaboration - rebuilding existing dashboards and alerts in Grafana and working with multiple engineering teams to ensure full monitoring coverage.
We are looking for someone who is comfortable working across both infrastructure and application code and can independently drive a cross-team observability migration.
Responsibilities:
- Lead the migration of monitoring, alerting, and dashboards from New Relic and Datadog to Grafana.
- Design and implement the target observability architecture (metrics, logs, traces) and the infrastructure-as-code (Terraform) needed to support it.
- Audit existing New Relic and Datadog dashboards, alerts, and integrations to build a complete migration inventory and ensure feature/coverage parity in Grafana.
- Make application-level code changes across services to add, adjust, or replace instrumentation (metrics exporters, logging, tracing libraries) as needed for the new stack.
- Work with AWS infrastructure, including EKS, Kubernetes, and Docker, to deploy and operate the new observability tooling.
- Configure AWS networking and access as needed to support the observability platform (VPCs, security groups, IAM).
- Contribute to CI/CD pipelines and deployment automation using GitHub Actions to support the rollout.
- Collaborate with development, platform, and infrastructure teams to coordinate the cutover from New Relic/Datadog to Grafana with minimal disruption.
- Identify and remediate any security issues (secrets, keys, dependencies) encountered in the course of this work.
- Document the new observability architecture, migration procedures, and dashboard/alert mappings.
- Use AI-assisted development tools, including Codex, to improve engineering productivity and delivery speed.
Requirements:
- Strong hands-on experience with Grafana, including dashboard design, alerting, and data source configuration (e.g. Prometheus, Loki, Tempo, or similar).
- Practical experience migrating away from commercial observability platforms such as New Relic and/or Datadog.
- Full-stack development experience — comfortable reading, modifying, and instrumenting application code across front-end and back-end services, not just infrastructure.
- Solid experience with AWS in production environments, including EKS and Kubernetes.
- Practical experience with Docker and containerized workloads.
- Strong Terraform and infrastructure-as-code skills.
- Experience with GitHub Actions or comparable CI/CD tools.
- Ability to work across multiple teams, services, and codebases to coordinate a cross-cutting migration.
- Fluency with Codex or comparable AI-assisted coding tools.
Nica to have:
- Prior experience running a similar New Relic/Datadog-to-Grafana (or equivalent) observability migration end to end.
- Grafana-specific certifications or demonstrated community contributions (plugins, dashboards, etc.).
- Experience with the broader Prometheus/Loki/Tempo (LGTM) ecosystem.
- AWS certification, such as AWS Solutions Architect Associate/Professional.
- Experience with RDS and other AWS data services.
- Prior contract or consulting experience with clearly defined deliverables.
- At least an Upper-Intermediate level of English
We offer*:
- Flexible working format - remote, office-based or flexible
- A competitive salary and good compensation package
- Personalized career growth
- Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
- Active tech communities with regular knowledge sharing
- Education reimbursement
- Memorable anniversary presents
- Corporate events and team buildings
- Other location-specific benefits
*not applicable for freelancers
Similar jobs
Est. 120,000 USD
N-iX is looking for Lead DevOps Engineer (AWS Migration) - to join the team About the client Our client is a leading European online car marketplace, serving over 30 million monthly users across 18 countries. As a Lead D…
Est. 245,000 USD
Principal Observability Platform Engineer – Nscale About Nscale Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscal…
Est. 225,000 USD
About Nscale Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale simplifies AI development while enabling superior…
Est. 240,000 USD
About Nscale Nscale is taking on the hyperscalers by building a vertically integrated GenAI cloud platform. We own the data centres, software, and applications that power today's AI stack using sustainable technology sol…
Est. 95,000 EUR
Please note that the job is only available from the locations outlined. We are looking for a Senior Software Engineer to help us take REDAPL, our Referential Data Platform, to the next level. REDAPL is Datadog's main pla…
Est. 180,000 USD
We are a team of engineers that translate our real-world experience to help our user communities solve problems. With a focus on service management, helping teams respond to incidents, run on-call, and automate their ope…
Est. 216,000 USD
This Engineering Manager will lead the Data Visualizations Explorations team within the Graphing organization, setting product direction, and coaching and developing team members. They will staff and drive projects that…
Est. 80,000 EUR
As a Staff Engineer on the Data Platform Experience team, you'll help shape how Datadog engineering teams build, operate, and evolve products on the Observability Data Platform. You'll lead the design and delivery of sha…
Est. 195,000 USD
Senior Observability Platform Engineer – Nscale About Nscale Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale s…
Est. 193,500 USD
We are a team of engineers that translate our real-world experience to help our user communities solve problems. With a focus on AI-accelerated workflows and next-generation developer ecosystems, you will have the opport…
Est. 225,000 USD
The Datadog for Startups (DDFS) program helps the next generation of fast-scaling companies adopt best-in-class observability and security from day one. We're looking for the technical engine of this program - someone wh…
Est. 216,000 USD
Datadog's Application Performance Monitoring (APM) provides deep visibility into the health, performance, and lifecycle of modern distributed applications, tracing requests from end-user devices (web and mobile) through…
Est. 84,000 USD
Are you ready to shape the future of fintech and financial inclusion in Latin America with cutting-edge technology and deep business impact? Join us as a DataOps Engineer and help power our next-generation regional data…
Est. 80,000 EUR
N-iX is a global software development service company that helps businesses across the globe create next-generation software products. Founded in 2002, we unite 2,400+ tech-savvy professionals across 40+ countries, worki…
Est. 80,000 EUR
We are looking for a Senior Software Engineer to help us take REDAPL, our Referential Data Platform, to the next level. REDAPL is Datadog's main platform for tracking our customers' infrastructure resources and relations…
Est. 173,500 USD
As a Key Accounts Enterprise Sales Engineer, you will provide technical expertise through sales presentations, product demonstrations, and supporting technical evaluations (POVs). Sales Engineers help qualify and close o…
Est. 120,000 EUR
As a Staff Engineer on the Data Platform Experience team, you'll help shape how Datadog engineering teams build, operate, and evolve products on the Observability Data Platform. You'll lead the design and delivery of sha…
Est. 225,000 USD
Datadog's Forward Deployed Engineering function is in an active growth phase, and the FDE Lead will play a central role in shaping what comes next. Working in close partnership with the Head of Datadog for Startups and F…
We are Datadog's in-house product experts. The Technical Solutions team enables Datadog's worldwide growth by educating potential clients and ensuring that existing customers are happy and successful. We share our techni…
As an Enterprise Sales Engineer, you'll partner with Sales to help customers understand the value of Datadog and how it solves their most important technical and business challenges. You'll lead technical evaluations, de…
Est. 95,000 EUR
The Team As Datadog’s in-house product experts, the Technical Escalation Engineering (TEE) team plays a critical role in driving our global success. We enable our customers, from the world’s most innovative startups to t…
Est. 151,500 USD
Reports to: Head of Developer Growth · New York, NY NOTE: This is not a traditional marketing role, it's a growth role that reports into the Product org. We're hiring someone to build our developer audience the way the b…
Est. 310,500 USD
The Dashboards product is Datadog's unified single-pane-of-glass for metrics, logs, and traces—a comprehensive treasure trove of observability data. We are transforming Dashboards into an AI-native control surface and th…
Est. 274,500 USD
Role Summary: Datadog is seeking a Staff Software Engineer to help shape the future of our Bring Your Own Cloud (BYOC) Logs offering by unifying observability pipelines with log management software that customers deploy…
Est. 99,000 USD
As Datadog’s in-house product experts, the Technical Escalation Engineering (TEE) team plays a critical role in driving our global success. We enable our customers, from the world’s most innovative startups to the larges…
Est. 120,000 EUR
We are looking for a Senior Software Engineer to help us take REDAPL, our Referential Data Platform, to the next level. REDAPL is Datadog's main platform for tracking our customers' infrastructure resources and relations…
We are Datadog's in-house product experts. The technical solutions team enables Datadog's worldwide growth by educating potential clients and ensuring that existing customers are happy and successful. The Partner Solutio…
Est. 120,000 DKK
As an Enterprise Sales Engineer, you will provide technical expertise through sales presentations, product demonstrations, and supporting technical evaluations (POVs). Sales Engineers help qualify and close opportunities…
Est. 333,000 USD
Datadog’s Cloud Observability group is one of the core data retrieval and processing groups powering our foundational product, Infrastructure Monitoring. The group’s scope includes integration with all major hyperscalers…
Est. 80,000 EUR
The Team: We are Datadog's in-house product experts. The technical solutions team enables Datadog's worldwide growth by educating potential clients and ensuring that existing customers are happy and successful. We share…