Staff Software Engineer, Observability

CoreWeaveSunnyvale, CA
4h$188,000 - $250,000Hybrid

About The Position

We are seeking a highly experienced Staff Software Engineer to lead our efforts in building, maintaining, and optimizing highly scalable, reliable, and secure systems. The Observability team is responsible for deploying and maintaining critical infrastructure at CoreWeave including our logging, tracing, and metrics platforms as well as the pipelines that feed them.

Requirements

  • 7+ years of experience in Software Engineering, Site Reliability Engineering, DevOps, or a related field.
  • Deep expertise across all observability pillars using tools like ClickHouse, Elastic, Loki, Victoria Metrics, Prometheus, Thanos and/or Grafana.
  • Expertise in Kubernetes, containerization, and microservices architectures.
  • Proven track record of leading incident management and post-mortem analysis.
  • Excellent problem-solving, analytical, and communication skills.

Nice To Haves

  • Experience running and scaling observability tools as a cloud provider.
  • Experience administering large-scale kubernetes clusters.
  • Deep understanding of data-streaming systems.

Responsibilities

  • Lead and mentor engineers, fostering a culture of collaboration and continuous improvement.
  • Scale logging, tracing, and metrics platforms to support a global datacenter footprint.
  • Develop and refine monitoring and alerting to enhance system reliability.
  • Advise engineers across CoreWeave on optimal usage of Observability systems.
  • Automate interactions with CoreWeave’s Compute Infrastructure layer.
  • Manage production clusters and ensure development teams follow best practices for deployments.

Benefits

  • Medical, dental, and vision insurance - 100% paid for by CoreWeave
  • Company-paid Life Insurance
  • Voluntary supplemental life insurance
  • Short and long-term disability insurance
  • Flexible Spending Account
  • Health Savings Account
  • Tuition Reimbursement
  • Ability to Participate in Employee Stock Purchase Program (ESPP)
  • Mental Wellness Benefits through Spring Health
  • Family-Forming support provided by Carrot
  • Paid Parental Leave
  • Flexible, full-service childcare support with Kinside
  • 401(k) with a generous employer match
  • Flexible PTO
  • Catered lunch each day in our office and data center locations
  • A casual work environment
  • A work culture focused on innovative disruption
© 2024 Teal Labs, Inc
Privacy PolicyTerms of Service