Remote Jobs RockRemote Jobs Rock

Senior Site Reliability Engineer for Fuse team

🕒 7 days ago
KubernetesGCPGoPython

📜 Description

  • Own and improve the reliability posture of Fuse services, APIs, and storage systems.
  • Establish SLIs, SLOs, and error budgets for customer-facing APIs and asynchronous jobs.
  • Build end-to-end observability across Data Hub item collections.
  • Create actionable dashboards and alerts using Grafana and Prometheus.
  • Improve capacity planning and autoscaling using workload telemetry.
  • Lead incident investigation and facilitate blameless incident reviews.

🛠️ Requirements

  • Strong hands-on experience operating services on Kubernetes in a major cloud environment, ideally GCP.
  • Experience with observability and incident diagnosis for distributed systems.
  • Proficiency in Go or Python.
  • Experience operating relational databases, preferably PostgreSQL.
  • Familiarity with large-scale data or indexing systems such as Bigtable or Elasticsearch.
  • Experience with CI/CD, Infrastructure as Code, and deployment automation.
  • Ability to work effectively in a distributed, remote-first team.

Benefits

  • Restricted Stock Units or Stock Options are granted depending on a team member’s role, seniority, and location.*
  • Everyone gets to participate in the company's success through the company performance bonus.*
🕒 4 days ago

As a Senior Site Reliability Engineer, you will define and measure reliability for a GPU-accelerated AI platform, owning service-level indicators and objectives while collaborating across teams.

Site Reliability EngineeringKubernetesAPI DevelopmentObservability
Fivetran

Senior Site Reliability Engineer

Fivetran👥 10,000+ employees🏢 Software Development🤝 B2B
🔥 17 hours ago

As a Senior Site Reliability Engineer, you will enhance the performance and reliability of Fivetran's infrastructure while collaborating with cross-functional teams to ensure robust data pipelines.

KubernetesPostgreSQLArgocdTerraform

Trusted by Remote Workers