Remote Jobs RockRemote Jobs Rock

Senior Site Reliability Engineer (SRE)

📅 Apr 10, 2025
Site Reliability EngineeringProblem AnalysisInfrastructure DevelopmentDevOps Practices

📜 Description

  • Demonstrate proficiency in problem analysis
  • Collaborate with teams to offer pragmatic solutions and be accountable for execution
  • Create, design, develop, and operate the shared infrastructure platform
  • Perpetually improve the developer experience by treating the platform as a product
  • Train and support software engineers on DevOps & SRE practices
  • Maintain platform security, compliance, and cost control

🛠️ Requirements

  • Significant site reliability engineering experience
  • Ability to design pragmatic and simple architectures
  • Commitment to code simplicity, quality, security, and performance

✨ Benefits

  • Collaborative work environment that values innovation and creativity
  • Competitive salary and benefits package
  • Professional development and career growth opportunities
Full job description
At Swile, we believe that good products can help reduce friction in daily professional life and boost employee satisfaction. Today, we provide innovative solutions in various areas such as Fintech, Travel, HR, and Employee Benefits to more than 5.5 million users in 85,000 companies in France and Brazil.
Your role as a Senior Site Reliability Engineer (SRE) centers around creatively solving problems, ensuring a balance between speed, reliability, pragmatism and excellence. It's less about the specific technologies used and more about crafting innovative solutions to business and engineering challenges.

Key Responsibilities:

As a Senior Site Reliability Engineer, your mission will be to:
  • Demonstrate proficiency in problem analysis
  • Collaborate with a variety of teams to offer pragmatic solutions to real world problems and be accountable for their end to end execution
  • Create, design, develop and operate the shared infrastructure platform to support Swile's growth, new products and international development
  • Perpetually improve the developer experience by considering the platform as a product (simplifying stuff, improving the time to market, automating everything that can be and that makes software engineers life better)
  • Train, coach and support software engineers to make teams autonomous on DevOps & SRE practices (to enable our “you build it, you run it” philosophy at scale)
  • Take care of platform security, compliance and keep costs under control
  • Obtain a good understanding of the business to provide relevant solutions to users and clients
  • Stay up-to-date on new technologies and architectures.
  • Demonstrate good judgment in their potential Swile applications
  • Provide positive, constructive and qualitative feedback during code reviews.
  • Mentor and promote tech growth within the team
  • ✨It will be a perfect match if you:

  • We welcome individuals with entrepreneurial backgrounds as well as those from established organizations.
  • At Swile, we believe that delivering impactful products requires engineers to understand the needs of users and clients as well as the code itself.
  • Have significant site reliability engineering experience
  • Value code simplicity, quality, security, performance and maintainability
  • Can design pragmatic & simple architectures to solve problems at scale
  • Want to work in a fast, high-growth startup environment that respects its engineers and customers
  • If you are a future responsible Swiler: you share our commitment to the environment, diversity, fairness and inclusion and are prepared to work every day to improve individual and collective performance.
  • ⚙️ Our Tech stack

  • You do not need to be familiar with our technical stack or any specific functional area, but we have a strong willingness to learn and adapt quickly.
  • Ruby/Rails, Typescript/React/Node.js, Android(Kotlin), iOS(Swift), Java, Golang, Python
  • AWS, Terraform, Kubernetes, PostgreSQL, Kafka, Redis, Snowflake, Github, Datadog
  • 💡 What’s in it for you ?

  • An opportunity to join an international tech company and a team of talented engineers
  • A collaborative work environment that values innovation and creativity
  • Competitive salary and benefits package
  • Professional development and career growth opportunities
  • Okta

    Senior Site Reliability Engineer

    🕒 11 days ago
    Okta👥 10,000+ employees🏢 Software Development

    As a Senior Site Reliability Engineer, you will build and operate a Kubernetes-based platform, enhancing internal workflows and ensuring reliability for AI-driven processes.

    KubernetesSite Reliability EngineeringPlatform EngineeringInfrastructure Engineering
    Crunchyroll

    Staff Site Reliability Engineer

    🕒 20 days ago
    Crunchyroll👥 1001 - 5000 employees🏢 Entertainment

    The Staff Site Reliability Engineer will enhance the reliability, scalability, performance, and security of Crunchyroll's data platforms while driving SRE practices and collaborating across teams.

    Site Reliability EngineeringPlatform EngineeringInfrastructure EngineeringKubernetes
    Crunchyroll

    Senior Site Reliability Engineer

    🕒 22 days ago
    Crunchyroll👥 1001 - 5000 employees🏢 Entertainment

    As a Senior Site Reliability Engineer, you will enhance the reliability, scalability, and security of Crunchyroll's data platforms while driving modern SRE practices and cross-functional collaboration.

    Site Reliability EngineeringKubernetesGoogle Cloud PlatformInfrastructure As Code
    Braze

    Senior Site Reliability Engineer I

    🕒 25 days ago
    Braze👥 10,000+ employees🏢 Software Development🤝 B2B

    As a Senior Site Reliability Engineer, you will ensure the reliability and uptime of internal services, collaborating with engineering teams to enhance infrastructure and automation.

    Site Reliability EngineeringInfrastructure As CodeChefTerraform
    Lambda

    Senior HPC Engineer - Fleet Engineering

    🕒 28 days ago
    Lambda👥 501 - 1000 employees🏢 Technology

    As a Senior Site Reliability Engineer, you will build and operate monitoring systems, automate HPC cluster management, and troubleshoot complex infrastructure issues.

    Site Reliability EngineeringHPC EngineeringDevOpsAI Infrastructure

    Trusted by Remote Workers