Remote Jobs RockRemote Jobs Rock

Senior Site Reliability Engineer

📅 Mar 19
DockerKubernetesTerraformAWS

📜 Description

  • Build and maintain cloud infrastructure for hosting applications and platforms.
  • Guide deployments using technologies like Docker, Kubernetes, and Terraform.
  • Work within a DevOps model alongside product development teams.
  • Manage and scale web applications and data platforms.
  • Create reusable and immutable infrastructure with Terraform.
  • Participate in on-call rotations and conduct root cause analysis.

🛠️ Requirements

  • 5+ years of professional experience
  • Take full ownership of significant system components with responsibility for their reliability and performance
  • Manage Lifecycle of core project infrastructure, from design through to deployment, maintenance, and performance optimization
  • Familiar with industry standards and devops best practices
  • Experience supporting the overall platform in on-call rotations.
  • Able to operate with minimal supervision

✨ Benefits

  • Every Teikametrics employee is eligible for company equity
  • Remote Work Flexibility - Work from home or from our offices, with flexible remote options
  • Broadband reimbursement
  • Group Medical Insurance – Coverage of INR 7,50,000 per annum for a family
  • Crèche benefit
Full job description
ABOUT THE COMPANY
Teikametrics is revolutionizing retail through our patented Artificial Retail Intelligence platform. Our proprietary orchestration layer functions as prompt intelligence, verticalizing AI for Amazon, Walmart, TikTok, and emerging marketplace use cases. For more information, visit www.teikametrics.com.
ABOUT THE ROLE
Teikametrics is looking for a Site Reliability Engineer at Bengaluru, India to help us build and maintain our cloud infrastructure for hosting Teikametrics applications and platforms, in addition to helping build internal DevOps tools and best practices required for efficient software development and deployment.
This highly visible role will help us guide deployments using the latest technologies such as Docker, Kubernetes, and Terraform while having an enormous impact on our entire organization.
You will work within a DevOps model alongside product development teams, designing, deploying, and managing automation tools that increase predictability, improve efficiency, and reduce operational cost.
ABOUT THE TEAM
We currently host services and infrastructure across AWS in tandem with third party providers and must address the security and scaling challenges that come with this. Our daily role involves:
  • Managing, and scaling web applications and data platforms

  • Building tools to implement devops and security best practices

  • Creating reusable and immutable infrastructure with Terraform

  • Continuously improve our infrastructure with monitoring, logging, and alerting

  • Developing authentication and gateway solutions for our infrastructure and applications

  • Investigating, identifying application issues and advising development teams on design, deployment and infrastructure choices.

  • Participating in on-call rotations, post-mortems and root cause analysis (RCA)

Press Reference about Teika Teikametrics’ Marketplace Optimization Platform, Announces Artificial Retail Intelligence (ARI): An AI-Powered Tool Designed to Drive Cross-Marketplace Success   The job description is representative of typical duties and responsibilities for the position and is not all-inclusive. Other duties and responsibilities may be assigned in accordance with business needs. We are proud to be an equal opportunity employer. A background check will be conducted after a conditional offer of employment is extended. #LI-Remote Beware of Recruitment Scams Teikametrics will never ask you to communicate via WhatsApp, Telegram, or other messaging apps during our hiring process. We do not request payment, financial information, or purchases of equipment during recruitment. All legitimate job openings are posted on our official careers page at teikametrics.com/careers, and our recruiters will only contact you through LinkedIn or official company email addresses (@teikametrics.com).   If you're unsure about a communication you've received, please contact us directly at careers@teikametrics.com to verify.

WHO YOU ARE

  • 5+ years of professional experience

  • Take full ownership of significant system components with responsibility for their reliability and performance

  • Manage Lifecycle of core project infrastructure, from design through to deployment, maintenance, and performance optimization

  • Familiar with industry standards and devops best practices

  • Experience supporting the overall platform in on-call rotations.

  • Able to operate with minimal supervision

  • HOW YOU'LL SPEND YOUR TIME

  • Managing deployment infrastructure and automations including Github, CI/CD pipelines,and other deployment tooling

  • Experience managing workflows and pipelines using tools like CircleCI, Argo Workflows etc

  • Cloud computing providers such as AWS

  • Hands-on experience with Kubernetes (EKS, GKE) or similar container orchestration platforms

  • Infrastructure as code tools such as Terraform

  • Experience coding with at least one language such as Bash, Python required

  • Hands-on experience with authentication and authorization technologies required

  • Automation experience of cloud environments

  • Containerization technologies and tools such as Docker

  • Monitoring tools such as Datadog, Opensearch, Sentry

  • WHAT CAN HELP YOU STAND OUT

  • Good at using AI agents and writing project specific standard guidelines

  • Experience operating data pipelines with Databricks, Kafka

  • Experience with Java, Javascript

  • Experience with managing infrastructure costs and budgets

  • Experience with databases like AWS RDS/Postgres

  • WE'VE GOT YOU COVERED

  • Every Teikametrics employee is eligible for company equity
  • Remote Work Flexibility - Work from home or from our offices, with flexible remote options
  • Broadband reimbursement
  • Group Medical Insurance – Coverage of INR 7,50,000 per annum for a family
  • Crèche benefit
  • Delinea

    Senior Site Reliability Engineer - FedRAMP

    🕒 19 days ago
    Delinea👥 11 - 50 employees🏢 Computer Software

    As a Senior Site Reliability Engineer, you will ensure the availability and performance of production SaaS services in Azure and AWS, focusing on reliability, incident response, and automation.

    Site Reliability EngineeringDevOpsAzureAWS
    Camunda

    Senior Site Reliability Engineer

    🕒 8 days ago
    Camunda👥 10,000+ employees🏢 Software Development🤝 B2B

    As a Senior Site Reliability Engineer, you'll design and maintain a Kubernetes-based multi-cloud platform, enhance monitoring tools, and mentor engineers while ensuring system reliability and scalability.

    KubernetesTerraformPrometheusGrafana
    Replit

    Staff Site Reliability Engineer

    🕒 11 days ago
    Replit👥 10,000+ employees🏢 Software Development🤝 B2B

    As a Staff Site Reliability Engineer, you will enhance the reliability and performance of Replit's infrastructure by implementing automation, leading incident management, and mentoring the engineering team.

    Site Reliability EngineeringDevOpsSystems EngineeringInfrastructure Engineering
    Bitdeer Technologies Group

    Sr GPU Cloud K8S Expert (SRE SME)

    🕒 18 days ago
    Bitdeer Technologies Group👥 201 - 500 employees🏢 Computer Software

    Design, deploy, and operate the control plane for an AI-operated GPU cloud, ensuring automated remediation and efficient management of Kubernetes clusters optimized for GPU workloads.

    KubernetesGPU WorkloadsNvidia GPU OperatorTopology-aware Scheduling
    Okta

    Senior Site Reliability Engineer

    🕒 11 days ago
    Okta👥 10,000+ employees🏢 Software Development

    As a Senior Site Reliability Engineer, you will build and operate a Kubernetes-based platform, enhancing internal workflows and ensuring reliability for AI-driven processes.

    KubernetesSite Reliability EngineeringPlatform EngineeringInfrastructure Engineering

    Trusted by Remote Workers