Remote Jobs RockRemote Jobs Rock

Engineering Manager, Cloud Infrastructure

πŸ•’ yesterday
Engineering ManagementInfrastructure As Code (iac)KubernetesNetworking

πŸ“œ Description

  • Own the cloud-platform roadmap, translating product and security needs into actionable outcomes.
  • Build self-service infrastructure interfaces to enable internal teams to provision resources independently.
  • Ensure platform availability, manage upgrades, and handle incident remediation.
  • Stay technically engaged by reviewing designs and debugging complex issues.
  • Coach and develop engineers, setting clear expectations and delegating ownership.

πŸ› οΈ Requirements

  • Demonstrated engineering management. You have led and developed engineers, made prioritization and performance decisions, hired thoughtfully, and delivered through a teamβ€”not only acted as its strongest individual contributor.
  • Software-oriented infrastructure depth. You have built and operated cloud platforms or distributed systems and can reason across infrastructure code, Kubernetes, networking, service identity, and stateful dependencies.
  • Safe-change and production judgment. You have owned consequential migrations and incidents, can explain failure modes and rollback limits, and know when simplifying a system is better than adding another platform.
  • Platform-product and engineering judgment. You understand internal customers, create interfaces other teams adopt, and make clear tradeoffs among reliability, developer autonomy, engineering effort, and workload efficiency.
  • Experience with multi-tenant, cellular, regional, or dedicated enterprise infrastructure.
  • Familiarity with GCP/GKE, Terraform or similar IaC systems, Cloudflare, Envoy/Istio, SPIFFE/SPIRE, and managed data services.
  • Experience with large-fleet rightsizing, infrastructure consolidation, or migrating CI compute without disrupting developer workflows.
  • A track record using AI tools to increase engineering output while preserving production safeguards.

✨ Benefits

  • πŸ’° Competitive Salary & Equity
  • πŸ’Ή 401(k) Program with a 4% match ( US Only )
  • Dental
  • Vision and Life Insurance
  • 🩼 Short Term and Long Term Disability
  • 🚼 Paid Parental
  • Medical
  • Caregiver Leave
  • 🏝 Flexible Time Off (FTO) + Holidays
  • πŸš— Commuter Benefits ( In-Office & US Only )
Full job description

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation.

About the Role

Replit enables people to build software with AI. The infrastructure underneath that experience must make it straightforward to launch services, isolate workloads, and run reliable systems at scale.

We're hiring a hands-on Engineering Manager to lead Cloud Infrastructure: the shared infrastructure as code (IaC), networking, storage, compute, and service mesh platforms that Replit's product and platform teams depend on. You'll lead and grow an existing engineering team building and operating these foundations, including Kubernetes, shared edge networking, service mesh and workload identity.

This is a platform-building role with production accountability. You should be comfortable going deep on a design or incident while developing a team that does not depend on you for every decision.

What You'll Do

  • Own the cloud-platform roadmap. Lead the team's IaC, networking, storage, compute, and service mesh platforms. Translate product, platform, reliability, and security needs into sequenced outcomes, balancing foundational investment, lifecycle work, and delivery commitments against the team's capacity.

  • Make infrastructure repeatable and self-service. Build maintained IaC interfaces for services, cells, connectivity, identities, and shared resources. Enable internal customer teams to provision infrastructure without bespoke coordination or dependence on individual experts.

  • Operate what the team builds. Own platform availability, upgrades, isolation, recovery, and incident remediation. Maintain clear SLOs, sustainable on-call coverage, and primary and backup owners for critical systems.

  • Stay technically engaged. Review designs and production changes, debug difficult failure modes, and use AI coding toolsβ€”including Replitβ€”to prototype and automate. Apply rigorous review and verification to AI-generated infrastructure changes.

  • Build and grow a high-ownership engineering team. Coach engineers, develop technical leaders, set clear expectations, manage performance, and hire against agreed needs. Delegate meaningful ownership as the team grows.

What You'll Bring

  • Demonstrated engineering management. You have led and developed engineers, made prioritization and performance decisions, hired thoughtfully, and delivered through a teamβ€”not only acted as its strongest individual contributor.

  • Software-oriented infrastructure depth. You have built and operated cloud platforms or distributed systems and can reason across infrastructure code, Kubernetes, networking, service identity, and stateful dependencies.

  • Safe-change and production judgment. You have owned consequential migrations and incidents, can explain failure modes and rollback limits, and know when simplifying a system is better than adding another platform.

  • Platform-product and engineering judgment. You understand internal customers, create interfaces other teams adopt, and make clear tradeoffs among reliability, developer autonomy, engineering effort, and workload efficiency.

Nice to Have

  • Experience with multi-tenant, cellular, regional, or dedicated enterprise infrastructure.

  • Familiarity with GCP/GKE, Terraform or similar IaC systems, Cloudflare, Envoy/Istio, SPIFFE/SPIRE, and managed data services.

  • Experience with large-fleet rightsizing, infrastructure consolidation, or migrating CI compute without disrupting developer workflows.

  • A track record using AI tools to increase engineering output while preserving production safeguards.

Full-Time Employee Benefits Include:

πŸ’° Competitive Salary & Equity

πŸ’Ή 401(k) Program with a 4% match (US Only)

βš•οΈ Health, Dental, Vision and Life Insurance

🩼 Short Term and Long Term Disability

🚼 Paid Parental, Medical, Caregiver Leave

🏝 Flexible Time Off (FTO) + Holidays

πŸš— Commuter Benefits (In-Office & US Only)

πŸ“± Monthly Wellness Stipend

πŸ§‘β€πŸ’» Autonomous Work Environment

πŸ–₯ In Office Set-Up Reimbursement (In-Office Only)

πŸš€ Quarterly Team Gatherings

β˜• In Office Amenities (In-Office Only)

Want to learn more about what we are up to?

Interviewing + Culture at Replit

To achieve our mission of making programming more accessible around the world, we need our team to be representative of the world. We welcome your unique perspective and experiences in shaping this product. We encourage people from all kinds of backgrounds to apply, including and especially candidates from underrepresented and non-traditional backgrounds.

Vapi

Engineering Manager, Trust & Safety

πŸ•’ 2 days ago
VapiπŸ‘₯ 51 - 200 employees🏒 Technology

Lead a team of engineers at Vapi, focusing on billing, fraud, and outbound telephony systems while prioritizing people management and service reliability.

Engineering ManagementCoachingPerformance ManagementSystem Design
Unzer

Engineering Manager (m/f/d)

πŸ•’ 4 days ago
UnzerπŸ‘₯ 201 - 500 employees🏒 Financial Services

As an Engineering Manager, you will lead and mentor a team of engineers, ensuring the delivery and quality of product areas while driving continuous improvement in engineering practices.

LeadershipSoftware EngineeringJavaKotlin
Snowflake

Engineering Manager - Cloud Efficiency

πŸ•’ 7 days ago
SnowflakeπŸ‘₯ 10,000+ employees🏒 Software Development

As an Engineering Manager for Cloud Efficiency, you will lead a team to optimize Snowflake's cloud infrastructure, driving insights and managing costs effectively.

Software EngineeringEngineering ManagementDistributed SystemsData Engineering
ASSYSTEM

Principal EC&I Engineer

πŸ•’ 2 days ago
ASSYSTEMπŸ‘₯ 5001 - 10,000 employees🏒 Management Consulting

As an Engineering Manager for the Electrical, Control, and Instrumentation (EC&I) department, you will lead projects for the design and delivery of advanced SMR nuclear power plants.

Electrical EngineeringControl EngineeringInstrumentationProject Management
ASSYSTEM

Principal EC&I Engineer

πŸ•’ 2 days ago
ASSYSTEMπŸ‘₯ 5001 - 10,000 employees🏒 Management Consulting

As an Engineering Manager in the EC&I department, you will lead the design and delivery of advanced SMR nuclear power plants, ensuring project success through meticulous planning and collaboration.

Electrical EngineeringControl EngineeringInstrumentationProject Management

Trusted by Remote Workers