As a Senior DevOps / Infrastructure Engineer, you will own the infrastructure, deployment, and operational systems supporting high-scale products across a rapidly evolving technology ecosystem.
Senior DevOps / Infrastructure Engineer
📜 Description
- Own infrastructure architecture and operations across production applications and services, including cloud, Kubernetes, databases, deployment systems, and blockchain infrastructure.
- Design, build, and operate highly available AWS environments, including Kubernetes clusters, networking, storage, scaling, upgrades, and related infrastructure services.
- Manage infrastructure as code using Terraform and GitOps practices, while designing and maintaining CI/CD pipelines that support safe deployments and reliable rollbacks.
- Operate PostgreSQL databases in production, with responsibility for availability, replication, backups, recovery, and performance optimization.
- Build comprehensive observability across infrastructure, applications, databases, and blockchain nodes, including monitoring and alerting capable of identifying incidents early.
- Diagnose complex production incidents across the full technology stack, from Kubernetes and networking through databases and blockchain systems.
🛠️ Requirements
- 5+ years of professional experience in DevOps, SRE, platform engineering, infrastructure engineering, or a closely related discipline.
- Proven track record of building and operating production infrastructure for highly available systems where reliability directly impacts users or business operations.
- Deep hands-on experience with Kubernetes, AWS, and Linux.
- Strong infrastructure-as-code experience with Terraform, combined with GitOps practices and CI/CD pipeline design.
- Solid networking fundamentals, including DNS, load balancing, TLS, and service discovery.
- Strong production PostgreSQL experience covering replication, backups, recovery, performance, and availability.
✨ Benefits
- Opportunity to own infrastructure supporting high-scale applications and distributed systems.
- Broad technical scope spanning cloud infrastructure, Kubernetes, databases, CI/CD, observability, security, and blockchain.
- Significant ownership and direct product impact within a small, highly technical engineering team.
- Opportunity to work on infrastructure supporting a decentralized exchange designed for high-throughput and low-latency trading.
- Exposure to complex distributed systems and production environments operating under demanding reliability and security requirements.
- Competitive compensation.
- Equity opportunity.
- Collaboration with a globally distributed engineering team working across multiple regions and time zones.
Full job description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior DevOps / Infrastructure Engineer based in United States .
As a Senior DevOps / Infrastructure Engineer, you will own the infrastructure, deployment, and operational systems supporting high-scale products across a rapidly evolving technology ecosystem. You will design, build, and operate production infrastructure end to end, spanning AWS, Kubernetes, CI/CD, databases, observability, security, and distributed systems. The role is highly hands-on and gives you significant ownership over architectural decisions and operational reliability. You will support demanding user-facing and financial technology products operating under high availability, security, throughput, and latency requirements. Experience with blockchain infrastructure is particularly valuable, while cloud security or SecOps expertise is a strong advantage. Working as the infrastructure owner during US hours, you will provide critical operational coverage across a globally distributed engineering organization.
Accountabilities:
- Own infrastructure architecture and operations across production applications and services, including cloud, Kubernetes, databases, deployment systems, and blockchain infrastructure.
- Design, build, and operate highly available AWS environments, including Kubernetes clusters, networking, storage, scaling, upgrades, and related infrastructure services.
- Manage infrastructure as code using Terraform and GitOps practices, while designing and maintaining CI/CD pipelines that support safe deployments and reliable rollbacks.
- Operate PostgreSQL databases in production, with responsibility for availability, replication, backups, recovery, and performance optimization.
- Build comprehensive observability across infrastructure, applications, databases, and blockchain nodes, including monitoring and alerting capable of identifying incidents early.
- Operate blockchain infrastructure such as nodes, RPC services, and indexers, monitoring systems at the protocol level for synchronization, connectivity, and transaction processing.
- Design disaster recovery and failover strategies for critical systems while strengthening security through access controls, secrets management, vulnerability management, monitoring, and related safeguards.
- Diagnose complex production incidents across the full technology stack, from Kubernetes and networking through databases and blockchain systems.
- Identify recurring operational bottlenecks and replace manual procedures with reliable automation.
- Contribute to architectural decisions and long-term infrastructure strategy, treating infrastructure as a product with defined reliability, maintenance, and user requirements.
- 5+ years of professional experience in DevOps, SRE, platform engineering, infrastructure engineering, or a closely related discipline.
- Proven track record of building and operating production infrastructure for highly available systems where reliability directly impacts users or business operations.
- Deep hands-on experience with Kubernetes, AWS, and Linux.
- Strong infrastructure-as-code experience with Terraform, combined with GitOps practices and CI/CD pipeline design.
- Solid networking fundamentals, including DNS, load balancing, TLS, and service discovery.
- Strong production PostgreSQL experience covering replication, backups, recovery, performance, and availability.
- Experience with high-availability architecture, incident response, root-cause analysis, and capacity planning.
- Hands-on experience operating blockchain infrastructure such as nodes, RPC services, indexers, validators, or blockchain data pipelines.
- Understanding of how distributed blockchain networks operate and how they can fail under real-world conditions.
- Strong ownership mindset, with the ability to take responsibility for problems from architecture through production.
- Comfortable designing for high load and partial failure and automating recurring operational tasks.
- Strong diagnostic and problem-solving skills across multiple infrastructure layers.
- Security or SecOps experience, including production environment hardening, AWS security services, and IAM design, is a strong asset.
- Experience with bare-metal operations, Helm or ArgoCD, messaging systems such as Kafka or NATS, or observability tools including Prometheus, Grafana, Loki, or OpenTelemetry is an asset.
- Experience operating blockchain infrastructure at significant scale, particularly within Ethereum, Polkadot, or Layer 2 ecosystems, is a strong plus.
- Exposure to financial, trading, or other latency-sensitive systems is beneficial.
- Experience using AI tools in infrastructure workflows is a plus.
- Familiarity with Go, Rust, or TypeScript is an additional advantage.
- Must be based in a US time zone (UTC-8 to UTC-5) while physically located in Canada.
- Fully remote position for candidates based in Canada and working within the required US time-zone range.
- Opportunity to own infrastructure supporting high-scale applications and distributed systems.
- Broad technical scope spanning cloud infrastructure, Kubernetes, databases, CI/CD, observability, security, and blockchain.
- Significant ownership and direct product impact within a small, highly technical engineering team.
- Opportunity to work on infrastructure supporting a decentralized exchange designed for high-throughput and low-latency trading.
- Exposure to complex distributed systems and production environments operating under demanding reliability and security requirements.
- Competitive compensation.
- Equity opportunity.
- Collaboration with a globally distributed engineering team working across multiple regions and time zones.
Requirements
Benefits
Similar jobs
Search more DevSecOps Engineer jobsAs a Senior DevOps Engineer, you will build and manage scalable cloud infrastructure, optimize CI/CD processes, and enhance operational reliability for a growing AdTech platform.
Senior DevOps/Platform Engineer
As a Senior DevOps/Platform Engineer, you will lead the migration of over 2,000 GitHub repositories, ensuring seamless CI/CD integration and platform stability.
Senior Observability/DevOps Engineer
Drive the migration of a large-scale Datadog environment, ensuring operational readiness and integration with AWS and Kubernetes while leveraging automation tools.
Senior DevOps/Platform Engineer
As a Senior DevOps / Platform Engineer, you'll lead the migration of over 2,000 GitHub repositories, ensuring smooth transitions and platform stability through automation and collaboration.
