Own the security of Alloy's AWS environment, partnering with the Infrastructure team to integrate security into cloud infrastructure and manage security incidents.
Cloud Operations Engineer
📜 Description
- Successfully co-ordinate with a global team of Cloud Operations Engineers to ensure uptime guarantees.
- Help scale the Cloud Operations Engineering team with new processes and tools.
- Assist in scoping, designing, and deploying systems to reduce Mean Time to Resolve for incidents.
- Monitor and detect customer-facing incidents on the Atlas platform and assist in their resolution.
- Automate routine monitoring and troubleshooting tasks.
- Diagnose live incidents and differentiate between platform and usage issues.
🛠️ Requirements
- Experience with being an on call DevOps, SRE, or Cloud Operations engineer (at least 8 years)
- Expertise with Linux system administration, configuration, troubleshooting
- Experience in monitoring, system performance data collection and analysis, and reporting
- Expertise with networking technologies like DNS, TCP/IP, etc
- Knowledge of database operations and concepts.
- Familiarity with Amazon Web Services and other Cloud infrastructure platforms (e.g. GCP, Azure)
- Capability to write small programs/scripts to solve both short-term systems problems
- A CS/CE degree or equivalent experience
- At least 1 of the following programming languages: Java, Go, Python, Javascript
- A keen interest in learning new things
Full job description
We are looking to speak to candidates who are based in Bengaluru for our hybrid working model.
Responsibilities
- Successfully co-ordinate with a global team of Cloud Operations Engineers who are tasked with ensuring our uptime guarantees to our Atlas customer base
- Help scale the worldwide Cloud Operations Engineering team with the strategic implementation of new processes and tools
- Assist in scoping, designing and deploying systems that reduce Mean Time to Resolve for customer incidents
- Monitor and detect emerging customer-facing incidents on the Atlas platform; assist in their proactive resolution
- Automate routine monitoring and troubleshooting tasks
- Diagnose live incidents, differentiate between platform issues versus usage issues, and take the next steps toward resolution
- Cooperate with our product management and cloud engineering organizations by identifying areas for improvement in the management applications powering the Atlas infrastructure
- Inform executive leadership and escalation management personnel of major outages
- Coordinate and participate in a weekly on-call rotation, where you will handle short term customer incidents (from direct surveillance or through alerts via our Technical Services Engineers)
Requirements
- Experience with being an on call DevOps, SRE, or Cloud Operations engineer (at least 8 years)
- Expertise with Linux system administration, configuration, troubleshooting
- Experience in monitoring, system performance data collection and analysis, and reporting
- Expertise with networking technologies like DNS, TCP/IP, etc
- Knowledge of database operations and concepts.
- Familiarity with Amazon Web Services and other Cloud infrastructure platforms (e.g. GCP, Azure)
- Capability to write small programs/scripts to solve both short-term systems problems
- A CS/CE degree or equivalent experience
- At least 1 of the following programming languages: Java, Go, Python, Javascript
- A keen interest in learning new things
Special requirements
- Willingness to work in 1st Shift (7:30 am IST to 3:30 pm IST).
- Willingness to work from Wednesday - Sunday or vice versa
Nice To Have
- MongoDB
- Splunk
- Kubernetes
About MongoDB
MongoDB is built for change, empowering our customers and our people to innovate at the speed of the market. We have redefined the data platform for the AI era, enabling builders to create, transform, and disrupt industries with software. MongoDB’s unified data platform, the most widely available, globally distributed data platform on the market, helps organizations modernize legacy workloads, embrace innovation, and unleash AI. Our cloud-native platform, MongoDB Atlas, is the only globally distributed, multi-cloud data platform and is available across AWS, Google Cloud, and Microsoft Azure.
With offices worldwide and over 67,000 customers, including AI-native startups and approximately 75% of the Fortune 100, relying on MongoDB for their most important applications, we’re powering the next era of software.
Our compass at MongoDB is our Leadership Commitment, guiding how and why we make decisions, show up for each other, and win. It’s what makes us MongoDB.
To drive the personal growth and business impact of our employees, we’re committed to developing a supportive and enriching culture for everyone. From employee affinity groups, to fertility assistance and a generous parental leave policy, we value our employees’ wellbeing and want to support them along every step of their professional and personal journeys. Learn more about what it’s like to work at MongoDB, and help us make an impact on the world!
MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request an accommodation due to a disability, please inform your recruiter.
MongoDB is an equal opportunities employer.
| Requisition ID | 3273541776 |
Similar jobs
Search more Cloud Engineer jobsSenior Cloud Architect / Cloud Engineer
Design and implement enterprise-level cloud architectures for MS Azure and M365, providing technical guidance and collaborating with cross-functional teams to meet project objectives.
Senior Cloud Architect / Cloud Engineer
Cloud Engineers and Architects will develop enterprise-level cloud architectures and provide technical guidance for various ICT projects, ensuring adherence to government standards.
Senior Cloud Architect / Cloud Engineer
Seeking an experienced Cloud Engineer or Architect to develop enterprise-level cloud architectures for MS Azure and M365, ensuring adherence to government standards and providing technical guidance.
As a Senior Cloud Security Engineer, you will secure our multi-account AWS and Kubernetes environment by building guardrails and managing vulnerabilities.
