← Back to jobs

Cloud Operations Engineer

Company
MongoDB
Location
Remote
Work type
Full Time
Posted
2026-08-10

Job description

Responsibilities
Successfully coordinate with a global team of Cloud Operations Engineers who are tasked with ensuring our uptime guarantees to our Atlas customer base
Help scale the worldwide Cloud Operations Engineering team with the strategic implementation of new processes and tools
Assist in scoping, designing and deploying systems that reduce Mean Time to Resolve for customer incidents
Monitor and detect emerging customer-facing incidents on the Atlas platform; assist in their proactive resolution
Automate routine monitoring and troubleshooting tasks
Diagnose live incidents, differentiate between platform issues versus usage issues, and take the next steps toward resolution
Cooperate with our product management and cloud engineering organizations by identifying areas for improvement in the management applications powering the Atlas infrastructure
Inform executive leadership and escalation management personnel of major outages
Coordinate and participate in a weekly on-call rotation, where you will handle short term customer incidents (from direct surveillance or through alerts via our Technical Services Engineers)

Requirements
Experience with being an on call DevOps, SRE, or Cloud Operations engineer (at least 2 years)
Expertise with Linux system administration, configuration, troubleshooting
Experience in monitoring, system performance data collection and analysis, and reporting
Expertise with networking technologies like DNS, TCP/IP, etc
Knowledge of database operations and concepts
Familiarity with Amazon Web Services and other Cloud infrastructure platforms (e.g. GCP, Azure)
Capability to write small programs/scripts to solve both short-term systems problems
A CS/CE degree or equivalent experience
At least 1 of the following programming languages: Java, Go, Python, Javascript
A keen interest in learning new things

Special requirements
Be a U.S. citizen or U.S. national. Due to the nature of this role and access requirements for FedRAMP environments, this position requires work to be performed on U.S. soil
Willingness and ability to participate in pager duty rotations during nights, weekends and holidays (approximately one out of every six weeks), at least during an initial ramping period (and potentially permanently)
Willingness and ability to work 2nd shift weekend hours (3pm - 12am EST; Saturday - Wednesday)

Nice To Have
MongoDB
Splunk
Kubernetes
Benefits include
Competitive salary, equity, pension and health insurance
Regular performance, compensation and development reviews
20 weeks Maternity & Paternity leave to spend time with new arrivals

MongoDB’s base salary range for this role in the U.S. is:
$90,000—$176,000 USD
Skills Required
On-call DevOps, SRE, or Cloud Operations experience (at least 2 years)
Linux system administration, configuration, troubleshooting
Monitoring, system performance data collection, analysis, and reporting
Networking technologies (DNS, TCP/IP)
Knowledge of database operations and concepts
Familiarity with AWS and other cloud platforms (GCP, Azure)
Ability to write small programs/scripts to solve systems problems
CS/CE degree or equivalent experience
Proficiency in at least one: Java, Go, Python, JavaScript
Willingness and ability to participate in pager duty rotations nights, weekends, holidays
Willingness and ability to work 2nd shift weekend hours (3pm - 12am EST; Saturday - Wednesday)
Be a US Person (U.S. citizen, U.S. national, lawful permanent resident, asylee, or refugee)
A keen interest in learning new things
Experience supporting FedRAMP or government customers (SLED/federal)
Experience with MongoDB
Experience with Splunk
Experience with Kubernetes

Original source