Expoint - all jobs in one place

Finding the best job has never been easier

Limitless High-tech career opportunities - Expoint

MongoDB Cloud Operations Engineer FedRamp 
United States, New York, New York 
779824011

24.06.2024
Responsibilities
  • Successfully coordinate and collaborate with a global team of Cloud Operations Engineers who are tasked with ensuring our uptime guarantees to our Atlas customer base
  • Help scale the worldwide Cloud Operations Engineering team with the strategic implementation and refinement of new processes and tools
  • Assist in scoping, designing and deploying systems that reduce Mean Time to Resolve for customer incidents
  • Monitor and detect emerging customer-facing incidents on the Atlas platform; assist in their proactive resolution
  • Automate routine monitoring and troubleshooting tasks
  • Diagnose live incidents, differentiate between platform issues versus usage issues, and take the next steps toward resolution
  • Assist in performing root cause analysis after incident recovered; identifying any breakdowns in processes or workflows that contributed to the event and what changes need to be made to prevent similar events
  • Contribute to documentation of corner case scenarios, troubleshooting workflows and SOPs
  • Work alongside our product management, cloud engineering and support organizations by identifying areas for improvement in the management applications powering the Atlas infrastructure
  • Inform executive leadership and escalation management personnel of major outages
  • Coordinate with the team to handle short term customer incidents (proactively from automated monitoring or through reactive alerts via our Technical Services team)
  • Work First Shift: 7am - 4pm EST
Requirements
  • 2+ years experience with being an on call DevOps, SRE, or Cloud Operations engineer
  • Expertise with Linux system administration, configuration, troubleshooting
  • Experience in monitoring, system performance data collection and analysis, and reporting
  • Knowledge of database operations and concepts
  • Expertise with networking technologies like DNS, TCP/IP, etc
  • Familiarity with Amazon Web Services and other Cloud infrastructure platforms (e.g. GCP, Azure)
  • Knowledgeable about a wide range of web and internet technologies
  • Capability to write small programs/scripts to solve both short-term systems problems
  • A CS/CE degree or equivalent experience
  • At least 1 of the following programming languages: Java, Go, Python, Javascript
  • A keen interest in learning new things
Special requirements
  • Be a US Person on US soil (i.e. U.S. citizen, U.S. national, lawful permanent resident, asylee, or refugee living and working in the United States)
  • Willingness and ability to participate in pager duty rotations during nights, weekends and holidays (approximately one out of every six weeks), at least during an initial ramping period (and potentially permanently)
$176,000 USD