The point where experts and best companies meet
Share
What you’ll be doing:
Design highly available and scalable systems to meet the demands of our HPC clusters
Evaluate new and innovative technologies as the landscape evolves
Continuously improve infrastructure provisioning and management using automation
Support a globally distributed, multi-cloud hybrid environment - AWS, GCP and On-prem
Build strong cross functional relationships and align with partners across various business units
Ensure the highest level of up-time and Quality of Service (QoS) to our users through operational excellence
Participate in team's on-call rotation and be a contact for service incidents
:
5+ years of experience in design, implementation, and delivery of large engineering projects
Comfortable with at least two of the following programming languages: Golang, Java, C/C++, Scala, Python, Elixir.
Understands scalability challenges and performance of server-side code. Able to craft and develophorizontally-scalable,resilient andperforming-under-loadsystems.
Versatile technologist with experience in full software development lifecycle – from inception and design to deployment, operation, and iterative development.
Proficient in cloud computing and are hands-on in at least one cloud platform: GCP, AWS, or Azure.
Proficient in modern CI/CD techniques, GitOps and Infrastructure as Code(IaC)
Strong work ethic and a passion for problem solving
B.S. degree in Computer Science or related technical field (or equivalent experience)
Detail oriented with great communication and collaboration skills
Ways to stand out from the crowd:
Prior experience building solutions for HPC clusters based on Slurm or Kubernetes
Strong understanding of Linux operation system and TCP/IP fundamentals
These jobs might be a good fit