Senior Software Engineer - HPC

NVIDIA
Westford, MA, US
Full-time

NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 fueled the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing.

More recently, GPU deep learning ignited modern AI and enabled the next era of computing. NVIDIA is a learning machine that constantly evolves by adapting to new opportunities that are hard to address, that matters to the world, and that only we can address.

This is our life’s work, to amplify human imagination and intelligence, and expand what is possible. We’re seeking strategic, bold, hard-working, and creative individuals who are passionate about helping us tackle challenges no one else can solve.

Make the choice to join us today.

We are looking for a Senior Software Engineer to join our mission to continue improving our HPC infrastructure. Our team builds and operates sophisticated infrastructure to enable business critical services and AI applications.

You will be working with a team of passionate and skilled engineers that are continuously working to provide better tools to build and manage this infrastructure.

Ideal candidate is strong in software development, designing and creating reliable distributed systems, and has the ability to implement well thought out long term maintenance strategy.

What you’ll be doing :

Design highly available and scalable systems to meet the demands of our HPC clusters

Evaluate new and innovative technologies as the landscape evolves

Continuously improve infrastructure provisioning and management using automation

Support a globally distributed, multi-cloud hybrid environment - AWS, GCP and On-prem

Build strong cross functional relationships and align with partners across various business units

Ensure the highest level of up-time and Quality of Service (QoS) to our users through operational excellence

Participate in team's on-call rotation and be a contact for service incidents

What we need to see :

10+ years of experience in design, implementation, and delivery of large engineering projects

Comfortable with at least two of the following programming languages : Golang, Java, C / C++, Scala, Python, Elixir.

Understands scalability challenges and performance of server-side code. Able to craft and develop horizontally-scalable, resilient and performing-under-load systems.

Versatile technologist with experience in full software development lifecycle from inception and design to deployment, operation, and iterative development.

Proficient in cloud computing and are hands-on in at least one cloud platform : GCP, AWS, or Azure.

Proficient in modern CI / CD techniques, GitOps and Infrastructure as Code(IaC)

Strong work ethic and a passion for problem solving

B.S. degree in Computer Science or related technical field (or equivalent experience)

Detail oriented with great communication and collaboration skills

Ways to stand out from the crowd :

Prior experience building solutions for HPC clusters based on Slurm or Kubernetes

Strong understanding of Linux operation system and TCP / IP fundamentals

The base salary range is 180,000 USD - 339,250 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

You will also be eligible for equity and . NVIDIA accepts applications on an ongoing basis.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

30+ days ago
Related jobs
Promoted
Raytheon
Belmont, Massachusetts

In this role, you will be joining a team where our software engineers and architects are developing and maintaining advanced ground station software. We bring the strength of more than 100 years of experience and renowned engineering expertise to meet the needs of today’s mission and stay ahead of t...

Abbott
Burlington, Massachusetts

The is responsible for executing and maintaining software quality engineering methodologies and providing quality engineering support for software. Senior Quality Engineer, Software. This includes, but is not limited to, review and approval of software test case protocols and reports, review of soft...

Fresenius Kidney Care
Lawrence, Massachusetts

Successfully completed studies in computer science, computer engineering, electrical engineering, software engineering, medical technology, or similar field of specialization. Senior Software Development Engineer. As a medical device Software Test Engineer in Fresenius Medical Care's Home business u...

Randstad Digital
Woburn, Massachusetts

Senior Software Engineers for Randstad Digital, LLC. ...

ZoomInfo
Waltham, Massachusetts

As a Senior Software Engineer on the Product Foundations Team, you will get to design our next generation provisioning & authorization systems, supporting the growth of our platform and the creation of an application marketplace. Serve as a Senior Engineer working on large data pipelines, to transfe...

The Resource Technology Partners
Waltham, Massachusetts

Validation / Lead Reliability Engineer. The Validation / Lead Reliability Engineer will lead and be an intrinsic part of a dynamic, collaborative team that believes deeply in the importance of what we are doing and that we can achieve it. BS in Math / Statistics / Engineering fields. ...

Capital One
Pepperell, Massachusetts

Main Street (21020), United States of America, Cambridge, MassachusettsSenior Software Engineer, Full StackThe Card CTO & Premium Products, Marketing, and Experiences (PPMX) Tech team delivers experiences for both our customers and our associates across Card Tech. We are seeking Full Stack Software ...

Randstad Digital
Woburn, Massachusetts

Must have Master's degree or foreign equivalent in Computer Science, Computer Engineering, Electronic Engineering, or a related field and 2 years experience in the proffered position, or as a Developer, Software Engineer, Technology Analyst, Programmer Analyst, or related occupation. Senior Software...

Cumulus
Waltham, Massachusetts

As a Senior Full Stack Software Engineer, you will focus on improving the functionality of Cumulus's web and mobile apps. As a Senior Full Stack Software Engineer, your focus will be on improving the functionality of Cumulus's web and mobile apps. As a Senior Full Stack Software Engineer, you will d...

Synopsys
Burlington, Massachusetts

Looking for a Senior R&D engineer in ProtoCompiler R&D with a specialization in Static Timing Analysis to join our team in Sunnyvale/Marlborough. Incorporate advanced software engineering tools and processes related to documentation and coding practices, memory and runtime profiling, coverag...