Platform Engineer - System Infrastructure Engineer

Baysystemsinc

Berkeley (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Baysystemsinc is seeking experienced software engineers in Berkeley, California, to lead API development and software engineering projects that automate supercomputing resources. This role involves collaboration with researchers and cross-functional teams in a hybrid work environment.

The ideal candidate has extensive experience in high-performance computing and is proficient in programming languages like C, Python, and advanced container technologies. Join us to make a significant impact in the field!

Qualifications

  • 8+ years of related experience with a Bachelor’s degree, or 6+ years with a Master’s degree.
  • 2+ years of experience with API and web services software development.
  • Familiarity with designing and building API interfaces.

Responsibilities

  • Develop and maintain a broad portfolio of software projects.
  • Build and support API endpoints for complex workflows.
  • Troubleshoot complex technical problems with team members.
  • Research and lead implementation of new technologies.
  • Present developments to NERSC staff and the HPC community.

Skills

API development
C, shell, and Python programming
Git and CI/CD pipelines
AI (machine learning) tools
Database administration (MongoDB, MySQL, PostgreSQL)
Container technologies (Docker, Kubernetes)
Software engineering best practices
Communication skills

Education

Bachelor’s degree or equivalent experience

Tools

Docker
Kubernetes
MongoDB
MySQL
PostgreSQL

Job description

If you are unable to complete this application due to a disability, contact this employer to ask for an accommodation or an alternative application process.

Temporary Berkeley, California, Berkeley, CA, US

7 days ago Requisition ID: 1418

1 year contract - Extension/Conversion based on job performance

Hybrid - 1 to 4 days onsite - Berkeley, California

In this exciting role, you will work on API development and other software engineering projects to help automate the use of supercomputing resources and introduce cloud-native and AI tools and techniques for researchers to use at massive scale. You’ll join a group of systems and software engineers and will routinely work with other groups across NERSC on a variety of projects. You’ll also collaborate with our counterparts at peer scientific facilities, also operated by the Department of Energy Office of Science, on a national program to pool together vast computational and storage resources through the development of APIs, distributed services, and community standards and best practices.

Job Responsibilities
  • Work with a team to develop and maintain a broad portfolio of software projects.
  • Build, refine and support API endpoints and integration to backend systems to enable automation for complex workflows.
  • Troubleshoot and solve complex technical problems with other team members
  • Develop and refactor scripts and other code.
  • Coordinate small project teams or other initiatives (such as the rollout of a new service or system, or a major equipment or software upgrade).
  • Work with vendors to prioritize efforts and enhance their technologies to meet user needs
  • Work with researchers to deploy services using Spin, our container cloud platform based on Kubernetes.
  • Collaborate within NERSC and across the DOE community to develop APIs and services, integrate them into the new NERSC supercomputer Doudna, the NERSC data center environment, and across multiple DOE facilities.
  • Present developments to NERSC staff and the broader HPC community at science conferences and industry meetings.
  • Analyze and solve complex technical problems requiring in-depth evaluation of variable factors
  • Work at a higher level of independence while carrying out work assignments.
  • Research, select, and lead the implementation of new technologies.
  • Develop team strategy and project plans.
  • Provide leadership and technical guidance to group members and other colleagues at NERSC.
  • Recommend and lead system improvement efforts that enhance system performance, reliability, and security.
  • Identify and evaluate emerging HPC technologies and features that could introduce novel capabilities or enhance existing system performance and utility.
  • Represent NERSC in technical or user advocacy groups to influence the HPC and DOE community to meet user needs.
Requirments
  • Typically, 8+ years of related experience with a Bachelor’s degree; alternatively, 6+ years with a Master’s degree; or equivalent career experience.
  • 2+ years of experience with API and web services software development on Linux systems in a high-performance computing, cloud computing, or hyper-scale environment.
  • Familiarity with designing and building API interfaces to compute, storage, or other backend systems.
  • Experience with some or all of our key technologies:
    • C, shell, and Python programming languages
    • Git, runners, and complex CI/CD pipelines
    • Using and developing AI (or machine learning) tools and services
    • Database administration and optimization (such as MongoDB, MySQL or PostgreSQL)
    • Container technology (such as Docker or Kubernetes)
  • Working knowledge of software engineering best practices for performance and security.
  • Ability to resolve complex issues in creative and effective ways and derive technical solutions in a collaborative environment to meet end user requirements or needs.
  • Demonstrated ability to work independently as well as collaboratively in large projects, and contribute to an active and respectful intellectual environment.
  • Creative, positive, and collaborative work style.
  • Excellent oral and written communication skills.
  • Experience with OpenAPI and other API frameworks.
  • Experience deploying and managing virtualization and/or container technologies
  • Ability to lead and coordinate projects.
  • Ability to analyze and resolve significant and unique issues requiring evaluation of multiple intangible factors.
  • Ability to exercise independent judgment in methods, techniques and evaluation criteria for obtaining results.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC/AI Performance Specialist
HPC/AI Performance Specialist

Lawrence Berkeley National Laboratory • Berkeley (CA)

On-site
USD 156,000 - 192,000
Health and retirement benefits
Tuition Assistance Program
Pet insurance
System Infrastructure / Platform Engineer, HPC Technology Department
System Infrastructure / Platform Engineer, HPC Technology Department

Berkeley Lab • Berkeley (CA)

Hybrid
USD 156,000 - 192,000
Exceptional health benefits
Tuition Assistance Program
Paid vacation and sick time
+2
System Infrastructure / Platform Engineer
System Infrastructure / Platform Engineer

LTD Global • South Carolina

On-site
USD 120,000 - 180,000
HPC/AI Performance Specialist
HPC/AI Performance Specialist

Berkeley Lab • Berkeley (CA)

Hybrid
USD 139,000 - 236,000
Exceptional health and retirement benefits
Tuition Assistance Program
Pet insurance
+2
System Infrastructure / Platform Engineer
System Infrastructure / Platform Engineer

Ltd Global • Berkeley (CA)

On-site
USD 120,000 - 150,000
Director, National Energy Research Scientific Computing Center (NERSC)
Director, National Energy Research Scientific Computing Center (NERSC)

Koitecc Solutions • Berkeley (CA)

On-site
USD 375,000 - 440,000
HPC Scientific Support Engineer
HPC Scientific Support Engineer

Berkeley Lab • Berkeley (CA)

Hybrid
USD 156,000 - 219,000
Cybersecurity Group Lead
Cybersecurity Group Lead

Lawrence Berkeley National Laboratory • Berkeley (CA)

Hybrid
USD 203,000 - 249,000
Health and retirement benefits
Tuition assistance program
Winter holiday shutdown
+2
HPC Scientific Support Engineer
HPC Scientific Support Engineer

Lawrence Berkeley National Laboratory • Berkeley (CA)

On-site
USD 156,000 - 192,000
HPC Storage Systems Group Leader
HPC Storage Systems Group Leader

Lawrence Berkeley National Laboratory • Berkeley (CA)

On-site
USD 203,000 - 249,000
Exceptional health benefits
Tuition Assistance Program
Parental bonding leave
+1