| n | Job Description: | n | This is a two-year contract of employment, inclusive of benefits. |
| This is a two-year contract of employment, inclusive of benefits. The Academic Research Services tea | | |
| m at UCSF is seeking an Storage Systems Engineer (SYS ADM 4) to serve as a technical resource in the | | |
| design, deployment, and operation of large-scale(multi petabytes) research storage and data infrast | | |
| ructure. This role will work in close partnership with the FAC Storage Lead to support UCSF’s evolvi | | |
| ng research ecosystem, including Physical/Virtual Compute, CoreHPC, the Research Analysis Environmen | | |
| t (RAE), and large institutional storage initiatives. This position is primarily responsible for arc | | |
| hitecture, implementation, and lifecycle management for the Facility for Advanced Computing (FAC), s | | |
| torage and systems, including support for large storage environments based on ZFS, NFS v3 and v4 & S | | |
| AMBA(SMB) NSF-funded infrastructure, and OS Nexus–aligned data platforms. The role ensures seamless | | |
| integration between storage systems and the Physical/Virtual Compute as well as CoreHPC compute clu | | |
| ster, enabling performant, reliable, and scalable data access for AI, data science, and computationa | | |
| l research workloads. The Storage Systems Engineer will: Work with the FAC Storage lead to continue | | |
| supporting the design and evolution of storage architecture on ZFS across on-prem and hybrid environ | | |
| ments, including ZFS, NFS, SMB VAST, parallel filesystems, and enterprise storage platforms Develop | | |
| and maintain data movement strategies and tooling (e.g., rsync, rclone, Globus, NFS, SMB workflows) | | |
| to support large-scale data ingestion, migration, and lifecycle management Ensure tight integration | | |
| between storage and Physical/Virtual Compute as well as CoreHPC compute cluster systems, optimizing | | |
| throughput, latency, and reliability for distributed workloads Support and scale storage systems ba | | |
| cking major institutional initiatives (FAC storage(ZFS), OS Nexus integration) Collaborate closely w | | |
| ith DevOps, networking, and security teams to deliver cohesive research infrastructure solutions Des | | |
| ign and implement monitoring, performance tuning, and capacity planning strategies for storage and d | | |
| ata systems Troubleshoot complex issues across storage, networking, and compute boundaries Participa | | |
| te in system upgrades, migrations, and expansion efforts with minimal disruption to researchers Prov | | |
| ide guidance to researchers on data organization, transfer strategies, and performance optimization | | |
| Evaluate and recommend emerging storage technologies and architectures This role may lead storage-fo | | |
| cused projects and contribute to cross-functional initiatives that improve the scalability, usabilit | | |
| y, and reliability of UCSF’s research computing ecosystem. Department Overview Academic Research Sys | | |
| tems (ARS) serves the needs of the UCSF research community by providing an integrated repository of | | |
| HIPAA compliant clinical and life sciences data and a centralized, secure, professionally managed in | | |
| frastructure for the storage and management of research data. ARS empowers medical scientific invest | | |
| igations by offering secure computing environments, data capture, management and analysis tools, and | | |
| support services which meet researchers’ needs. The Research Infrastructure team of the Academic Re | | |
| search Service (ARS) focuses on large scale research platform support, high performance computationa | | |
| l and storage services for UCSF researchers so they can address complex computational, AI, and data | | |
| science problems. | | |
| | | |
| t | Qualifications: | t | The Academic Research Services team at UCSF is seeking an Storage Systems Engineer (SYS ADM 4) to s |
| | | erve as a technical resource in the design, deployment, and operation of large-scale(multi petabytes |
| | | ) research storage and data infrastructure. This role will work in close partnership with the FAC St |
| | | orage Lead to support UCSF’s evolving research ecosystem, including Physical/Virtual Compute, CoreHP |
| | | C, the Research Analysis Environment (RAE), and large institutional storage initiatives. |
| REQUIRED QUALIFICATIONS - Bachelor's degree in a related area, such as computer science or engineeri | | |
| ng, and 6+ years of experience with storage infrastructure support and management, or 10+ years of r | | |
| elated experience with large-scale storage systems - Demonstrated skill (5 years +) deploying, manag | | |
| ing, and troubleshooting ZFS (or similar) InfiniBand-based clusters - Strong knowledge of ZFS, high- | | |
| performance parallel filesystems, and storage such as GPFS, Lustre, Vast, DDN, etc - Advanced knowle | | |
| dge of computer security best practices and policies, including demonstrated experience securing res | | |
| earch cyberinfrastructure systems to meet NIST 800-171 / 800-223, HIPAA, or IS-3 requirements - Know | | |
| ledge of HPC job scheduler system design and operation, such as SLURM or PBS, - Ability to elicit an | | |
| d communicate technical and non-technical information in a clear and concise manner. - Self-motivate | | |
| d and works independently and as part of a team. Demonstrates problem-solving skills. Able to learn | | |
| effectively and meet deadlines. - Understanding of system performance monitoring and actions that ca | | |
| n be taken to improve or correct performance. - Demonstrated advanced knowledge, skills, and abiliti | | |
| es associated with system problem identification and resolution. Experience with design, configurati | | |
| on, operation, repair, and tuning of technology systems. - Advanced experience writing and editing t | | |
| he most complex scripts used to perform system maintenance and administration. - Demonstrated testin | | |
| g and test planning skills. Demonstrated ability to create automated testing. - Ability to write tec | | |
| hnical documentation in a clear and concise manner. Ability to develop runbooks defining complex tec | | |
| hnical processes in a clear and concise manner PREFERRED QUALIFICATIONS - Expert knowledge of Virtua | | |
| l Machines, Bare Metal Servers & HPC systems infrastructure design - Knowledge of the design, develo | | |
| pment and application of technology and systems to meet business needs. - General knowledge of other | | |
| areas of IT. E.g., Active Directory, Domain Controllers, Network Infrastructure. - Demonstrated ski | | |
| lls associated with adapting equipment and technology to serve user needs. Demonstrated comprehensiv | | |
| e understanding of how system management actions affect other systems, system users and dependent/re | | |
| lated functions. - Professional certification in enterprise storage technologies (e.g., NetApp, Dell | | |
| EMC PowerScale, IBM Storage Scale, VAST, Pure Storage) | | |
| | | This position is primarily responsible for architecture, implementation, and lifecycle management f |
| | | or the Facility for Advanced Computing (FAC), storage and systems, including support for large stora |
| | | ge environments based on ZFS, NFS v3 and v4 & SAMBA(SMB) NSF-funded infrastructure, and OS Nexus-ali |
| | | gned data platforms. The role ensures seamless integration between storage systems and the Physical/ |
| | | Virtual Compute as well as CoreHPC compute cluster, enabling performant, reliable, and scalable data |
| | | access for AI, data science, and computational research workloads. |
| | | |
| | | The Storage Systems Engineer will: |
| | | |
| | | * |
| | | |
| | | Work with the FAC Storage lead to continue supporting the design and evolution of storage architect |
| | | ure on ZFS across on-prem and hybrid environments, including ZFS, NFS, SMB VAST, parallel filesystem |
| | | s, and enterprise storage platforms |
| | | * |
| | | |
| | | Develop and maintain data movement strategies and tooling (e.g., rsync, rclone, Globus, NFS, SMB wo |
| | | rkflows) to support large-scale data ingestion, migration, and lifecycle management |
| | | * |
| | | |
| | | Ensure tight integration between storage and Physical/Virtual Compute as well as CoreHPC compute cl |
| | | uster systems, optimizing throughput, latency, and reliability for distributed workloads |
| | | * |
| | | |
| | | Support and scale storage systems backing major institutional initiatives (FAC storage(ZFS), OS Nex |
| | | us integration) |
| | | * |
| | | |
| | | Collaborate closely with DevOps, networking, and security teams to deliver cohesive research infras |
| | | tructure solutions |
| | | * |
| | | |
| | | Design and implement monitoring, performance tuning, and capacity planning strategies for storage a |
| | | nd data systems |
| | | * |
| | | |
| | | Troubleshoot complex issues across storage, networking, and compute boundaries |
| | | * |
| | | |
| | | Participate in system upgrades, migrations, and expansion efforts with minimal disruption to resear |
| | | chers |
| | | * |
| | | |
| | | Provide guidance to researchers on data organization, transfer strategies, and performance optimiza |
| | | tion |
| | | * |
| | | |
| | | Evaluate and recommend emerging storage technologies and architectures |
| | | |
| | | This role may lead storage-focused projects and contribute to cross-functional initiatives that imp |
| | | rove the scalability, usability, and reliability of UCSF’s research computing ecosystem. |
| | | |
| | | Department Overview |
| | | |
| | | Academic Research Systems (ARS) serves the needs of the UCSF research community by providing an int |
| | | egrated repository of HIPAA compliant clinical and life sciences data and a centralized, secure, pro |
| | | fessionally managed infrastructure for the storage and management of research data. ARS empowers med |
| | | ical scientific investigations by offering secure computing environments, data capture, management a |
| | | nd analysis tools, and support services which meet researchers’ needs. |
| | | |
| | | The Research Infrastructure team of the Academic Research Service (ARS) focuses on large scale rese |
| | | arch platform support, high performance computational and storage services for UCSF researchers so t |
| | | hey can address complex computational, AI, and data science problems. |