Back to dashboard

Storage Systems Engineer

Job ID 3270 | Run 20260928-184624

Changed Fields

FieldPreviousCurrent
IAPStaff Plan (target potential payout of $900, maximum of $1,800)Tier C Plan (target potential payout of 3.5%, maximum of 5%)

Job Description Diff


Previous Job Description
Current Job Description
f1Job Description:f1Job Description:
t2This is a two-year contract of employment, inclusive of benefits. The Academic Research Services teat2Certain terms and conditions of employment for this position, including the rate of pay, benefits, e
>m at UCSF is seeking an Storage Systems Engineer (SYS ADM 4) to serve as a technical resource in the>tc., are currently subject to negotiation with the appropriate union This is a two-year contract of 
> design, deployment, and operation of large-scale(multi petabytes) research storage and data infrast>employment, inclusive of benefits. The Academic Research Services team at UCSF is seeking an Storage
>ructure. This role will work in close partnership with the FAC Storage Lead to support UCSF’s evolvi> Systems Engineer (SYS ADM 4) to serve as a technical resource in the design, deployment, and operat
>ng research ecosystem, including Physical/Virtual Compute, CoreHPC, the Research Analysis Environmen>ion of large-scale(multi petabytes) research storage and data infrastructure. This role will work in
>t (RAE), and large institutional storage initiatives. This position is primarily responsible for arc> close partnership with the FAC Storage Lead to support UCSF’s evolving research ecosystem, includin
>hitecture, implementation, and lifecycle management for the Facility for Advanced Computing (FAC), s>g Physical/Virtual Compute, CoreHPC, the Research Analysis Environment (RAE), and large institutiona
>torage and systems, including support for large storage environments based on ZFS, NFS v3 and v4 & S>l storage initiatives. This position is primarily responsible for architecture, implementation, and 
>AMBA(SMB) NSF-funded infrastructure, and OS Nexus–aligned data platforms. The role ensures seamless >lifecycle management for the Facility for Advanced Computing (FAC), storage and systems, including s
>integration between storage systems and the Physical/Virtual Compute as well as  CoreHPC compute clu>upport for large storage environments based on ZFS, NFS v3 and v4 & SAMBA(SMB) NSF-funded infrastruc
>ster, enabling performant, reliable, and scalable data access for AI, data science, and computationa>ture, and OS Nexus–aligned data platforms. The role ensures seamless integration between storage sys
>l research workloads. The Storage Systems Engineer will: Work with the FAC Storage lead to continue >tems and the Physical/Virtual Compute as well as  CoreHPC compute cluster, enabling performant, reli
>supporting the design and evolution of storage architecture on ZFS across on-prem and hybrid environ>able, and scalable data access for AI, data science, and computational research workloads. The Stora
>ments, including ZFS, NFS, SMB VAST, parallel filesystems, and enterprise storage platforms Develop >ge Systems Engineer will: Work with the FAC Storage lead to continue supporting the design and evolu
>and maintain data movement strategies and tooling (e.g., rsync, rclone, Globus, NFS, SMB workflows) >tion of storage architecture on ZFS across on-prem and hybrid environments, including ZFS, NFS, SMB 
>to support large-scale data ingestion, migration, and lifecycle management Ensure tight integration >VAST, parallel filesystems, and enterprise storage platforms Develop and maintain data movement stra
>between storage and Physical/Virtual Compute as well as  CoreHPC compute cluster systems, optimizing>tegies and tooling (e.g., rsync, rclone, Globus, NFS, SMB workflows) to support large-scale data ing
> throughput, latency, and reliability for distributed workloads Support and scale storage systems ba>estion, migration, and lifecycle management Ensure tight integration between storage and Physical/Vi
>cking major institutional initiatives (FAC storage(ZFS), OS Nexus integration) Collaborate closely w>rtual Compute as well as  CoreHPC compute cluster systems, optimizing throughput, latency, and relia
>ith DevOps, networking, and security teams to deliver cohesive research infrastructure solutions Des>bility for distributed workloads Support and scale storage systems backing major institutional initi
>ign and implement monitoring, performance tuning, and capacity planning strategies for storage and d>atives (FAC storage(ZFS), OS Nexus integration) Collaborate closely with DevOps, networking, and sec
>ata systems Troubleshoot complex issues across storage, networking, and compute boundaries Participa>urity teams to deliver cohesive research infrastructure solutions Design and implement monitoring, p
>te in system upgrades, migrations, and expansion efforts with minimal disruption to researchers Prov>erformance tuning, and capacity planning strategies for storage and data systems Troubleshoot comple
>ide guidance to researchers on data organization, transfer strategies, and performance optimization >x issues across storage, networking, and compute boundaries Participate in system upgrades, migratio
>Evaluate and recommend emerging storage technologies and architectures This role may lead storage-fo>ns, and expansion efforts with minimal disruption to researchers Provide guidance to researchers on 
>cused projects and contribute to cross-functional initiatives that improve the scalability, usabilit>data organization, transfer strategies, and performance optimization Evaluate and recommend emerging
>y, and reliability of UCSF’s research computing ecosystem. Department Overview Academic Research Sys> storage technologies and architectures This role may lead storage-focused projects and contribute t
>tems (ARS) serves the needs of the UCSF research community by providing an integrated repository of >o cross-functional initiatives that improve the scalability, usability, and reliability of UCSF’s re
>HIPAA compliant clinical and life sciences data and a centralized, secure, professionally managed in>search computing ecosystem. Department Overview Academic Research Systems (ARS) serves the needs of 
>frastructure for the storage and management of research data. ARS empowers medical scientific invest>the UCSF research community by providing an integrated repository of HIPAA compliant clinical and li
>igations by offering secure computing environments, data capture, management and analysis tools, and>fe sciences data and a centralized, secure, professionally managed infrastructure for the storage an
> support services which meet researchers’ needs. The Research Infrastructure team of the Academic Re>d management of research data. ARS empowers medical scientific investigations by offering secure com
>search Service (ARS) focuses on large scale research platform support, high performance computationa>puting environments, data capture, management and analysis tools, and support services which meet re
>l and storage services for UCSF researchers so they can address complex computational, AI,  and data>searchers’ needs. The Research Infrastructure team of the Academic Research Service (ARS) focuses on
> science problems.> large scale research platform support, high performance computational and storage services for UCSF
 > researchers so they can address complex computational, AI,  and data science problems.
33
4Qualifications:4Qualifications:
5REQUIRED QUALIFICATIONS - Bachelor's degree in a related area, such as computer science or engineeri5REQUIRED QUALIFICATIONS - Bachelor's degree in a related area, such as computer science or engineeri
>ng, and 6+ years of experience with storage infrastructure support and management, or 10+ years of r>ng, and 6+ years of experience with storage infrastructure support and management, or 10+ years of r
>elated experience with large-scale storage systems - Demonstrated skill (5 years +) deploying, manag>elated experience with large-scale storage systems - Demonstrated skill (5 years +) deploying, manag
>ing, and troubleshooting ZFS (or similar) InfiniBand-based clusters - Strong knowledge of ZFS, high->ing, and troubleshooting ZFS (or similar) InfiniBand-based clusters - Strong knowledge of ZFS, high-
>performance parallel filesystems, and storage such as GPFS, Lustre, Vast, DDN, etc - Advanced knowle>performance parallel filesystems, and storage such as GPFS, Lustre, Vast, DDN, etc - Advanced knowle
>dge of computer security best practices and policies, including demonstrated experience securing res>dge of computer security best practices and policies, including demonstrated experience securing res
>earch cyberinfrastructure systems to meet NIST 800-171 / 800-223, HIPAA, or IS-3 requirements - Know>earch cyberinfrastructure systems to meet NIST 800-171 / 800-223, HIPAA, or IS-3 requirements - Know
>ledge of HPC job scheduler system design and operation, such as SLURM or PBS, - Ability to elicit an>ledge of HPC job scheduler system design and operation, such as SLURM or PBS, - Ability to elicit an
>d communicate technical and non-technical information in a clear and concise manner. - Self-motivate>d communicate technical and non-technical information in a clear and concise manner. - Self-motivate
>d and works independently and as part of a team. Demonstrates problem-solving skills. Able to learn >d and works independently and as part of a team. Demonstrates problem-solving skills. Able to learn 
>effectively and meet deadlines. - Understanding of system performance monitoring and actions that ca>effectively and meet deadlines. - Understanding of system performance monitoring and actions that ca
>n be taken to improve or correct performance. - Demonstrated advanced knowledge, skills, and abiliti>n be taken to improve or correct performance. - Demonstrated advanced knowledge, skills, and abiliti
>es associated with system problem identification and resolution. Experience with design, configurati>es associated with system problem identification and resolution. Experience with design, configurati
>on, operation, repair, and tuning of technology systems. - Advanced experience writing and editing t>on, operation, repair, and tuning of technology systems. - Advanced experience writing and editing t
>he most complex scripts used to perform system maintenance and administration. - Demonstrated testin>he most complex scripts used to perform system maintenance and administration. - Demonstrated testin
>g and test planning skills. Demonstrated ability to create automated testing. - Ability to write tec>g and test planning skills. Demonstrated ability to create automated testing. - Ability to write tec
>hnical documentation in a clear and concise manner. Ability to develop runbooks defining complex tec>hnical documentation in a clear and concise manner. Ability to develop runbooks defining complex tec
>hnical processes in a clear and concise manner PREFERRED QUALIFICATIONS - Expert knowledge of Virtua>hnical processes in a clear and concise manner PREFERRED QUALIFICATIONS - Expert knowledge of Virtua
>l Machines, Bare Metal Servers & HPC systems infrastructure design - Knowledge of the design, develo>l Machines, Bare Metal Servers & HPC systems infrastructure design - Knowledge of the design, develo
>pment and application of technology and systems to meet business needs. - General knowledge of other>pment and application of technology and systems to meet business needs. - General knowledge of other
> areas of IT. E.g., Active Directory, Domain Controllers, Network Infrastructure. - Demonstrated ski> areas of IT. E.g., Active Directory, Domain Controllers, Network Infrastructure. - Demonstrated ski
>lls associated with adapting equipment and technology to serve user needs. Demonstrated comprehensiv>lls associated with adapting equipment and technology to serve user needs. Demonstrated comprehensiv
>e understanding of how system management actions affect other systems, system users and dependent/re>e understanding of how system management actions affect other systems, system users and dependent/re
>lated functions. - Professional certification in enterprise storage technologies (e.g., NetApp, Dell>lated functions. - Professional certification in enterprise storage technologies (e.g., NetApp, Dell
> EMC PowerScale, IBM Storage Scale, VAST, Pure Storage)> EMC PowerScale, IBM Storage Scale, VAST, Pure Storage)