More from this employer
Nashville, USA
Position Summary The Controls Systems Operator position is part of the Vanderbilt University Maintenance and Operations (VU) and is a key individual contributor responsible for analyzing systems alarms and failures; dete
Nashville, USA
The Vanderbilt University Institute of National Security is seeking a Director, to establish and lead a world-class AI-augmented gaming Center for the Institute of National Security at Vanderbilt University. The Director
Nashville, USA
ABOUT THE WORK UNIT The Meydan Lab is interested in the mechanisms of protein synthesis, and the role of translation regulation and ribosomes in human health. We use various biochemistry techniques, sequencing of ribosom
New York City, USA
Please excuse any formatting issues. This is a known systems issue we are working on fixing. Position Summary: The Development Coordinator will play a key role in supporting the New York Development Team within Vanderbil
Nashville, USA
The Program Specialist is responsible for managing the department administrative office front desk, travel and reimbursements, clerical support for classrooms, procurement/purchasing role, and manage undergrad student wo
Nashville, USA
Nashville, USA
Nashville, USA
Nashville, USA
Nashville, USA
Nashville, USA
New York City, USA
Nashville, USA
Base Career helps you apply smarter for this job.
The Systems Engineer is the hands-on individual responsible for the design, security, and dayto- day operation of the Wicked Problems Lab's research computing environment — spanning AWS cloud, on-premises GPU/AI compute, databases, and systems security.
This is a practitioner role for someone who provisions infrastructure, configures and hardens systems, and unblocks researchers directly, with a high degree of autonomy in a fastmoving national security research setting.
About the Work Unit: The Wicked Problems Lab applies advanced computing, AI, and data science to highconsequence national security problems, with partnerships across the Five Eyes and Indo- Pacific communities.
The environment is small, technically sophisticated, and fast-paced, and its systems must be secure, dependable, and built to handle sensitive research data.
Key Functions and Expected Performance: - Cloud & on-prem infrastructure: Design, architect, and operate the Lab's hybrid environment — AWS (EC2, VPC, S3, IAM) under least-privilege and on-premises GPU/AI compute systems; scale storage and GPU capacity to meet workload demand; engineer for resilience through monitoring, backup, and disaster recovery; own workload placement across cloud and local systems for cost, performance, and data sensitivity. - Systems configuration & administration: Administer Linux servers end to end; manageconfiguration as code (Terraform, Ansible, scripting) for reproducible, documentedenvironments; own SSH/key lifecycle and access across a distributed fleet. - Data & database management: Stand up, secure, tune, and back up research databases (PostgreSQL, NoSQL, and comparable relational/vector stores); build and maintain reliable ETL and datatransfer/movement pipelines with attention to integrity, throughput, and reproducibility. - Security & compliance: Harden systems against sophisticated threats — patch/vulnerability management, segmentation, secrets management, and endpoint detection and response; enforce access control and data-handling appropriate to sensitive research; comply with University and partner security requirements. - Research enablement: Serve as first point of contact for researchers' systems needs and AI/developer tooling; advocate for Vanderbilt's core values; stay current with cloud, GPU/AI, and security technologies; other duties as needed. - Networking & traffic analysis: Configure and troubleshoot network infrastructure (routers, switches, VPN, segmentation); capture and analyze network traffic at the packet level; deploy and tune intrusion detection/prevention systems to monitor for and investigate anomalous activity.
Supervisory Relationships: This position has no supervisory responsibility and reports administratively and functionally to the Director of the Wicked Problems Lab.
Education and Certifications - Bachelor's in Computer Science/Engineering or related field is necessary; equivalent experience may substitute.
Relevant AWS/Linux/security certifications preferred.
Experience and Skills: - 4+ years hands-on systems/infrastructure administration is necessary, with demonstrated command of AWS and strong Linux administration. - Experience operating GPU compute for AI/ML (CUDA stack, model serving/fine-tuning) is necessary; experience with enterprise-class NVIDIA GPU systems preferred. - Infrastructure-as-code (Terraform/Ansible), scripting (Bash/Python), database administration (PostgreSQL, NoSQL, or comparable), ETL/data-movement pipelines, and working systems/network security knowledge are necessary. - U.S. citizenship and ability to obtain/maintain a U.S. security clearance are preferred.
The Systems Engineer is the hands-on individual responsible for the design, security, and dayto- day operation of the Wicked Problems Lab's research computing environment — spanning AWS cloud, on-premises GPU/AI compute, databases, and systems security.
This is a practitioner role for someone who provisions infrastructure, configures and hardens systems, and unblocks researchers directly, with a high degree of autonomy in a fastmoving national security research setting.
About the Work Unit: The Wicked Problems Lab applies advanced computing, AI, and data science to highconsequence national security problems, with partnerships across the Five Eyes and Indo- Pacific communities.
The environment is small, technically sophisticated, and fast-paced, and its systems must be secure, dependable, and built to handle sensitive research data.
Key Functions and Expected Performance: - Cloud & on-prem infrastructure: Design, architect, and operate the Lab's hybrid environment — AWS (EC2, VPC, S3, IAM) under least-privilege and on-premises GPU/AI compute systems; scale storage and GPU capacity to meet workload demand; engineer for resilience through monitoring, backup, and disaster recovery; own workload placement across cloud and local systems for cost, performance, and data sensitivity. - Systems configuration & administration: Administer Linux servers end to end; manageconfiguration as code (Terraform, Ansible, scripting) for reproducible, documentedenvironments; own SSH/key lifecycle and access across a distributed fleet. - Data & database management: Stand up, secure, tune, and back up research databases (PostgreSQL, NoSQL, and comparable relational/vector stores); build and maintain reliable ETL and datatransfer/movement pipelines with attention to integrity, throughput, and reproducibility. - Security & compliance: Harden systems against sophisticated threats — patch/vulnerability management, segmentation, secrets management, and endpoint detection and response; enforce access control and data-handling appropriate to sensitive research; comply with University and partner security requirements. - Research enablement: Serve as first point of contact for researchers' systems needs and AI/developer tooling; advocate for Vanderbilt's core values; stay current with cloud, GPU/AI, and security technologies; other duties as needed. - Networking & traffic analysis: Configure and troubleshoot network infrastructure (routers, switches, VPN, segmentation); capture and analyze network traffic at the packet level; deploy and tune intrusion detection/prevention systems to monitor for and investigate anomalous activity.
Supervisory Relationships: This position has no supervisory responsibility and reports administratively and functionally to the Director of the Wicked Problems Lab.
A private research university teaching and creating knowledge.
Visit company websiteJobs and hiring trendsSkip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
Education and Certifications - Bachelor's in Computer Science/Engineering or related field is necessary; equivalent experience may substitute.
Relevant AWS/Linux/security certifications preferred.
Experience and Skills: - 4+ years hands-on systems/infrastructure administration is necessary, with demonstrated command of AWS and strong Linux administration. - Experience operating GPU compute for AI/ML (CUDA stack, model serving/fine-tuning) is necessary; experience with enterprise-class NVIDIA GPU systems preferred. - Infrastructure-as-code (Terraform/Ansible), scripting (Bash/Python), database administration (PostgreSQL, NoSQL, or comparable), ETL/data-movement pipelines, and working systems/network security knowledge are necessary. - U.S. citizenship and ability to obtain/maintain a U.S. security clearance are preferred.