General Atomics (GA), and its affiliated companies, is one of the world’s leading resources for high-technology systems development ranging from the nuclear fuel cycle to remotely piloted aircraft, airborne sensors, and advanced electric, electronic, wireless and laser technologies.
With consultative direction within the Core Services group, this position is responsible for architecting, optimizing, and maintaining a high-velocity, on-premises Linux data pipeline supporting real-time, AI-driven machine learning models for the DIII-D national fusion reactor program.
As a hands-on technical practitioner, you will manage bare-metal deployments, orchestrate high-availability parallel storage arrays, and troubleshoot ultra-low-latency network protocols across a heavily segmented, secure infrastructure. This role focuses on deep, OS-level infrastructure and physical data center operations, bridging the gap between raw hardware capabilities and mission-critical scientific computing.
- Bare-Metal & OS Administration: Plan, manage, and optimize day-to-day operations of on-premises, physical x86 and GPU server infrastructure running Red Hat Enterprise Linux (RHEL).
- High-Velocity Networking: Provision and troubleshoot core network sharing protocols (NFS) and ultra-low-latency networking hardware supporting 100GbE backbones and RDMA over Converged Ethernet (RoCEv2).
- Operational Technology (OT) Security: Maintain and fortify a heavily firewalled, push-only segmented internal network architecture, ensuring strict "Deny All Inbound" rules protect sensitive reactor control systems.
- Storage Architecture Lifecycle: Build, configure, and maintain high-density NVMe storage tiers and software-defined, parallel filesystems (BeeGFS/ZFS) to handle massive multi-gigabyte payload dumps with zero ingestion bottlenecks.
- Performance & I/O Benchmarking: Conduct granular system-level I/O benchmarking, diagnose deep kernel-level platform anomalies, and implement kernel parameter tuning to maximize data velocity and prevent data loss.
- Automation & Provisioning: Develop and maintain automated infrastructure workflows and configuration management frameworks using Bash shell scripting and Ansible to streamline bare-metal node deployment and environment validation.
- Data Center Operations: Manage physical data center footprints, including high-density server rack layouts, equipment delivery, hardware diagnostics, power infrastructure/UPS, and asset lifecycle tracking.
- Collaborative Consultation: Act as a technical infrastructure expert, guiding the development of innovative solutions to unique computing challenges for data acquisition systems and scientific teams.
- Vendor & Planning Coordination: Analyze new hardware architectures, engineer custom hardware Bill of Materials (BOM) alongside OEMs/ODMs, and represent the organization as a primary technical contact with suppliers.
- Compliance & Safety: Observe all laws, regulations, and facility safety obligations, ensuring that all system upgrades and physical data center modifications are designed with established personnel operating procedures properly considered.
We recognize and appreciate the value and contributions of individuals with diverse backgrounds and experiences and welcome all qualified individuals to apply.