Responsibilities
- Assess Isilon and NetApp cluster capacity and performance for node replacement or upgrade needs.
- Manage integration of new Isilon and NetApp nodes, ensuring compatibility and minimal downtime.
- Oversee upgrade and replacement of HPC, Kubernetes, Greenplum, Impala, and GPU server hardware.
- Utilize Impala Insight IQ and NetApp tools for in-depth analysis to diagnose and resolve throughput, latency, and capacity issues.
- Leverage OneFS and ONTAP troubleshooting tools and logs to identify and rectify system errors and hardware malfunctions.
- Address performance bottlenecks and system failures in HPC, Kubernetes, Greenplum, GPU, Impala, and NetApp environments.
- Implement continuous monitoring with Isilon OneFS event logs, NetApp ONTAP tools, SNMP alerts, and Grafana.
- Schedule regular maintenance using Isilon and NetApp software suites for optimal storage system performance.
- Monitor and maintain HPC, Kubernetes, Greenplum, GPU, Impala, and NetApp server environments, applying updates and performing health checks.
- Apply Impala and NetApp-specific patches and firmware updates to nodes and system software.
- Test new software updates in a sandbox environment before cluster deployment.
- Manage and coordinate software updates and patch installations across HPC, Kubernetes, Greenplum, GPU, Impala, and NetApp systems.
- Provide specialized support and consultation for optimizing Isilon and NetApp storage solutions.
- Offer technical support for optimizing HPC, Kubernetes, Greenplum, and GPU server configurations.
- Assist with design, configuration, and optimization of NetApp storage architectures.
- Manage phased decommissioning of aging Isilon and NetApp hardware, ensuring compliance.
- Coordinate replacement of EOL hardware for HPC, Kubernetes, Greenplum, GPU, Isilon, and NetApp.
- Ensure Impala, NetApp, HPC, Kubernetes, Greenplum, and GPU storage solutions are integrated and compatible with IT infrastructure.
- Develop and implement strategies for effective integration of storage and processing technologies.
- Maintain detailed documentation of system configurations, upgrades, and troubleshooting activities.
- Conduct training sessions or create guides for internal teams on system operations and best practices.
- Plan and test disaster recovery procedures for Impala, NetApp, HPC, Kubernetes, Greenplum, and GPU servers.
- Ensure backup and recovery processes are in place and regularly tested.
- Develop capacity planning strategies to proactively scale storage and compute resources.
- Develop and maintain automation scripts for routine tasks like monitoring, data migration, or maintenance.
- Utilize automation tools to streamline repetitive tasks.
- Ensure compliance with organizational security standards and regulations during maintenance, upgrades, and troubleshooting.
- Implement or coordinate security best practices and hardening measures for Isilon, NetApp, HPC, Kubernetes, Greenplum, and GPU servers.
- Participate in data center activities, including racking and initial installation of servers and storage equipment.
- Ensure proper cabling, power connections, and network configurations during initial setup.
Requirements
- 4–6 years of hands-on Linux support experience in enterprise environments
- Strong Linux administration and troubleshooting skills (RHEL/CentOS/Ubuntu)
- Experience with system monitoring, log analysis, and production support
- Basic scripting knowledge (Bash/Python preferred)
- Understanding of networking fundamentals (TCP/IP, DNS, SSH)
- Familiarity with ETL/data pipeline support and batch job monitoring
- Experience with cloud environments (AWS/Azure/GCP) is a plus
- Knowledge of Docker/Kubernetes is preferred
- Strong problem-solving and communication skills
Skills
- Linux support
- Linux administration
- Linux troubleshooting
- RHEL
- CentOS
- Ubuntu
- System monitoring
- Log analysis
- Production support
- Bash scripting
- Python scripting
- Networking fundamentals
- TCP/IP
- DNS
- SSH
- ETL support
- Data pipeline support
- Batch job monitoring
- AWS
- Azure
- GCP
- Docker
- Kubernetes
- Problem-solving
- Communication
- Isilon OneFS
- NetApp ONTAP
- Impala Insight IQ
- SNMP alerts
- Grafana
- HPC
- Greenplum
- GPU servers
- Impala
Experience Level
- 4–6 years of hands-on Linux support experience
Benefits
- Comprehensive health insurance
- 401K
- PTO
- Sick days
- End year bonus
