S
SBT
Posted 94 days agoVerified live 2d ago

Cloud Operation Center Engineer - Bilingual in Korean

Brief overview

Irvine, CA, USAIn-person

About the company

Leaders in the semiconductor industry know that recruiting people with specialized skills in a competitive global market is a significant challenge.

Job description

Summary

SBT is seeking a Cloud Operation Center Engineer who is bilingual in Korean. This role is responsible for the operation and support of Private Cloud infrastructure, focusing on ensuring stable operations of an OpenStack-based environment and performing incident analysis and automation initiatives.

Responsibilities

  • Create, modify, and manage Palo Alto firewall policies
  • Review traffic logs and validate firewall policy changes
  • Perform network troubleshooting using packet capture / traffic dumps
  • Diagnose issues based on NAT, security policies, and session states
  • Configure and manage Citrix ADC VPX-based load balancers (LB/CLB)
  • Manage VIPs, services, service groups, and health checks
  • Resolve LB incidents such as backend server failures and connection issues
  • Analyze logs using nstrace, tcpdump, and related tools
  • Provision, deploy, and manage OpenStack instances
  • Manage volumes, shared volumes, networks, and security groups
  • Troubleshoot instance boot failures, network issues, and volume attachment problems
  • Support compute node failures and live migration operations
  • Operate Dell servers used as OpenStack compute nodes
  • Monitor server health via iDRAC and perform basic maintenance
  • Handle hardware failures and coordinate with Dell support
  • Support hardware replacement and vendor service activities
  • Manage NetApp ONTAP storage systems
  • Monitor nodes, SVMs, volumes, and aggregates
  • Analyze performance metrics such as latency, IOPS, and throughput
  • Respond to storage incidents and performance degradation issues
  • Perform Linux system administration and incident troubleshooting
  • Handle root password changes, mount issues, repository issues, etc
  • Resolve boot failures, filesystem issues, and network interface problems
  • Analyze system logs and systemd service issues
  • Operate monitoring platforms such as Zabbix, Grafana, and Prometheus
  • Monitor servers, networks, storage, OpenStack, LB, and firewall systems
  • Manage alerts, dashboards, metrics, and triggers
  • Perform root cause analysis using logs, metrics, and alerts
  • Automate deployments and updates using Ansible
  • Automate repetitive operational tasks
  • Build workflow automation using Microsoft Teams / Power Automate
  • Automate alerts, reporting, approval, and request workflows
  • Develop scripts using Python, Shell, or PowerShell
  • Leverage AI coding tools for scripting, log analysis, and documentation

Skills

  • Experience in Linux OS administration and troubleshooting
  • Understanding of TCP/IP, routing, NAT, and firewall policies
  • Hands-on experience with Palo Alto or similar firewalls
  • Experience with Citrix ADC VPX or similar load balancers
  • Experience with OpenStack or private cloud environments
  • Experience with NetApp ONTAP or enterprise storage systems
  • Experience with Dell servers and iDRAC-based operations
  • Packet capture / traffic dump troubleshooting experience
  • Experience with Ansible or scripting for automation
  • Ability to perform root cause analysis using logs, metrics, and network data
  • Experience with Zabbix, Grafana, or Prometheus
  • OpenStack component experience (Nova, Neutron, Cinder, Glance)
  • NetApp ONTAP CLI and performance tuning experience
  • Citrix ADC nstrace / tcpdump troubleshooting experience
  • Microsoft Teams / Power Automate workflow automation experience
  • Scripting skills in Python, Shell, or PowerShell
  • Understanding of REST API, Webhook, JSON, YAML
  • Git-based version control experience
  • Experience with AI-assisted development tools (ChatGPT, Copilot, etc.)
  • Experience using AI for RCA, log analysis, and automation

Qualifications

Must Haves

  • Experience in Linux OS administration and troubleshooting
  • Understanding of TCP/IP, routing, NAT, and firewall policies
  • Hands-on experience with Palo Alto or similar firewalls
  • Experience with Citrix ADC VPX or similar load balancers
  • Experience with OpenStack or private cloud environments
  • Experience with NetApp ONTAP or enterprise storage systems
  • Experience with Dell servers and iDRAC-based operations
  • Packet capture / traffic dump troubleshooting experience
  • Experience with Ansible or scripting for automation
  • Ability to perform root cause analysis using logs, metrics, and network data

Nice to Haves

  • Experience with Zabbix, Grafana, or Prometheus
  • OpenStack component experience (Nova, Neutron, Cinder, Glance)
  • NetApp ONTAP CLI and performance tuning experience
  • Citrix ADC nstrace / tcpdump troubleshooting experience
  • Microsoft Teams / Power Automate workflow automation experience
  • Scripting skills in Python, Shell, or PowerShell
  • Understanding of REST API, Webhook, JSON, YAML
  • Git-based version control experience
  • Experience with AI-assisted development tools (ChatGPT, Copilot, etc.)
  • Experience using AI for RCA, log analysis, and automation

More jobs like this