Calling all originals: At Levi Strauss & Co., you can be yourself — and be part of something bigger. We’re a company of people who like to forge our own path and leave the world better than we found it. Who believe that what makes us different makes us stronger. So add your voice. Make an impact. Find your fit — and your future.
As a Platform Operations Engineer, you will be responsible for maintaining the reliability, performance, and operational health of Levi Strauss & Co.'s enterprise infrastructure platforms. This role focuses on proactive monitoring, incident response, Oracle database operations, Linux system administration, performance optimization, and automation of recurring operational tasks.
You will work within a global 24x7 follow-the-sun support model and partner with infrastructure, cloud, security, and application teams to ensure platform stability and business continuity.
Infrastructure Operations
Monitor and maintain enterprise production environments to ensure high availability and operational stability.
Investigate and resolve infrastructure alerts, incidents, and service degradation events.
Participate in a 24x7 operational support model and incident response process.
Perform root cause analysis and implement preventative measures to reduce recurring issues.
Performance & Capacity Management
Investigate and remediate high CPU utilization events across Linux and enterprise platforms.
Analyze workload patterns and identify opportunities to improve platform efficiency and resource utilization.
Address memory utilization issues and capacity constraints.
Support proactive capacity planning and infrastructure health initiatives.
Oracle Database Operations
Monitor Oracle database environments and respond to operational alerts.
Troubleshoot Oracle errors, database performance issues, and alert log events.
Manage tablespace growth, storage utilization, and database capacity concerns.
Support database maintenance, operational health checks, and performance tuning activities.
Monitoring & Automation
Utilize monitoring tools such as Oracle Enterprise Manager (OEM) and other enterprise monitoring platforms.
Identify recurring alert patterns and develop automated remediation solutions where appropriate.
Improve operational efficiency through scripting and automation.
Contribute to continuous improvement initiatives focused on reducing operational toil and incident volume.
Collaboration & Service Management
Work closely with global operations teams, cloud engineers, database administrators, and application support teams.
Follow ITIL-based incident, problem, and change management processes.
Exercise sound escalation judgment and engage subject matter experts when necessary.
Support knowledge sharing and documentation of operational procedures and troubleshooting practices.