Job description
Collaborates with appropriate stakeholders to determine user requirements for a set of features. Reviews work items to deepen knowledge of product features in partnership with appropriate stakeholders and executes project plans, release plans, and work items Design and implement E2E scenarios and performance testing, profiling, and optimization strategies for storage systems. Develop benchmarking and automation tools to evaluate system performance, validate end-to-end scenarios, and improve productivity Analyze system bottlenecks, latency issues, and resource utilization across compute, storage, and networking layers. Perform root cause analysis of complex issues and work with the component team to resolve issues and enhance the overall product quality Define key performance metrics (KPIs) and provide data-driven recommendations for scaling and tuning Leverages subject-matter expertise of cross-product features with appropriate stakeholders to drive multiple group's project plans, release plans, and work items Acts as a Designated Responsible Individual (DRI) and guides other engineers by developing and following the playbook, working on call to monitor system/product/service for degradation, downtime, or interruptions, alerting stakeholders about status and initiates actions to restore system/product/service for simple and complex problems when appropriate Bachelor's Degree in Computer Science or related technical field AND 4+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python Master's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python OR Bachelor's Degree in Computer Science or related technical field AND 8+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python OR equivalent experience Experience in performance and system engineering, distributed systems, or large-scale cloud services