04 ago
|
Red Hat
|
Italia
The Red Hat Performance and Scale Engineering team is looking for an AI Performance Engineer to join us in the PSAP - Performance and Scale for AI Platforms team. As recent advances in AI technologies have taken the world by storm, Red Hat has jointly engineered an enterprise grade platform - OpenShift AI, based on open source AI technologies, to help enterprises leverage the full potential of these transformative AI technologies. As part of this team, you will be responsible for the performance and scalability assessments of OpenShift AI platform - that includes but not limited to notebooks as a service, data science pipelines, model serving stack, feature store, edge AI, and a distributed model training stack. Our goal is to make OpenShift AI the platform of choice for Red Hat’s enterprise customers for leveraging AI technologies. You will help us achieve those goals through targeted improvements in the performance and scalability of the OpenShift AI platform.
Per essere preso/a in considerazione per un colloquio, la preghiamo di assicurarsi che la sua candidatura sia pienamente in linea con le specifiche del lavoro riportate di seguito.
You will be required to formulate and execute performance test plans. You will investigate cloud infrastructure, on-prem hardware, RHEL, OpenShift, and OpenShift AI performance tuning knobs. In addition, you will triage and potentially fix performance issues, create new benchmarking tests and automation tools as needed, and socialize performance results on a regular basis. This role needs an engineer that thinks creatively, adapts to rapid change, and has the willingness to learn and apply new technologies. You will be joining a vibrant open source culture, and helping promote performance and innovation in this Red Hat engineering team.
The border mission of the Performance and Scale team is to establish performance and scale leadership of the Red Hat product and cloud services portfolio. The scope includes component level,
system and solution analysis and targeted enhancements. The team collaborates with engineering, product management, product marketing and customer support as well as hardware and software partners.
Any experience with AI technologies and frameworks is mandatory for the role.
What you will do
-
- Execute performance and scalability benchmarks against various components of the OpenShift AI platform to drive improvements and detect regressions
- Develop tools and automation to aid the performance benchmarking work
- Collaborate with other teams to resolve performance issues
- Triage, debug, and solve customer cases related to AI performance
- Submit performance benchmarking results to industry consortia
- Publish results, conclusions, recommendations and best practices via internal test reports, presentations, and external blogs to support our partners and customers.
- Participate in internal and external conferences about your work and results
What you will bring
-
- Experience in running performance tests, data capture, data analysis, and visualization
- Experience with AI technologies and frameworks (pytorch, transformers, etc)
- Experience with systems performance engineering and metrics collection tools such as iostat, vmstat, sar, perf, and prometheus. xysqume
- Programming experience in Python or willingness to learn
- Experience working with the Linux operating system (RHEL, Fedora or CentOS preferred)
- Excellent written and verbal language skills in English
Equal Opportunity Policy (EEO)
Red Hat is proud to be an equal opportunity workplace and an affirmative action employer. We review applications for employment without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, citizenship, age, veteran status, genetic information, physical or mental disability, medical condition, marital status, or any other basis prohibited by law.
#J-18808-Ljbffr
📌 Software Engineer - Performance and Scale Engineering (Spain, Italy, Portugal and Czech Republic) (Italia)
🏢 Red Hat
📍 Italia