HPC / AI Specialist (Torino)

HPC / AI Specialist (Torino)

18 set
|
The Italian Institute of Artificial Intelligence (AI4I)
|
Torino

18 set

The Italian Institute of Artificial Intelligence (AI4I)

Torino

ppThe Italian Institute of Artificial Intelligence (AI4I) is seeking a hands-on HPC / AI Specialist to support and optimize the compute infrastructure powering the AI Foundry and Deployment activities. /ppAt AI4I, you will work on a state-of-the-art AI computing environment built with best-in-class technologies acquired over the past year, including next-generation GPU systems such as NVIDIA B200 accelerators and high-performance distributed storage solutions such as VAST. This infrastructure is designed to support AI training, fine-tuning, and inference workloads for both research and industrial deployment. /ppLeonardo hosts the physical hardware infrastructure and delivers agreed infrastructure services in partnership with AI4I. In this role, you will focus on optimizing performance and providing direct technical support to internal users and clients running AI workloads, while contributing to the continuous evolution and improvement of the system design. /ppYou will act as a key interface between the machine infrastructure and the teams executing AI workflows, ensuring efficient, stable, and predictable operations. /ppbHybrid work /b: Flexible arrangements may be negotiated /ppThe position will remain open until it is filled and multiple candidates may be hired. /ppbAbout The Role /b /ppAs HPC / AI Specialist at AI4I, you will operate at the intersection of infrastructure operations and applied AI execution. You will ensure that engineers, researchers, and deployment teams can efficiently run training, fine-tuning, inference, and data-intensive pipelines on shared compute resources. /ppThis is a cross-unit role shared 50% with the new Deployment Unit, working closely with both infrastructure and client-facing teams. /ppbYou Will Work Closely With /b /pulliAI engineers and ML / GenAI teams running training, fine-tuning, and inference workloads /liliCloud / DevOps engineers operating the private cloud /liliThe Deployment Unit supporting industrial AI clients /liliHardware vendors and infrastructure partners /li /ulpThis role is strongly operational and performance-oriented, with a focus on workload efficiency, system tuning, AI workload optimization, and user support. /ppbKey Responsibilities /b /pulliOperate and maintain Linux-based HPC clusters supporting AI training, fine-tuning, and inference workloads /liliManage GPU and CPU compute environments, including workload scheduling, resource isolation,



and performance tuning /liliSupport distributed and software-defined storage systems used for large-scale datasets /liliAct as the primary technical interface between infrastructure operations and internal or external users running AI workflows /liliProvide hands-on technical support for AI workload optimization, including distributed training and parameter-efficient fine-tuning of foundation models on HPC infrastructure /liliSupport foundation model fine-tuning workflows (including parameter-efficient approaches), including configuration of data pipelines, checkpoints, runtime settings, and GPU memory optimization in HPC environments /liliOptimize resource utilization and workload performance across multi-tenant environments /liliSupport containerized workloads running on shared compute infrastructure /liliMonitor system health, performance, and capacity; troubleshoot user-facing production issues /liliContribute to the continuous improvement and evolution of system architecture in collaboration with infrastructure teams /liliSupport internal users and clients with debugging, environment setup, and best practices for scalable AI execution /li /ulpbRequired Qualifications /b /pulliSolid background in CPU and GPU architectures and performance characteristics /liliExperience operating HPC clusters or large-scale compute environments /liliHands-on experience with distributed and software-defined storage systems (e.g., VAST or equivalent) /liliExperience with workload managers and job schedulers (e.g., Slurm or equivalent) /liliExperience troubleshooting performance bottlenecks in compute or storage environments /liliPractical understanding of AI training and fine-tuning workloads, including GPU memory management, batch sizing strategies, distributed execution constraints, checkpointing, and data pipeline performance in HPC or large-scale compute environments /liliScripting and automation skills (Bash, Python, or equivalent) /liliExperience supporting shared infrastructure with uptime and operational responsibility /li /ulpbAdditional Strengths /b /pulliExperience supporting AI / ML training workloads in production environments /liliExperience with parameter-efficient fine-tuning workflows and runtime optimization in shared HPC environments /liliFamiliarity with foundation model adaptation workflows and large-scale training constraints /liliFamiliarity with containerized execution environments /liliExperience operating multi-tenant compute environments /liliExperience with monitoring and observability systems /liliNetworking fundamentals for high-throughput environments /liliExperience collaborating with engineering or deployment teams in production settings /li /ulpbKey Performance Metrics /b /pulliCluster availability and operational stability /liliGPU and CPU utilization efficiency /liliWorkload performance and scheduling effectiveness /liliTime required to debug and resolve user issues /liliTime required to onboard new workloads and users /li /ulpbWhat We Offer /b /pulliA collaborative environment with engineers and researchers working on real industrial AI deployments /liliDirect impact: your infrastructure will run daily AI workloads and production systems /liliAn office at the epicenter of tech: OGR Torino technology hub /liliCompetitive compensation and access to advanced computing infrastructure /liliSalary range: €40,000 – €60,000 gross/year (depending on experience) /li /ulpbAbout Us /b /ppThe Italian Institute for Artificial Intelligence (AI4I) was founded as a research institute to perform transformative, application-oriented research in Artificial Intelligence, driving innovation and industrial progress. The Institute is designed to engage and empower gifted, entrepreneurial, and ambitious researchers who are committed to generating real-world impact at the intersection of science, technology, and industrial transformation. /ppCompetitive salaries, performance-based incentives, access to dedicated high-performance computing resources, state-of-the-art laboratories, and strong industrial collaborations are among the distinctive features that define AI4I. The Institute fosters a dynamic international environment and an ecosystem that supports the creation and growth of innovative startups. /ppAI4I’s mission is to advance scientific research, technology transfer, and, more broadly, Italy’s innovation capacity, promoting positive impact across industry, services, and public administration. To achieve this, the Institute contributes to building a research and innovation infrastructure that leverages AI methods, with a special focus on manufacturing processes and the broader Industry 4.0 value chain. /ppAI4I also maintains strategic relationships with leading organizations in Italy and abroad, including Competence Centers and European Digital Innovation Hubs (EDIHs), positioning itself as an attractive destination for researchers, companies, and startups seeking collaboration and impact. /p /p #J-18808-Ljbffr

📌 HPC / AI Specialist (Torino)
🏢 The Italian Institute of Artificial Intelligence (AI4I)
📍 Torino

Candidati a questo annuncio

Mostra le tue capacità professionali all'azienda, compila il form e lascia un tocco personale nella lettera di presentazione, aiuterà il recruiter nella scelta del candidato.

Iscriviti a questa job alert:

Ricevi via email le nuove offerte di lavoro per: hpc / ai specialist (torino) / torino

Iscriviti a questa job alert:

Ricevi via email le nuove offerte di lavoro per: hpc / ai specialist (torino) / torino