Overview
In this role you will design and implement data ingestion and processing pipelines to enable a data-driven transformation of manufacturing data.
You will join the AMS team to support global incident handling and enhancement requests for ingestion processes, collaborating with the Milan team to align requirements and loading pipelines.
You will deploy and operate data and ML workloads in AWS, moving from prototype to production while ensuring quality and observability.
This position offers hands-on work with cutting-edge data tech in a globally connected, innovation-driven environment.
Retribuzione / Benefits
Buoni pasto
Flexible working hours
Hybrid Remote Working
Employee benefits
Responsabilità
Design and build streaming data pipelines using Kafka/MSK and stream processing (Flink/Spark Streaming)
Develop batch transformations into a medallion lakehouse on S3 (Bronze/Silver/Gold) with Iceberg/Parquet
Manage data storage across relational (Aurora/RDS), NoSQL (DynamoDB),
and lakehouse layers
Design, deploy, and manage AWS infrastructure for data/ML workloads via IaC (EKS, Fargate, Lambda, VPC)
Migrate prototypes to production with refactoring, optimization, CI/CD, and best practices
Ensure observability, data quality, and maintenance of pipelines in production
Requisiti fondamentali
Master's in Computer Science/Engineering
Strong SQL and Python
Streaming pipelines with Kafka/MSK and Flink or Spark Streaming
AWS hands-on: EKS, Fargate, Lambda, S3, Aurora/RDS, DynamoDB, VPC
Big-data stack (Spark, Hive, HDFS) and lakehouse formats (Iceberg/Parquet)
Git, CI/CD, Docker, IaC (Terraform), REST APIs; agile/scrum
Strong communication
Curiosity about learning new technologies
Good organizational and time management skills
Kafka/MSK
Flink
Spark Streaming
📌 Software Data Engineer (Aws) (Bari)
🏢 Pirelli
📍 Bari