Deep Learning Scientist – Foundation Models (Roma)

Deep Learning Scientist – Foundation Models (Roma)

16 ago
|
Translated
|
Roma

16 ago

Translated

Roma

ph3About Translated /h3 pTranslated is on a mission to allow everyone to understand and be understood, in their own language. We are a technology-powered professional translation provider. We partner with over 200,000 professional translators worldwide, in 200 languages. Our 310,000 clients range from private individuals who need their CV translated to some of the world's largest companies, such as Uber and Airbnb. /p pOur progress is largely powered by our ability to leverage scientific progress and realize the best synergy between humans and machines. We invest heavily in RD, including Large Language Models, Machine Translation, expressive speech synthesis, and privacy-preserving training. We operate as a science-driven company, where scientific innovations quickly make their way into production and have a measurable impact on our products and operations. /p h3We are looking for a Deep Learning Scientist to join our Foundation Models team, working on the development of large-scale textual language models from scratch. /h3 pThe ideal candidate has a strong interest in Large Language Models, large-scale deep learning, and experimental research, and is excited about pushing the state of the art in multilingual language modeling. /p h3The Project: Foundation Models /h3 pTranslated is building its own generation of multilingual Foundation Models, with the ambition of advancing the state of the art in open Large Language Models, particularly in multilingual capabilities. /p pOur work covers the complete LLM pre-training lifecycle: from data selection and curation to model architecture, scaling experiments, large-scale distributed training, evaluation, and continued training. /p pA central research question for the team is how to build models that perform strongly not only in English, but across a broad range of languages, including languages that are traditionally underrepresented in large-scale pre-training datasets. /p ul lilarge-scale language model pre-training /li limultilingual language modeling /li lidata curation, filtering, scoring, and mixture design /li liscaling laws and training dynamics /li limodel architecture and optimization /li lidistributed and multi-GPU training /li lievaluation and benchmarking of Large Language Models /li licontinued pre-training and mid-training strategies /li /ul pYou will work in the Foundation Models team, a multidisciplinary group of scientists and engineers responsible for designing, training, and evaluating Translated’s next generation of Large Language Models.



/p pThe team works closely with Translated's AI Research and Engineering organizations, combining experimental research with the engineering required to train models at scale on large GPU clusters and HPC infrastructure. /p h3Responsibilities /h3 ul lidesign and conduct research on Large Language Model pre-training /li lidesign experiments, implement them in code, run them at scale, and analyze their results /li liinvestigate model architectures, optimization strategies, training dynamics, and scaling behavior /li liresearch and develop multilingual training strategies /li liwork on data selection, quality, composition, and data mixture experiments for LLM pre-training /li lievaluate models across multilingual and general-purpose benchmarks /li limonitor and benchmark the state of the art in Large Language Models /li lirun experiments on large-scale GPU and HPC infrastructure /li litranslate research hypotheses into measurable experiments and actionable decisions for large-scale training /li /ul h3Requirements /h3 ul li3+ years of research/industry experience in a relevant area of Deep Learning, Machine Learning, Natural Language Processing, or Large Language Models /li listrong understanding of modern deep learning and Transformer-based language models /li liexcellent programming skills in Python /li liexperience designing and running machine learning experiments /li lifamiliarity with GPU-based training environments and Unix/Linux systems /li liability to analyze experimental results and make research decisions based on empirical evidence /li liinterest in large-scale experimental research and language model pre-training /li liability to follow, understand, and reproduce recent scientific literature /li liexcellent written and spoken English /li liability to work effectively with both researchers and engineers /li /ul h3Bonus points if...



/h3 ul liyou have direct experience pre-training or continuing the pre-training of Large Language Models /li liyou have experience with distributed and multi-GPU training /li liyou have worked with large-scale training frameworks such as Megatron Bridge /li liyou have experience optimizing GPU utilization, throughput, memory consumption, or large-scale training stability /li liyou have worked on multilingual NLP or multilingual language models /li liyou have experience with large-scale dataset curation, filtering, deduplication, or data mixture design /li liyou have experience with LLM evaluation and benchmarking /li liyou have experience working with HPC environments /li /ul h3Our Offer /h3 pDepending on expertise, the offered salary typically ranges between €40.000,00 and €55.000,00. Compensation generally grows quickly as experience leads to greater contributions. Exceptionally strong candidates may be offered salaries above this range to acknowledge their higher potential impact. /p pWe offer a flexible smart-working policy, while requiring regular presence at our Rome headquarters. /p h3Headquarter /h3 pTranslated is hosted at Pi Campus, a working environment immersed in nature where six luxury villas in Rome, Italy, have been converted into functional offices designed to foster talent growth. Pi Campus is also a venture firm created by Translated to reinvest part of its profits into promising AI startups. /p h3Benefits and Perks /h3 ul liAt Translated, we see our people as athletes, and Pi Campus as the place where they can reach their full potential. /li liThis vibrant environment fosters talent aggregation and continuous growth of mind, body, and spirit. /li liWe nurture and support the team's personal and professional development every day through growth-oriented initiatives such as personalized learning paths and networking events with industry leaders. /li liHealth-oriented initiatives including massages, sauna, workouts in our gym and swimming pool; /li liRelaxation rooms, open meeting rooms, a cafeteria, and fully equipped kitchens in every villa. /li liIn case of need, every team member has access to psychological, legal, and financial support. /li liLearn more about our company: /li /ul h3Diversity Statement /h3 pAt Translated, we proudly embrace and celebrate each individual’s unique qualities, regardless of race, sexual orientation, gender identity, or any other differences. We recognize that diverse perspectives empower us to overcome challenges, foster innovation, and drive excellence. As an inclusive and equal-opportunity employer, we are committed to cultivating an environment where everyone feels welcome, valued, and supported to achieve their full potential. /p /p #J-18808-Ljbffr

📌 Deep Learning Scientist – Foundation Models (Roma)
🏢 Translated
📍 Roma

Candidati a questo annuncio

Mostra le tue capacità professionali all'azienda, compila il form e lascia un tocco personale nella lettera di presentazione, aiuterà il recruiter nella scelta del candidato.

Iscriviti a questa job alert:

Ricevi via email le nuove offerte di lavoro per: deep learning scientist – foundation models (roma) / roma

Iscriviti a questa job alert:

Ricevi via email le nuove offerte di lavoro per: deep learning scientist – foundation models (roma) / roma