Tether Operations Limited is seeking an expert in AI model serving and inference optimization based in Italy. The role involves designing high-performance model architectures, ensuring low latency, and optimizing memory usage on resource-constrained devices. Candidates should possess a PhD in a related field and have extensive experience in GPU kernel writing and inference optimization. Di seguito, troverà un'analisi completa di tutti i requisiti per i potenziali candidati e le istruzioni su come candidarsi. In bocca al lupo The position requires a deep understanding of cutting-edge AI techniques, with responsibilities that encompass building robust inference pipelines and monitoring performance in live environments. xrdztoy Join us to contribute to innovative AI systems #J-18808-Ljbffr
📌 Edge AI Inference Architect: Low-Latency Model Serving (Italia)
🏢 Altro
📍 Italia