This job is no longer available
This job expired on 19/07/2026. It no longer accepts applications.
AI Research Engineer – Model Compression & Quantization (Remote)
tether · Roma
Job description
About the role
Join Tether’s AI model team to drive innovation in model serving and inference architectures. You will focus on optimizing model deployment and inference strategies to deliver responsive, efficient, and scalable performance across diverse applications.
Key responsibilities
- Design and optimize model serving pipelines for both resource‑efficient and complex multi‑modal models.
- Develop inference frameworks that handle text, image, and audio data at scale.
- Collaborate with cross‑functional teams to integrate optimized models into production systems.
- Research and implement state‑of‑the‑art compression and quantization techniques.
Required profile
- Strong background in advanced model architectures and AI research.
- Proven experience designing scalable model serving and inference solutions.
- Excellent English communication skills and ability to work remotely with a global team.
Required skills
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Italia.
Salaries by job title
- Barista 79
- Badante 37
- pizzaiolo 34
- District manager contratti luce e gas 31
- Cameriere 30
- cuoco 23
- Collaboratore Telefonico Smart Working 22
- Cameriera 19
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
tether
Roma