AI Research Engineer (Model Compression & Quantization)
About this role
AI Research Engineer (Model Compression & Quantization)
As a member of our AI research team, you will drive innovation in model compression and efficient deployment for advanced multimodal AI systems, including large language models (LLMs) and vision-language models (VLMs). Your work will focus on reducing model footprint and computational cost while preserving accuracy, enabling high-performance AI to run efficiently across resource-constrained edge devices. You will apply and advance compression techniques such as quantization, knowledge distillation, and pruning to streamline complex multimodal architectures that integrate text, images, and audio.
Read the full description on TalentApply
Create a free account to see the complete job description, how well your CV matches this role, and apply in one click.
AI rewrites and formats your CV so it reads well and gets past screeners.
Get your match percentage for this exact role before you spend time applying.
Send a polished application in one click — no retyping the same details.
Follow every application in one place instead of digging through your inbox.
Free account · No card required
Your next opportunity starts here
Prepare, apply, track, interview and get hired — all from one platform, with AI in your corner.