Applied Scientist, Alexa Edge AI (Vision, Audio and Multimodal Models), Bengaluru

Application ends: October 13, 2026
Apply Now

Job Description

Amazon’s Alexa Edge AI team is hiring an applied scientist for its newly established team in Bengaluru. The role covers computer vision, audio and speech processing, and multimodal semantic understanding, for models deployed both on devices and in the cloud. As a founding member of the new site, the scientist also helps hire and onboard the team and set its technical practices.

At a glance

  • Role: Applied Scientist, Alexa Edge AI
  • Posted: 11 August 2026
  • Focus: Computer vision, audio and speech ML, multimodal fusion, edge deployment
  • Location: Bengaluru, Karnataka, India
  • Job ID: 10498825
  • Apply: open until filled (re-checked 13 October 2026)

About the Alexa Edge AI role

  • Design and build deep learning models for computer vision, audio understanding and multimodal fusion across visual, audio and text inputs
  • Own the full ML lifecycle: problem framing, data strategy and annotation design, experiments, evaluation, optimisation and deployment
  • Apply state-of-the-art techniques and publish at venues such as CVPR, NeurIPS, ICASSP, ICCV and ACL
  • Help hire, onboard and set technical practices for the new Bengaluru site

What Amazon is looking for

  • A PhD, or a master’s degree and at least 3 years of experience in CS, CE, ML or a related field
  • At least 1 year of building models for business applications
  • Programming in Java, C++, Python or a related language
  • Experience building speech recognition, machine translation or natural language processing systems

Nice to have

  • Knowledge of standard speech and machine learning techniques
  • Patents or publications at top peer-reviewed venues
  • A PhD or work experience in computer vision, audio or multimodal language models
  • Distributed training, model compression and inference optimisation such as pruning, quantisation and distillation

Pay and location

Amazon’s posting does not state a salary; pay is discussed during the hiring process. The role is based in Bengaluru, India.

How to apply

Apply through the official Amazon job posting. Amazon’s posting does not give a closing date, so the role is open until filled; ResearchJobs.in will re-check this listing on 13 October 2026. Amazon lists the minimum and preferred qualifications on the posting; read both before applying. Check the posting for location and work-authorisation details before you apply.

See all our artificial intelligence jobs, or browse more research jobs on ResearchJobs.in.

Hiring institution: Amazon

Official advertisement: www.amazon.jobs

How to prepare for this application

  • Multimodal basics: revise audio-visual fusion, contrastive pre-training and how multimodal LLMs connect encoders to a language model.
  • Edge constraints: be ready to talk about running models under tight memory and power budgets.
  • Speech experience: ASR, MT or NLP systems are a basic requirement, so describe one you built end to end.
  • Founding-team mindset: prepare examples of setting up processes or mentoring in a new team.
  • Leadership Principles: Amazon interviews lean on them; prepare two or three STAR stories for each.

About Amazon in India

Amazon runs large science and engineering teams in India, including in Bengaluru, Hyderabad, Chennai, Pune and Gurugram. Applied scientists and data scientists work on machine learning problems across retail, logistics, advertising, devices, payments and cloud services.

We send one confirmation email first. Every alert has an unsubscribe link.