AI Researcher, TAO Multi-Modal Model Development

NVIDIA·Hanoi, Hanoi, Vietnam | Ho Chi Minh City, Ho Chi Minh City, Vietnam·posted 99d ago · last seen 39m ago

Track this application

Get Started Free

Match score against your CV

Get Started Free

Tailor your resume to this job

Get Started Free

About interviewing at NVIDIA

Recruiter screen, a technical screen mixing resume deep-dive with live coding, a hiring-manager conversation, then a panel of three to five 45–60 minute rounds: coding, systems design under hardware constraints, a domain deep-dive, and behavioral. Highly team-specific — you interview directly with the team — with C++ depth expected almost universally and decisions sometimes taking five-plus weeks after the panel.

Read the full NVIDIA interview process →

Description

NVIDIA is seeking a motivated AI Model Development Researcher to join the TAO — Train, Adapt, Optimize — Multi-Modal Model Development team in Hanoi or Ho Chi Minh City, Vietnam. In this role, you will contribute to the development, adaptation, optimization, and evaluation of advanced AI models within the NVIDIA frameworks. You will work on cutting-edge areas such as multi-modal learning, vision-language models, image segmentation, foundation model adaptation, and scalable deep learning workflows.

You will collaborate with engineers, researchers, and cross-functional teams to build practical AI solutions that can be integrated into production pipelines, NVIDIA SDKs, and real-world customer use cases. This is an excellent opportunity for an early-career engineer/scientist who is passionate about machine learning, deep learning, vision-language models, and building high-quality AI software.

What you'll be doing:

  • Develop and fine-tune multi-modal AI models using NVIDIA’s TAO Toolkit and deep learning frameworks.

  • Contributes to the design and implementation of vision-language models (VLMs) and universal segmentation systems.

  • Conduct experiments and benchmarking to evaluate model accuracy, robustness, and scalability.

  • Collaborate with cross-functional teams to integrate your research into production-level pipelines and NVIDIA SDKs.

  • Participate in research discussions, code reviews, and technical documentation to share insights and improve methodologies.

What we need to see:

  • BS or MS in Electrical Engineering, Computer Engineering, Computer Science, or a related field (or equivalent experience).

  • 2+ years of experience in machine learning, deep learning, or computer vision model development.

  • Strong Python programming skills and proficiency with PyTorch or similar frameworks.

  • Solid understanding of neural network architectures, transformers, and multi-modal learning techniques.

  • Excellent problem-solving abilities, attention to detail, and a collaborative mindset.

  • Familiarity with vision-language models, image segmentation, or large-scale pretraining is a strong plus.


Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/


NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

More engineering roles at NVIDIA