
Search by job, company or skills
We are looking for a Solutions Architect to help customers and partners in South East Asia embrace NVIDIA technologies to build, deploy, and scale vision AI and video AI solutions. This is a highly technical role that requires deep expertise in VLM AI models and software engineering practices. We need a passionate, hard-working, expert and creative individual to help us pursue the many opportunities in this region.
A Solution Architect is the first line of technical expertise between NVIDIA and our customers, as well as our partners. Your duties will vary from solutions design, training/workshops, troubleshooting, project coordination, industry and marketing speaking engagements, customer relationship management and more.
What you'll be doing:
Drive the adoption of NVIDIA's vision AI and video AI capabilities.
Provide hands-on technical leadership to put VLM (Visual Language Models) and WFM (World Foundation Models) into applications for computer vision, document processing, and video analytics.
Support solution development and technical validation around NVIDIA vision AI and video AI technologies
Work with field, product, and engineering teams to translate customer requirements into deployable architectures, technical feedback, and roadmap input.
Deliver demos, workshops, technical reviews, and partner enablement sessions that accelerate adoption of NVIDIA AI technologies.
What we need to see:
BS, MS, or PhD in Computer Science, Electrical Engineering, Computer Engineering, Physics, Mathematics, Machine Learning, or a related field, or equivalent experience.
5+ years of proven experience in computer vision, vision AI, video AI, or world foundation models.
Hands-on experience in building and scaling production VLM and multi-modal AI systems.
Strong Python skills and experience with PyTorch.
Experience in the production trade-offs around latency, quality, and scale.
Familiarity with Docker, Kubernetes, Helm, or equivalent infrastructure tooling.
Ability to build and explain end-to-end AI pipelines, including model selection, customization, deployment, observability, and integration into customer applications.
Strong communication skills and the ability to work effectively across customers, partners, product management, and engineering.
Ways to stand out from the crowd:
Hands-on experience with NVIDIA AI technologies such as NVIDIA Metropolis + NVIDIA DeepStream SDK for live video pipelines, NVIDIA NIM for VLMs for serving, and Nemotron + Cosmos models.
Experience supporting model adaptation using NVIDIA TAO Toolkit + NVIDIA NeMo frameworks.
Experience benchmarking AI agents for quality, latency, reliability, safety, and user experience.
Experience working with enterprise partners to productionize voice AI applications in domains such as contact centers, digital humans, assistants, or industry-specific voice workflows.
Regarded as a top choice among technology companies, NVIDIA delivers highly competitive salaries and a well-rounded benefits package. As you consider your future, explore what we can provide for you and your family at
Job ID: 153248555
Skills:
software engineering practices , Computer Vision, Pytorch, Docker, Helm, Python, Kubernetes, NVIDIA TAO Toolkit, vision AI, video AI, NVIDIA DeepStream SDK, VLM AI models, NVIDIA NeMo, NVIDIA Metropolis, NVIDIA NIM
Skills:
software engineering practices , Kubernetes, Pytorch, Python, Docker, Helm, Computer Vision, VLM AI models, NVIDIA DeepStream SDK, video AI, NVIDIA NIM, NVIDIA Metropolis, NVIDIA NeMo frameworks, NVIDIA TAO Toolkit, vision AI
Skills:
Java, Cloud Services, Data Analytics, Networking, Rest Apis, Database Architecture, Infrastructure Architecture, Python, AWS
Skills:
Google Cloud, Pvs, Linux, Azure, Kubernetes, AWS, Aws S3, Azure Blob, KMS Key Vault encryption, Nutanix NKP, immutable storage policies, Service Mesh, Red Hat OpenShift, public cloud security, IAM roles, SUSE Rancher, CSI, cni
Skills:
Gpu, Infrastructure Management, Servers, Network Design, Ipmi, Devops, MLops, Linux, It Networking, Ethernet, Kubernetes, network storage, batch orchestrators, Base Command Manager, infiniband, RedFish, container runtimes, Cumulus Linux, Network Operator, AI solutions, NVIDIA systems, Systems Management, cloud native stacks, SLURM, RoCE fabrics, HPC Infrastructure design, DCGM, ip networking