All jobs

Senior Solution Architect – AI Development

100% Remote Full-time Open now

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

NVIDIA leads the computing future, guided by our dedication to innovation and quality. As a NIM Solution Architect, you will help build AI computing, applying NVIDIA’s advanced technologies to optimize models, develop AI workflows, and support customers with advanced solutions.

What You’ll Be Doing:

  • Drive the implementation and deployment of NVIDIA Inference Microservice (NIM) solutions

  • Apply NVIDIA NIM Factory Pipeline to package optimized models (including LLM, VLM, Retriever, CV, OCR, etc.) into containers, providing standardized API access for on-prem or cloud deployment

  • Refine NIM tools for the community, aiding them in building high-performing NIMs

  • Build and implement agentic AI tailored to customer business scenarios using NIMs

  • Deliver technical projects, demos, and client support tasks as directed by the Solution Architecture Leadership

  • Provide technical support and mentorship to customers, facilitating the adoption and implementation of NVIDIA technologies and products

  • Collaborate with multi-functional teams to develop and broaden our AI solutions portfolio

  • Be an internal advocate for NVIDIA software and total solutions within the technical community

  • Position yourself as an inspiring leader in the industry by incorporating NVIDIA technology, especially inference services, into LHA, business partners, and the broader community and Assist in supporting the NVAIE team and driving NVAIE business in China

What We Need To See:

  • 5+ years of experience.

  • Bachelor’s or equivalent experience in Computer Science, Artificial Intelligence, or a relevant field.

  • Proven experience in deploying and optimizing large language models. Proficiency in at least one inference framework (e.g., TensorRT, ONNX Runtime, PyTorch)

  • Strong programming skills in Python or C++. Familiarity with mainstream inference engines (e.g., vLLM, SGLang)

  • Experience with DevOps/MLOps, including Docker, Git, and CI/CD practices

  • Excellent problem-solving skills and the ability to solve complex technical issues

  • Proven ability to collaborate effectively across diverse, global teams, adapting communication styles while maintaining clear, constructive professional interactions

  • Experience in architectural build for field LLM project. Expertise in model optimization techniques, particularly using TensorRT

  • Knowledge of AI workflow development and implementation, and experience with cluster resource management tools. Familiarity with agile development methodologies

  • CUDA optimization experience and extensive experience in crafting and deploying large-scale HPC and enterprise computing systems

apply to this job

You might also like

Freelance Mathematics Expert with Python Expertise – AI Trainer

100% Remote Full-time

Client Security Analyst

100% Remote Full-time

Senior Technical Writer

100% Remote Full-time

Technical Writer

100% Remote Full-time

Principal Software Engineer Dynamo

100% Remote Full-time

[Remote] System Hardware Engineer Intern (5G/6G System R&D)

100% Remote Full-time

Experienced Remote Case Manager RN – Aetna One Advocacy, Arizona – Competitive Salary & Comprehensive Benefits

100% Remote Full-time

Remote Entry-Level Data Entry Analyst – Aetna Healthcare Data Management & Reporting (Work‑From‑Home)

100% Remote Full-time

Account Manager - Affiliate Marketing

100% Remote Full-time

COMMISSION-ONLY — Sponsorship & Affiliate Manager (Partnership Revenue)

100% Remote Full-time

Principal Data Scientist, Generative AI. Remote or Hybrid East Coast United States

100% Remote Full-time

Experienced Remote Data Entry Clerk – Flexible, Lucrative Opportunity for Detail-Oriented Individuals

100% Remote Full-time

Experienced and Ambitious Entry Level Customer Support Associate for Dynamic Telecommunications Services at blithequark

100% Remote Full-time

Senior Analyst: Policy

100% Remote Full-time

Experienced Live Chat Agent – Remote Full-Time Customer Service Representative for Dynamic Online Support Team

100% Remote Full-time

Sales Development Rep (Remote in US)

100% Remote Full-time

Experienced Full Stack Retail Manager – Customer Experience & Sales Development in Chattanooga, TN at arenaflex

100% Remote Full-time

[Remote-Position] REMOTE Data Analyst

100% Remote Full-time

Virtual Assistant Data Entry Jr (Part-Time) – Amazon Store

100% Remote Full-time

Remote Leasing Agent – Flexible Work, Commission-Based!

100% Remote Full-time