Chine

NIM Solution Architect

NIM Solution Architect

Emploi Chine

Titre du poste : NIM Solution Architect

Entreprise : Nvidia

Description du poste : NVIDIA is leading company of AI computing. At NVIDIA, our employees are passionate about AI, HPC , VISUAL, GAMING. Our Solution Architect team is more focusing to bring NVIDIA new technology into difference industries. We help to design the architecture of AI computing platform, analysis the AI and HPC applications to deliver our value to customers. This role will be instrumental in leveraging NVIDIA’s cutting-edge technologies to optimize open-source and proprietary large models, create AI workflows, and support our customers in implementing advanced AI solutions.What you’ll be doing:Drive the implementation and deployment of NVIDIA Inference Microservice (NIM) solutionsUse NVIDIA NIM Factory Pipeline to package optimized models (including LLM, VLM, Retriever, CV, OCR, etc.) into containers providing standardized API access for on-prem or cloud deploymentRefine NIM tools for the community, help the community to build their performant NIMsDesign and implement agentic AI tailored to customer business scenarios using NIMsDeliver technical projects, demos and client support tasks as directed by the Solution Architecture LeadershipProvide technical support and guidance to customers, facilitating the adoption and implementation of NVIDIA technologies and productsCollaborate with cross-functional teams to enhance and expand our AI solutions portfolioBe an internal champion for NVIDIA software and total solutions in technical communityBe an industry thought leader on integrating NVIDIA technology especially inference services into LHA, business partners and whole communityAssist in supporting NVAIE team and driving NVAIE business in ChinaWhat we need to see:3+ years working experience with Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, or a related fieldProven experience in deploying and optimizing large language modelsProficiency in at least one inference framework (e.g., TensorRT, ONNX Runtime, PyTorch)Strong programming skills in Python or C++Familiarity with main stream inference engines (e.g., vLLM, SGLang)Experience with DevOps/MLOps such as Docker, Git, and CI/CD practicesExcellent problem-solving skills and ability to troubleshoot complex technical issuesDemonstrated ability to collaborate effectively across diverse, global teams, adapting communication styles while maintaining clear, constructive professional interactionsWays to stand out from the crowd:Experience in architectural design for field LLM projectsExpertise in model optimization techniques, particularly using TensorRTKnowledge of AI workflow design and implementation, experience on cluster resource management tools. Familiarity with agile development methodologiesCUDA optimization experience, extensive experience designing and deploying large scale HPC and enterprise computing systems

Salaire attendu :

Localisation : 上海市

Date du poste : Thu, 01 May 2025 03:40:29 GMT

Postulez dès maintenant !

Artia13

Depuis 1998, je poursuis une introspection constante qui m’a conduit à analyser les mécanismes de l’information, de la manipulation et du pouvoir symbolique. Mon engagement est clair : défendre la vérité, outiller les citoyens, et sécuriser les espaces numériques. Spécialiste en analyse des médias, en enquêtes sensibles et en cybersécurité, je mets mes compétences au service de projets éducatifs et sociaux, via l’association Artia13. On me décrit comme quelqu’un de méthodique, engagé, intuitif et lucide. Je crois profondément qu’une société informée est une société plus libre.

Artia13 has 14418 posts and counting. See all posts by Artia13