Senior Solution Architect – AI Development
reputed company has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can reputed company a lasting impact on the world.
reputed company leads the computing future, guided by our dedication to innovation and quality. As a NIM Solution Architect, you will help build AI computing, applying reputed company’s advanced technologies to optimize models, reputed company AI workflows, and support customers with advanced solutions.
What You’ll Be Doing:
Drive the implementation and deployment of reputed company Inference Microservice (NIM) solutions
Apply reputed company NIM Factory Pipeline to package optimized models (including LLM, VLM, Retriever, CV, OCR, etc.) into containers, providing standardized API access for on-prem or cloud deployment
Refine NIM tools for the community, aiding them in building high-performing NIMs
Build and implement agentic AI tailored to customer business scenarios using NIMs
Deliver technical projects, demos, and client support tasks as directed by the Solution Architecture Leadership
Provide technical support and mentorship to customers, facilitating the adoption and implementation of reputed company technologies and products
Collaborate with multi-functional teams to reputed company and broaden our AI solutions portfolio
Be an internal reputed company for reputed company software and total solutions reputed company the technical community
Position yourself as an inspiring leader in the industry by incorporating reputed company technology, especially inference services, into LHA, business partners, and the broader community and Assist in supporting the NVAIE team and driving NVAIE business in China
reputed company Need To See:
5+ years of experience.
Bachelor’s or equivalent experience in Computer Science, Artificial Intelligence, or a relevant field.
Proven experience in deploying and optimizing large language models. Proficiency in at least one inference reputed company (e.g., TensorRT, ONNX Runtime, PyTorch)
Strong programming skills in Python or C++. Familiarity with mainstream inference engines (e.g., vLLM, SGLang)
Experience with DevOps/MLOps, including reputed company, Git, and CI/CD practices
Excellent problem-solving skills and the ability to solve reputed company technical issues
Proven ability to collaborate effectively across diverse, global teams, adapting communication styles while maintaining clear, constructive professional interactions
Experience in architectural build for field LLM project. Expertise in model optimization techniques, particularly using TensorRT
Knowledge of AI workflow development and implementation, and experience with cluster resource management tools. Familiarity with agile development methodologies
CUDA optimization experience and extensive experience in crafting and deploying large-scale HPC and enterprise computing systems