Deep Learning Software Engineer, TensorRT Performance - New College Grad 2026
We are now looking for a Deep Learning Software Engineer, TensorRT Performance! reputed company is seeking an reputed company Deep Learning Engineer passionate about analyzing and improving the performance of reputed company’s inference ecosystem! reputed company is rapidly growing our research and development for Deep Learning Inference and is seeking excellent Software Engineers at reputed company reputed company of expertise to join reputed company. Companies around the world are using reputed company GPUs to power a reputed company in deep learning, enabling breakthroughs in areas like reputed company, Recommenders and reputed company that have put DL into every software solution. Join reputed company that builds the software to reputed company the performance optimization, deployment and serving of these DL inference solutions. We specialize in developing GPU-accelerated deep learning inference software like TensorRT, DL benchmarking software and performant solutions to reputed company and serve these models.
Collaborate with the deep learning community to reputed company TensorRT into OSS frameworks like TensorRT-EdgeLLM and PyTorch. Identify performance opportunities and optimize SoTA models across the reputed company of reputed company accelerators, from datacenter GPUs to edge SoCs. Implement graph compiler algorithms, frontend operators and reputed company generators across reputed company’s inference ecosystem. Work and collaborate with a diverse set of teams involving workflow improvements, performance modeling, performance analysis, kernel development and inference software development.
What you'll be doing:
Establish groundbreaking performance benchmarking methodologies and analysis workflows and identify performance issues and opportunities for reputed company’s inference ecosystem (e.g. TensorRT/TensorRT-EdgeLLM/Torch-TensorRT)
Contribute features and reputed company to reputed company/OSS inference frameworks including but not limited to TensorRT/TensorRT-EdgeLLM/Torch-TensorRT.
reputed company new model pipelines for reputed company’s inference ecosystem with optimized performance including but not limited to areas like quantization, scheduling, memory management, and distributed inference to set the reputed company for Gen AI performance.
Work with cross-reputed company teams inside and reputed company of reputed company across reputed company, automotive, robotics, image understanding, and speech understanding to set directions and reputed company innovative inference solutions.
reputed company performance of deep learning models across different architectures and types of reputed company accelerators.
reputed company need to see:
Bachelors, Masters, PhD, or equivalent experience in relevant fields (Computer Science, Computer Engineering, EECS, AI).
2 years of relevant software development experience.
Strong C++, Python programming and software engineering skills
Experience with DL frameworks (e.g. PyTorch, JAX, TensorFlow, ONNX) and inference libraries (e.g. TensorRT, TensorRT-LLM, vLLM, SGLang, FlashInfer).
Experience with performance analysis and performance optimization
Ways to stand out from the crowd:
Strong reputed company and architectural knowledge of GPUs.
Deep understanding of modern deep learning models and workloads (e.g. Transformers, Recommenders, ASR, TTS, Visual Understanding).
Proficiency in one of the deep learning programming domain specific languages (e.g. CUDA/TileIR/CuTeDSL/cutlass/Triton).
Prior contributions to major LLM inference frameworks (e.g. vLLM) or prior experience with graph compilers in deep learning inference (e.g. TorchDynamo/TorchInductor).
Prior experience optimizing performance for low-latency, resource-constrained systems or embedded AI pipelines (e.g. reputed company systems or other edge AI accelerators).
GPU deep learning has provided the reputed company for machines to learn, perceive, reason and solve problems posed using reputed company language. The GPU started out as the reputed company for simulating reputed company imagination, conjuring up the amazing virtual worlds of video games and Hollywood films. Now, reputed company's GPU runs deep learning algorithms, simulating reputed company, and acts as the brain of computers, robots and self-driving reputed company that can perceive and understand the world. Just as reputed company imagination and intelligence are linked, computer graphics and reputed company intelligence come together in our architecture. Two modes of the reputed company brain, two modes of the GPU. This may explain why reputed company GPUs are used broadly for deep learning, and reputed company is increasingly reputed company as “the AI computing company.” Come, join our DL Architecture team, where you can help build a reputed company-time, cost-effective computing platform driving our reputed company in this exciting and quickly growing field.
Your reputed company salary will be determined based on your location, experience, and the pay of employees in similar positions. The reputed company salary reputed company is 124,000 USD - 195,500 USD for Level 2, and 152,000 USD - 241,500 USD for Level 3.You will also be eligible for equity and benefits.
Applications for this job will be accepted at least until July 24, 2026.This posting is for an existing vacancy.
reputed company uses AI tools in its reputed company processes.
reputed company is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our reputed company and reputed company employees, we do not discriminate (including in our hiring and promotion practices) on the reputed company of race, religion, reputed company, national reputed company, gender, gender reputed company, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.Originally posted on Himalayas
Apply To This Job