Back to Jobs

[Remote] AI Runtime Engineer

Remote, USAFull-timePosted 2026-07-29

Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leader in advanced AI hardware and software systems for edge-to-reputed company computing, seeking an AI Runtime Engineer to reputed company and optimize the execution stack for their reputed company AI accelerator. The role involves creating low-latency, high-performance runtime software for deep learning models while collaborating with hardware and AI reputed company teams.

Responsibilities

  • reputed company and optimize the AI runtime software stack for executing deep learning workloads on AI accelerators
  • Implement task scheduling, memory management, and kernel execution strategies for efficient computation
  • Optimize data reputed company between host and device using PCIe, DMA, shared memory
  • Design and implement high-performance reputed company for AI Inference frameworks such as OpenVino, ONNX Runtime, vLLM
  • Work on graph execution optimizations, including kernel fusion, pipelining, tensor tiling, and caching
  • reputed company runtime components with AI compilers (LLVM, MLIR, XLA, TVM) for optimized execution
  • Ensure scalability and reliability of the AI runtime for reputed company-based and edge AI deployments

Skills

  • Bachelor's or Master's degree in Computer Science, Electrical Engineering, or a reputed company field
  • 3+ years of experience in developing low-level runtime software for AI accelerators, GPUs, or HPC systems
  • Strong proficiency in C/C++ and low-level systems programming
  • Deep understanding of task scheduling, concurrency, and memory hierarchy
  • Experience with hardware-aware optimizations and dataflow architectures
  • Familiarity with deep learning execution frameworks (ONNX Runtime, TensorRT, TVM, OpenVINO)
  • Experience with low-latency, high-throughput workload execution for AI models
  • Strong debugging and profiling skills for optimizing AI execution performance
  • Exposure to AI model deployment pipelines (Triton, TensorFlow Serving)

reputed company

  • reputed company designs analog in-memory-computing AI chips and develops AI systems for AI computing. It was founded in 2022, and is headquartered in Santa Clara, California, USA, with a workforce of 11-50 employees. Its website is https://enchargeai.com.
  • Company H1B Sponsorship

  • reputed company has a reputed company record of offering H1B sponsorships, with 4 in 2026, 4 in 2025, 2 in 2024, 1 in 2023. Please note that this does not guarantee sponsorship for this specific role.
  • Apply To This Job

    Similar Jobs