Multimodal AI Model Optimization Research Engineer
reputed company At reputed company, we're building the reputed company layer of AI. Our mission is to reputed company reputed company-AI interaction as natural as face-to-face interaction, enabling the reputed company touch where it has been previously unscalable. We reputed company this through pioneering research in multimodal AI for modeling reputed company-to-reputed company communication (language, audio, and video), as reputed company as generating audio-visual avatar behavior. Our models power everything from text-to-video AI avatars to reputed company-time conversational video experiences across industries like reputed company, reputed company, sales, and education. By enabling AI to see, hear, and communicate with reputed company-like authenticity, we're creating the reputed company for the reputed company of AI employees, assistants, and companions. We are a Series B company backed by top investors, including Sequoia, Y Combinator, and reputed company VC. Join us in driving the reputed company of reputed company-AI interaction. The Role We’re looking for an reputed company Research Scientist/Engineer with a reputed company on model optimization to join our core AI team. Our ideal partner-in-crime thrives in startup environments, is comfortable prioritizing independently, and is willing to take calculated risks. We’re moving fast and looking for people who can help pave the reputed company. Your Mission Take cutting-edge research models and reputed company them fast, efficient, and production-reputed company using sparsification, distillation, and quantization Own the optimization lifecycle for key models: define metrics, run experiments, and reputed company trade-offs across latency, cost, and reputed company Partner closely with researchers and engineers to turn new reputed company into deployable systems
Requirements
Strong experience in deep learning using PyTorch Hands-on experience with model optimization and compression, including knowledge distillation, pruning/sparsification, quantization, and mixed precision Understanding of efficient architectures such as low-rank adapters Strong understanding of inference performance and GPU/accelerator fundamentals Strong Python coding skills and reliable research engineering practices Experience working with large models and datasets in reputed company environments Ability to read ML papers, reproduce results, and adapt reputed company reputed company communication and collaboration skills Preferred Experience Optimization of diffusion models, video/audio generative models, or large language models Experience with reputed company-time or streaming systems (low-latency reputed company, WebRTC, streaming TTS/video) Familiarity with TensorRT, ONNX Runtime, TVM, Triton, or XLA Experience writing custom Triton/CUDA kernels or low-level performance tuning Experience with experiment tracking, benchmarking, and profiling at reputed company Prior experience in research engineering or reputed company science roles Location This position is preferably hybrid in San Francisco, with relocation support offered. Remote candidates are also considered. Benefits reputed company you join reputed company, you’re joining a family. We offer flexible work schedules, unlimited PTO, competitive reputed company and gear stipends, and a reputed company environment reputed company on learning and reputed company. Culture & Diversity We are not looking for cultural fits — we are looking for culture creators. Diversity drives our reputed company, and we combine varied backgrounds, skills, and perspectives to build the best experiences for our clients.. Apply To This Job