Back to Jobs

Senior Research Scientist

San Francisco · Hybrid

San FranciscoHybrid$190,000 - $250,000Posted 8 months ago
ML/AI EngineersSenior (5-8 years)

About the Role

Overview

This role is with a well-funded AI infrastructure company building next-generation multimodal AI models and a high-performance training and serving platform. The team is pushing the boundaries of large-scale AI systems and translating cutting-edge research into real-world, production-grade applications.

About the Role

We are seeking an exceptional Senior Research Scientist specializing in advanced AI and machine learning to drive foundational research and scientific innovation. In this role, you will lead research initiatives across large language models, generative modeling, optimization, and scalable training systems.

You will work hands-on with modern ML frameworks, conduct large-scale experiments, and collaborate closely with engineering teams to bring impactful research into production. This role is ideal for a researcher who is passionate about advancing the state of the art while seeing their work deployed in real systems.

Equal Opportunity

This employer is an equal opportunity organization. All qualified applicants will be considered without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, or disability.

About the Company

Compensation: $190K–$250K base + equity

Responsibilities

  • Lead research in advanced ML areas such as LLMs, generative AI, foundation models, diffusion models, and novel Transformer architectures
  • Design, implement, and evaluate new ML algorithms using frameworks like PyTorch and JAX
  • Run large-scale distributed training experiments across multi-GPU and accelerator-based systems
  • Drive performance improvements through framework debugging, profiling, and training pipeline optimization
  • Produce high-quality research output, including publications, internal research reports, patents, and reproducible code
  • Collaborate with engineering and product teams to transition research prototypes into scalable production systems
  • Stay current with the latest AI research and integrate state-of-the-art techniques into the company’s technical roadmap
  • Mentor junior researchers and contribute to building a strong, collaborative research culture

Required Qualifications

  • PhD in Computer Science, Machine Learning, AI, Mathematics, or a related field
  • 5+ years of academic or industry research experience (flexible for exceptional new PhDs from top programs)
  • Strong publication record in top-tier venues such as NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, or similar
  • Strong programming skills in Python with deep experience in PyTorch and/or JAX
  • Experience with performance profiling, debugging, and optimization of ML systems
  • Hands-on experience with large-scale distributed training

Nice to Have

  • Deep experience with the JAX / Flax / XLA stack
  • Exposure to efficient inference, model serving, or production ML systems
  • Research or engineering experience with diffusion models or other generative techniques
  • Experience collaborating closely with cross-functional engineering teams
  • Contributions to open-source ML frameworks or research libraries

Benefits & Perks

  • Medical, dental, and vision insurance
  • 401(k) plan
  • Daily lunch, snacks, and beverages
  • Flexible time off
  • Competitive salary and equity

Interested in this role?

Apply now and hear back within 48 hours

Join & Apply

Already have an account? Sign in

Posted 8 months ago

San Francisco

Similar Jobs

San Francisco · Hybrid

A well-funded AI infrastructure company is building next-generation multimodal foundation models alongside a high-efficiency serving platform. With deep industry backing and close collaboration with hardware partners, the team is scaling rapidly to deliver the full stack powering frontier AI models and real-world, high-performance applications.

$200,000 - $275,000Apply
San Francisco · Hybrid

Overview This role is with a rapidly growing AI infrastructure company building next-generation multimodal AI systems and a high-performance training and serving platform. The team is well-funded and works closely with leading accelerator partners to push the limits of performance for large-scale AI workloads. About the Role We are looking for a GPU Kernel Engineer who is passionate about extracting maximum performance from modern accelerators. In this role, you will design, implement, and optimize custom GPU kernels that power large-scale AI training and inference systems. You will work across the full hardware–software stack, from low-level kernel development to integrating optimized operations into high-level machine learning frameworks. This role is ideal for engineers who thrive at the intersection of GPU programming, systems engineering, and cutting-edge AI workloads. Compensation: $190K–$250K base + equity Equal Opportunity This employer is an equal opportunity organization. All qualified applicants will be considered without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, or disability.

$190,000 - $250,000Apply
San Francisco · Hybrid

Overview This role is with a fast-growing AI infrastructure company building next-generation multimodal models and a high-performance model training and serving platform. The team is backed by significant funding and works closely with leading accelerator partners to build the full software stack powering frontier AI systems. About the Role We are seeking a highly skilled Distributed Training Engineer to design, optimize, and maintain the software stack that enables large-scale AI training workloads. You will work across the entire machine learning infrastructure—from low-level CUDA/ROCm runtimes to high-level frameworks like JAX and PyTorch—ensuring systems are fast, stable, and scalable. This role is ideal for engineers who enjoy deep systems work, debugging complex hardware–software interactions, and optimizing performance across the full ML stack. You will play a critical role in enabling the training and deployment of large language models and generative AI systems. Compensation: $190K–$250K base + equity

$190,000 - $250,000Apply

Hiring

First candidates by day three.

One form, one call, a written search plan inside 24 hours. You pay nothing until someone signs.

Start a search

Looking

Roles that never hit a job board.

One profile covers every search we run. We only reach out when the role clears your bar — always free, never mass mail.

👋Hey — hiring engineers? Click me for jokes.