A well-funded AI infrastructure company is building next-generation multimodal foundation models and a high-efficiency serving platform. With deep industry backing and close collaboration with hardware partners, the team is scaling rapidly to deliver the full stack powering real-time, production-grade AI applications.
Software Engineer, Fullstack
San Francisco · Hybrid
About the Role
Overview
A well-funded AI infrastructure company is building next-generation multimodal foundation models and a high-efficiency serving platform. With deep industry backing and close collaboration with hardware partners, the team is scaling rapidly to deliver the full stack powering real-time AI applications and developer-facing platforms.
About the Role
This role focuses on building the developer-facing surfaces of a large-scale AI platform. You will design and implement fast, intuitive interfaces that allow users to interact with large language models, manage usage, and integrate APIs seamlessly into their own products.
You’ll collaborate closely with ML researchers and systems engineers, gain exposure to how large AI models are deployed and served at scale, and play a key role in delivering an excellent developer experience (DX). This role is ideal for engineers who enjoy fullstack ownership, care deeply about performance and UX, and want exposure to the full AI stack.
Equal Opportunity
We are an equal opportunity employer and value diversity at our company. All qualified applicants will be considered without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, disability, veteran status, or any other protected characteristic.
Responsibilities
- Design and build a low-latency, chat-style interface for users to interact with LLMs
- Build and maintain a developer console for API key management, usage tracking, and budget controls
- Integrate payments and billing systems to support subscription and usage-based pricing
- Develop a dynamic documentation portal for API references and guides
- Build and maintain CLI and SDK wrappers to simplify API integration for users
- Develop secure backend APIs and low-latency streaming solutions
- Collaborate closely with backend, ML, and infrastructure engineers to deliver end-to-end features
- Ensure high performance, reliability, and responsiveness across frontend and backend systems
Required Qualifications
- Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience
- 3+ years of software engineering experience with a focus on frontend or fullstack development
- Strong proficiency in TypeScript and Python
- Solid understanding of responsive design principles and UX fundamentals
- Experience building modern web interfaces and client-side applications
- Strong collaboration and communication skills across engineering and ML teams
- Comfortable working in a fast-moving, high-ownership, in-office environment
Nice to Have
- Experience with ML systems engineering or inference engines (e.g., vLLM, SGLang, TRT-LLM)
- Experience building real-time or streaming interfaces using WebSockets or Server-Sent Events (SSE)
- Experience integrating billing systems such as Stripe, including webhooks and subscription logic
- Strong documentation mindset with experience improving developer experience (DX)
- Ability to read and understand backend Python codebases (e.g., FastAPI, model-serving systems)
Benefits & Perks
- Medical, dental, and vision insurance
- 401(k)
- Daily meals and snacks
- Flexible time off
- Competitive compensation and equity
Interested in this role?
Apply now and hear back within 48 hours
Join & ApplyAlready have an account? Sign in
Posted 8 months ago
San Francisco
Similar Jobs
This role is with a rapidly growing, venture-backed AI software company building an intelligent operating layer for complex, integration-heavy enterprises. The platform helps organizations plan, deploy, and optimize their teams through advanced analytics, automation, and AI-driven workflows. The company works in real-world, security-constrained environments—including hybrid, on-prem, and restricted systems—and is focused on shipping durable, production-grade enterprise software.
Overview This role is with a well-funded AI infrastructure company building next-generation multimodal models and a high-performance model serving platform. The team is scaling rapidly to deliver production-grade systems that power real-time AI applications, working closely across research, systems, and infrastructure. About the Role This is a rare opportunity to help architect and lead the development of a next-generation model serving platform—the core engine that brings highly efficient multimodal foundation models into production. As a senior technical leader, you will both build critical components yourself and guide other engineers, shaping architectural decisions, engineering standards, and execution quality. You’ll work across the full AI stack, from GPU execution and optimized runtimes to distributed serving, scheduling, and APIs that power low-latency, real-time inference. This role is ideal for engineers who enjoy deep systems work, thrive on ownership, and want to lead the development of foundational AI infrastructure. Equal Opportunity This employer is an equal opportunity organization. All qualified applicants will be considered without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, or disability. Compensation: $230K–$300K base + equity
