
Principal Cloud Backend Engineer
Hybrid
Full Time
#Software Engineering
#Cloud
#AI
#Go
#Rust
#C++
#Kubernetes
#Docker
#AWS
#GCP
#Azure
#SQL
#NoSQL
We are currently witnessing the rise of pervasive AI, a transformative era where organizations leverage generative technology to unlock hidden value, streamline operations, and drive innovation at scale. At SambaNova Systems, we have developed the SambaNova Suite, the industry's first full-stack generative AI platform. By integrating our proprietary SN40L chip with state-of-the-art open-source models, we provide enterprise and government clients with a secure, high-performance environment that allows them to retain ownership of their fine-tuned models. We are looking for a Principal Cloud Backend Engineer to join our team in a hybrid capacity within the United States to help build the economic and technical backbone of our inference services.
Key outcomes
- Lead the technical vision and architectural strategy for our inference serving and monetization infrastructure.
- Design and implement fault-tolerant, highly available systems capable of scaling to meet significant demand.
- Develop flexible monetization frameworks, including systems for usage metering, quota management, and complex entitlement enforcement.
- Build clean, robust APIs to integrate our platform with external billing and payment providers.
- Architect distributed systems that handle real-time rate limiting, fair-share scheduling, and multi-tenant resource management.
- Collaborate with Product, Finance, and Go-To-Market teams to translate business requirements into scalable technical solutions.
- Mentor engineering staff and establish best practices regarding code quality, observability, and testing for financial data pipelines.
Requirements
- Over 10 years of professional software engineering experience, with a focus on large-scale, distributed backend systems in the cloud.
- At least 5 years serving in a Principal or Lead Engineer capacity, with a proven history of delivering business-critical platforms.
- Expert proficiency in Go, Rust, or C++, including a deep understanding of systems programming and performance optimization.
- Extensive hands-on experience with cloud-native technologies such as Kubernetes and Docker, alongside major providers like AWS, GCP, or Azure.
- Strong background in both SQL and NoSQL databases, specifically in modeling data for high-throughput, low-latency applications.
- Solid foundation in event-driven architecture, microservices, and API design using REST or gRPC.
- Excellent communication skills and the ability to drive technical consensus across diverse teams.
Preferred qualifications
- Direct experience building or extending platforms for usage-based billing, subscription management, or entitlements.
- Experience working with billing providers such as Stripe or Metronome.
- Background in AI/ML infrastructure, particularly in platforms designed for serving and scaling model deployment pipelines.
Compensation
We provide a comprehensive total rewards package that includes base salary, equity, and a robust suite of benefits. Our insurance offerings include medical, dental, vision, disability, and life insurance. We also provide an employee assistance program to support your well-being. Employees enjoy flexible working hours and the ability to work in a hybrid environment.
How to apply
We invite you to submit your resume and a cover letter for consideration. In your cover letter, please detail your experience architecting large-scale systems, with a specific focus on any work involving monetization, billing, or entitlements. We are interested in learning about the challenges you faced regarding reliability and accuracy and how you successfully addressed them.





