Member of Technical Staff, Full-Stack Engineering
Build and own production systems across the stack—from developer-facing web applications and APIs to databases, Kubernetes infrastructure, and the cloud systems powering large-scale image, video, and world-model inference.
What you’ll do
- Build and own full-stack product experiences for TensorScale’s AI inference platform, from frontend interfaces to backend services and infrastructure
- Design and implement reliable backend services and APIs in Go and/or Python
- Build fast, polished developer-facing web applications using React / Next.js
- Design and operate cloud infrastructure running production AI workloads on Kubernetes
- Manage infrastructure as code using Terraform and improve how we provision, deploy, and operate services across environments
- Design, operate, and optimize production databases such as PostgreSQL for high-throughput, reliable workloads
- Build systems for authentication, billing, usage metering, job orchestration, model deployment, and API lifecycle management
- Improve reliability, observability, scalability, and operational simplicity across the production stack
- Debug production issues across application, database, networking, Kubernetes, and cloud infrastructure layers
- Work closely with ML systems engineers to turn high-performance image, video, and world-model inference systems into reliable developer products
Minimum qualifications
- 3+ years of experience building and operating production software systems
- Strong backend engineering skills in Go and/or Python
- Hands-on experience with Kubernetes in production environments
- Experience managing cloud infrastructure with Terraform or equivalent infrastructure-as-code tooling
- Proficiency with at least one production relational database, preferably PostgreSQL
- Experience building modern web applications with React and/or Next.js
- Strong understanding of APIs, distributed systems, networking, and production system reliability
- Ability to independently own features and systems from design through deployment and operation
- Strong debugging instincts across the application and infrastructure stack
- Comfort working onsite with the team in the Bay Area
Preferred qualifications
- Experience building developer platforms, API products, cloud infrastructure, or AI infrastructure
- Deep familiarity with PostgreSQL performance, schema design, migrations, replication, and operational best practices
- Experience with Kubernetes networking, autoscaling, scheduling, ingress, service discovery, and cluster operations
- Experience operating infrastructure across AWS, GCP, or other major cloud providers
- Familiarity with observability stacks such as Prometheus, Grafana, OpenTelemetry, or equivalent
- Experience with message queues, caching systems, and asynchronous job processing
- Familiarity with GPU infrastructure, AI inference workloads, or distributed compute systems
- Experience building usage metering, billing, rate limiting, authentication, or other infrastructure for developer-facing APIs
- Experience working in an early-stage startup where engineers own systems end-to-end
To apply, email careers@tensorscale.io.