Senior Software Engineer Full Stack Serverless
Job description
You will build core Serverless product features across frontend and backend, including dashboards, logs, observability, configuration, and usage experiences. You will design APIs, improve performance, reliability, and scalability, and own features from design through production iteration. You will work on reusable systems that support enterprise customers.
Responsibilities
- Build and maintain Serverless UI features
- Design and implement backend APIs
- Improve performance, reliability, and scalability of customer-facing systems
- Align product features with platform capabilities
- Own features from design through production and iteration
Requirements
- 5+ years of frontend and backend experience
- Ability to switch between UI, backend, and performance work
- Proficiency with TypeScript, Python, Postgres, and Next.js
- Experience owning production features end-to-end
- Experience with developer platforms or infrastructure-adjacent products
- Familiarity with production observability tooling
- Experience in generative AI inference, training, and compute
- Background in distributed systems, container orchestration, or cloud-native architectures
- Experience with real-time systems, streaming logs, or high-throughput data pipelines
- Exposure to Kubernetes, Prometheus, Datadog, or gRPC
Benefits
- Equity
- Health, dental, and vision insurance
- Regular team events and offsites
Skills mentioned

fal.ai is the generative media cloud trusted by over 1,000,000 developers and leading companies worldwide. The platform provides access to 600+ production-ready image, video, audio, and 3D models, all accessible through a unified API. With the world's fastest inference engine that's up to 10x faster than alternatives, fal enables applications to scale from prototype to 100M+ daily inference calls with 99.99% uptime. The platform powers AI features for major enterprises including Canva, Perplexity, Quora's Poe, and Adobe. Founded in 2021 by Burkay Gur and Gorkem Yurtseven, fal.ai delivers on-demand serverless GPUs and dedicated compute clusters featuring the latest NVIDIA hardware including H100, H200, and B200 chips. The company has developed the fastest inference for models like SDXL and Whisper, supporting hundreds of millions of customers. With SOC 2 compliance, enterprise-grade security, and pricing starting at $1.2 per hour for premium GPUs, fal enables developers to build, deploy, and train custom AI models at scale.
Apply for this job
Use the application link supplied with this listing to apply to fal. Check the destination before entering personal information.
