Principal Software Engineer
Job description
Overview
Working at Atlassian India
Atlassian's mission "to unleash the potential of every team" is the guiding light behind what we do. Our products, including Jira, Confluence and Bitbucket, help teams everywhere work better together, from NASA to Cochlear.
Our office is in Bengaluru, but eligible candidates can work remotely across India: from home, from an office, or somewhere in between. We call this TEAM Anywhere.
Your future team: KubePlat Team
The KubePlat group in Core Engineering owns Atlassian's entire Kubernetes ecosystem. Every Atlassian cloud product, including Jira, Confluence, Loom and Rovo, runs on the clusters, platform services and deployment systems we build and operate. Our mission is to provide Atlassian a secure, reliable, cost-efficient and globally distributed compute platform, as the foundation for building great products.
What we own:
- Kubernetes infrastructure: hundreds of production clusters (EKS and GKE) across AWS, GCP and regulated cloud environments.
- Fleet management: automated, cloud-agnostic lifecycle management for clusters and platform components, with safe progressive rollouts.
- Kubernetes Platform-as-a-Service: an opinionated, multi-tenant platform for running Atlassian's microservices.
- Cloud resource management: self-serve, Kubernetes-driven provisioning of the cloud resources that services depend on.
- Secure sandboxed compute: strongly isolated environments for running untrusted code and AI-agent workloads.
- Software delivery: artifact management, deployment orchestration and automated verification for thousands of deployments every day.
- AI/ML infrastructure: the GPU compute, high-bandwidth networking and parallel filesystem ecosystem on Kubernetes that powers model training and AI inference behind Atlassian's AI features.
We are in the middle of a major transformation: moving to a multi-cloud, cell-based architecture, meeting FedRAMP and data-residency requirements, and building a reliable, cloud-native foundation for AI training and inference at scale.
Responsibilities
As a Principal Engineer, you will set the technical direction for Atlassian's Kubernetes, compute and AI/ML infrastructure platform. You will lead engineers across several teams on our hardest problems, and take large, cross-team initiatives from design to launch.
What you'll do
- Define the architecture and multi-year roadmap for a multi-cloud, cell-based Kubernetes fleet that scales to thousands of clusters.
- Design Kubernetes-native control planes (controllers, operators, CRDs and policy) that give engineers self-serve compute, networking and cloud resources.
- Raise platform reliability and security: reduce blast radius, build safe progressive rollouts, define SLOs, isolate workloads and support compliance (for example FedRAMP).
- Lead compute efficiency through autoscaling, bin packing, capacity planning and cost optimisation.
- Build the AI/ML infrastructure ecosystem on Kubernetes: GPU compute at scale, high-bandwidth, low-latency networking, and high-performance parallel filesystems for training and inference.
- Influence engineering and product leaders globally, set architectural standards, and mentor senior engineers.
Qualifications
- 10+ years building and operating large-scale cloud infrastructure or platform-as-a-service systems.
- Deep, hands-on Kubernetes expertise from running it in production at scale: internals, networking, multi-tenancy, and building operators and CRDs.
- Strong experience with AWS and/or GCP (EKS/GKE) and distributed systems design.
- Strong programming skills in one or more languages such as Go, Python, Java or similar.
- Platform engineering experience with Infrastructure as Code, GitOps/CD (for example Terraform, Crossplane, ArgoCD) and observability.
- A track record of leading technical direction across teams, with excellent communication and mentoring skills.
Nice to have
- Experience with Kubernetes fleet management, cell-based architectures, sandboxing (gVisor, Firecracker) or regulated environments.
- Experience providing AI/ML infrastructure: GPU clusters, high-bandwidth networking (for example RDMA, InfiniBand, AWS EFA, GPUDirect) or parallel filesystems (for example Lustre, Amazon FSx for Lustre, GCP Parallelstore).
- Contributions to CNCF open-source projects.
Skills mentioned

The Future of Atlassian starts with you. At Atlassian, our future is rooted in helping teams unleash their potential by building tools that inspire collaboration and facilitate growth — interested in what’s next? We’re looking for people who believe that we can accomplish so much more together than apart. Let’s build our future, together We’ve got an ambitious road ahead. Join our team and help us shape the future. Going virtual together We're entering a new era of work, one that will see more of our day-to-day tasks taking place in a virtual environment, and we're adapting to embrace this change. Into the cloud Our products help teams of all sizes to do amazing things. As we bring our product suite to the cloud we're innovating every day. Atlassian is for everyone It's our mission to unleash the potential in every team, and we know that teams perform best when they are diverse and every team member feels that they belong. It's the unique contributions of all Atlassians that drive our success, and we're committed to building a culture where everyone has the opportunity to do meaningful work and be recognized for their efforts. To that end, we are committed to providing an environment free of discrimination for everyone. Perks of the trade Foundation leave: We love to pay it forward, so you get five paid days a year to volunteer at your favorite charity. Be the change in your community. Ownership: We're all eligible for equity awards, and we have big plans. That means when we play together, we win together. Work/life balance: Here at Atlassian, life is good. We have flexible hours, loads of time off, awesome events, and a generous relocation program. Modern Health: Our global partnership provides app-based mental health support at no cost. Learning budget: Growth is important. It’s why we provide a dedicated learning budget to support everyone’s growth & development. Workspace your way: Whether you’re based in an office or working remotely, we provide the tools or financial support to make your workspace work for you.
Apply for this job
Use the application link supplied with this listing to apply to Atlassian. Check the destination before entering personal information.
