Job Description

About Onehouse
Onehouse is a mission-driven company dedicated to freeing data from data platform lock-in. We deliver the industry’s most interoperable data lakehouse through a cloud-native managed service built on Apache Hudi. Onehouse enables organizations to ingest data at scale with minute-level freshness, centrally store it, and make available to any downstream query engine and use case (from traditional analytics to real-time AI / ML).

We are a team of self-driven, inspired, and seasoned builders that have created large-scale data systems and globally distributed platforms that sit at the heart of some of the largest enterprises out there including Uber, Snowflake, AWS, Linkedin, Confluent and many more. Riding off a fresh $35M Series B backed by Craft, Greylock and Addition Ventures, we're now at $68M total funding and looking for rising talent to grow with us and become future leaders of the team. Come help us build the world's best fully managed and self-optimizing data lake platform!

The Community You Will Join
When you join Onehouse, you're joining a team of passionate professionals tackling the deeply technical challenges of building a 2-sided engineering product. Our engineering team serves as the bridge between the worlds of open source and enterprise: contributing directly to and growing Apache Hudi (already used at scale by global enterprises like Uber, Amazon, ByteDance etc) and concurrently defining a new industry category - the transactional data lake. The Cloud Infrastructure team is the glue that binds all of this together. We enable our data platform and infrastructure teams to be successful by acting as the foundation of a truly cloud native product.
The Impact You Will Drive:
  • As a Backend Engineer on the Kubernetes Infrastructure team, you'll design and build the software that powers Onehouse's cloud platform. You'll work on Kubernetes-native infrastructure including Operators, Controllers, and platform automation that manage distributed data systems at scale. Your work will directly influence how our platform provisions, upgrades, monitors, and operates customer deployments across AWS, GCP, and Azure.

    This is a software engineering role focused on extending Kubernetes not operating clusters as a DevOps or Site Reliability Engineer.
  • A Typical Day:
  • Design and develop Kubernetes Operators and Controllers using Go and controller-runtime.

  • Build Kubernetes-native automation for provisioning, upgrading, scaling, and managing distributed systems.

  • Develop Custom Resource Definitions (CRDs) and reconciliation logic to automate platform lifecycle management.

  • Build backend services that integrate with Kubernetes APIs and cloud infrastructure.

  • Improve the reliability, scalability, and observability of infrastructure software running on Kubernetes.

  • Collaborate with backend, data platform, and frontend engineers to deliver platform capabilities.

  • Design software that powers Spark and AI workloads on Kubernetes.

  • Participate in architecture discussions and contribute to long-term platform design.

  • Debug complex distributed systems running in Kubernetes production environments.

  • Maintain high engineering standards through code reviews, automated testing, and observability.

  • What You Bring to the Table:
  • 5+ years of experience building backend or infrastructure software.
  • Strong programming skills in Go (preferred) or Java
  • Experience developing Kubernetes Operators using controller-runtime.
  • Hands-on experience with:
  • Kubernetes Controllers
    Custom Resource Definitions (CRDs)
    reconciliation loops
    Operator pattern
  • Strong understanding of Kubernetes APIs and Kubernetes architecture.
  • Experience building distributed systems or cloud-native backend services.
  • Experience debugging production systems running on Kubernetes.
  • Familiarity with one or more cloud platforms (AWS, GCP, or Azure).
  • Strong software engineering fundamentals including testing, debugging, code reviews, and system design.
  • Nice to Have

  • Built production Kubernetes Operators.
  • Contributed to Kubernetes ecosystem projects.
  • Designed software for large-scale cloud platforms
  • Experience building control planes for distributed systems
  • Passion for open source and cloud-native technologies.
  • What This Role Is Not

    This is not primarily a DevOps, Site Reliability Engineering (SRE), or Kubernetes cluster administration role.

    While you'll work closely with cloud infrastructure and production systems, your primary focus will be designing and building Kubernetes-native software rather than day-to-day cluster operations.