Skip to main content
Bma Group logo

Principal Compute, Kubernetes Orchestration & BCDR Engineer

Bma Group
2 days ago
Full-time
On-site
Delaware, United States

Job Description

Job Description: Principal Compute, Kubernetes Orchestration & BCDR Engineer

Overview

We are seeking a motivated and skilled Principal Compute, Kubernetes Orchestration & BCDR Engineer to join our innovative and forward-thinking team. This role offers a unique opportunity to design, implement, and manage Kubernetes-based orchestration platforms and advance our organization's Business Continuity and Disaster Recovery (BCDR) strategy.

As a Principal Engineer, you will play a key role in shaping the future state of our IT infrastructure through modern orchestration methods and robust resilience solutions, all while upholding our organization's commitment to fostering diversity, inclusion, and equality.

You will collaborate with multidisciplinary teams, work on cutting-edge technology projects, and create scalable and resilient systems to support enterprise business operations.

Responsibilities

As a Principal Compute, Kubernetes Orchestration & BCDR Engineer, your responsibilities will include but are not limited to:

  • Kubernetes Orchestration:

    • Design, implement, and manage Kubernetes clusters, ensuring scalability, high availability, and security.
    • Implement container orchestration solutions to support microservices architecture.
    • Automate application deployment, scaling, and management on Kubernetes.
  • BCDR Strategy:

    • Develop and maintain a robust Business Continuity and Disaster Recovery (BCDR) strategy.
    • Implement policies, processes, and technologies to guarantee operational resilience and minimize downtime during incidents.
    • Perform regular testing of disaster recovery plans and ensure compliance with industry standards.
  • Compute Engineering:

    • Optimize infrastructure and compute resources to meet organizational workload demands.
    • Develop robust systems for containerized application hosting and lifecycle management.
    • Monitor resource utilization and resolve operational issues in a timely manner.
  • Collaboration & Leadership:

    • Partner with cross-functional teams, including DevOps, Security, and IT, to ensure seamless delivery of services.
    • Mentor and guide junior engineers, fostering an inclusive and supportive team environment where everyone's voice is valued.
    • Advocate for infrastructure automation, continuous improvement, and adoption of modern cloud-native practices.
  • Technical Innovation:

    • Research and implement emerging Kubernetes tools, practices, and technologies.
    • Drive the adoption of advanced orchestration strategies tailored to address evolving business needs.

Qualifications

Required Skills and Experience:

  • Proven expertise in Kubernetes, containerization, and orchestration platforms.
  • Strong foundation in IT infrastructure, cloud-native software engineering, and compute systems.
  • Experience with Business Continuity and Disaster Recovery (BCDR) strategy implementation and management.
  • Hands-on experience working with CI/CD pipelines and infrastructure as code (IaC) tools (e.g., Terraform, Ansible).
  • Proficiency in programming/scripting languages such as Python, Go, or Bash.
  • Familiarity with cloud platforms such as AWS, Azure, or Google Cloud.
  • Exceptional problem-solving, communication, and analytical skills.

Preferred Qualifications:

  • Certifications: Kubernetes Certified Administrator (CKA) or Kubernetes Certified Developer (CKAD).
  • Experience building highly available, fault-tolerant distributed systems.
  • Knowledge of networking principles, including load balancing, DNS, and TCP/IP.
  • Commitment to promoting diversity, equity, and inclusion in professional environments.

Education:

  • Bachelor's or Master’s degree in Computer Science, IT, or a related field, or equivalent work experience.

Day-to-Day

Your typical day will involve:

  • Building and managing Kubernetes clusters to orchestrate containerized workloads across development, staging, and production environments.
  • Collaborating with application development teams to containerize applications and optimize CI/CD pipelines.
  • Monitoring system health and performance, proactively addressing potential risks or issues.
  • Conducting regular BCDR tests and participating in incident management to restore operations efficiently in case of disruptions.
  • Hosting knowledge-sharing sessions to upskill team members and promote inclusive collaboration.
  • Exploring and testing innovative tools and technologies to enhance system resilience, scalability, and security.

Why Join Us?

We are an equal-opportunity employer committed to building a diverse and inclusive workplace where everyone feels valued and supported. By joining our team, you will work in a collaborative, growth-focused environment on meaningful projects that drive operational excellence and innovation. We celebrate uniqueness and empower people to bring their authentic selves to work every day.

If you are passionate about emerging technologies, robust system designs, and fostering inclusive environments, we can't wait to meet you.

Apply now and help shape the future of modern IT engineering with us!