Remote Jobs RockRemote Jobs Rock

Staff Software Engineer , Anywhere Cloud - AI Systems & Runtimes

πŸ•’ 6 days ago
KubernetesAI/ML SystemsPythonGo

πŸ“œ Description

  • Design and implement scalable application services wrapping AI capabilities for enterprise use.
  • Lead deployment of inference servers using KServe, KubeRay, or Knative for AI workloads.
  • Build internal tooling and SDKs to enhance team agility and integrate Foundation Models into product features.
  • Architect robust Retrieval-Augmented Generation pipelines and prompt management services.
  • Collaborate with UI engineers, UX designers, and Product Management to ensure usability of the AI platform.
  • Ensure AI workloads are secure and optimized for GPU resource scheduling within Kubernetes.

πŸ› οΈ Requirements

  • Bachelor’s degree with 6+ years of software engineering experience, including 2+ years focused on AI/ML systems.
  • Expert proficiency in Python and strong competence in a systems language like Go or Rust/C++.
  • Deep understanding of LLM deployment challenges and runtimes.
  • Familiarity with quantization techniques to optimize model size/speed.
  • Experience building complex workflows using tools like LangChain or LlamaIndex.

✨ Benefits

  • Generous PTO Policy and support for work-life balance with Unplugged Days.
  • Flexible work-from-home policy.
  • Mental and physical wellness programs.
  • Phone and internet reimbursement program.
  • Access to continued career development.
Full job description

Join Cloudera's Anywhere Cloud team as a Staff Software Engineer to lead the architecture and delivery of a cloud-native AI platform, bridging AI research and production-grade environments.

Description

  • Design and implement scalable application services wrapping AI capabilities for enterprise use.
  • Lead deployment of inference servers using KServe, KubeRay, or Knative for AI workloads.
  • Build internal tooling and SDKs to enhance team agility and integrate Foundation Models into product features.
  • Architect robust Retrieval-Augmented Generation pipelines and prompt management services.
  • Collaborate with UI engineers, UX designers, and Product Management to ensure usability of the AI platform.
  • Ensure AI workloads are secure and optimized for GPU resource scheduling within Kubernetes.

Requirements

  • Bachelor’s degree with 6+ years of software engineering experience, including 2+ years focused on AI/ML systems.
  • Expert proficiency in Python and strong competence in a systems language like Go or Rust/C++.
  • Deep understanding of LLM deployment challenges and runtimes.
  • Familiarity with quantization techniques to optimize model size/speed.
  • Experience building complex workflows using tools like LangChain or LlamaIndex.

Benefits

  • Generous PTO Policy and support for work-life balance with Unplugged Days.
  • Flexible work-from-home policy.
  • Mental and physical wellness programs.
  • Phone and internet reimbursement program.
  • Access to continued career development.
Harvey

Staff Software Engineer, Model Infrastructure

HarveyπŸ‘₯ 10,000+ employees🏒 Software Development🀝 B2B
πŸ•’ yesterday

As a Staff Software Engineer on the Model Infrastructure team, you'll lead the design and development of systems that ensure high availability and operational excellence for AI inference at Harvey.

Software EngineeringDistributed SystemsCloud InfrastructureNetworking
Harvey

Senior Software Engineer, Model Infrastructure

HarveyπŸ‘₯ 10,000+ employees🏒 Software Development🀝 B2B
πŸ•’ yesterday

As a Staff Software Engineer on the Model Infrastructure team, you'll design and develop systems that ensure high availability and operational excellence for AI requests at Harvey.

Software EngineeringDistributed SystemsCloud InfrastructureNetworking

Trusted by Remote Workers