Skip to main content

Eloquent AI

AI Engineer, AIOps & Infrastructure

San FranciscofulltimemidAdded 1 month ago

About this role

Eloquent AI seeks a Senior Software Engineer to design and operate scalable infrastructure for deploying autonomous AI agents in production. You'll automate LLMOps and MLOps workflows, optimize GPU workloads, and ensure enterprise-grade reliability for large-scale AI systems across financial services.

What you'll do

  • Design and build scalable ML infrastructure for production AI agent deployment
  • Automate LLMOps and MLOps workflows including model training, fine-tuning, and monitoring
  • Optimize GPU and cloud compute workloads to improve efficiency and reduce latency
  • Develop Kubernetes-based solutions and custom operators for ML orchestration
  • Implement system observability, logging, and performance tracking for AI models
  • Participate in on-call rotations ensuring 24/7 reliability of critical systems

What they're looking for

  • Kubernetes and containerized workload management
  • Cloud platforms (AWS, GCP, Azure)
  • Python development for ML/AI services
  • ML model deployment and inference optimization
  • Distributed computing and system architecture
  • GPU workload optimization
  • Vector databases and RAG architectures
  • Monitoring, logging, and observability tools
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Eloquent AI

Eloquent AI builds autonomous AI agents and systems powered by large language models to handle complex enterprise workflows in financial services and insurance. The company is hiring front-end engineers, full-stack AI engineers, senior infrastructure engineers, and AI specialists to develop conversational interfaces, scalable backends, production deployment systems, and multimodal AI agents.

View all jobs at Eloquent AI

Likely interview questions

  • Walk us through your experience designing and operating Kubernetes clusters at scale. How have you handled GPU workload orchestration and resource optimization?
  • Describe a time you automated an MLOps or LLMOps workflow. What tools did you use, and what were the key challenges you solved?