Skip to main content

Bjak

Applied AI Engineer

United States (Remote)fulltimemidAdded 1 month ago

About this role

A1 is seeking an Applied AI Engineer to build production-ready AI features for a smart assistant platform. You'll work across the full stack—from model optimization to system design to user experience—ensuring AI workflows are reliable, fast, and actually work in the real world.

What you'll do

  • Design and iterate on prompts, tools, memory, and agent workflows to shape model behavior
  • Build end-to-end AI features from model selection through production deployment
  • Debug issues across model, orchestration, infrastructure, and UX layers
  • Create lightweight evaluation frameworks to measure real-world performance
  • Optimize for latency, cost, and production reliability
  • Collaborate with product and engineering to translate ambiguous problems into working systems

What they're looking for

  • Python and production-quality code
  • PyTorch or JAX
  • LLM APIs and model fine-tuning experience
  • ML inference and serving (e.g., vLLM)
  • Vector databases
  • Machine learning fundamentals and neural networks
  • Full-stack debugging across abstraction layers
  • Problem-solving in ambiguous, fast-moving environments
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Bjak

Bjak is a Southeast Asian fintech super app offering insurance, payments, savings, wallets, and investment products through a unified platform. The company is hiring full stack engineers, backend engineers, iOS developers, and Android engineers to build scalable features and reliable systems across mobile and web products.

Website
bjak.com
View all jobs at Bjak

Likely interview questions

  • Walk us through a time you deployed an ML model to production. What went wrong, and how did you debug it across the stack?
  • Describe your experience with prompt engineering or fine-tuning LLMs. How did you measure whether your changes actually improved real-world performance?