OUR SECTORS

At USA Tech Recruit, our sectors cover a wide range of industries within the field of technology.

Submit vacancy
Looking for tech jobs in Europe?
Visit European Tech Recruit
Looking for tech jobs globally?
Visit Tech Recruit

Job search

Our sectors

Client services

About us

Looking for tech jobs in Europe?
Visit European Tech Recruit
Looking for tech jobs globally?
Visit Tech Recruit

Client services

At European Recruitment, our sectors cover a wide
range of industries within the field of technology

Submit Vacancy

About us

At European Recruitment, our sectors cover a wide
range of industries within the field of technology

Submit Vacancy

Client services

Learn about what client services we offer at USA Tech Recruit and browse though our success stories.

Submit vacancy
Looking for tech jobs in Europe?
Visit European Tech Recruit
Looking for tech jobs globally?
Visit Tech Recruit
Looking for tech jobs in Europe?
Visit Tech Recruit
Looking for tech jobs globally?
Visit Tech Recruit

Our Sectors

At European Recruitment, our sectors cover a wide range of industries within the field of technology

Submit Vacancy

About us

Learn more about USA Tech Recruit's story, mission and values, meet our team, and read about our commitment to DE&I.

Submit vacancy
Looking for tech jobs in Europe?
Visit European Tech Recruit
Looking for tech jobs globally?
Visit Tech Recruit
>
Looking for tech jobs in Europe?
Visit European Tech Recruit
Looking for tech jobs globally?
Visit Tech Recruit

Our Sectors

At European Recruitment, our sectors cover a wide range of industries within the field of technology

Submit Vacancy

Backend Engineer, AI (Agent Systems)

Recruitment Consultant
Henry Paget
Posted
28 days ago

Role

As a Backend Engineer, AI, you own the inference and orchestration layer that powers every AI interaction in the product. Your work sits between models and users, where latency, correctness, reliability, and cost directly impact real-world experience.

You will build and operate production systems that turn model capability into fast, stable, observable APIs used across mobile and desktop clients.

 

Focus

  • Build and operate backend systems that serve AI-powered features in production.

  • Design inference pipelines, orchestration layers, and service boundaries around models.

  • Own production concerns: monitoring, logging, alerting, and incident response.

  • Optimize latency and throughput across inference, caching, batching, and streaming.

 

Ideal Experiences

  • Strong backend engineering fundamentals in production environments.

  • Experience running high-throughput, low-latency services.

  • Familiarity with AI inference patterns (LLMs, embeddings, multimodal).

  • Comfortable debugging distributed systems under load.

  • Bias toward shipping and learning from production behavior.

 

Outcomes

  • Backend systems run reliably at scale, handling production AI traffic with low latency and high throughput.

  • APIs are stable, clear, and support seamless integration with frontend and ML systems.

  • Production incidents are quickly detected, diagnosed, and resolved, minimizing user impact.

  • Iterative improvements based on real usage continuously increase system performance and reliability.

 

Tech Stack

  • Python

  • NodeJs

  • Pytorch

  • OpenAI / Anthropic / open-source LLMs

  • SQl & noSQL

  • Kubernetes

  • Docker

 

 

Industry
Contract Type
Permanent
Location
United States
City
San Francisco
Work Model
remote

Apply Now

By applying to this role, you acknowledge that we may collect, store, and process your personal data on our systems.

For more information, please refer to our
Privacy Notice

    Name
    Email
    Phone
    Location
    Message

    Upload CV:

    Choose file

    Formats: Word, PDF (max. size: 20MB)

    Subscribe for industry highlights.

    Send Application

     

    Other relevant jobs

    Submit CV
    Submit Vacancy
    Cookie Settings
    We use cookies to enhance your experience and analyze site traffic and movements. Read our cookie policy here.