Senior AI Engineer

New
T
TMTGSocial Media
Remote — US onlyFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Experience
5+ years
Required Skills
PythonMachine LearningGoLLM

Requirements

  • 5+ years of backend engineering experience in a statically typed language (Go, Java, Kotlin, C#).
  • Proven experience shipping production LLM-backed features (RAG, streaming, tool use).
  • Hands-on experience with LLM serving stacks including routing, failover, token streaming, and cost metering.
  • Expertise in retrieval systems, vector search, and embedding pipelines.
  • Strong API design instincts and ability to own services from schema to deployment.
  • Pragmatic approach to AI evaluation and monitoring failure modes.
  • Familiarity with Go (strongly preferred) and Python.
  • Experience with LLM gateways or serving infrastructure (e.g., LiteLLM, vLLM, TGI).
  • Knowledge of vector databases (e.g., Qdrant, pgvector).
  • Experience with fine-tuning open-weight models (LoRA) and related eval disciplines.
  • Knowledge of content moderation or safety tooling.

Responsibilities

  • Architect and build core AI infrastructure for a platform serving millions of users.
  • Implement LLM-backed features including retrieval-augmented generation (RAG), streaming responses, and tool usage.
  • Develop systems for model routing, failover, token streaming, and cost/usage metering.
  • Build and maintain retrieval systems including vector search, embedding pipelines, and context assembly.
  • Design and implement instrumentation for monitoring failure modes like hallucinations and provider outages.
  • Define evaluation processes to measure AI answer quality, relevance, and safety using repeatable tests.
  • Own services end-to-end, from API design and schema to deployment and on-call support.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now