Senior AI Engineer
New
T
TMTGSocial Media
Remote — US onlyFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- PythonMachine LearningGoLLM
Requirements
- 5+ years of backend engineering experience in a statically typed language (Go, Java, Kotlin, C#).
- Proven experience shipping production LLM-backed features (RAG, streaming, tool use).
- Hands-on experience with LLM serving stacks including routing, failover, token streaming, and cost metering.
- Expertise in retrieval systems, vector search, and embedding pipelines.
- Strong API design instincts and ability to own services from schema to deployment.
- Pragmatic approach to AI evaluation and monitoring failure modes.
- Familiarity with Go (strongly preferred) and Python.
- Experience with LLM gateways or serving infrastructure (e.g., LiteLLM, vLLM, TGI).
- Knowledge of vector databases (e.g., Qdrant, pgvector).
- Experience with fine-tuning open-weight models (LoRA) and related eval disciplines.
- Knowledge of content moderation or safety tooling.
Responsibilities
- Architect and build core AI infrastructure for a platform serving millions of users.
- Implement LLM-backed features including retrieval-augmented generation (RAG), streaming responses, and tool usage.
- Develop systems for model routing, failover, token streaming, and cost/usage metering.
- Build and maintain retrieval systems including vector search, embedding pipelines, and context assembly.
- Design and implement instrumentation for monitoring failure modes like hallucinations and provider outages.
- Define evaluation processes to measure AI answer quality, relevance, and safety using repeatable tests.
- Own services end-to-end, from API design and schema to deployment and on-call support.
View Full Description & ApplyYou'll be redirected to the employer's site