Staff Machine Learning Engineer

New
P
PrimerArtificial Intelligence
Location: Pasadena, California, United States; Remote; San Francisco, California, United States; Washington, District of Columbia, United StatesFull-TimeStaff
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Experience
6+ years
Required Skills
PythonMachine LearningPyTorchNLPDistributed Systems

Requirements

  • BS, MS, or PhD in computer science, a related field, or equivalent practical experience.
  • 6+ years building production backend software with a track record of shipping and operating ML-driven functionality.
  • Mastery of data structures and algorithms.
  • Hands-on depth with LLMs and agentic systems (prompt and context engineering, tool use, retrieval and RAG).
  • Experience with the broader ML toolkit such as PyTorch.
  • Experience defining evaluation harnesses to measure and improve model quality.
  • Fluency authoring production APIs in Python (Rust a plus).
  • Experience with knowledge graphs, information retrieval at scale, or LLM fine-tuning and post-training (Bonus).
  • High agency and a bias to action in ambiguous, fast-moving environments.

Responsibilities

  • Set technical and product direction across LLM, agentic, and NLP systems.
  • Design and build distributed, agentic systems at company-wide scale including retrieval-grounded reasoning.
  • Develop pipelines for entity recognition, relation extraction, semantic search, and knowledge-graph generation.
  • Package, deploy, and operate low-latency, high-concurrency inference systems using tools like Triton and vLLM.
  • Build evaluation harnesses and labeled-data flywheels to measure and improve model performance.
  • Partner with cross-functional teams to shape technical roadmaps and drive root cause analysis for operational issues.
  • Raise the engineering bar by establishing better patterns and practices for the team.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now