Staff Machine Learning Engineer
New
P
PrimerArtificial Intelligence
Location: Pasadena, California, United States; Remote; San Francisco, California, United States; Washington, District of Columbia, United StatesFull-TimeStaff
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 6+ years
- Required Skills
- PythonMachine LearningPyTorchNLPDistributed Systems
Requirements
- BS, MS, or PhD in computer science, a related field, or equivalent practical experience.
- 6+ years building production backend software with a track record of shipping and operating ML-driven functionality.
- Mastery of data structures and algorithms.
- Hands-on depth with LLMs and agentic systems (prompt and context engineering, tool use, retrieval and RAG).
- Experience with the broader ML toolkit such as PyTorch.
- Experience defining evaluation harnesses to measure and improve model quality.
- Fluency authoring production APIs in Python (Rust a plus).
- Experience with knowledge graphs, information retrieval at scale, or LLM fine-tuning and post-training (Bonus).
- High agency and a bias to action in ambiguous, fast-moving environments.
Responsibilities
- Set technical and product direction across LLM, agentic, and NLP systems.
- Design and build distributed, agentic systems at company-wide scale including retrieval-grounded reasoning.
- Develop pipelines for entity recognition, relation extraction, semantic search, and knowledge-graph generation.
- Package, deploy, and operate low-latency, high-concurrency inference systems using tools like Triton and vLLM.
- Build evaluation harnesses and labeled-data flywheels to measure and improve model performance.
- Partner with cross-functional teams to shape technical roadmaps and drive root cause analysis for operational issues.
- Raise the engineering bar by establishing better patterns and practices for the team.
View Full Description & ApplyYou'll be redirected to the employer's site