AI Inference Engineer
New
E
EvergridArtificial Intelligence
Remote (US-based)Full-TimeMiddle
SalaryCompetitive base salary, Performance-based commission, Equity
Apply NowOpens the employer's application page
Job Details
- Experience
- 1+ years of professional experience
- Required Skills
- PythonSoftware Development
Requirements
- Bachelor’s, Master’s, or PhD in Computer Science, Engineering, Mathematics, or a related field (or equivalent practical experience).
- 1+ years of professional experience in a fast-paced engineering environment.
- Experience with SGlang, vLLM, and other inference engines & schedulers.
- Strong experience writing production-level software, with a preference for Python.
- Familiarity with AI/ML pipelines and the lifecycle of model development, deployment, and monitoring.
- Strong communication skills, particularly when discussing complex technical topics.
- Experience building, deploying, or optimizing AI/ML systems is highly valued.
- Comfort working directly with customers and owning outcomes end-to-end.
Responsibilities
- Design, build, and maintain production-grade software systems and inference services, with a strong emphasis on Python.
- Own customer engagements end-to-end: problem framing, evaluation, proof-of-concept, production deployment, and monitoring.
- Collaborate directly with customer engineering teams across sales, implementation, and expansion phases.
- Turn ambiguous objectives into clear technical specifications and well-scoped PoCs.
- Optimize AI/ML inference pipelines for latency, throughput, reliability, and cost.
- Contribute improvements to Evergrid's inference stack, tooling, and platform capabilities.
- Operate with high ownership, acting as engineer, technical lead, and execution driver for customer-facing initiatives.
View Full Description & ApplyYou'll be redirected to the employer's site