Senior Staff Software Engineer, Serving

New
J
JobgetherAdTech Infrastructure
Full-time remote work available across approved US statesFull-TimeStaff
SalaryBase salary of $255,000–$310,000 for the San Francisco Bay Area, New York City, and Los Angeles/Orange County; $234,000–$285,000 for Seattle/Olympia, Austin, San Diego, Santa Barbara, and Boston; $219,000–$266,000 for other cities and towns within approved states.
Apply NowOpens the employer's application page

Job Details

Experience
10+ years
Required Skills
GoDistributed Systems

Requirements

  • 10+ years of professional software engineering experience, with a strong track record working on complex, large-scale systems.
  • Strong Computer Science fundamentals, including data structures, algorithms, distributed systems, and system architecture.
  • M.S. or higher in Computer Science or a related field, or equivalent professional experience.
  • Strong ability to design and operate highly scalable, reliable, and performance-sensitive production systems.
  • Experience working with machine learning infrastructure, model serving, inference systems, or other high-performance computing environments is highly valuable.
  • Strong analytical and performance-engineering skills, with the ability to identify bottlenecks and optimize systems across multiple layers of the stack.
  • Ability to collaborate effectively with machine learning engineers and other technical teams to translate research and modeling requirements into production-ready systems.
  • Experience with Go is a plus.
  • Comfortable taking ownership of technically complex projects from architecture and implementation through production operations and continuous improvement.

Responsibilities

  • Design, build, and operate high-throughput, low-latency services that process bid requests, execute real-time bidding logic, and respond to ad exchanges.
  • Improve the performance, scalability, and reliability of systems handling millions of requests per second while operating within strict latency requirements.
  • Develop and optimize GPU-powered inference services using TensorRT to efficiently execute neural network models in production.
  • Profile complete inference pipelines and identify opportunities to optimize GPU utilization, batching, memory access, serialization, networking, and request handling.
  • Partner with machine learning engineers to productionize new model architectures and features, translating modeling requirements into efficient, reliable serving implementations.
  • Design and operate large-scale feature store fleets providing low-latency access to real-time and precomputed features.
  • Build benchmarking and performance-analysis tools that enable engineers to compare models, evaluate latency and throughput trade-offs, and detect performance regressions before deployment.
  • Improve experimentation infrastructure that supports offline training runs, candidate model evaluation, and model leaderboards.
  • Own engineering changes throughout their complete lifecycle, including system design, implementation, testing, deployment, observability, capacity planning, incident response, and ongoing optimization.
View Full Description & ApplyYou'll be redirected to the employer's site
Base salary of $255,000–$310,000 for the San Francisco Bay Area, New York City, and Los Angeles/Orange County; $234,000–$285,000 for Seattle/Olympia, Austin, San Diego, Santa Barbara, and Boston; $219,000–$266,000 for other cities and towns within approved states.
Apply Now