Suggested rewrite: Led a cross-functional initiative that improved [business outcome] by [measurable result], demonstrating experience relevant to this role...
Senior Staff Software Engineer, Serving
Review the role, location requirements, compensation details, and application process before deciding whether this opportunity fits your next career move.
- Remote from
- USA
- Salary
- USD 219k–310k / yr
- Department
- Software Engineering
- Employment
- Full Time
- Experience
- Senior
- Published
- Apply before
- 30 Oct 2026
- Listing views
- 22
- Application actions
- 2
Make your next move.
Prepare your resume, explore your fit, and draft a cover letter for this opportunity.
The role, at a glance.
This is a senior staff-level software engineering role on Liftoff’s Accelerate Serving team, responsible for mission-critical real-time ad-serving infrastructure. The engineer will build and operate high-throughput, low-latency bid-request and bidding services while improving scalability, reliability, observability, and capacity planning. A major focus is GPU-powered neural-network inference using TensorRT, including profiling bottlenecks in batching, memory, networking, serialization, and GPU utilization. The role also partners closely with machine learning engineers to productionize models, operate low-latency feature stores, and improve experimentation and benchmarking systems.
Role DNA
A quick view of the complexity, pace, ownership and collaboration implied by the job description.
Pace & Pressure
5/5Autonomy Level
5/5Communication Load
4/5Salary analysis
Estimated compensation compared with the broader US market for similar roles.
Core skills
Skills and capabilities most closely associated with this opportunity.
Sample interview questions
I would first segment latency by request type, model version, host, and deployment window, then use distributed tracing and service metrics to isolate time spent in queueing, serialization, network calls, feature retrieval, batching, GPU execution, and response handling. I would compare the affected path with a known-good baseline, reproduce the issue under representative load, and validate a fix with canary deployment and latency-percentile guardrails.
I would balance GPU utilization and throughput against p95/p99 request latency, queue wait time, model shape constraints, and traffic variability. The design should use bounded batching windows, enforce per-request deadlines, and expose controls for batch size and concurrency so the system can adapt without violating exchange latency requirements.
I would prioritize predictable low-latency reads, horizontal scalability, replication, and graceful degradation. The design would include local or regional caching, clear data freshness guarantees, resilient fallback behavior for missing features, capacity headroom, and comprehensive monitoring for availability, tail latency, staleness, and hot-key behavior.
I would establish offline quality and performance benchmarks, validate serving compatibility, and run shadow traffic or replay tests before exposing the model to production decisions. I would then use a staged rollout with experiment cohorts, monitor business metrics alongside latency, error rate, GPU utilization, and resource cost, and retain an immediate rollback path.
I begin by aligning stakeholders on the problem, measurable success criteria, risks, and system constraints. I create an architecture and rollout plan, break work into independently deliverable milestones, communicate trade-offs clearly, and ensure operational ownership through testing, dashboards, runbooks, incident readiness, and post-launch optimization.
About this role.
Liftoff is a leading AI-powered performance marketing platform for the mobile app economy. Our end-to-end technology stack helps app marketers acquire and retain high-value users, while enabling publishers to maximize revenue across programmatic and direct demand.
Liftoff’s solutions, including Accelerate, Direct, Monetize, Intelligence, and Vungle Exchange, support over 6,600 mobile businesses across 74 countries in sectors such as gaming, social, finance, ecommerce, and entertainment. Founded in 2012 and headquartered in Redwood City, CA, Liftoff has a diverse, global presence.
At Liftoff, we’re solving one of the core problems faced by every mobile app: growth. To do so, we build Machine Learning and Big Data-driven technology that can accurately predict which apps a user will like, and connect them in a compelling way. Our systems operate at a scale unseen outside of the largest Internet companies — processing over 12 million requests per second and interacting with over 5 billion unique users each day.
The Serving team builds and maintains the mission-critical infrastructure responsible for serving ads across Liftoff’s product line.
As an engineer on Liftoff’s Accelerate Serving team, you will:
- Design, build, and operate the high-throughput, low-latency services that receive bid requests, execute Liftoff’s bidding logic, and respond to ad exchanges in real time.
- Improve the performance, scalability, and reliability of systems that process millions of requests per second under strict latency constraints.
- Develop and optimize GPU-powered inference services that execute neural network models using TensorRT.
- Profile end-to-end inference pipelines to identify and eliminate bottlenecks in GPU utilization, batching, memory access, serialization, networking, and request handling.
- Partner with machine learning engineers to productionize new model architectures and features, translating modeling requirements into efficient and reliable serving implementations.
- Design and operate large feature store fleets that provide low-latency access to real-time and precomputed features.
- Develop benchmarking and performance-analysis tools that help engineers compare models, understand latency and throughput trade-offs, and identify regressions before deployment.
- Improve experimentation tooling that allows machine learning teams to orchestrate offline training runs, evaluate candidate models, and maintain model leaderboards.
- Own changes through their full lifecycle—from system design and implementation to testing, deployment, observability, capacity planning, incident response, and continued optimization.
Requirements:
- Strong core Computer Science fundamentals (data structures, algorithms, system architecture)
- 10+ years of industry experience
- M.S.. or higher in Computer Science (or equivalent work experience)
- Experience with Go is a plus
Location:
This role is eligible for full-time remote work in one of our entities: CA, CO, ID, IL, FL, GA, MA, MI, MN, MO, NJ, NV, NY, OR, PA, TX, UT, and WA
We are a remote-first company with US hubs in Redwood City, Los Angeles, and New York City.
Travel Expectations:
We offer several opportunities for in-person team gatherings, including but not limited to project meetings, regional meetups, and company-wide events. We expect our employees to attend these gatherings at least once per quarter. These gatherings provide essential opportunities for collaboration, communication, and team building.
Compensation:
Liftoff offers all employees a full compensation package that includes equity and health/vision/dental benefits associated with your country of residence. Base compensation will vary based on the candidate’s location and experience.
The following are our base salary ranges for this role:
- SF Bay Area, NYC, Los Angeles/Orange County: $255,000 – $310,000
- Seattle/Olympia, Austin, San Diego, Santa Barbara, Boston: $234,000 – $285,000
- All other cities and towns in our approved states: $219,000 – $266,000
#LI-EL1
Liftoff offers a fast-paced, collaborative, and innovative work environment where employees are empowered to grow and make an impact. We’re shaping the future of the mobile app ecosystem—join us and help accelerate what’s next.
Liftoff’s compensation strategy includes competitive salaries, equity, and benefits designed to support employee well-being and performance. We benchmark compensation based on role, level, and location to ensure fairness and market alignment. Benefits may include medical coverage, wellness stipends, and additional perks based on your country of residence.
Liftoff is an equal opportunity employer. We are committed to creating an inclusive environment for all employees and applicants regardless of race, ethnicity, national origin, age, marital status, disability, sexual orientation, gender identity, religion, veteran status, or any other characteristic protected by applicable law.
Agency and Third Party Recruiter Notice:
Liftoff does not accept unsolicited resumes from individual recruiters or third-party recruiting agencies in response to job postings. No fee will be paid to third parties who submit unsolicited candidates directly to our hiring managers or Recruiting Team. All candidates must be submitted via our Applicant Tracking System by approved Liftoff vendors who have been expressly requested to make a submission by our Recruiting Team for a specific job opening. No placement fees will be paid to any firm unless such a request has been made by the Liftoff Recruiting Team and such a candidate was submitted to the Liftoff Recruiting Team via our Applicant Tracking System.
This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.
Apply now.
Follow the employer’s application method and review Jobicy’s safety guidance before sharing personal information.
Continue on the employer website
Protect your personal information and never pay to secure an interview or job offer. .
