layiq
worthy; deserving; fitting; suitable.
A role, opportunity, or path that merits attention, time, and pursuit.
Loading LAYIQ…Job opportunity
Designworks Talent
Bellevue, Washington, United States
Source: Designworks Talent careers · View original posting
From Designworks Talent's posting. “We” and “our” refer to the employer.
Inference Engineer
Hybrid | Bellevue, WA (downtown)
Senior and Staff (multiple roles available)
Build the Inference Platform Powering Next-Generation AI Applications
A well-funded, rapidly growing AI infrastructure company is building a next-generation cloud platform designed to power the full lifecycle of artificial intelligence. The organization is developing a comprehensive AI infrastructure, platform, and services portfolio that supports the full spectrum of AI workloads—including large-scale compute, model training, fine-tuning, inference, and emerging agentic AI applications.
Backed by significant long-term investment, the company combines the speed, ownership, and innovation of a startup with the stability and resources of an established parent organization. Engineering teams are intentionally lean, highly collaborative, and AI-native, leveraging modern tooling and automation to build infrastructure capable of supporting the industry's most demanding AI workloads.
We're seeking
Inference Engineers to build and operate the model-serving systems behind a next-generation AI inference platform. This team focuses on delivering high-throughput, low-latency, reliable inference experiences that enable customers to consume advanced AI capabilities through production-scale APIs.
The Opportunity
This is a foundational engineering role focused on building the systems that bring AI models from research environments into reliable production services. You'll work on the infrastructure layer responsible for serving large models efficiently, optimizing performance, and ensuring reliability as usage scales.
You'll collaborate closely with GPU performance, AI training infrastructure, platform engineering, and operations teams to solve complex challenges around model serving, latency optimization, resource efficiency, and production reliability.
This opportunity is ideal for engineers who enjoy working at the intersection of distributed systems, machine learning infrastructure, GPU computing, and large-scale production systems.
Compensation
Competitive base pay for Bellevue market
Certain roles are eligible for additional rewards, including merit increases, annual bonus, and long term incentives. These awards are allocated based on individual performance
U.S. based employees have access to medical, dental, and vision insurance, a 401(k) plan and company match, employees also receive per calendar year, paid holidays.
Location
Hybrid role based in the Bellevue, WA area.
Approximately three days per week in the office.
Candidates elsewhere in the U.S. who are open to relocation are encouraged to apply.
U.S. work authorization is required. Visa sponsorship is not currently available.
Why Join?
Build the inference platform powering the next generation of AI applications.
Work directly on large-scale model serving, GPU optimization, and production AI systems.
Solve complex challenges around latency, throughput, reliability, and cost efficiency.
Join early enough to influence architecture, tooling, and engineering practices.
Collaborate with a highly experienced team building critical AI infrastructure from the ground up.
Enjoy the ownership and technical impact of a startup environment backed by significant long-term investment.
LAYIQ is an independent job-discovery service. This listing does not imply a partnership with or endorsement by the employer. Review the original posting for current details and availability.
Employer posted: