layiq
worthy; deserving; fitting; suitable.
A role, opportunity, or path that merits attention, time, and pursuit.
Loading LAYIQ…Job opportunity
United States
Source: OpenRouter careers · View original posting
From OpenRouter's posting. “We” and “our” refer to the employer.
OpenRouter is the leading AI routing and infrastructure layer that enterprises use to access, manage, and optimize the best large language models across providers—without lock-in, capacity constraints, or unnecessary cost. We power the most advanced AI teams in the world by giving them the flexibility to move fast, scale confidently, and stay future-proof as models evolve.
As enterprise adoption of AI accelerates, OpenRouter sits at the center of how organizations operationalize LLMs across research, product, and production workloads.
OpenRouter routes almost a billion requests and more than 20 trillion tokens a day, across 80+ providers and thousands of endpoints. Every one of those providers can degrade, rate-limit, change behavior, or go down without warning. Our customers count on us to absorb that chaos so their apps never notice.
We're hiring our first AI Inference SRE to own the operational health of our provider supply. You'll make sure every endpoint we route to is fast, correct, and available, and that we detect and route around problems before customers do. You'll sit on the Provider Operations team, reporting to the Provider Operations Manager.
4+ years in SRE, production engineering, or infrastructure roles running high-traffic, customer-facing systems.
Strong with observability tooling and practice: metrics, tracing, logs, SLOs/error budgets, alerting that is always actionable.
Capable software engineer who prefers writing tools over executing runbooks. TypeScript and/or Python.
Experienced with distributed systems failure modes: timeouts, retries, backpressure, partial outages, noisy neighbors.
Calm, clear incident commander who communicates well with external partners under pressure.
Understands, or is eager to learn deeply, how LLM inference is served: streaming, tool calling, prompt caching, throughput/latency tradeoffs, and how provider APIs differ.
Nice to Have
Experience at an inference provider, model lab, GPU cloud, or API gateway/CDN company.
Experience with our stack: TypeScript, Cloudflare Workers, Postgres, ClickHouse, GCP, Vercel.
Background in routing, load balancing, or traffic management systems.
Experience with evals or synthetic monitoring for ML systems.
LAYIQ is an independent job-discovery service. This listing does not imply a partnership with or endorsement by the employer. Review the original posting for current details and availability.
Employer posted:
Anduril · Aberdeen, Maryland, United States; Colorado Springs, Colorado, United States; San Antonio, Texas, United States; Tampa, Florida, United States
Anduril · Costa Mesa, California, United States; Huntsville, Alabama, United States; Irvine, California, United States; Reston, Virginia, United States; Seattle, Washing…
Anduril · Costa Mesa, California, United States; Huntsville, Alabama, United States; Irvine, California, United States; Reston, Virginia, United States; Seattle, Washing…
Blew & Associates, P.A. · Orlando, FL, US; Tampa, FL, US
General Dynamics Information Technology · USA LA Home Office (LAHOME)
NiSource · Merrillville IN-SLC