layiq
worthy; deserving; fitting; suitable.
A role, opportunity, or path that merits attention, time, and pursuit.
Loading LAYIQ…Job opportunity
Exiger
Jersey City, New Jersey, United States; McLean, Virginia, United States; Richmond, Virginia, United States
Source: Exiger careers · View original posting
From Exiger's posting. “We” and “our” refer to the employer.
Exiger is the AI partner for supply chain and procurement automation. Its centralized
1EXIGER.AI platform allows organizations to manage their entire operating network, from parts to suppliers, regulators, and customers, through an autonomous agentic workforce that learns, adapts, and augments with every decision. The AI-native platform, combined with the largest supply chain knowledge asset, empowers 550+ global customers, including 150 Fortune 500 and 80+ government and defense industrial organizations, to act first.
Exiger is FedRAMP® authorized and a 2x Leader in Gartner® Magic Quadrant™ for Supplier Risk Management.
Site Reliability Engineer
Location: U.S. (Hybrid)
This role requires U.S. citizenship and eligibility for a U.S. security clearance.
Exiger is transforming how governments and global enterprises manage supply chain, defense, and geopolitical risk. Our AI-powered platform equips the world's most important institutions with the intelligence they need to protect critical infrastructure, secure national interests, and make data-driven operational decisions.
From identifying counterfeit parts in defense supply chains to anticipating geopolitical risk exposure, Exiger enables mission owners to act with clarity and confidence in complex, high-stakes environments.
This is our first dedicated Site Reliability Engineering hire and a founding role. You will help stand up the SRE function at Exiger: setting the standards, tooling, and practices that keep 1Exiger reliable for our 550+ customers, including Fortune 500 companies and U.S. government agencies. You will own reliability across the full service lifecycle, from design and capacity planning through deployment, monitoring, and incident response, and build the automation that lets the platform scale without scaling headcount.
Because you are first, we need someone who has practiced SRE before and can bring the playbook, not learn it on the job.
You will use your expertise in coding, algorithms, complexity analysis, and large-scale distributed system design to solve the reliability challenges that are unique to operating a mission-critical AI platform in regulated and government environments.
SRE's culture of intellectual curiosity, problem solving and openness is key to its success. Our organization brings together people with a wide variety of backgrounds, experiences and perspectives. We encourage them to collaborate, think big and take risks in a blame-free environment. We promote self-direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow.
Bachelor's or Master's degree in Computer Science, a related field, or equivalent practical experience.
6 years of experience in software or systems engineering, including at least 4 years in a dedicated Site Reliability Engineering, production engineering, or platform reliability role. As our first SRE hire, you must have practiced SRE before and be ready to establish the function.
4 years of experience designing, analyzing, and troubleshooting large-scale distributed systems.
Strong grounding in Unix/Linux internals (filesystems, processes, system calls) and networking fundamentals (TCP/IP, DNS, routing, load balancing).
Hands-on experience establishing core SRE practices from the ground up: SLIs, SLOs, and error budgets, monitoring and observability, capacity planning, and automation that removes repetitive manual work.
A rigorous, empirical mindset: you form hypotheses, measure outcomes, and make metrics-driven decisions rather than relying on intuition or anecdote.
Experience with chaos engineering or fault-injection testing (for example game days, Chaos Monkey, Gremlin, or LitmusChaos) to validate system resilience.
Proven incident management experience: on-call ownership, leading response under pressure, and driving blameless postmortems to root cause.
Experience in troubleshooting and supporting applications like web services, data storage, databases, and data pipelines, with Linux/Unix or other operating systems.
Familiarity with cloud platforms (AWS) and secure system integration.
Comfort integrating AI coding assistants (such as Claude and Codex) into your daily engineering workflow.
Ability to translate ambiguous mission problems into structured technical solutions.
Ability to operate independently in dynamic, high-stakes environments.
Willingness to travel as needed to support customer engagements.
4 years of experience programming in Go or C (Java also welcome), with the ability to debug, optimize, and automate rather than just script.
Experience supporting ML or data platforms in production.
Familiarity with data warehouses such as Snowflake, Redshift and/or Apache Iceberg.
Experience operating in FedRAMP or other regulated or government environments.
High-performance culture rooted in accountability, collaboration, and a shared commitment to excellence.
®
Magic Quadrant™ for Supplier Risk Management, twice selected as one of Fast Company's 'Brands That Matter,' and recipient of the Third Party Risk Association's Innovator Award, Exiger's technology has been recognized by leading analyst evaluations and 50+ awards. Learn more at
Exiger.com and follow Exiger on LinkedIn.
At Exiger, our values define how we work and why we lead. We are mission-inspired, imagination-driven, trust-anchored, and compassion-focused—committed to building technology that makes the world safer, more transparent, and more resilient.
All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability or protected veteran status, or any other legally protected basis, in accordance with applicable law.
Exiger’s hybrid work policy is periodically reviewed and adjusted to align with evolving business needs.
LAYIQ is an independent job-discovery service. This listing does not imply a partnership with or endorsement by the employer. Review the original posting for current details and availability.
Employer posted: