Careers at LuminosAI

Build the standard for reliable AI agents.

We're solving the hardest problem in autonomous AI: turning unpredictable agent behavior, legal risk, and brand liability into rigorous, automated quality assurance.

How We Work

Rigorous evaluation. Frictionless QA.

We bring together ML researchers, legal scholars, and evaluation engineers who believe quality assurance for AI agents must be empirical, automated, and entirely handled for our customers.

01

Dynamic Trajectory Evaluation

We don't settle for static prompts, superficial vibe-checks, or HHH frameworks. We rigorously evaluate multi-step agent workflows, tool execution, and complex reasoning sequences to catch unintended behaviors and performance downsides before they reach production.

02

Tailored Eval Generation

AI creators should be focused on the problems their agents solve, not anticipating every way they might fail. The hardest part of agent QA isn't just writing the code for tests—it's figuring out which negative behaviors actually need testing in the first place. Instead of leaving our customers to guess at their blind spots, we build platforms that automatically uncover vulnerabilities and generate highly targeted evaluations.

03

Frictionless Quality Assurance

Shipping autonomous agents shouldn't require our customers to build an internal QA department from scratch. We build comprehensive evaluation infrastructure that integrates directly into their workflows, making sophisticated downside protection and regulatory compliance completely hands-off.

Join Us

Open Positions

Explore our current opportunities and find where you can make an immediate impact.

Sales•US Remote•Full-Time

Account Executive

Partner with enterprise and mid-market AI teams and compliance leaders to demonstrate the Luminos.AI platform, automate AI oversight, and drive revenue.

Legal Engineering & Applied AI•US Remote•Part-Time

Fellowship Program (Spring 2027)

Work directly with senior lawyers and engineers at the frontier of AI agent oversight, designing evaluation frameworks and testing generative systems.

Don't see your role? Say hello anyway.

We're always looking for exceptional engineers, researchers, and risk practitioners. Send us your background.