A precise cable-stayed bridge extending across still water towards the horizon.

Confidence, built on evidence.

Restore faithin your work.

Independent, research-backed evaluation for AI used in consequential work.
Built around your model, your use case, and the decision you need to stand behind.

A bridge between capability and confidence.

Your model.
Your questions.

Every investigation starts with the work your AI needs to do. We design the tests around your system, industry and operating conditions.

Your AI system

Oneirix

Evidence for your decision

  • Original research & proprietary benchmarks
  • Bespoke testing & verification
  • Internal computation & failure analysis
The subject

Your AI
system

Oneirix
  • Original research & proprietary benchmarks
  • Bespoke testing & verification
  • Internal computation & failure analysis
The outcome

Evidence for
your decision

Bespoke testing.
Better decisions.

Buying a company, developing a model, choosing a system or investigating a failure. We shape the evaluation around what you need to know.

Your systemYour use caseYour domain

  1. 01

    Scope

    Define the decision, domain and model access.

  2. 02

    Test

    Build an evaluation around your real use case.

  3. 03

    Investigate

    Examine behaviour, failures and their possible causes.

  4. 04

    Report

    Explain the findings, limits and next steps.

Evidence for
your decision.

Our research informs tailored tests and our own benchmarks. Where model access allows, we measure internal computation during training and inference to investigate behaviour beyond outputs.

An independent report connects the findings to your decision, with clear recommendations and limits on what the evidence establishes.

Oneirix

Evaluation report

Illustrative contents
  1. 01System, use case and domain
  2. 02Questions and evaluation scope
  3. 03Methods, benchmarks and access
  4. 04Findings and failure evidence
  5. 05Limitations and open questions
  6. 06Recommendations and next steps
Explore what we offer

Built in the lab.
Tested in the world.

Oneirix's local research workstation, showing its GPU, cooling system and model evaluation hardware.

Local R&D

Model analysis and evaluation infrastructure.

Delta Road research equipment in a vehicle: a laptop running GNSS diagnostics alongside a connected receiver board.

Field research R&D in progress

Real-world vision and sensor evaluation.

Hands-on technical R&D, from model behaviour to real deployment environments.

Research that
gets to work.

01 / Discover

Research

Discover failure modes conventional evaluation misses.

02 / Apply

Evaluation

Design a bespoke investigation around your model, use case and decision.

03 / Re-evaluate

Verification

Test specific claims and revisit performance as models and conditions change.

Original methods. Applied to your questions.

Our research and evaluation tools provide the foundation. Each engagement applies them to a specific system and context, producing evidence you can use and methods we can continue to develop.

A compounding research cycleA repeating research cycle: new failure modes lead to new tests, a proprietary failure corpus, better assurance, and further discovery.NEW FAILUREMODESNEW TESTSPROPRIETARYFAILURE CORPUSBETTERASSURANCE
Rohan Maskrey, founder of Oneirix.

Founder.

Rohan Maskrey

Founder, Oneirix

Independent AI researcher

Research focused on failure dynamics, recurrent computation and AI assurance.

Raising pre-seed funding.

Work you can
stand behind.

Tell us what you need to understand about your model or AI workflow. We'll shape an independent investigation around your questions.

Work with Oneirix