Research & benchmarks

Inspect the evidence behind the algorithms.

Controlled figures, benchmark methodology, model boundaries, and reproducible product visuals for technical evaluation, not customer field-performance claims.

BenchmarksMethodologyClaim boundaries
Registered lidar point cloud of a San Francisco intersection
Real lidar point cloud · Daniel L. Lu · CC BY 4.0Source + license ↗
EvidenceControlledClearly labeled
FiguresReproducibleSource and model lineage
TiersPulse → ApexRuntime through forensic depth
BoundaryNo field claimUntil deployment-specific validation
01 / INTERACTIVE DIAGRAM

Benchmark snapshot

Controlled detection metrics — not field performance.

NADIR VISUAL · SEED DATAILLUSTRATIVE · NOT FIELD PERFORMANCE
Nominal Caution Critical

Illustrative seed-data visualization. Validate on your representative fleet sample before operational decisions.

02 / DETAILS

Two ways to read this page

Plain language for fleet and investor conversations. Technical detail for engineering diligence.

Plain-language summary

NADIR helps fleets detect likely ADAS miscalibration after repair using data you already export.

Shadow mode means advisory tiers and signed records — not automatic vehicle changes.

  • Start with a four-week $750 proof pilot.
  • Success = MTTFF-RE within seven days on most repair events.
  • Expand only after a written pass/fail readout.
03 / EVIDENCE STANDARD

Controlled benchmarks, not customer logos.

Figures on this site illustrate detector behavior on seed and simulation data.

Mobileye overlay study compares residual scoring against reference paths under documented scenarios.

Do not extrapolate figures to field accuracy without your own pilot readout.

  • BenchmarkMobileye overlay scorecard
  • FiguresDetection survival · FPR strata
  • ReproPublic paper + JSON scorecard
04 / RESEARCH

Controlled benchmarks and reproducibility

Figures illustrate methodology — not field performance guarantees.

Compares residual scoring against reference paths under windshield and yaw perturbations.

Public scorecard JSON enables third-party reproduction of charts.

05 / DETECTION SCIENCE

Cross-modal residual scoring

How Pulse turns telematics into tiers after repair.

A green ADAS dashboard means self-check passed — not that alignment matches factory spec.

When camera lane proxy and radar range-rate disagree for days after glass work, that pattern elevates tier.

Single spikes from weather or harsh braking decay; persistence drives CAUTION and CRITICAL.

06 / GOVERNANCE

Shadow mode and product boundaries

Every public NADIR deployment starts advisory. These boundaries protect fleet legal teams and set honest expectations for insurers.

Shadow mode means NADIR scores telematics, assigns NOMINAL / CAUTION / CRITICAL tiers, and exports signed evidence — without writing calibration parameters to the vehicle.

Your maintenance team decides whether to pull a unit, schedule a bay recalibration, or return it to route. NADIR does not override OEM safety controllers.

This is intentional: post-repair detection pilots should prove alert quality before anyone discusses automation or correction paths.

  • Console CRITICAL rows are recommendations, not work orders.
  • Weekly digest emails summarize CAUTION trends without paging on-call at 2 a.m.
  • Sample EVBs document what the model saw — useful for QA, not a legal verdict.
Actuation
None
Default mode
Shadow
ECU writes
Not supported in pilots
07 / PILOT KPI

MTTFF-RE — the one pilot metric

Mean Time To First Flag after Repair Event. One number for pass, fail, or stop.

When a vehicle leaves a glass or body shop, start a clock at repair order close.

MTTFF-RE measures hours until NADIR assigns CAUTION or CRITICAL in the post-repair observation window.

If most events flag within seven days, your maintenance team gets actionable lead time before silent misalignment becomes a safety conversation.

  • Median MTTFF-RE ≤ 168 hours is the default pilot target.
  • Pass rate ≥ 70% of in-cohort events is the companion statistic.
  • Events without telemetry are excluded — never counted as passes.
08 / REFERENCE TABLE

Reference matrix — research

Dense field reference for engineering diligence on this topic.

CategoryField / controlValueNotes
Pilotmodeshadow_advisoryNo ECU writes
PilotkpiMTTFF-REMedian ≤168h target
Pilotcohort50–150 VINsShadow Proof SKU
IngesttelemetryPOST /v1/telemetry/batchCSV/JSONL
IngestrepairPOST /v1/repair-eventsRO anchor
ScoretierNOMINAL|CAUTION|CRITICALPolicy-bound
Scoremodelpulse_residual_*Versioned
ProductqueueConsole CRITICAL rowsOperator-facing
ProductdigestWeekly CSV/emailCAUTION summary
EvidencesignatureHMAC-SHA256Integrity only
EvidenceverifyPOST /v1/evidence/verifyRecompute digest
SecuritytransitTLS 1.2+API only
SecurityauthBearer + API keyTenant scoped
BoundaryasilNot claimedAdvisory analytics
BoundaryactuationNoneShadow default
Commercialproof$750 flat4 weeks
Commercialops$750/moAfter pass
Commercialregional$15k+150–500 VINs
09 / GLOSSARY

Shared vocabulary

Terms used across NADIR pilots, API docs, and readouts.

MTTFF-RE

Mean Time To First Flag after Repair Event — hours from RO close to first CAUTION+ tier.

RO closed Monday 6pm, first CAUTION Wednesday 2pm → MTTFF-RE = 44 hours.

Shadow mode

Advisory scoring and evidence export without ECU writes or automated dispatch.

Console shows CRITICAL; maintenance decides whether to pull the unit.

EVB

Evidence Vehicle Bundle — canonical signed JSON record of tier transition lineage.

Includes model_id, policy_id, frame hashes, repair_event_id.

Pulse tier

Low-latency residual scoring tier used in public shadow pilots.

Batch scoring within minutes of telematics upload.

Cross-modal residual

Disagreement metric between sensor modalities (e.g., camera lane vs radar range-rate).

Persistent elevation after glass repair triggers CAUTION.

RO close

Timestamp when repair order is marked complete — starts post-repair observation window.

Must be UTC with timezone documented in data contract.

Policy ID

Version identifier for tier threshold rules separate from model weights.

policy_post_repair_glass_v3.2

Quarantine batch

Ingest batch rejected for schema errors — not silently partially scored.

Clock regression rows excluded with explicit error report.

Exhibit A

Week-four pilot chart plotting MTTFF-RE per repair event.

Median line vs 168h threshold.

Coverage exclusion

Repair event omitted from KPI due to missing telematics — not counted as pass.

Listed in readout appendix with reason no_telemetry.

Reason codes

Machine-readable tags explaining why tier was assigned.

CAMERA_RADAR_DIVERGENCE, POST_REPAIR_PERSISTENCE.

Adapter manifest

Versioned field mapping from fleet telematics export to NADIR schema.

geotab_pulse_v1.4.json
10 / DETAILS

Common questions

Straight answers for buyers and reviewers.

Common questions

Straight answers for buyers and reviewers.

12 / METHODOLOGY

A benchmark is only useful when lineage is complete.

NADIR separates source scenario, detector configuration, threshold policy, metric computation, and final interpretation.

NADIR / LAYERSSENSOR
RELIABILITY
SYSTEM
COMPOSABLE · VERSIONED · TRACEABLE
01
SOURCE

Scenario

Dataset, synthetic perturbation, or Calibration Lab event.

02
CONFIG

Manifest

Sensor expectations, segment, tier, and model version.

03
EXECUTE

Run

Residuals, uncertainty, gates, and alerts.

04
MEASURE

Metrics

Detection time, recall, false positives, and stability.

05
REPORT

Evidence

Artifacts, assumptions, limitations, and reproducible IDs.

13 / EVIDENCE LEVELS

Keep controlled benchmarks separate from field claims.

The evaluation surface explicitly labels what each evidence type can support.

CAPABILITY
Controlled replayCalibration LabField deployment
Best for
RepeatabilityInteractive causalityOperational validity
Ground truth
Dataset / injectedScenario-definedOften partial
Environment
HistoricalSimulated browser labReal operations
Claim strength
Algorithm comparisonMechanism demonstrationDeployment-specific outcome
Current public use
AvailableAvailableDesign-partner stage
Open the demo lab →
14 / EVALUATION WORKFLOW

Review a result without losing the assumptions.

A technical evaluation should be able to move backward from the chart to the exact source and configuration.

01
FIGURE

Read

Identify metric, units, sample, and comparison.

02
LINEAGE

Trace

Open scenario, manifest, model, and threshold.

03
RUN

Reproduce

Run replay or Calibration Lab configuration.

04
REVIEW

Challenge

Inspect limitations and alternative explanations.

05
NEXT

Decide

Choose whether field validation is warranted.

Discuss technical validation →

Evaluate NADIR on a defined scenario.

Bring a dataset, sensor path, or operating question and preserve every assumption.