Behavioral fingerprint
Compare response distributions against versioned reference samples. Experimental; no validated fingerprint ships in v0.2.0.
Methodology / mt-bench-0.1.0
ModelTrust evaluates externally observable behavior. Black-box identification can estimate consistency with a reference model, but it cannot cryptographically prove which weights served a response.
Interpretation rule
Report statistical confidence and capability integrity—never a binary “real vs fake” verdict.
A provider may route, quantize, wrap, fine-tune, rate-limit, or otherwise alter an endpoint. Observed capability can differ without establishing deception or hidden model identity.
1 IMPLEMENTED · 3 EXPERIMENTAL · 4 PLACEHOLDERS
Compare response distributions against versioned reference samples. Experimental; no validated fingerprint ships in v0.2.0.
Measure retention across bounded, contamination-aware reasoning tasks. Experimental.
Test adherence to conflicting, nested, and structured instructions. Placeholder.
Inspect schema adherence, argument validity, and call consistency. Placeholder.
Measure recall, position sensitivity, and degradation under longer contexts. Placeholder.
Compare provider-reported usage with independent token estimates and response shape. Experimental.
Measure bounded request completion time at the verification server. Implemented for one control request.
Aggregate success, error, and timeout rates across repeated samples. Placeholder.
Benchmarks will be versioned. Future reports should publish sampling rules, reference windows, scoring transformations, and confidence intervals without exposing active anti-gaming items.
Providers may pay for tests, never rankings. A future ModelTrust router must remain separated from verification scoring and cannot receive preferential methodology treatment.