We have a new paper on model weight proofs that run on every request inside the inference engine releasing soon
In v1, we were relying on logprobs to verify the correct model is loaded via the expected fingerprint
Alongside slashable bonds for operators, i’d argue that was already very strong but it will be greatly improved in v2 which we are releasing within a month
For v2, we don’t have to rerun inference on a validator GPU in the same way you have to with logprobs - instead we have proofs that can be validated very quickly on a CPU with less overhead
Preview of benchmarks attached