>Any data of mine or my company's going over the wire to these models needs to stay verifiably private.
I don't think this is possible without running everyting locally and the data not leaving the machine (or possibly local network) you control.
I don't think this is possible without running everyting locally and the data not leaving the machine (or possibly local network) you control.
Using cryptographic primitives and hardware root of trust (even GPU trusted execution which NVIDIA now supports for nvlink) you can basically attest to certain compute operations. Of which might be confidential inference.
My company, EQTY Lab, and others like Edgeless Systems or Tinfoil are working hard in this space.
https://pasteboard.co/k1hjwT7pWI6x.png
reach out if interested in collab.