Seems easy to find: https://ceph.io/en/news/blog/2024/ceph-a-journey-to-1tibps/
Ceph: 68 nodes, 2x100Gbps Mellanox and 10x 14TiB NVMe SSDs per node, 504 clients, 1TiB/s of FIO random read workload
I also assume that the batch size (block size) is different enough that this alone would make a big difference.
Ceph cluster achieves 1 TiB/s / 1.7 TiB/s = 0.58% of theoretical throughput.
3FS cluster achieves 6.6 TiB/s / 9 TiB/s = 0.73% of theoretical throughput.