HNHacker News
TopNewBestAskShowJobs

suprasam

107 karma · joined October 2, 2018

submissionscomments
suprasam··on Show HN: Open Lustre in the cloud, with ZFS OSTs on object storage not disks
Thanks. The part that surprised me most was how well ZFS fits object storage and extended to pNFS too!
suprasam··on Show HN: Open Lustre in the cloud, with ZFS OSTs on object storage not disks
This work was presented at LUG 2026 https://www.opensfs.org/wp-content/uploads/16-ZettaLane_LUG_...
suprasam··on Native ZFS VDEV for Object Storage (OpenZFS Summit)
ZeroFS doesn't exploit ZFS strengths with no native ZFS support, just an afterthought with NBD + SlateDB LSM Good for small burst workloads where everything kept it in memory for LSM batch writes. Once compaction hits all bets off with performance and not sure about crash consistency since it is playing with fire. ZFS special vdev + ZIL on ssd is much safer. No need for LSM. MayaNAS ZFS metadata at SSD speed and large blocks get throughput from high latency S3 at network speed.
suprasam··on Native ZFS VDEV for Object Storage (OpenZFS Summit)
Yes that is the value prop. Cheap S3 instead of expensive EBS.

  EBS limitations:
  - Per-instance throughput caps
  - Pay for full provisioned capacity whether filled or not

 S3:
  - Pay only for what you store
  - No per-instance bandwidth limits as long as you have network optimized instance
suprasam··on Native ZFS VDEV for Object Storage (OpenZFS Summit)
For RDBMS pages on object storage - you might be thinking of Neon.tech. They built a custom page server for PostgreSQL that stores pages directly on S3.
suprasam··on Native ZFS VDEV for Object Storage (OpenZFS Summit)
I hope you are not having that massive storage storage on public-cloud then you would need MayaNAS to reduce storage costs. For S3 as frontend use MinIO gateway - serves S3 API from your ZFS filesystem
suprasam··on Native ZFS VDEV for Object Storage (OpenZFS Summit)
It is all part of ZFS architecture with two tiers: - Special vdev (SSD): All metadata + small blocks (configurable threshold, typically <128KB) - Object storage: Bulk data only If the workload is randomized 4K small data blocks - that's SSD latency, not S3 latency.
suprasam··on Native ZFS VDEV for Object Storage (OpenZFS Summit)
Yes, this is a core use case ZFS fits nicely. See slide 31 "Multi-Cloud Data Orchestration" in the talk.

Not only backup but also DR site recovery.

  The workflow:

  1. Server A (production): zpool on local NVMe/SSD/HD
  2. Server B (same data center): another zpool backed by objbacker.io → remote object storage (Wasabi, S3, GCS)
  3. zfs send from A to B - data lands in object storage

  Key advantage: no continuously running cloud VM. You're just paying for object storage (cheap) not compute (expensive). Server B is in your own data center - it can be a VM too.
For DR, when you need the data in cloud:

  - Spin up a MayaNAS VM only when needed
  - Import the objbacker-backed pool - data is already there
  - Use it, then shut down the VM
suprasam··on [dead]
New ZFS file system on object storage from https://www.zettalane.com
suprasam··on MayaNAS: ZFS on S3 object storage for high-throughput
Experience very high-throughput > 3GB/s, even on a single VM instance with no traditional disk resources other than S3 object storage.
suprasam··on Another ZFS Port on Linux
Demo regarding unified ARC & Pagecache https://youtu.be/be0ph4b9vUE
suprasam··on Announce Crossmeta FUSE for Windows
The fun part is you can even develop from Linux using MinGW32 Cross Compile environment which produces native windows programs.

Crossmeta FUSE also includes sshfs, fuse-nfs for remote file access and s3backer to connect to any S3 compatible cloud storage.

All Crossmeta File systems fully visible to Windows Subsystem for Linux (WSL) but integration can be better with your help by voting on https://wpdev.uservoice.com/forums/266908-command-prompt-con...

suprasam··on Announce Crossmeta FUSE for Windows
It has NO low level fuse api https://github.com/billziss-gh/winfsp/issues/94

Newer FUSE version 3 mostly address their limitations within Linux implementation. It adds support for concurrent writes and large transfer beyond 4KB. Crossmeta FUSE is look alike of Linux FUSE but the implementation does not inherit those limitations. It requires to do large concurrent transfer 128KB from Windows NT cache manager. I have to work the sshfs code to handle this since it was not exposed to such large transfer.

suprasam··on Announce Crossmeta FUSE for Windows
Crossmeta kernel has read/write support for XFS and EXT2/3/4 It has readonly support for reiserfs,Apple HFS+ All kernel drivers can be unloaded when not in use.

From FUSE it has support for sshfs and NFS client V3 and V4.

suprasam··on Announce Crossmeta FUSE for Windows
The alternative project has no FUSE low-level API support. This is required to implement any real file system with FUSE on windows. Also they have dependency on cygwin project if your code requires some POSIX semantics, which they all do since most of the FUSE projects are from Linux.

Crossmeta is just NOT about FUSE, it can also provide POSIX capabilities similar to Microsoft WSL. More here https://github.com/crossmeta/sys/