Microsoft Storage Spaces Direct virtualises all storage devices attached to a server to create a highly available, scalable, clustered software-defined storage solution. Each server runs Windows Server Datacentre edition and manages all storage attached to that server — including NVMe, SSD, SAS and SATA drives — to create a high-performance clustered pool of storage that can provide Hyper-V VMs with flexible storage at a fraction of the cost of traditional SAN or NAS arrays.
In essence, Microsoft Storage Spaces Direct is a hyper-converged infrastructure with the ability to provide up to 13.7 million IOPS. The architecture allows you to deploy a cluster in under 15 minutes, and is completely self-healing, with the ability to add more nodes as and when required — with support for 16 servers and over 400 drives, for up to 1PB of storage per cluster.
Microsoft Storage Spaces Direct features a built-in server-side cache to maximise storage performance — a large, persistent, real-time read and write cache. The cache is configured automatically when Storage Spaces Direct is enabled; in most cases, no manual management whatsoever is required. How the cache works depends on the types of drives present.
Resilient File System (ReFS) is Microsoft's newest file system, designed to take advantage of new drive technologies, scale efficiently to large data sets across diverse workloads, provide data protection and self-healing against possible data corruption.
Once these tiers are configured, ReFS uses them to deliver fast storage for hot data and capacity-efficient storage for cold data. All writes occur in the performance tier, and large chunks of data that remain in the performance tier are efficiently moved to the capacity tier in real time. If using a hybrid deployment (mixing flash and HDD drives), the cache in Storage Spaces Direct helps accelerate reads, reducing the effect of data fragmentation characteristic of virtualised workloads. Otherwise, in an all-flash deployment, reads also occur in the performance tier.
In addition to providing resiliency improvements, ReFS introduces new features for performance-sensitive and virtualised workloads. Real-time tier optimisation, block cloning, and sparse VDL are good examples of the evolving capabilities of ReFS, designed to support dynamic and diverse workloads. Mirror-accelerated parity delivers both high performance and capacity-efficient storage for your data.
To deliver both high performance and capacity efficient storage, ReFS divides a volume into two logical storage groups, known as tiers. These tiers can have their own drive and resiliency types, allowing each tier to optimise for either performance or capacity.
ReFS is designed to support extremely large data sets — millions of terabytes — without negatively impacting performance, achieving greater scale than prior file systems. Storage Spaces Direct utilises ReFS to provide storage management across a maximum of 16 nodes and up to 400 drives. ReFS supports data deduplication with volumes up to 64TB and file sizes up to 1TB.
Supported drives: NVMe, SSD, SAS, SATA, SMR (Shingled Magnetic Recording)
Deploying ReFS on Storage Spaces with shared SAS enclosures is suitable for hosting archival data and storing user documents. Integrity streams, online repair, and alternate data copies enable ReFS and Storage Spaces to jointly detect and correct corruptions within both metadata and data. Storage Spaces deployments can also utilise block cloning and the scalability offered in ReFS.
Considering Storage Spaces Direct for your Hyper-V environment? Talk to us for free, no-obligation advice.
Talk to a Specialist