[Twg] Lustre Requirements and Roadmap
Christopher J. Morrone
morrone2 at llnl.gov
Thu Jan 27 21:56:34 UTC 2011
John,
Thanks for creating that list!
I think that for the metadata performance, we should use labels other
than "min" and "max". After all, we would never say "no, we won't
accept a filesystem that is capable of more than than this". :)
Also, the requirements are only useful if we can give target dates for
the requirements. So maybe the "min" column becomes Q3/2012, and the
max column Q1/2014. I am just pulling those dates out of thin air, they
could be something else.
I think we should split the aggregate and single client file creates/sec
requirements into two separate requirement lines.
For backend storage:
I think that we should take out the "framework to enable alternatives to
ldiskfs" line. That is a design decision rather than a requirement. I
think it is certainly necessary given the requirements, but if we are
trying to distinguish between requirements and architectural decisions,
that line should be removed.
Rather than "no performance impact", say "low performance impact". Zero
impact is probably impossible. :) At some point we need to be clear on
exactly what quilifies as "low". One part of that might be:
- The solution must perform filesystem integrity checking and repair
on-line.
I.E., There may be no required downtime just for consistence checking
and repair.
What is the rationale for the "direct I/O mode" requirement? Not that I
am opposed to it...but is that really a requirement, or a solution?
For end-to-end integrity, I would suggest that T10 PI is not sufficient,
since it only concerns itself with the path between the HBA and the
backend storage (disks/flash). I think that we need to define where the
"ends" are, and they should really extend up into memory on the lustre
servers. T10 DIF + Oracle's DIX might be a better example. Examples
are not necessarily what we want here.
Chris
On 01/27/2011 09:24 AM, John Carrier wrote:
> The following is an incomplete list of requirements to motivate further
> discussion :
>
> * metadata performance
>
> GOAL: improve file system scalability and interactive
> performance
>
> requirements: min max
> - # files in file system 100 billion 1 trillion
> - # files in directory 50 million 10 billion
> - file creates / sec 100 thousand 30 thousand
> (aggregate) (single client)
> - directory lisings / sec
> - open files per process - 100 thousand
> - file system capacity 30 PB 100 PB
> - # clients 30 thousand ?00 thousand
>
>
> * backend storage
>
> GOAL: provide reliable, scalable backing store for
> Lustre servers
>
> requirements:
> - large LUNs (min 32 TB)
> - end-to-end data integrity (T10 PI or equivalent)
> - no performance impact for file system repair
> - framework to enable alternatives to ldiskfs
> - direct I/O mode
> - ??
>
>
>
>
> _______________________________________________
> twg mailing list
> twg at lists.opensfs.org
> http://lists.opensfs.org/listinfo.cgi/twg-opensfs.org
> .
>
More information about the Twg
mailing list