linux-mm.kvack.org archive mirror
 help / color / mirror / Atom feed
From: David Lang <david@lang.hm>
To: Chris Mason <clm@fb.com>
Cc: "akpm@linux-foundation.org" <akpm@linux-foundation.org>,
	"linux-kernel@vger.kernel.org" <linux-kernel@vger.kernel.org>,
	"linux-ide@vger.kernel.org" <linux-ide@vger.kernel.org>,
	"lsf-pc@lists.linux-foundation.org"
	<lsf-pc@lists.linux-foundation.org>,
	"linux-mm@kvack.org" <linux-mm@kvack.org>,
	"linux-scsi@vger.kernel.org" <linux-scsi@vger.kernel.org>,
	"rwheeler@redhat.com" <rwheeler@redhat.com>,
	"James.Bottomley@hansenpartnership.com"
	<James.Bottomley@hansenpartnership.com>,
	"linux-fsdevel@vger.kernel.org" <linux-fsdevel@vger.kernel.org>,
	"mgorman@suse.de" <mgorman@suse.de>
Subject: Re: [Lsf-pc] [LSF/MM TOPIC] really large storage sectors - going beyond 4096 bytes
Date: Wed, 22 Jan 2014 18:46:11 -0800 (PST)	[thread overview]
Message-ID: <alpine.DEB.2.02.1401221836330.13577@nftneq.ynat.uz> (raw)
In-Reply-To: <1390421691.1198.43.camel@ret.masoncoding.com>

On Wed, 22 Jan 2014, Chris Mason wrote:

> On Wed, 2014-01-22 at 11:50 -0800, Andrew Morton wrote:
>> On Wed, 22 Jan 2014 11:30:19 -0800 James Bottomley <James.Bottomley@hansenpartnership.com> wrote:
>>
>>> But this, I think, is the fundamental point for debate.  If we can pull
>>> alignment and other tricks to solve 99% of the problem is there a need
>>> for radical VM surgery?  Is there anything coming down the pipe in the
>>> future that may move the devices ahead of the tricks?
>>
>> I expect it would be relatively simple to get large blocksizes working
>> on powerpc with 64k PAGE_SIZE.  So before diving in and doing huge
>> amounts of work, perhaps someone can do a proof-of-concept on powerpc
>> (or ia64) with 64k blocksize.
>
>
> Maybe 5 drives in raid5 on MD, with 4K coming from each drive.  Well
> aligned 16K IO will work, everything else will about the same as a rmw
> from a single drive.

I think this is the key point to think about here. How will these new hard drive 
large block sizes differ from RAID stripes and SSD eraseblocks?

In all of these cases there are very clear advantages to doing the writes in 
properly sized and aligned chunks that correspond with the underlying structure 
to avoid the RMW overhead.

It's extremely unlikely that drive manufacturers will produce drives that won't 
work with any existing OS, so they are going to support smaller writes in 
firmware. If they don't, they won't be able to sell their drives to anyone 
running existing software. Given the Enterprise software upgrade cycle compared 
to the expanding storage needs, whatever they ship will have to work on OS and 
firmware releases that happened several years ago.

I think what is needed is some way to be able to get a report on how man RMW 
cycles have to happen. Then people can work on ways to reduce this number and 
measure the results.

I don't know if md and dm are currently smart enough to realize that the entire 
stripe is being overwritten and avoid the RMW cycle. If they can't, I would 
expect that once we start measuring it, they will gain such support.

David Lang

--
To unsubscribe, send a message with 'unsubscribe linux-mm' in
the body to majordomo@kvack.org.  For more info on Linux MM,
see: http://www.linux-mm.org/ .
Don't email: <a href=mailto:"dont@kvack.org"> email@kvack.org </a>

  reply	other threads:[~2014-01-23  2:46 UTC|newest]

Thread overview: 59+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2013-12-20  9:30 LSF/MM 2014 Call For Proposals Mel Gorman
2014-01-06 22:20 ` [LSF/MM TOPIC] [ATTEND] persistent memory progress, management of storage & file systems Ric Wheeler
2014-01-06 22:32   ` faibish, sorin
2014-01-07 19:44     ` Joel Becker
2014-01-21  7:00 ` LSF/MM 2014 Call For Proposals Michel Lespinasse
2014-01-22  3:04 ` [LSF/MM TOPIC] really large storage sectors - going beyond 4096 bytes Ric Wheeler
2014-01-22  5:20   ` Joel Becker
2014-01-22  7:14     ` Hannes Reinecke
2014-01-22  9:34   ` [Lsf-pc] " Mel Gorman
2014-01-22 14:10     ` Ric Wheeler
2014-01-22 14:34       ` Mel Gorman
2014-01-22 14:58         ` Ric Wheeler
2014-01-22 15:19           ` Mel Gorman
2014-01-22 17:02             ` Chris Mason
2014-01-22 17:21               ` James Bottomley
2014-01-22 18:02                 ` Chris Mason
2014-01-22 18:13                   ` James Bottomley
2014-01-22 18:17                     ` Ric Wheeler
2014-01-22 18:35                       ` James Bottomley
2014-01-22 18:39                         ` Ric Wheeler
2014-01-22 19:30                           ` James Bottomley
2014-01-22 19:50                             ` Andrew Morton
2014-01-22 20:13                               ` Chris Mason
2014-01-23  2:46                                 ` David Lang [this message]
2014-01-23  5:21                                   ` Theodore Ts'o
2014-01-23  8:35                               ` Dave Chinner
2014-01-23 12:55                                 ` Theodore Ts'o
2014-01-23 19:49                                   ` Dave Chinner
2014-01-23 21:21                                   ` Joel Becker
2014-01-22 20:57                             ` Martin K. Petersen
2014-01-22 18:37                     ` Chris Mason
2014-01-22 18:40                       ` Ric Wheeler
2014-01-22 18:47                       ` James Bottomley
2014-01-23 21:27                         ` Joel Becker
2014-01-23 21:34                           ` Chris Mason
2014-01-23  8:27                     ` Dave Chinner
2014-01-23 15:47                       ` James Bottomley
2014-01-23 16:44                         ` Mel Gorman
2014-01-23 19:55                           ` James Bottomley
2014-01-24 10:57                             ` Mel Gorman
2014-01-30  4:52                               ` Matthew Wilcox
2014-01-30  6:01                                 ` Dave Chinner
2014-01-30 10:50                                 ` Mel Gorman
2014-01-23 20:34                           ` Dave Chinner
2014-01-23 20:54                         ` Christoph Lameter
2014-01-23  8:24                 ` Dave Chinner
2014-01-23 20:48             ` Christoph Lameter
2014-01-22 20:47           ` Martin K. Petersen
2014-01-23  8:21         ` Dave Chinner
2014-01-22 15:14     ` Chris Mason
2014-01-22 16:03       ` James Bottomley
2014-01-22 16:45         ` Ric Wheeler
2014-01-22 17:00           ` James Bottomley
2014-01-22 21:05             ` Jan Kara
2014-01-23 20:47     ` Christoph Lameter
2014-01-24 11:09       ` Mel Gorman
2014-01-24 15:44         ` Christoph Lameter
2014-01-22 15:54   ` James Bottomley
2014-03-14  9:02 ` Update on LSF/MM [was Re: LSF/MM 2014 Call For Proposals] James Bottomley

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=alpine.DEB.2.02.1401221836330.13577@nftneq.ynat.uz \
    --to=david@lang.hm \
    --cc=James.Bottomley@hansenpartnership.com \
    --cc=akpm@linux-foundation.org \
    --cc=clm@fb.com \
    --cc=linux-fsdevel@vger.kernel.org \
    --cc=linux-ide@vger.kernel.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=linux-mm@kvack.org \
    --cc=linux-scsi@vger.kernel.org \
    --cc=lsf-pc@lists.linux-foundation.org \
    --cc=mgorman@suse.de \
    --cc=rwheeler@redhat.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox