Commits · 53b03744e5699832e6c5b04f2ec506d8b0c50c38 · matisse / android_kernel_samsung_matisse

30 Sep, 2006 12 commits

[PATCH] cfq-iosched: Kill O(N) runtime of cfq_resort_rr_list() · 53b03744

Jens Axboe authored 18 years ago


Currently it scales with number of processes in that priority group,
which is potentially not very nice as it's called quite often.
Basically we always need to do tail inserts, except for the case of a
new process. So just mark/detect a queue as such.
Signed-off-by: Jens Axboe <axboe@suse.de>

53b03744

[PATCH] Make sure all block/io scheduler setups are node aware · b5deef90

Jens Axboe authored 18 years ago


Some were kmalloc_node(), some were still kmalloc(). Change them all to
kmalloc_node().
Signed-off-by: Jens Axboe <axboe@suse.de>

b5deef90

[PATCH] Audit block layer inlines · 1ea25ecb

Jens Axboe authored 18 years ago


Kill a few inlines that bring in too much code to more than one location
Shrinks kernel text by about 300 bytes on 32-bit x86.
Signed-off-by: Jens Axboe <axboe@suse.de>

1ea25ecb

[PATCH] cfq-iosched: use new io context counting mechanism · 4050cf16

Jens Axboe authored 18 years ago


It's ok if the read path is a lot more costly, as long as inc/dec is
really cheap. The inc/dec will happen for each created/freed io context,
while the reading only happens when a disk queue exits.
Signed-off-by: Jens Axboe <axboe@suse.de>

4050cf16

[PATCH] cfq-iosched: kill cfq_exit_lock · fc46379d

Jens Axboe authored 18 years ago


cfq_exit_lock is protecting two things now:

- The per-ioc rbtree of cfq_io_contexts

- The per-cfqd linked list of cfq_io_contexts

The per-cfqd linked list can be protected by the queue lock, as it is (by
definition) per cfqd as the queue lock is.

The per-ioc rbtree is mainly used and updated by the process itself only.
The only outside use is the io priority changing. If we move the
priority changing to not browsing the rbtree, we can remove any locking
from the rbtree updates and lookup completely. Let the sys_ioprio syscall
just mark processes as having the iopriority changed and lazily update
the private cfq io contexts the next time io is queued, and we can
remove this locking as well.
Signed-off-by: Jens Axboe <axboe@suse.de>

fc46379d

[PATCH] cfq-iosched: cleanups, fixes, dead code removal · 89850f7e

Jens Axboe authored 18 years ago


A collection of little fixes and cleanups:

- We don't use the 'queued' sysfs exported attribute, since the
  may_queue() logic was rewritten. So kill it.

- Remove dead defines.

- cfq_set_active_queue() can be rewritten cleaner with else if conditions.

- Several places had cfq_exit_cfqq() like logic, abstract that out and
  use that.

- Annotate the cfqq kmem_cache_alloc() so the allocator knows that this
  is a repeat allocation if it fails with __GFP_WAIT set. Allows the
  allocator to start freeing some memory, if needed. CFQ already loops for
  this condition, so might as well pass the hint down.

- Remove cfqd->rq_starved logic. It's not needed anymore after we dropped
  the crq allocation in cfq_set_request().

- Remove uneeded parameter passing.
Signed-off-by: Jens Axboe <axboe@suse.de>

89850f7e

[PATCH] Drop useless bio passing in may_queue/set_request API · cb78b285
Jens Axboe authored 18 years ago
```
It's not needed for anything, so kill the bio passing.
Signed-off-by: Jens Axboe <axboe@suse.de>
```
cb78b285

[PATCH] cfq-iosched: kill crq · 5e705374

Jens Axboe authored 18 years ago


Get rid of the cfq_rq request type. With the added elevator_private2, we
have enough room in struct request to get rid of any crq allocation/free
for each request.
Signed-off-by: Jens Axboe <axboe@suse.de>

5e705374

[PATCH] cfq-iosched: remove the crq flag functions/variable · 5380a101

Jens Axboe authored 18 years ago


There's just one flag currently (SYNC), and that one can be grabbed from
the request.
Signed-off-by: Jens Axboe <axboe@suse.de>

5380a101

[PATCH] cfq-iosched: convert to using the FIFO elevator defines · 95e8810b
Jens Axboe authored 18 years ago
```
Signed-off-by: Jens Axboe <axboe@suse.de>
```
95e8810b
[PATCH] cfq-iosched: migrate to using the elevator rb functions · 21183b07
Jens Axboe authored 18 years ago
```
This removes the rbtree handling from CFQ.
Signed-off-by: Jens Axboe <axboe@suse.de>
```
21183b07

[PATCH] elevator: move the backmerging logic into the elevator core · 9817064b

Jens Axboe authored 18 years ago


Right now, every IO scheduler implements its own backmerging (except for
noop, which does no merging). That results in duplicated code for
essentially the same operation, which is never a good thing. This patch
moves the backmerging out of the io schedulers and into the elevator
core. We save 1.6kb of text and as a bonus get backmerging for noop as
well. Win-win!
Signed-off-by: Jens Axboe <axboe@suse.de>

9817064b

21 Aug, 2006 1 commit

[PATCH] cfq_cic_link: fix usage of wrong cfq_io_context · be33c3a6

Oleg Nesterov authored 18 years ago


Obviously, cfq_cic_link() shouldn't free a just allocated cfq_io_context?
The dead key is from __cic, so drop that.
Signed-off-by: Oleg Nesterov <oleg@tv-sign.ru>
Signed-off-by: Jens Axboe <axboe@suse.de>

be33c3a6

25 Jul, 2006 1 commit

[PATCH] cfq-iosched: don't use a hard jiffies value, translate from msecs · 44eb1231

Jens Axboe authored 18 years ago


The CIC_SEEKY() test really wants to use the minimum of either:

- 2 msecs (not jiffies)

- or, the pending slice time

So code it like that.
Signed-off-by: Jens Axboe <axboe@suse.de>

44eb1231

30 Jun, 2006 1 commit

Remove obsolete #include <linux/config.h> · 6ab3d562

Jörn Engel authored 18 years ago

Signed-off-by: Jörn Engel <joern@wohnheim.fh-wedel.de>
Signed-off-by: Adrian Bunk <bunk@stusta.de>

6ab3d562

23 Jun, 2006 6 commits

[PATCH] rbtree: support functions used by the io schedulers · dd67d051

Jens Axboe authored 18 years ago


They all duplicate macros to check for empty root and/or node, and
clearing a node. So put those in rbtree.h.
Signed-off-by: Jens Axboe <axboe@suse.de>

dd67d051

[PATCH] cfq-iosched: rq update fixes · fd61af03

Jens Axboe authored 18 years ago


- Remember to set ->last_sector so that the cfq_choose_req() logic
  works correctly.

- Remove redundant call to cfq_choose_req()
Signed-off-by: Jens Axboe <axboe@suse.de>

fd61af03

[PATCH] cfq-iosched: many performance fixes · caaa5f9f

Jens Axboe authored 18 years ago


This is a collection of patches that greatly improve CFQ performance
in some circumstances.

- Change the idling logic to only kick in after a request is done and we
  are deciding what to do. Before the idling included the request service
  time, so it was hard to adjust. Now it's true think/idle time.

- Take advantage of TCQ/NCQ/queueing for seeky sync workloads, but keep
  it in control for sync and sequential (or close to) workloads.

- Expire queues immediately and move on to other busy queues, if we are
  not going to idle after the current one finishes.

- Don't rearm idle timer if there are no busy queues. Just leave the
  system idle.
Signed-off-by: Jens Axboe <axboe@suse.de>

caaa5f9f

[PATCH] cfq-iosched: correctly set ioprio on both targets · 35e6077c

Jens Axboe authored 18 years ago


Patch originally from Vasily Tarasov <vtaras@sw.ru>

If you set io-priority of process 1 using sys_ioprio_set system call by
another process 2 (like ionice do), then cfq_init_prio_data() function
sets priority of process 2 (current) on queue of process 1 and clears
the flag, that designates change of ioprio.  So the process  1 will work
like with priority of process 2.

I propose not to call cfq_init_prio_data() on io-priority change, but
only mark queue as queue with changed prority.  Every time when new
request comes cfq-scheduler checks for this flag and atomaticaly changes
priority of queue to new value.
Signed-off-by: Jens Axboe <axboe@suse.de>

35e6077c

[PATCH] Kill PF_SYNCWRITE flag · b31dc66a

Jens Axboe authored 18 years ago

A process flag to indicate whether we are doing sync io is incredibly
ugly. It also causes performance problems when one does a lot of async
io and then proceeds to sync it. Part of the io will go out as async,
and the other part as sync. This causes a disconnect between the
previously submitted io and the synced io. For io schedulers such as CFQ,
this will cause us lost merges and suboptimal behaviour in scheduling.

Remove PF_SYNCWRITE completely from the fsync/msync paths, and let
the O_DIRECT path just directly indicate that the writes are sync
by using WRITE_SYNC instead.
Signed-off-by: Jens Axboe <axboe@suse.de>

b31dc66a

[PATCH] cfq-iosched: Don't set the queue batching limits · 271f18f1

Jens Axboe authored 18 years ago


We cannot update them if the user changes nr_requests, so don't
set it in the first place. The gains are pretty questionable as
well. The batching loss has been shown to decrease throughput.
Signed-off-by: Jens Axboe <axboe@suse.de>

271f18f1

20 Jun, 2006 1 commit

Fix up CFQ scheduler for recent rbtree node shrinkage · 6b41fd17

Linus Torvalds authored 18 years ago


The color is now in the low bits of the parent pointer, and initializing
it to 0 happens as part of the whole memset above, so just remove the
unnecessary RB_CLEAR_COLOR.
Signed-off-by: Linus Torvalds <torvalds@osdl.org>

6b41fd17

14 Jun, 2006 1 commit

[PATCH] cfq-iosched: fix crash in do_div() · 553698f9

Jens Axboe authored 18 years ago


We don't clear the seek stat values in cfq_alloc_io_context(), and if
->seek_mean is unlucky enough to be set to -36 by chance, the first
invocation of cfq_update_io_seektime() will oops with a divide by zero
in do_div().

Just memset the entire cic instead of filling invididual values
independently.
Signed-off-by: Jens Axboe <axboe@suse.de>
Signed-off-by: Linus Torvalds <torvalds@osdl.org>

553698f9

08 Jun, 2006 1 commit

[PATCH] elevator switching race · bc1c1169

Jens Axboe authored 18 years ago


There's a race between shutting down one io scheduler and firing up the
next, in which a new io could enter and cause the io scheduler to be
invoked with bad or NULL data.

To fix this, we need to maintain the queue lock for a bit longer.
Unfortunately we cannot do that, since the elevator init requires to be
run without the lock held.  This isn't easily fixable, without also
changing the mempool API.  So split the initialization into two parts,
and alloc-init operation and an attach operation.  Then we can
preallocate the io scheduler and related structures, and run the attach
inside the lock after we detach the old one.

This patch has survived 30 minutes of 1 second io scheduler switching
with a very busy io load.
Signed-off-by: Jens Axboe <axboe@suse.de>
Signed-off-by: Linus Torvalds <torvalds@osdl.org>

bc1c1169

01 Jun, 2006 5 commits

[PATCH] cfq-iosched: busy_rr fairness fix · b52a8348

Jens Axboe authored 18 years ago


Now that we select busy_rr for possible service, insert entries at the
back of that list instead of at the front.
Signed-off-by: Jens Axboe <axboe@suse.de>

b52a8348

[PATCH] cfq-iosched: fix bug in timer handling for the idle class · ae818a38

Jens Axboe authored 18 years ago


There's a small window from when the timer is entered and we grab
the queue lock, where cfq_set_active_queue() could be rearming the
timer for us. Seen in the wild on a 12-way ppc box. Fix this by
just using mod_timer(), which will do the right thing for us.
Signed-off-by: Jens Axboe <axboe@suse.de>

ae818a38

[PATCH] cfq-iosched: Detect hardware queueing · 25776e35

Jens Axboe authored 18 years ago


If the hardware is doing real queueing, decide that it's worthless to
idle the hardware. It does reasonable simultaneous io in that case
anyways, and the idling hurts some work loads.
Signed-off-by: Jens Axboe <axboe@suse.de>

25776e35

[PATCH] cfq-iosched: Detect idle process issuing async request · 12e9fddd

Jens Axboe authored 18 years ago


If we are anticipating a sync request from this process and we are
waiting for that and see an async request come in, expire that slice
and move on.
Signed-off-by: Jens Axboe <axboe@suse.de>

12e9fddd

[PATCH] cfq-iosched: check busy queues before deciding we are idle · e0de0206

Jens Axboe authored 18 years ago


For just one busy queue (like async write out), we often overlooked
that we could queue more io and decided we were idle instead. This causes
us quite a bit of performance loss.
Signed-off-by: Jens Axboe <axboe@suse.de>

e0de0206

30 May, 2006 1 commit

[PATCH] cfq-iosched: fixup locking and ->queue_list list management · 3793c65c

Jens Axboe authored 18 years ago


- Drop cic from the list when seen as dead.
- Fixup the locking, just use a simple spinlock.
Signed-off-by: Jens Axboe <axboe@suse.de>
Signed-off-by: Linus Torvalds <torvalds@osdl.org>

3793c65c

21 Apr, 2006 1 commit

[RBTREE] Change rbtree off-tree marking in I/O schedulers. · 3db3a445

David Woodhouse authored 19 years ago

They were abusing the rb_color field to mark nodes which weren't currently
on the tree. Fix that to use the same method as eventpoll did -- setting
the parent pointer to point back to itself. And use the appropriate
accessor macros for setting and reading the parent.
Signed-off-by: David Woodhouse <dwmw2@infradead.org>

3db3a445

18 Apr, 2006 3 commits

[PATCH] cfq: Further rbtree traversal and cfq_exit_queue() race fix · be3b0753

OGAWA Hirofumi authored 19 years ago


In current code, we are re-reading cic->key after dead cic->key check.
So, in theory, it may really re-read *after* cfq_exit_queue() seted NULL.

To avoid race, we copy it to stack, then use it. With this change, I
guess gcc will assign cic->key to a register or stack, and it wouldn't
be re-readed.
Signed-off-by: OGAWA Hirofumi <hirofumi@mail.parknet.co.jp>
Signed-off-by: Jens Axboe <axboe@suse.de>

be3b0753

[PATCH 2/2] cfq: fix cic's rbtree traversal · dbecf3ab

OGAWA Hirofumi authored 19 years ago


When queue dies, we set cic->key=NULL as dead mark. So, when we
traverse a rbtree, we must check whether it's still valid key. if it
was invalidated, drop it, then restart the traversal from top.
Signed-off-by: OGAWA Hirofumi <hirofumi@mail.parknet.co.jp>
Signed-off-by: Jens Axboe <axboe@suse.de>

dbecf3ab

[PATCH 1/2] iosched: fix typo and barrier() · fba82272

OGAWA Hirofumi authored 19 years ago


On rmmod path, cfq/as waits to make sure all io-contexts was
freed. However, it's using complete(), not wait_for_completion().

I think barrier() is not enough in here. To avoid the following case,
this patch replaces barrier() with smb_wmb().

	cpu0			visibility			cpu1
	                [ioc_gnone=NULL,ioc_count=1]

ioc_gnone = &all_gone		NULL,ioc_count=1
atomic_read(&ioc_count)		NULL,ioc_count=1
wait_for_completion()		NULL,ioc_count=0	atomic_sub_and_test()
				NULL,ioc_count=0	if ( && ioc_gone)
						    [ioc_gone==NULL,
						    so doesn't call complete()]
			   &all_gone,ioc_count=0
Signed-off-by: OGAWA Hirofumi <hirofumi@mail.parknet.co.jp>
Signed-off-by: Jens Axboe <axboe@suse.de>

fba82272

28 Mar, 2006 3 commits

[BLOCK] cfq-iosched: seek and async performance fixes · 206dc69b

Jens Axboe authored 19 years ago


Detect whether a given process is seeky and if so disable (mostly) the
idle window if it is. We still allow just a little idle time, just enough
to allow that process to submit a new request. That is needed to maintain
fairness across priority groups.

In some cases, we could setup several async queues. This is not optimal
from a performance POV, since we want all async io in one queue to perform
good sorting on it. It also impacted sync queues, as async io got too much
slice time.
Signed-off-by: Jens Axboe <axboe@suse.de>

206dc69b

[PATCH] cfq-iosched: small cfq_choose_req() optimization · e8a99053

Andreas Mohr authored 19 years ago


this is a small optimization to cfq_choose_req() in the CFQ I/O scheduler
(this function is a semi-often invoked candidate in an oprofile log):
by using a bit mask variable, we can use a simple switch() to check
the various cases instead of having to query two variables for each check.
Benefit: 251 vs. 285 bytes footprint of cfq_choose_req().
Also, common case 0 (no request wrapping) is now checked first in code.
Signed-off-by: Andreas Mohr <andi@lisas.de>
Signed-off-by: Jens Axboe <axboe@suse.de>

e8a99053

[PATCH] [BLOCK] cfq-iosched: change cfq io context linking from list to tree · e2d74ac0

Jens Axboe authored 19 years ago


On setups with many disks, we spend a considerable amount of time
looking up the process-disk mapping on each queue of io. Testing with
a NULL based block driver, this costs 40-50% reduction in throughput
for 1000 disks.
Signed-off-by: Jens Axboe <axboe@suse.de>

e2d74ac0

26 Mar, 2006 1 commit

[PATCH] mempool: use mempool_create_slab_pool() · 93d2341c

Matthew Dobson authored 19 years ago


Modify well over a dozen mempool users to call mempool_create_slab_pool()
rather than calling mempool_create() with extra arguments, saving about 30
lines of code and increasing readability.
Signed-off-by: Matthew Dobson <colpatch@us.ibm.com>
Signed-off-by: Andrew Morton <akpm@osdl.org>
Signed-off-by: Linus Torvalds <torvalds@osdl.org>

93d2341c

18 Mar, 2006 2 commits
- [PATCH] fix rmmod problems with elevator attributes, clean them up · e572ec7e
  Al Viro authored 19 years ago
  
  e572ec7e
- [PATCH] elevator_t lifetime rules and sysfs fixes · 3d1ab40f
  Al Viro authored 19 years ago
  
  3d1ab40f