* [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2
[not found] ` <20060217205204.GC2889@ccure.user-mode-linux.org>
@ 2006-02-27 16:27 ` Anthony Brock
2006-03-06 17:10 ` Blaisorblade
0 siblings, 1 reply; 6+ messages in thread
From: Anthony Brock @ 2006-02-27 16:27 UTC (permalink / raw)
To: user-mode-linux-devel
Sorry for taking so long to get back on this. We've been busy with some
other issues. Some comments are inline below in addition to the output
from a new crash.
Quoting Jeff Dike <jdike@addtoit.com>:
> On Wed, Feb 15, 2006 at 01:09:33PM -0800, Anthony Brock wrote:
> Also, this is x86_64, correct? If so, try
> echo 0 > /proc/sys/kernel/vsyscall64
Both host and client are 32-bit. The host is a dual AMD Opteron, but
we're running the 32-bit (SMP) version of Linux.
> As for the OOM cases, the data that BB asked for would still be
> helpful in diagnosing them.
I'll try to get that to you later this week.
In the mean time, we've experienced a second crash with the
2.6.15.1-bs1 kernel. It looks to be very similar to the previous crash.
Only, in this case, I issued the command "uml_mconsole www-prod-1 sysrq
s" in preparation for a backup (after issuing a "uml_mconsole $1 stop"
command). I am including the first dump below. The rest of the dump can
be found at:
http://www.coherenthosting.com/prj/uml/miscellaneous/error.txt
The time stamp at the top is in UTC, and specifies when the crash started.
Tony
-- time-stamp -- Feb/27/06 16:03:00 --
SysRq : Emergency Sync
events/0: page allocation failure. order:0, mode:0x20
0829f510: [<0808a1f3>] __alloc_pages+0x274/0x286
0829f550: [<0808c7e5>] kmem_getpages+0x4a/0x9f
0829f564: [<0808d2b4>] cache_grow+0x96/0x122
0829f598: [<0808d493>] cache_alloc_refill+0x153/0x186
0829f5bc: [<0808d62d>] kmem_cache_alloc+0x34/0x44
0829f5d0: [<0808d134>] alloc_slabmgmt+0x20/0x45
0829f5e4: [<0808d2cc>] cache_grow+0xae/0x122
0829f614: [<0808d493>] cache_alloc_refill+0x153/0x186
0829f638: [<0808d6dc>] __kmalloc+0x5b/0x6d
0829f654: [<08171704>] __alloc_skb+0x52/0x129
0829f670: [<0806097e>] uml_net_rx+0x1e/0x13c
0829f690: [<08060ac0>] uml_net_interrupt+0x1f/0x91
0829f6d4: [<0808595b>] handle_IRQ_event+0x24/0x54
0829f6f8: [<080859de>] __do_IRQ+0x53/0x91
0829f70c: [<08057bc8>] do_IRQ+0x20/0x2c
0829f714: [<08057d5d>] sigio_handler+0xa5/0xcf
0829f72c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
0829f750: [<0806792e>] sig_handler+0xe/0x11
0829f860: [<08057cdf>] sigio_handler+0x27/0xcf
0829f87c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
0829f8a0: [<0806792e>] sig_handler+0xe/0x11
0829f9e8: [<0805a694>] set_signals+0xa4/0xc0
0829fa28: [<0805a694>] set_signals+0xa4/0xc0
0829fa64: [<0805a694>] set_signals+0xa4/0xc0
0829fa8c: [<0805a694>] set_signals+0xa4/0xc0
0829fae0: [<0805a641>] set_signals+0x51/0xc0
0829fb80: [<081607e6>] handle_sysrq+0x1c/0x20
0829fb94: [<0806258b>] mconsole_sysrq+0x5d/0x63
0829fbc0: [<0806220f>] mconsole_stop+0x40/0x78
0829fbcc: [<08061e1d>] mc_work_proc+0x38/0x4f
0829fbdc: [<0807c175>] worker_thread+0x188/0x205
0829fbe4: [<08061de5>] mc_work_proc+0x0/0x4f
0829fc08: [<0806bde4>] default_wake_function+0x0/0x12
0829fc2c: [<0807bfed>] worker_thread+0x0/0x205
0829fc30: [<0807f21b>] kthread+0x74/0x9a
0829fc4c: [<0807f1a7>] kthread+0x0/0x9a
0829fc54: [<0806784e>] run_kernel_thread+0x30/0x3b
0829fc64: [<0807f1a7>] kthread+0x0/0x9a
0829fc70: [<08067838>] run_kernel_thread+0x1a/0x3b
0829fce8: [<0806b8e4>] schedule_tail+0x23/0xb5
0829fcfc: [<0805de14>] new_thread_handler+0x7f/0xa7
0829fd00: [<0807f1a7>] kthread+0x0/0x9a
Mem-info:
DMA per-cpu:
cpu 0 hot: low 0, high 18, batch 3 used:2
cpu 0 cold: low 0, high 6, batch 1 used:5
DMA32 per-cpu: empty
Normal per-cpu: empty
HighMem per-cpu: empty
Free pages: 376kB (0kB HighMem)
Active:2985 inactive:8684 dirty:4 writeback:0 unstable:0 free:94
slab:3428 mapped:933 pagetables:48
DMA free:376kB min:1024kB low:1280kB high:1536kB active:11940kB
inactive:34736kB present:65536kB pages_scanned:0 all_unreclaimable?
no
lowmem_reserve[]: 0 0 0 0
DMA32 free:0kB min:0kB low:0kB high:0kB active:0kB inactive:0kB
present:0kB pages_scanned:0 all_unreclaimable? no
lowmem_reserve[]: 0 0 0 0
Normal free:0kB min:0kB low:0kB high:0kB active:0kB inactive:0kB
present:0kB pages_scanned:0 all_unreclaimable? no
lowmem_reserve[]: 0 0 0 0
HighMem free:0kB min:128kB low:128kB high:128kB active:0kB inactive:0kB
present:0kB pages_scanned:0 all_unreclaimable? no
lowmem_reserve[]: 0 0 0 0
DMA: 0*4kB 1*8kB 1*16kB 1*32kB 1*64kB 0*128kB 1*256kB 0*512kB 0*1024kB
0*2048kB 0*4096kB = 376kB
DMA32: empty
Normal: empty
HighMem: empty
Swap cache: add 0, delete 0, find 0/0, race 0+0
Free swap = 655340kB
Total swap = 655340kB
Free swap: 655340kB
16384 pages of RAM
0 pages of HIGHMEM
952 reserved pages
11194 pages shared
0 pages swap cached
-------------------------------------------------------
This SF.Net email is sponsored by xPML, a groundbreaking scripting language
that extends applications into web and mobile media. Attend the live webcast
and join the prime developer group breaking into this new coding territory!
http://sel.as-us.falkag.net/sel?cmd=lnk&kid=110944&bid=241720&dat=121642
_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
^ permalink raw reply [flat|nested] 6+ messages in thread
* Re: [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2
2006-02-27 16:27 ` [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2 Anthony Brock
@ 2006-03-06 17:10 ` Blaisorblade
2006-03-06 19:33 ` Anthony Brock
0 siblings, 1 reply; 6+ messages in thread
From: Blaisorblade @ 2006-03-06 17:10 UTC (permalink / raw)
To: user-mode-linux-devel; +Cc: Anthony Brock, Jeff Dike
On Monday 27 February 2006 17:27, Anthony Brock wrote:
> Sorry for taking so long to get back on this. We've been busy with some
> other issues. Some comments are inline below in addition to the output
> from a new crash.
> In the mean time, we've experienced a second crash with the
> 2.6.15.1-bs1 kernel. It looks to be very similar to the previous crash.
> Only, in this case, I issued the command "uml_mconsole www-prod-1 sysrq
> s" in preparation for a backup (after issuing a "uml_mconsole $1 stop"
> command).
It's a particular scenario, and I've been doubtly about it, but indeed it
seems that it's safe (at least by design).
> I am including the first dump below. The rest of the dump can
> be found at:
> http://www.coherenthosting.com/prj/uml/miscellaneous/error.txt
>
> The time stamp at the top is in UTC, and specifies when the crash started.
>
> Tony
> -- time-stamp -- Feb/27/06 16:03:00 --
> SysRq : Emergency Sync
> events/0: page allocation failure. order:0, mode:0x20
0x20 == __GFP_HIGH == GFP_ATOMIC.
> 0829f510: [<0808a1f3>] __alloc_pages+0x274/0x286
> 0829f550: [<0808c7e5>] kmem_getpages+0x4a/0x9f
> 0829f564: [<0808d2b4>] cache_grow+0x96/0x122
> 0829f598: [<0808d493>] cache_alloc_refill+0x153/0x186
> 0829f5bc: [<0808d62d>] kmem_cache_alloc+0x34/0x44
> 0829f5d0: [<0808d134>] alloc_slabmgmt+0x20/0x45
> 0829f5e4: [<0808d2cc>] cache_grow+0xae/0x122
> 0829f614: [<0808d493>] cache_alloc_refill+0x153/0x186
> 0829f638: [<0808d6dc>] __kmalloc+0x5b/0x6d
> 0829f654: [<08171704>] __alloc_skb+0x52/0x129
> 0829f670: [<0806097e>] uml_net_rx+0x1e/0x13c
> 0829f690: [<08060ac0>] uml_net_interrupt+0x1f/0x91
> 0829f6d4: [<0808595b>] handle_IRQ_event+0x24/0x54
> 0829f6f8: [<080859de>] __do_IRQ+0x53/0x91
> 0829f70c: [<08057bc8>] do_IRQ+0x20/0x2c
> 0829f714: [<08057d5d>] sigio_handler+0xa5/0xcf
> 0829f72c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
> 0829f750: [<0806792e>] sig_handler+0xe/0x11
> 0829f860: [<08057cdf>] sigio_handler+0x27/0xcf
> 0829f87c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
> 0829f8a0: [<0806792e>] sig_handler+0xe/0x11
> 0829f9e8: [<0805a694>] set_signals+0xa4/0xc0
> 0829fa28: [<0805a694>] set_signals+0xa4/0xc0
> 0829fa64: [<0805a694>] set_signals+0xa4/0xc0
> 0829fa8c: [<0805a694>] set_signals+0xa4/0xc0
> 0829fae0: [<0805a641>] set_signals+0x51/0xc0
> 0829fb80: [<081607e6>] handle_sysrq+0x1c/0x20
> 0829fb94: [<0806258b>] mconsole_sysrq+0x5d/0x63
> 0829fbc0: [<0806220f>] mconsole_stop+0x40/0x78
> 0829fbcc: [<08061e1d>] mc_work_proc+0x38/0x4f
> 0829fbdc: [<0807c175>] worker_thread+0x188/0x205
> 0829fbe4: [<08061de5>] mc_work_proc+0x0/0x4f
> 0829fc08: [<0806bde4>] default_wake_function+0x0/0x12
> 0829fc2c: [<0807bfed>] worker_thread+0x0/0x205
> 0829fc30: [<0807f21b>] kthread+0x74/0x9a
>
> Mem-info:
> DMA per-cpu:
> cpu 0 hot: low 0, high 18, batch 3 used:2
> cpu 0 cold: low 0, high 6, batch 1 used:5
> DMA32 per-cpu: empty
> Normal per-cpu: empty
> HighMem per-cpu: empty
> Free pages: 376kB (0kB HighMem)
> Active:2985 inactive:8684 dirty:4 writeback:0 unstable:0 free:94
> slab:3428 mapped:933 pagetables:48
> DMA free:376kB min:1024kB low:1280kB high:1536kB active:11940kB
> inactive:34736kB present:65536kB pages_scanned:0 all_unreclaimable?
> no
> lowmem_reserve[]: 0 0 0 0
There are 376kB of free pages, so it's strange that it can't satisfy this
request.
[...]
> DMA: 0*4kB 1*8kB 1*16kB 1*32kB 1*64kB 0*128kB 1*256kB 0*512kB 0*1024kB
> 0*2048kB 0*4096kB = 376kB
> Swap cache: add 0, delete 0, find 0/0, race 0+0
> Free swap = 655340kB
> Total swap = 655340kB
> Free swap: 655340kB
> 16384 pages of RAM
> 952 reserved pages
> 11194 pages shared
> 0 pages swap cached
--
Inform me of my mistakes, so I can keep imitating Homer Simpson's "Doh!".
Paolo Giarrusso, aka Blaisorblade (Skype ID "PaoloGiarrusso", ICQ 215621894)
http://www.user-mode-linux.org/~blaisorblade
___________________________________
Yahoo! Mail: gratis 1GB per i messaggi e allegati da 10MB
http://mail.yahoo.it
-------------------------------------------------------
This SF.Net email is sponsored by xPML, a groundbreaking scripting language
that extends applications into web and mobile media. Attend the live webcast
and join the prime developer group breaking into this new coding territory!
http://sel.as-us.falkag.net/sel?cmd=lnk&kid=110944&bid=241720&dat=121642
_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
^ permalink raw reply [flat|nested] 6+ messages in thread
* RE: [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2
@ 2006-03-06 19:04 John Dykstra
0 siblings, 0 replies; 6+ messages in thread
From: John Dykstra @ 2006-03-06 19:04 UTC (permalink / raw)
To: Blaisorblade, user-mode-linux-devel; +Cc: Jeff Dike
I've been seeing MM red zone messages on 2.6.16-fc3 with Jeff's patchset
from a couple of weeks ago. I haven't reported them since that's a
pretty old version by now. Should I post the next bunch I see?
-- John
> -----Original Message-----
> From: user-mode-linux-devel-admin@lists.sourceforge.net
> [mailto:user-mode-linux-devel-admin@lists.sourceforge.net] On
> Behalf Of Blaisorblade
> Sent: Monday, March 06, 2006 11:11 AM
> To: user-mode-linux-devel@lists.sourceforge.net
> Cc: Anthony Brock; Jeff Dike
> Subject: Re: [uml-devel] Re: [uml-user] Crash and stability
> problems with 2.6.15.3-bs2
>
>
> On Monday 27 February 2006 17:27, Anthony Brock wrote:
> > Sorry for taking so long to get back on this. We've been busy with
> > some other issues. Some comments are inline below in
> addition to the
> > output from a new crash.
>
> > In the mean time, we've experienced a second crash with the
> > 2.6.15.1-bs1 kernel. It looks to be very similar to the previous
> > crash. Only, in this case, I issued the command "uml_mconsole
> > www-prod-1 sysrq s" in preparation for a backup (after issuing a
> > "uml_mconsole $1 stop" command).
>
> It's a particular scenario, and I've been doubtly about it,
> but indeed it
> seems that it's safe (at least by design).
>
> > I am including the first dump below. The rest of the dump can
> > be found at:
>
> > http://www.coherenthosting.com/prj/uml/miscellaneous/error.txt
> >
> > The time stamp at the top is in UTC, and specifies when the crash
> > started.
> >
> > Tony
>
> > -- time-stamp -- Feb/27/06 16:03:00 --
> > SysRq : Emergency Sync
> > events/0: page allocation failure. order:0, mode:0x20
> 0x20 == __GFP_HIGH == GFP_ATOMIC.
>
> > 0829f510: [<0808a1f3>] __alloc_pages+0x274/0x286
> > 0829f550: [<0808c7e5>] kmem_getpages+0x4a/0x9f
> > 0829f564: [<0808d2b4>] cache_grow+0x96/0x122
> > 0829f598: [<0808d493>] cache_alloc_refill+0x153/0x186
> > 0829f5bc: [<0808d62d>] kmem_cache_alloc+0x34/0x44
> > 0829f5d0: [<0808d134>] alloc_slabmgmt+0x20/0x45
> > 0829f5e4: [<0808d2cc>] cache_grow+0xae/0x122
> > 0829f614: [<0808d493>] cache_alloc_refill+0x153/0x186
> > 0829f638: [<0808d6dc>] __kmalloc+0x5b/0x6d
> > 0829f654: [<08171704>] __alloc_skb+0x52/0x129
> > 0829f670: [<0806097e>] uml_net_rx+0x1e/0x13c
> > 0829f690: [<08060ac0>] uml_net_interrupt+0x1f/0x91
> > 0829f6d4: [<0808595b>] handle_IRQ_event+0x24/0x54
> > 0829f6f8: [<080859de>] __do_IRQ+0x53/0x91
> > 0829f70c: [<08057bc8>] do_IRQ+0x20/0x2c
> > 0829f714: [<08057d5d>] sigio_handler+0xa5/0xcf
> > 0829f72c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
> > 0829f750: [<0806792e>] sig_handler+0xe/0x11
> > 0829f860: [<08057cdf>] sigio_handler+0x27/0xcf
> > 0829f87c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
> > 0829f8a0: [<0806792e>] sig_handler+0xe/0x11
> > 0829f9e8: [<0805a694>] set_signals+0xa4/0xc0
> > 0829fa28: [<0805a694>] set_signals+0xa4/0xc0
> > 0829fa64: [<0805a694>] set_signals+0xa4/0xc0
> > 0829fa8c: [<0805a694>] set_signals+0xa4/0xc0
> > 0829fae0: [<0805a641>] set_signals+0x51/0xc0
> > 0829fb80: [<081607e6>] handle_sysrq+0x1c/0x20
> > 0829fb94: [<0806258b>] mconsole_sysrq+0x5d/0x63
> > 0829fbc0: [<0806220f>] mconsole_stop+0x40/0x78
> > 0829fbcc: [<08061e1d>] mc_work_proc+0x38/0x4f
> > 0829fbdc: [<0807c175>] worker_thread+0x188/0x205
> > 0829fbe4: [<08061de5>] mc_work_proc+0x0/0x4f
> > 0829fc08: [<0806bde4>] default_wake_function+0x0/0x12
> > 0829fc2c: [<0807bfed>] worker_thread+0x0/0x205
> > 0829fc30: [<0807f21b>] kthread+0x74/0x9a
>
> >
> > Mem-info:
> > DMA per-cpu:
> > cpu 0 hot: low 0, high 18, batch 3 used:2
> > cpu 0 cold: low 0, high 6, batch 1 used:5
> > DMA32 per-cpu: empty
> > Normal per-cpu: empty
> > HighMem per-cpu: empty
> > Free pages: 376kB (0kB HighMem)
> > Active:2985 inactive:8684 dirty:4 writeback:0 unstable:0 free:94
> > slab:3428 mapped:933 pagetables:48 DMA free:376kB min:1024kB
> > low:1280kB high:1536kB active:11940kB inactive:34736kB
> present:65536kB
> > pages_scanned:0 all_unreclaimable? no
> > lowmem_reserve[]: 0 0 0 0
>
> There are 376kB of free pages, so it's strange that it can't
> satisfy this
> request.
>
> [...]
> > DMA: 0*4kB 1*8kB 1*16kB 1*32kB 1*64kB 0*128kB 1*256kB
> 0*512kB 0*1024kB
> > 0*2048kB 0*4096kB = 376kB
>
> > Swap cache: add 0, delete 0, find 0/0, race 0+0
> > Free swap = 655340kB
> > Total swap = 655340kB
> > Free swap: 655340kB
> > 16384 pages of RAM
>
> > 952 reserved pages
> > 11194 pages shared
> > 0 pages swap cached
>
> --
> Inform me of my mistakes, so I can keep imitating Homer
> Simpson's "Doh!". Paolo Giarrusso, aka Blaisorblade (Skype ID
> "PaoloGiarrusso", ICQ 215621894)
> http://www.user-mode-linux.org/~blaisorblade
>
>
>
>
>
>
> ___________________________________
> Yahoo! Mail: gratis 1GB per i messaggi e allegati da 10MB
> http://mail.yahoo.it
>
>
>
> -------------------------------------------------------
> This SF.Net email is sponsored by xPML, a groundbreaking
> scripting language that extends applications into web and
> mobile media. Attend the live webcast and join the prime
> developer group breaking into this new coding territory!
http://sel.as-us.falkag.net/sel?cmd=lnk&kid=110944&bid=241720&dat=121642
_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
-------------------------------------------------------
This SF.Net email is sponsored by xPML, a groundbreaking scripting language
that extends applications into web and mobile media. Attend the live webcast
and join the prime developer group breaking into this new coding territory!
http://sel.as-us.falkag.net/sel?cmd_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
^ permalink raw reply [flat|nested] 6+ messages in thread
* Re: [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2
2006-03-06 17:10 ` Blaisorblade
@ 2006-03-06 19:33 ` Anthony Brock
2006-03-06 22:09 ` Blaisorblade
0 siblings, 1 reply; 6+ messages in thread
From: Anthony Brock @ 2006-03-06 19:33 UTC (permalink / raw)
To: Blaisorblade; +Cc: user-mode-linux-devel, Jeff Dike
> On Monday 06 March 2006 9:11, Blaisorblade wrote:
> On Monday 27 February 2006 17:27, Anthony Brock wrote:
> > In the mean time, we've experienced a second crash with the
> > 2.6.15.1-bs1 kernel. It looks to be very similar to the previous crash.
> > Only, in this case, I issued the command "uml_mconsole www-prod-1 sysrq
> > s" in preparation for a backup (after issuing a "uml_mconsole $1 stop"
> > command).
>
> It's a particular scenario, and I've been doubtly about it, but
> indeed it seems that it's safe (at least by design).
I apologize. However, I'm having difficulty understanding your meaning here.
> > Mem-info:
> > DMA per-cpu:
> > cpu 0 hot: low 0, high 18, batch 3 used:2
> > cpu 0 cold: low 0, high 6, batch 1 used:5
> > DMA32 per-cpu: empty
> > Normal per-cpu: empty
> > HighMem per-cpu: empty
> > Free pages: 376kB (0kB HighMem)
> > Active:2985 inactive:8684 dirty:4 writeback:0 unstable:0 free:94
> > slab:3428 mapped:933 pagetables:48
> > DMA free:376kB min:1024kB low:1280kB high:1536kB active:11940kB
> > inactive:34736kB present:65536kB pages_scanned:0 all_unreclaimable?
> > no
> > lowmem_reserve[]: 0 0 0 0
>
> There are 376kB of free pages, so it's strange that it can't satisfy
> this request.
I agree. However, this is the only instance where we've been able to
reproduce a crash with a non-bs2 patched kernel. Other than this, the
bs1 patches have been very stable.
What can we do to further diagnose this? Alternatively, should we try a
different guest kernel version or set of patches?
Tony
-------------------------------------------------------
This SF.Net email is sponsored by xPML, a groundbreaking scripting language
that extends applications into web and mobile media. Attend the live webcast
and join the prime developer group breaking into this new coding territory!
http://sel.as-us.falkag.net/sel?cmd=lnk&kid=110944&bid=241720&dat=121642
_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
^ permalink raw reply [flat|nested] 6+ messages in thread
* Re: [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2
2006-03-06 19:33 ` Anthony Brock
@ 2006-03-06 22:09 ` Blaisorblade
0 siblings, 0 replies; 6+ messages in thread
From: Blaisorblade @ 2006-03-06 22:09 UTC (permalink / raw)
To: Anthony Brock; +Cc: user-mode-linux-devel, Jeff Dike
On Monday 06 March 2006 20:33, Anthony Brock wrote:
> > On Monday 06 March 2006 9:11, Blaisorblade wrote:
> >
> > On Monday 27 February 2006 17:27, Anthony Brock wrote:
> > > In the mean time, we've experienced a second crash with the
> > > 2.6.15.1-bs1 kernel. It looks to be very similar to the previous crash.
> > > Only, in this case, I issued the command "uml_mconsole www-prod-1 sysrq
> > > s" in preparation for a backup (after issuing a "uml_mconsole $1 stop"
> > > command).
> >
> > It's a particular scenario, and I've been doubtly about it, but
> > indeed it seems that it's safe (at least by design).
>
> I apologize. However, I'm having difficulty understanding your meaning
> here.
I.e. I had doubt about "stop + sysrq s" being safe; today when I looked at the
code and at what I remembered of my doubt, I concluded I was wrong.
> > > Mem-info:
> > > DMA per-cpu:
> > > cpu 0 hot: low 0, high 18, batch 3 used:2
> > > cpu 0 cold: low 0, high 6, batch 1 used:5
> > > DMA32 per-cpu: empty
> > > Normal per-cpu: empty
> > > HighMem per-cpu: empty
> > > Free pages: 376kB (0kB HighMem)
> > > Active:2985 inactive:8684 dirty:4 writeback:0 unstable:0 free:94
> > > slab:3428 mapped:933 pagetables:48
> > > DMA free:376kB min:1024kB low:1280kB high:1536kB active:11940kB
> > > inactive:34736kB present:65536kB pages_scanned:0 all_unreclaimable?
> > > no
> > > lowmem_reserve[]: 0 0 0 0
> >
> > There are 376kB of free pages, so it's strange that it can't satisfy
> > this request.
>
> I agree. However, this is the only instance where we've been able to
> reproduce a crash with a non-bs2 patched kernel. Other than this, the
> bs1 patches have been very stable.
>
> What can we do to further diagnose this? Alternatively, should we try a
> different guest kernel version or set of patches?
>
> Tony
--
Inform me of my mistakes, so I can keep imitating Homer Simpson's "Doh!".
Paolo Giarrusso, aka Blaisorblade (Skype ID "PaoloGiarrusso", ICQ 215621894)
http://www.user-mode-linux.org/~blaisorblade
___________________________________
Yahoo! Mail: gratis 1GB per i messaggi e allegati da 10MB
http://mail.yahoo.it
-------------------------------------------------------
This SF.Net email is sponsored by xPML, a groundbreaking scripting language
that extends applications into web and mobile media. Attend the live webcast
and join the prime developer group breaking into this new coding territory!
http://sel.as-us.falkag.net/sel?cmd=lnk&kid=110944&bid=241720&dat=121642
_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
^ permalink raw reply [flat|nested] 6+ messages in thread
* RE: [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2
@ 2006-03-06 22:19 John Dykstra
0 siblings, 0 replies; 6+ messages in thread
From: John Dykstra @ 2006-03-06 22:19 UTC (permalink / raw)
To: Blaisorblade, user-mode-linux-devel; +Cc: Jeff Dike
Slab corruption: start=0922713c, len=512
Redzone: 0x5a2cf071/0x5a2cf071.
Last user: [<080b57ba>](free_pipe_info+0x3e/0x58)
160: 00 4e 68 08 6b 6b 6b 6b 6b 6b 6b 6b 6b 6b 6b 6b
Next obj: start=09227348, len=512
Redzone: 0x170fc2a5/0x170fc2a5.
Last user: [<080b57f2>](pipe_new+0x1e/0x64)
000: 18 10 17 09 18 10 17 09 00 00 00 00 02 00 00 00
010: 20 56 86 08 04 00 00 00 00 00 00 00 00 00 00 00
In the past when I've seen this, it has seemed associated with hostfs.
-- John
> -----Original Message-----
> From: Dykstra, John [MNBL1:0S03:EXCH]
> Sent: Monday, March 06, 2006 1:05 PM
> To: 'Blaisorblade'; user-mode-linux-devel@lists.sourceforge.net
> Cc: Jeff Dike
> Subject: RE: [uml-devel] Re: [uml-user] Crash and stability
> problems with 2.6.15.3-bs2
>
>
> I've been seeing MM red zone messages on 2.6.16-fc3 with
> Jeff's patchset from a couple of weeks ago. I haven't
> reported them since that's a pretty old version by now.
> Should I post the next bunch I see?
>
> -- John
>
> > -----Original Message-----
> > From: user-mode-linux-devel-admin@lists.sourceforge.net
> > [mailto:user-mode-linux-devel-admin@lists.sourceforge.net] On
> > Behalf Of Blaisorblade
> > Sent: Monday, March 06, 2006 11:11 AM
> > To: user-mode-linux-devel@lists.sourceforge.net
> > Cc: Anthony Brock; Jeff Dike
> > Subject: Re: [uml-devel] Re: [uml-user] Crash and stability
> > problems with 2.6.15.3-bs2
> >
> >
> > On Monday 27 February 2006 17:27, Anthony Brock wrote:
> > > Sorry for taking so long to get back on this. We've been busy with
> > > some other issues. Some comments are inline below in
> > addition to the
> > > output from a new crash.
> >
> > > In the mean time, we've experienced a second crash with the
> > > 2.6.15.1-bs1 kernel. It looks to be very similar to the previous
> > > crash. Only, in this case, I issued the command "uml_mconsole
> > > www-prod-1 sysrq s" in preparation for a backup (after issuing a
> > > "uml_mconsole $1 stop" command).
> >
> > It's a particular scenario, and I've been doubtly about it,
> > but indeed it
> > seems that it's safe (at least by design).
> >
> > > I am including the first dump below. The rest of the dump can be
> > > found at:
> >
> > > http://www.coherenthosting.com/prj/uml/miscellaneous/error.txt
> > >
> > > The time stamp at the top is in UTC, and specifies when the crash
> > > started.
> > >
> > > Tony
> >
> > > -- time-stamp -- Feb/27/06 16:03:00 --
> > > SysRq : Emergency Sync
> > > events/0: page allocation failure. order:0, mode:0x20
> > 0x20 == __GFP_HIGH == GFP_ATOMIC.
> >
> > > 0829f510: [<0808a1f3>] __alloc_pages+0x274/0x286
> > > 0829f550: [<0808c7e5>] kmem_getpages+0x4a/0x9f
> > > 0829f564: [<0808d2b4>] cache_grow+0x96/0x122
> > > 0829f598: [<0808d493>] cache_alloc_refill+0x153/0x186
> > > 0829f5bc: [<0808d62d>] kmem_cache_alloc+0x34/0x44
> > > 0829f5d0: [<0808d134>] alloc_slabmgmt+0x20/0x45
> > > 0829f5e4: [<0808d2cc>] cache_grow+0xae/0x122
> > > 0829f614: [<0808d493>] cache_alloc_refill+0x153/0x186
> > > 0829f638: [<0808d6dc>] __kmalloc+0x5b/0x6d
> > > 0829f654: [<08171704>] __alloc_skb+0x52/0x129
> > > 0829f670: [<0806097e>] uml_net_rx+0x1e/0x13c
> > > 0829f690: [<08060ac0>] uml_net_interrupt+0x1f/0x91
> > > 0829f6d4: [<0808595b>] handle_IRQ_event+0x24/0x54
> > > 0829f6f8: [<080859de>] __do_IRQ+0x53/0x91
> > > 0829f70c: [<08057bc8>] do_IRQ+0x20/0x2c
> > > 0829f714: [<08057d5d>] sigio_handler+0xa5/0xcf
> > > 0829f72c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
> > > 0829f750: [<0806792e>] sig_handler+0xe/0x11
> > > 0829f860: [<08057cdf>] sigio_handler+0x27/0xcf
> > > 0829f87c: [<0805e435>] sig_handler_common_skas+0xa5/0xbe
> > > 0829f8a0: [<0806792e>] sig_handler+0xe/0x11
> > > 0829f9e8: [<0805a694>] set_signals+0xa4/0xc0
> > > 0829fa28: [<0805a694>] set_signals+0xa4/0xc0
> > > 0829fa64: [<0805a694>] set_signals+0xa4/0xc0
> > > 0829fa8c: [<0805a694>] set_signals+0xa4/0xc0
> > > 0829fae0: [<0805a641>] set_signals+0x51/0xc0
> > > 0829fb80: [<081607e6>] handle_sysrq+0x1c/0x20
> > > 0829fb94: [<0806258b>] mconsole_sysrq+0x5d/0x63
> > > 0829fbc0: [<0806220f>] mconsole_stop+0x40/0x78
> > > 0829fbcc: [<08061e1d>] mc_work_proc+0x38/0x4f
> > > 0829fbdc: [<0807c175>] worker_thread+0x188/0x205
> > > 0829fbe4: [<08061de5>] mc_work_proc+0x0/0x4f
> > > 0829fc08: [<0806bde4>] default_wake_function+0x0/0x12
> > > 0829fc2c: [<0807bfed>] worker_thread+0x0/0x205
> > > 0829fc30: [<0807f21b>] kthread+0x74/0x9a
> >
> > >
> > > Mem-info:
> > > DMA per-cpu:
> > > cpu 0 hot: low 0, high 18, batch 3 used:2
> > > cpu 0 cold: low 0, high 6, batch 1 used:5
> > > DMA32 per-cpu: empty
> > > Normal per-cpu: empty
> > > HighMem per-cpu: empty
> > > Free pages: 376kB (0kB HighMem)
> > > Active:2985 inactive:8684 dirty:4 writeback:0 unstable:0 free:94
> > > slab:3428 mapped:933 pagetables:48 DMA free:376kB min:1024kB
> > > low:1280kB high:1536kB active:11940kB inactive:34736kB
> > present:65536kB
> > > pages_scanned:0 all_unreclaimable? no
> > > lowmem_reserve[]: 0 0 0 0
> >
> > There are 376kB of free pages, so it's strange that it can't
> > satisfy this
> > request.
> >
> > [...]
> > > DMA: 0*4kB 1*8kB 1*16kB 1*32kB 1*64kB 0*128kB 1*256kB
> > 0*512kB 0*1024kB
> > > 0*2048kB 0*4096kB = 376kB
> >
> > > Swap cache: add 0, delete 0, find 0/0, race 0+0
> > > Free swap = 655340kB
> > > Total swap = 655340kB
> > > Free swap: 655340kB
> > > 16384 pages of RAM
> >
> > > 952 reserved pages
> > > 11194 pages shared
> > > 0 pages swap cached
> >
> > --
> > Inform me of my mistakes, so I can keep imitating Homer
> > Simpson's "Doh!". Paolo Giarrusso, aka Blaisorblade (Skype ID
> > "PaoloGiarrusso", ICQ 215621894)
> > http://www.user-mode-linux.org/~blaisorblade
> >
> >
> >
> >
> >
> >
> > ___________________________________
> > Yahoo! Mail: gratis 1GB per i messaggi e allegati da 10MB
> > http://mail.yahoo.it
> >
> >
> >
> > -------------------------------------------------------
> > This SF.Net email is sponsored by xPML, a groundbreaking
> > scripting language that extends applications into web and
> > mobile media. Attend the live webcast and join the prime
> > developer group breaking into this new coding territory!
> http://sel.as-us.falkag.net/sel?cmd=lnk&kid=110944&bid=241720&
dat=121642
_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
-------------------------------------------------------
This SF.Net email is sponsored by xPML, a groundbreaking scripting language
that extends applications into web and mobile media. Attend the live webcast
and join the prime developer group breaking into this new coding territory!
http://sel.as-us.falkag.net/sel?cmd_______________________________________________
User-mode-linux-devel mailing list
User-mode-linux-devel@lists.sourceforge.net
https://lists.sourceforge.net/lists/listinfo/user-mode-linux-devel
^ permalink raw reply [flat|nested] 6+ messages in thread
end of thread, other threads:[~2006-03-06 22:56 UTC | newest]
Thread overview: 6+ messages (download: mbox.gz follow: Atom feed
-- links below jump to the message on this page --
[not found] <LEEIJGBJEEABAGOFFGAHEEKCDEAA.brocka@sterlingcgi.com>
[not found] ` <CBELIFGMBIEPIEOFCJFAGEMACAAA.brocka@sterlingcgi.com>
[not found] ` <20060215183943.GA7120@ccure.user-mode-linux.org>
[not found] ` <20060215130933.mn5mlmmh8hkowgsk@www.scgiservices.com>
[not found] ` <20060217205204.GC2889@ccure.user-mode-linux.org>
2006-02-27 16:27 ` [uml-devel] Re: [uml-user] Crash and stability problems with 2.6.15.3-bs2 Anthony Brock
2006-03-06 17:10 ` Blaisorblade
2006-03-06 19:33 ` Anthony Brock
2006-03-06 22:09 ` Blaisorblade
2006-03-06 19:04 John Dykstra
-- strict thread matches above, loose matches on Subject: below --
2006-03-06 22:19 John Dykstra
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox