Re: [PATCH tty v11 1/2] serial: 8250: Switch to nbcon console, take 2

From: Petr Mladek

Date: Tue Aug 18 2026 - 11:53:38 EST


On Tue 2026-08-18 13:19:12, Jon Hunter wrote:
> Hi Petr,
>
> On 18/08/2026 08:51, Petr Mladek wrote:
>
> ...
>
> > > This change is causing a boot regression for our Tegra20 and Tegra30
> > > platforms. Reverting this on top of -next fixes the issue. Previously with
> > > V5 I did not see a boot issue only an issue in suspend. So far I have not
> > > had chance to dig any further.
> >
> > Interesting.
> >
> > Another clue, mentioned in the v9 thread [1], is that the boot
> > regression does not happen with v11 when "keep_bootcon" option
> > is used.
> >
> > The "keep_bootcon" option causes that the boot console driver stays
> > registered even when the full featured driver gets registered
> > later.
> >
> > The most important effect is that the printk kthreads can't
> > be used as long as any boot console driver is registered.
> > All drivers need to be called in the legacy loop in this case.
> > There are two reasons for this:
> >
> > 1. Boot console drivers are synchronized only by
> > the legacy console_lock (console_sem). port->lock
> > is available only for the full featured driver.
> >
> > 2. There is no easy way to match boot console and
> > full featured console drivers working on the same
> > HW.
> >
> > So, the regression seems to happen when the printk kthreads
> > start being used.
> >
> > Jon, could you please share the full log when "keep_bootcon"
> > is used?
>
>
> Yes absolutely. You can find the boot log here [0]. So far nothing really
> stands out to me but let me know if you see anything.
>
> [0] https://pastebin.com/FhQVSqfy

Thanks for the log.

One or two things look strange/important to me.
But let me show all important parts:

[ 0.000000] Kernel command line: console=ttyS0,115200n8 console=tty1 earlycon ignore_loglevel root=/dev/nfs rw ip=192.168.99.2:192.168.99.1:192.168.99.1:255.255.255.0::eth0:off nfsroot=192.168.99.1:/home/ausvrl81292/nfsroot,tcp rootwait keep_bootcon

The last "console=" parameter is "console=tty1". It is a so called
preferred console. It has several effects:

+ it should get associated with /dev/console

+ it does not replay the log from the beginning when
registered. Only newer messages are shown.

+ Boot consoles should get unregistered when this console
gets registered (unless keep_bootcon is defined).

Note that "ttyS0" is _not_ the _preferred_console. As a result:

+ it will replay all messages when registered

+ boot console won't get unregistered when this one
is registered


Now, the ordering is:

[ 0.000000] earlycon: uart0 MMIO:0x70006300 (options '115200n8')
[ 0.000000] printk: legacy bootconsole [uart0] enabled

First, earlycon is registered thanks because of the "earlycon"
parameter.

[ 0.036461] Console: colour dummy device 80x30
[ 0.041020] printk: legacy console [tty1] enabled

Second, the graphical "tty1" gets registered because
of the "console=tty1" parameter.

Normally, the boot console should get unregistered at
this point. But it stays because of the "keep_bootcon"
parameter.

[ 0.645646] Serial: 8250/16550 driver, 4 ports, IRQ sharing disabled
[ 0.655230] printk: console [ttyS0] disabled

IMPORTANT: This is the weird thing! I do not understand why "ttySO"
gets disabled when it has not been registered yet.

[ 0.659964] 70006300.serial: ttyS0 MMIO32:0x70006300 (irq = 51, base_baud = 13500000) is a Tegra
[ 0.000000] Booting Linux on physical CPU 0x0
[ 0.669026] printk: console [ttyS0] enabled

The real "ttyS0" driver has been registered (added to
console_list) and the legacy loop started flushing
the messages in console_unlock().

The real "ttyS0" console driver started emitting messages
from the beginning.

The boot console driver emits only the newly added messages
"printk: console [ttyS0] enabled".

[ 0.000000] Linux version 7.2.0-next-20260817 (jonathanh@build-jonathanh-noble-20260527) (arm-buildroot-linux-gnueabihf-gcc.br_real (Buildroot 2021.11-11272-ge2962af) 13.2.0, GNU ld (GNU Binutils) 2.42) #15 SMP PREEMPT Tue Aug 18 04:56:56 UTC 2026
[ 0.000000] CPU: ARMv7 Processor [411fc090] revision 0 (ARMv7), cr=10c5387d

The real "ttyS0" console driver replays the entire log.

[ 0.659964] 70006300.serial: ttyS0 MMIO32:0x70006300 (irq = 51, base_baud = 13500000) is a Tegra
[ 0.669026] printk: console [ttyS0] enabled


And then all messages are emitted twice (by the boot console
driver and by the real console driver:

[ 1.574761] loop: module loaded
[ 1.574761] loop: module loaded
[ 1.584820] CAN device driver interface
[ 1.584820] CAN device driver interface

The real console driver would normally emit these messages
from the printk kthread. But it does it in the legacy loop
because the boot console driver is still registered.

[ 2.904178] ------------[ cut here ]------------
[ 2.904178] ------------[ cut here ]------------
[ 2.913507] WARNING: drivers/soc/tegra/pmc.c:4999 at tegra_pmc_enter_suspend_mode+0x168/0x178, CPU#0: swapper/0/0
[ 2.913507] WARNING: drivers/soc/tegra/pmc.c:4999 at tegra_pmc_enter_suspend_mode+0x168/0x178, CPU#0: swapper/0/0

IMPORTANT: This is a warning. It is printed with
NBCON_PRIO_EMERGENCY. These messages would normally
get flushed by nbcon_atomic_flush_pending() directly
from printk() using con->write_atomic(). It would
take over the console ownership from the kthread
when needed.

In this particular log, it is emitted from the legacy
loop in console_unlock() because the boot console
is still registered.


Summary:

Almost everything works as expected except for:

1. I am not sure why "printk: console [ttyS0] disabled" is printed.
It does not make any sense to me.

2. The WARNING would be handled with NBCON_PRIO_EMERGENCY.
I wonder if this warning happened also with "v5" of this
patchset.

Why is the WARNING important?

If the WARNING happened also with v5 of this patchset
then it tested emergency mode as well. But the system
booted with v5. So that a difference between v5 and v11
patchset might be important.

If The WARNING did _not_ happen with v5 then we probably did
not test the emergency mode in this version. So that
the problem might be in the emergency mode handling.



Ideas for testing:

1. I wonder if adding a WARN() with v5 of this patchset
would make v5 fail as well.

2. If wonder if boot_delay=10 makes any difference. It might
prevent some races.

Best Regards,
Petr