[patch 0/8] exec/exit: POSIX timer related bugfixes and related cleanups

From: Thomas Gleixner

Date: Fri Sep 04 2026 - 08:15:36 EST


Recent findings from Hyunwoo unearthed two bugs in handling POSIX timers on
exec().

The relevant patches, reports and discussions can be found here:

https://patch.msgid.link/aok1rdkBgZsynHZB@v4bel
https://patch.msgid.link/ao7Q8miiuLAPVnWv@v4bel

TLDR:

Both problems are related to non-leader exec(). POSIX CPU timers which are
targeted at tasks hold a pid reference of the target task, which is used to
look up the task in the related POSIX timer operations.

The non-leader exec() switches the TID of the old and the new leader, which
obviously invalidates these references for pid_task(PIDTYPE_PID) lookups.

This causes UAFs due to the resulting list corruptions or premature freeing
without removing the underlying POSIX CPU timers from the involved tasks.

The first issue which corrupts the signal pending list is solved by:

- Preventing the queueing of per task signals on a task which has
PF_EXITING set.

- Protecting the unlocked setting of PF_EXITING in exit_signals() with
sighand lock.

- Flushing all per task signals right in exit_signals()

The second issue which keeps the POSIX CPU timers queued on the new leader
is solved by:

- Moving the exec related POSIX timer cleanup right after de_thread()
which ensures that the timers queued in new_leader::posix_cputimers
are removed before the underlying POSIX timers are deleted.

After looking deeper at the exit() handling it turned out that the POSIX
timer cleanups can be done early in do_exit() instead of delaying them
until release_task().

The reason for this late cleanup is that POSIX CPU timers can be created,
rearmed and deleted as long as a task is visible, i.e. the pid is hashed
and sighand is not NULL. This allows to retrieve information from the timer
up to the point where the task is gone for real and that can't be changed
easily as that'd be a user visible change.

But once PF_EXITING is set on a task the task does not longer expire POSIX
CPU timers. So it makes no sense that the timers stay queued in
task::posix_cputimers after that point.

The only thing which needs to be prevented is that timers are requeued on
task::posix_cputimers once PF_EXITING is set or requeued on
signal::posix_cputimers when PF_EXITING is set and signal::live is zero,
which indicates that the thread group is dead.

With that solved the timers can be dequeued from task::posix_cputimer
pending when a task exits and from signal::posix_cputimer pending once the
threadgroup reaches the dead state, i.e. signal::live goes to zero.

The series applies on 7.3-rc1 and is avalaible from git:

git://git.kernel.org/pub/scm/linux/kernel/git/tglx/devel.git posix-timers

Thanks,

tglx
---
fs/exec.c | 18 ++++---
include/linux/posix-timers.h | 39 ++--------------
include/linux/sched/task.h | 1
kernel/exit.c | 24 +++------
kernel/signal.c | 70 ++++++++++++++++------------
kernel/time/posix-cpu-timers.c | 99 +++++++++++++++++++++++++++++++++++++----
kernel/time/posix-timers.c | 26 ++++++++--
kernel/time/posix-timers.h | 3 +
8 files changed, 179 insertions(+), 101 deletions(-)