Received: by 2002:a05:6359:6284:b0:131:369:b2a3 with SMTP id se4csp4738463rwb; Tue, 8 Aug 2023 13:03:50 -0700 (PDT) X-Google-Smtp-Source: AGHT+IEbC/c/gF5RuEhCybiqSMpcNiFCMkptHKD74PBHxGUfknnTXAIF0FptFz6c8mBefVAklC4R X-Received: by 2002:aa7:88d2:0:b0:687:4fa4:7f2e with SMTP id k18-20020aa788d2000000b006874fa47f2emr508371pff.13.1691525030280; Tue, 08 Aug 2023 13:03:50 -0700 (PDT) ARC-Seal: i=1; a=rsa-sha256; t=1691525030; cv=none; d=google.com; s=arc-20160816; b=zDXPN46/zKLFm4tph2gnY5QpTDjPacMfthdJvYLynJxKLtvXoSBS1Zj+gSTS4cwuhQ Q/0TEos7+Ab0Z9xHHTB6ZcvPUBR0b1wx0/H16IM5n1GAkkIgRp6dd/NuUYUytJFEEKEL m+w3yFcTEZ6+8p7DPQiGN09uxqadeh5ZOzvsIVuPc31p7DIw+EF3hzPlbMH8swD9lrrk k44RizIL08Mrto12X9F4E+HB1jmBUyJ8/PleLzL0GuTDHoEnF8KvS0LE8X25ydtWM5dE 8im1lb8Mb1Nl+S4xU2j/FgjsMqyWCF3joij63d/qNT4DxD7idC1SMK14Q2sE0gykmaZJ t+sg== ARC-Message-Signature: i=1; a=rsa-sha256; c=relaxed/relaxed; d=google.com; s=arc-20160816; h=list-id:precedence:content-transfer-encoding:mime-version :references:in-reply-to:message-id:date:subject:cc:to:from :dkim-signature; bh=65gjSxQ4NhG6oWq6CPJZFjDqu+AonGruTrkgt8zaBwQ=; fh=oQJfh+Vr4T6Dbl3+mbv1Iadpbiax+AEVFgc4RYmtzn0=; b=rBSU83joHK6xfJrpa5RhsM6XEZtMohdgfN5+klT9LA/HmY5/zW/HHxFZQv4qxA5cWr /kIMOFV4HI6Ri7Mil9ZB2gnJM6JdVfO7fAZGEHMUoHxNnYOC5nzCa3I9/Zp/j3Z4sUKu F7QDQ4yjgnStqvDmChqb6yrcpkawPvoV9vXC/bJNjk4jMeXOdBQN2K1yMv3o13Z+RUyA uMjozMSNrMfYFypwloUcRxFLOAmkMkFxwTrI/19fsg4c+A+CayDfap/9sleewQmaUtjr xZ2u3sV8wkod9saQ88dvWHys+KsziRko+F4vKX6W26Ro9vbrfGJacLdFKEwzsS3NpoKL RWNQ== ARC-Authentication-Results: i=1; mx.google.com; dkim=pass header.i=@collabora.com header.s=mail header.b="l+/6PNCb"; spf=pass (google.com: domain of linux-kernel-owner@vger.kernel.org designates 2620:137:e000::1:20 as permitted sender) smtp.mailfrom=linux-kernel-owner@vger.kernel.org; dmarc=pass (p=QUARANTINE sp=QUARANTINE dis=NONE) header.from=collabora.com Return-Path: Received: from out1.vger.email (out1.vger.email. [2620:137:e000::1:20]) by mx.google.com with ESMTP id bt25-20020a056a00439900b00679a6ce03f6si7792825pfb.59.2023.08.08.13.03.37; Tue, 08 Aug 2023 13:03:50 -0700 (PDT) Received-SPF: pass (google.com: domain of linux-kernel-owner@vger.kernel.org designates 2620:137:e000::1:20 as permitted sender) client-ip=2620:137:e000::1:20; Authentication-Results: mx.google.com; dkim=pass header.i=@collabora.com header.s=mail header.b="l+/6PNCb"; spf=pass (google.com: domain of linux-kernel-owner@vger.kernel.org designates 2620:137:e000::1:20 as permitted sender) smtp.mailfrom=linux-kernel-owner@vger.kernel.org; dmarc=pass (p=QUARANTINE sp=QUARANTINE dis=NONE) header.from=collabora.com Received: (majordomo@vger.kernel.org) by vger.kernel.org via listexpand id S232161AbjHHSiB (ORCPT + 99 others); Tue, 8 Aug 2023 14:38:01 -0400 Received: from lindbergh.monkeyblade.net ([23.128.96.19]:34772 "EHLO lindbergh.monkeyblade.net" rhost-flags-OK-OK-OK-OK) by vger.kernel.org with ESMTP id S232037AbjHHSha (ORCPT ); Tue, 8 Aug 2023 14:37:30 -0400 Received: from madras.collabora.co.uk (madras.collabora.co.uk [IPv6:2a00:1098:0:82:1000:25:2eeb:e5ab]) by lindbergh.monkeyblade.net (Postfix) with ESMTPS id 99CC014FDA; Tue, 8 Aug 2023 09:39:03 -0700 (PDT) Received: from localhost.localdomain (unknown [59.103.218.230]) (using TLSv1.3 with cipher TLS_AES_256_GCM_SHA384 (256/256 bits) key-exchange X25519 server-signature RSA-PSS (4096 bits) server-digest SHA256) (No client certificate requested) (Authenticated sender: usama.anjum) by madras.collabora.co.uk (Postfix) with ESMTPSA id 7C0AB66071FF; Tue, 8 Aug 2023 11:43:36 +0100 (BST) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/simple; d=collabora.com; s=mail; t=1691491421; bh=8GLPcDiIM54Mi7lYvOuYJH3K/M4uc0JKwxg63FI/vKw=; h=From:To:Cc:Subject:Date:In-Reply-To:References:From; b=l+/6PNCbz0uJwFktAVC2zKgUcdMUfX5EjHPHkvuzDBQQGObPhV6Gqpfa+csvaBUFO nVZLcSDkH3jJn1pvqMsn653dePEUTRwTgshcD9kCLJkl8CCAgRoNPwx4sjU3YfZLKk T0wdW83ckyCwsSPgigCyDZLkoka1qGV/qILxxx/UrxklXYBcHxbhhJbLk9Kxt49WzJ wkj8TUDQvEqeLyIsshDktKwZyaqK1ahNYmUINqZv6WrqbSyrA82N76bpm7mMDhOVGZ 9Oq+dO2dS7jDEMnN8VAv11vvSmtcd3PCj8FgwQE1DLJPUPiJ4boDbpbyjVDc0erSTC /1iMkLFc9srgg== From: Muhammad Usama Anjum To: Peter Xu , David Hildenbrand , Andrew Morton , =?UTF-8?q?Micha=C5=82=20Miros=C5=82aw?= , Andrei Vagin , Danylo Mocherniuk , Paul Gofman , Cyrill Gorcunov , Mike Rapoport , Nadav Amit Cc: Alexander Viro , Shuah Khan , Christian Brauner , Yang Shi , Vlastimil Babka , "Liam R . Howlett" , Yun Zhou , Suren Baghdasaryan , Alex Sierra , Muhammad Usama Anjum , Matthew Wilcox , Pasha Tatashin , Axel Rasmussen , "Gustavo A . R . Silva" , Dan Williams , linux-kernel@vger.kernel.org, linux-fsdevel@vger.kernel.org, linux-mm@kvack.org, linux-kselftest@vger.kernel.org, Greg KH , kernel@collabora.com Subject: [PATCH v27 3/6] fs/proc/task_mmu: Add fast paths to get/clear PAGE_IS_WRITTEN flag Date: Tue, 8 Aug 2023 15:43:06 +0500 Message-Id: <20230808104309.357852-4-usama.anjum@collabora.com> X-Mailer: git-send-email 2.39.2 In-Reply-To: <20230808104309.357852-1-usama.anjum@collabora.com> References: <20230808104309.357852-1-usama.anjum@collabora.com> MIME-Version: 1.0 Content-Transfer-Encoding: 8bit X-Spam-Status: No, score=-2.1 required=5.0 tests=BAYES_00,DKIM_SIGNED, DKIM_VALID,DKIM_VALID_AU,DKIM_VALID_EF,SPF_HELO_NONE,SPF_PASS, URIBL_BLOCKED autolearn=ham autolearn_force=no version=3.4.6 X-Spam-Checker-Version: SpamAssassin 3.4.6 (2021-04-09) on lindbergh.monkeyblade.net Precedence: bulk List-ID: X-Mailing-List: linux-kernel@vger.kernel.org Adding fast code paths to handle specifically only get and/or clear operation of PAGE_IS_WRITTEN, increases its performance by 0-35%. The results of some test cases are given below: Test-case-1 t1 = (Get + WP) time t2 = WP time t1 t2 Without this patch: 140-170mcs 90-115mcs With this patch: 110mcs 80mcs Worst case diff: 35% faster 30% faster Test-case-2 t3 = atomic Get and WP t3 Without this patch: 120-140mcs With this patch: 100-110mcs Worst case diff: 21% faster Signed-off-by: Muhammad Usama Anjum --- The test to measure the performance can be found: https://is.gd/FtSKcD 8 8192 3 1 0 and 8 8192 3 1 1 arguments have been used to produce the above mentioned results. --- fs/proc/task_mmu.c | 38 ++++++++++++++++++++++++++++++++++++++ 1 file changed, 38 insertions(+) diff --git a/fs/proc/task_mmu.c b/fs/proc/task_mmu.c index 6131dc04cceb3..2c3349b76b067 100644 --- a/fs/proc/task_mmu.c +++ b/fs/proc/task_mmu.c @@ -2107,6 +2107,43 @@ static int pagemap_scan_pmd_entry(pmd_t *pmd, unsigned long start, return 0; } + if (!p->vec_out) { + /* Fast path for performing exclusive WP */ + for (addr = start; addr != end; pte++, addr += PAGE_SIZE) { + if (pte_uffd_wp(ptep_get(pte))) + continue; + make_uffd_wp_pte(vma, addr, pte); + if (!flush) { + start = addr; + flush = true; + } + } + goto flush_and_return; + } + + if (!p->arg.category_anyof_mask && !p->arg.category_inverted && + p->arg.category_mask == PAGE_IS_WRITTEN && + p->arg.return_mask == PAGE_IS_WRITTEN) { + for (addr = start; addr < end; pte++, addr += PAGE_SIZE) { + unsigned long next = addr + PAGE_SIZE; + + if (pte_uffd_wp(ptep_get(pte))) + continue; + ret = pagemap_scan_output(p->cur_vma_category | PAGE_IS_WRITTEN, + p, addr, &next); + if (next == addr) + break; + if (~p->arg.flags & PM_SCAN_WP_MATCHING) + continue; + make_uffd_wp_pte(vma, addr, pte); + if (!flush) { + start = addr; + flush = true; + } + } + goto flush_and_return; + } + for (addr = start; addr != end; pte++, addr += PAGE_SIZE) { unsigned long categories = p->cur_vma_category | pagemap_page_category(p, vma, addr, ptep_get(pte)); @@ -2131,6 +2168,7 @@ static int pagemap_scan_pmd_entry(pmd_t *pmd, unsigned long start, } } +flush_and_return: if (flush) flush_tlb_range(vma, start, addr); -- 2.39.2