Received: by 10.192.165.148 with SMTP id m20csp2233484imm; Thu, 26 Apr 2018 07:52:43 -0700 (PDT) X-Google-Smtp-Source: AIpwx499AubjtPnx4lqXQUYqqH2WLUALl09QdtTofLD+6kzxhV+FRUG9OYITT+N+FUaa7sOxJSKS X-Received: by 10.98.138.68 with SMTP id y65mr32565811pfd.110.1524754363888; Thu, 26 Apr 2018 07:52:43 -0700 (PDT) ARC-Seal: i=1; a=rsa-sha256; t=1524754363; cv=none; d=google.com; s=arc-20160816; b=xRNX3SQfIpHAti1pFGG03lkTEfN15T+yr3B+G2raUn0LnCVMytO+1f0X4oagQyZ6F6 Nwmp8ItQRnwJdRdx+dx13DPfmroOagQUKDJWLM+21c8tGD95/6pXS7Epbqr/c7cuy8jl TjFpVVLUno7jc4jrydIVY7XPQFSFxDKIrDlMRgiNrNKF5RyVpYNQ0RaiHAuoCgL0+hWC xDgTV8JM0pRzsUcy305PB038uYXbiNXbL4zdgGyn3R1MIa8kyxW/KD8XZ758/EAP/zrx 0fFSAhOm4Wib/hHgl76DolvBcIzXObc3rVmS2mw1W4zOWDM0UaIP3lLwx5tKgphGJr2m Qtlg== ARC-Message-Signature: i=1; a=rsa-sha256; c=relaxed/relaxed; d=google.com; s=arc-20160816; h=list-id:precedence:sender:references:in-reply-to:message-id:date :subject:cc:to:from:dkim-signature:arc-authentication-results; bh=QrqVhQ8X+XDYoGAFMiMk5tKgDC6IfjR81fG2iZp9Gb8=; b=ypRNOIeYPDxyU5UjMb4Gd2PCNZ4iSDu/rFYKWt6NL5ekVg+8ClXsqfZwAjPZCie/ZX CCulY/YTfWQs9AAaaeJrbT4ewnrstIBaR/U+ZBq5QbAuLlnTLIYzH2dhed/Z8HG3mHMm hfigkr1Ck/w3bhBibWuJBmo3vnkVD7A2nuu7rgd0rP8+GnBRBJO4F90jpZHH56HKx4KT S4N50iS+eJUm/WsdknrY/KTSzPFIvW1IXmcR0oHl6Lcu9iLLFrarijDVmTmSpnTHlmv2 duOCvj6XhtixhJAYfDv1WTRZD6bG8ZZN1eiQiHdqnl3XmviJAlRY4dXf33bByuuvuDUm fQUw== ARC-Authentication-Results: i=1; mx.google.com; dkim=pass header.i=@google.com header.s=20161025 header.b=AeMYtzSN; spf=pass (google.com: best guess record for domain of linux-kernel-owner@vger.kernel.org designates 209.132.180.67 as permitted sender) smtp.mailfrom=linux-kernel-owner@vger.kernel.org; dmarc=pass (p=REJECT sp=REJECT dis=NONE) header.from=google.com Return-Path: Received: from vger.kernel.org (vger.kernel.org. [209.132.180.67]) by mx.google.com with ESMTP id e6si16049611pgn.473.2018.04.26.07.52.29; Thu, 26 Apr 2018 07:52:43 -0700 (PDT) Received-SPF: pass (google.com: best guess record for domain of linux-kernel-owner@vger.kernel.org designates 209.132.180.67 as permitted sender) client-ip=209.132.180.67; Authentication-Results: mx.google.com; dkim=pass header.i=@google.com header.s=20161025 header.b=AeMYtzSN; spf=pass (google.com: best guess record for domain of linux-kernel-owner@vger.kernel.org designates 209.132.180.67 as permitted sender) smtp.mailfrom=linux-kernel-owner@vger.kernel.org; dmarc=pass (p=REJECT sp=REJECT dis=NONE) header.from=google.com Received: (majordomo@vger.kernel.org) by vger.kernel.org via listexpand id S1756599AbeDZOvN (ORCPT + 99 others); Thu, 26 Apr 2018 10:51:13 -0400 Received: from mail-pf0-f196.google.com ([209.85.192.196]:43180 "EHLO mail-pf0-f196.google.com" rhost-flags-OK-OK-OK-OK) by vger.kernel.org with ESMTP id S1756461AbeDZOvG (ORCPT ); Thu, 26 Apr 2018 10:51:06 -0400 Received: by mail-pf0-f196.google.com with SMTP id j11so18571281pff.10 for ; Thu, 26 Apr 2018 07:51:06 -0700 (PDT) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=google.com; s=20161025; h=from:to:cc:subject:date:message-id:in-reply-to:references; bh=QrqVhQ8X+XDYoGAFMiMk5tKgDC6IfjR81fG2iZp9Gb8=; b=AeMYtzSNCzjxh6uY3K8vmRKejjZBmBrE3dxHrfUS11gWwckISB+DBfsPg7tX7mb/pW MgXUHqkrHYnGAzD2F2ILsPZbm+IRXfRHfbdse1tw+7AYlkWHH/DXnCmt3xz9nVxhB5KQ LsslvfUnWlfyVnHZ1Fn5YBzIhx7taMMK2WeIMwyTCI1D+/MmDfvbNYyYBvYRnbdGxF8O 9H2zsvsqOKhUK8LbA0QQgsDMeUfu/sqfcFuzrCa4agdtRCMS2aTI4+BzinvxIkuSu0bo yJaaYgxDWWmMPrKlaXKZBOtPFNgdz72jX4XgOOHbDqnYZHJQFwnMNufsN3zQ+0+6726R PrYA== X-Google-DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20161025; h=x-gm-message-state:from:to:cc:subject:date:message-id:in-reply-to :references; bh=QrqVhQ8X+XDYoGAFMiMk5tKgDC6IfjR81fG2iZp9Gb8=; b=is9pbbmrYWWOyx9r/ohx6/vk9MC3r7MqIuol2hr0HmMCqIGn0Wvg8ILiYfZdk4FLMb DXkOx7wBlMq5bDFpgYo1M32Dz5og0f9oA+8Q3/Mo/DLGhWtF0jB4uYn81zNbxA+KPl51 FG51SYl3mZmGvPSkyNpeh2A2BsaOK41YPvK/aTSQ/w5oJ/FjFnLF/MoT/19DgMjKLeIM t7szqaY5sgOqZzbiV14t4MBLuErFEt9rUbk1Q5Mn3UfcVNTjGQmQ2JYl+5BxHFrCTWp2 jvaQHucKerc2iGVT/52sYSBJQHGybEzwha7f9jveB/UnswFUwUXGEqWkgVZDA2PMGyGR cwxg== X-Gm-Message-State: ALQs6tAjkpdqJLpIfIwzaOdVg65fCCNfyjFrDkRJ1pXbkMYD5hBF8JOH /cw2ExwO9k7axyjLeNS1SfOkwA== X-Received: by 10.99.7.86 with SMTP id 83mr12667848pgh.211.1524754265650; Thu, 26 Apr 2018 07:51:05 -0700 (PDT) Received: from localhost ([2620:15c:2c4:201:f5a:7eca:440a:3ead]) by smtp.gmail.com with ESMTPSA id w19sm6218790pgv.59.2018.04.26.07.51.04 (version=TLS1_2 cipher=ECDHE-RSA-CHACHA20-POLY1305 bits=256/256); Thu, 26 Apr 2018 07:51:04 -0700 (PDT) From: Eric Dumazet To: "David S . Miller" Cc: netdev , Andy Lutomirski , linux-kernel , linux-mm , Ka-Cheong Poon , Eric Dumazet , Eric Dumazet , Soheil Hassas Yeganeh Subject: [PATCH v3 net-next 2/2] selftests: net: tcp_mmap must use TCP_ZEROCOPY_RECEIVE Date: Thu, 26 Apr 2018 07:50:56 -0700 Message-Id: <20180426145056.220325-3-edumazet@google.com> X-Mailer: git-send-email 2.17.0.484.g0c8726318c-goog In-Reply-To: <20180426145056.220325-1-edumazet@google.com> References: <20180426145056.220325-1-edumazet@google.com> Sender: linux-kernel-owner@vger.kernel.org Precedence: bulk List-ID: X-Mailing-List: linux-kernel@vger.kernel.org After prior kernel change, mmap() on TCP socket only reserves VMA. We have to use getsockopt(fd, IPPROTO_TCP, TCP_ZEROCOPY_RECEIVE, ...) to perform the transfert of pages from skbs in TCP receive queue into such VMA. struct tcp_zerocopy_receive { __u64 address; /* in: address of mapping */ __u32 length; /* in/out: number of bytes to map/mapped */ __u32 recv_skip_hint; /* out: amount of bytes to skip */ }; After a successful getsockopt(...TCP_ZEROCOPY_RECEIVE...), @length contains number of bytes that were mapped, and @recv_skip_hint contains number of bytes that should be read using conventional read()/recv()/recvmsg() system calls, to skip a sequence of bytes that can not be mapped, because not properly page aligned. Signed-off-by: Eric Dumazet Cc: Andy Lutomirski Cc: Soheil Hassas Yeganeh --- tools/testing/selftests/net/tcp_mmap.c | 64 +++++++++++++++----------- 1 file changed, 37 insertions(+), 27 deletions(-) diff --git a/tools/testing/selftests/net/tcp_mmap.c b/tools/testing/selftests/net/tcp_mmap.c index dea342fe6f4e88b5709d2ac37b2fc9a2a320bf44..77f762780199ff1f69f9f6b3f18e72deddb69f5e 100644 --- a/tools/testing/selftests/net/tcp_mmap.c +++ b/tools/testing/selftests/net/tcp_mmap.c @@ -76,9 +76,10 @@ #include #include #include -#include #include #include +#include +#include #ifndef MSG_ZEROCOPY #define MSG_ZEROCOPY 0x4000000 @@ -134,11 +135,12 @@ void hash_zone(void *zone, unsigned int length) void *child_thread(void *arg) { unsigned long total_mmap = 0, total = 0; + struct tcp_zerocopy_receive zc; unsigned long delta_usec; int flags = MAP_SHARED; struct timeval t0, t1; char *buffer = NULL; - void *oaddr = NULL; + void *addr = NULL; double throughput; struct rusage ru; int lu, fd; @@ -153,41 +155,46 @@ void *child_thread(void *arg) perror("malloc"); goto error; } + if (zflg) { + addr = mmap(NULL, chunk_size, PROT_READ, flags, fd, 0); + if (addr == (void *)-1) + zflg = 0; + } while (1) { struct pollfd pfd = { .fd = fd, .events = POLLIN, }; int sub; poll(&pfd, 1, 10000); if (zflg) { - void *naddr; + socklen_t zc_len = sizeof(zc); + int res; - naddr = mmap(oaddr, chunk_size, PROT_READ, flags, fd, 0); - if (naddr == (void *)-1) { - if (errno == EAGAIN) { - /* That is if SO_RCVLOWAT is buggy */ - usleep(1000); - continue; - } - if (errno == EINVAL) { - flags = MAP_SHARED; - oaddr = NULL; - goto fallback; - } - if (errno != EIO) - perror("mmap()"); + zc.address = (__u64)addr; + zc.length = chunk_size; + zc.recv_skip_hint = 0; + res = getsockopt(fd, IPPROTO_TCP, TCP_ZEROCOPY_RECEIVE, + &zc, &zc_len); + if (res == -1) break; + + if (zc.length) { + assert(zc.length <= chunk_size); + total_mmap += zc.length; + if (xflg) + hash_zone(addr, zc.length); + total += zc.length; } - total_mmap += chunk_size; - if (xflg) - hash_zone(naddr, chunk_size); - total += chunk_size; - if (!keepflag) { - flags |= MAP_FIXED; - oaddr = naddr; + if (zc.recv_skip_hint) { + assert(zc.recv_skip_hint <= chunk_size); + lu = read(fd, buffer, zc.recv_skip_hint); + if (lu > 0) { + if (xflg) + hash_zone(buffer, lu); + total += lu; + } } continue; } -fallback: sub = 0; while (sub < chunk_size) { lu = read(fd, buffer + sub, chunk_size - sub); @@ -228,6 +235,8 @@ void *child_thread(void *arg) error: free(buffer); close(fd); + if (zflg) + munmap(addr, chunk_size); pthread_exit(0); } @@ -371,7 +380,8 @@ int main(int argc, char *argv[]) setup_sockaddr(cfg_family, host, &listenaddr); if (mss && - setsockopt(fdlisten, SOL_TCP, TCP_MAXSEG, &mss, sizeof(mss)) == -1) { + setsockopt(fdlisten, IPPROTO_TCP, TCP_MAXSEG, + &mss, sizeof(mss)) == -1) { perror("setsockopt TCP_MAXSEG"); exit(1); } @@ -402,7 +412,7 @@ int main(int argc, char *argv[]) setup_sockaddr(cfg_family, host, &addr); if (mss && - setsockopt(fd, SOL_TCP, TCP_MAXSEG, &mss, sizeof(mss)) == -1) { + setsockopt(fd, IPPROTO_TCP, TCP_MAXSEG, &mss, sizeof(mss)) == -1) { perror("setsockopt TCP_MAXSEG"); exit(1); } -- 2.17.0.484.g0c8726318c-goog