From patchwork Mon Aug 16 09:45:44 2021 Content-Type: text/plain; charset="utf-8" MIME-Version: 1.0 Content-Transfer-Encoding: 7bit X-Patchwork-Submitter: Mikhail Nitenko X-Patchwork-Id: 29562 Delivered-To: ffmpegpatchwork2@gmail.com Received: by 2002:a05:6602:2a4a:0:0:0:0 with SMTP id k10csp1874827iov; Mon, 16 Aug 2021 02:46:05 -0700 (PDT) X-Google-Smtp-Source: ABdhPJzmEtr9hcolpzmH+He18irJ3bzWqafz7da3izC/+IWkJMiDD9d8ri/N+Yrc+B1yum0jhB+Y X-Received: by 2002:aa7:d504:: with SMTP id y4mr2595339edq.138.1629107165756; Mon, 16 Aug 2021 02:46:05 -0700 (PDT) ARC-Seal: i=1; a=rsa-sha256; t=1629107165; cv=none; d=google.com; s=arc-20160816; b=xxrD4a7hpBOl/fxYxkINPNzvzmNhlUboZ0cQ2dfzB3uptzkIRSNaLiM73AnkGqlQmH vme2TvLAlNyW2C0ZACdsiYU9zK7IhQORJlm3ienfTIKFGD4+XKwgaOuVZVm4d2PA0iMr O/dC2CssftKqOXIpMR4EgTH9wiJwmXHkkPcKmdNfMIym8nQCcpYwlL2G0H1C0NxgSm6L n+151I+ah5pOYyn0365MWVjdgt+xg1+fhbOM9HcrLjS9FMfNrizZp1e3hSZRUpGms9zZ 02FijDbp+NPRH5d1aOvzSP4BSQX5CQ+th7cb27H+PWlz6wE9PSddeJdaClESTznMot9x ZhFA== ARC-Message-Signature: i=1; a=rsa-sha256; c=relaxed/relaxed; d=google.com; s=arc-20160816; h=sender:errors-to:content-transfer-encoding:cc:reply-to :list-subscribe:list-help:list-post:list-archive:list-unsubscribe :list-id:precedence:subject:mime-version:message-id:date:to:from :dkim-signature:delivered-to; bh=BnzWk55vxaM6sJ+VoQwhS4bcnZA0xWXEJwYRvCkbKcE=; b=jkoIQHZylla27LIaV6OduVx4Yry3c8GeJUs62nEWEbmXS+zZ9abyTm9IPOJZ9/XS93 0Eyed7ilCE0W4QIpR0/p0OwUI0hweEh/CEHpeEymTnKGNH6Ob6g6zRCsXbdLpE6sGSkd UlDCBVdpxuIl04Kxk8e4Yg/28doAhd5rCTX3OVfLr4nO4+uhzZCCIvTDvsVrOEfbPm1t mPKfgwP6w18Xx/CRSAN52fjaicnd5kwPJzuRAzEKW1pB1zW7yORU6hKewO7VkrTJFnNG bFd+RfHJKkrbFoDZoqCNb3DVYXMK3vLx6SF0PEOhHIbuiIs8UOlhSs71J0a7+dqPWQx+ DIcQ== ARC-Authentication-Results: i=1; mx.google.com; dkim=neutral (body hash did not verify) header.i=@gmail.com header.s=20161025 header.b=oc+GhEAr; spf=pass (google.com: domain of ffmpeg-devel-bounces@ffmpeg.org designates 79.124.17.100 as permitted sender) smtp.mailfrom=ffmpeg-devel-bounces@ffmpeg.org; dmarc=fail (p=NONE sp=QUARANTINE dis=NONE) header.from=gmail.com Return-Path: Received: from ffbox0-bg.mplayerhq.hu (ffbox0-bg.ffmpeg.org. [79.124.17.100]) by mx.google.com with ESMTP id n4si9847396edy.3.2021.08.16.02.46.05; Mon, 16 Aug 2021 02:46:05 -0700 (PDT) Received-SPF: pass (google.com: domain of ffmpeg-devel-bounces@ffmpeg.org designates 79.124.17.100 as permitted sender) client-ip=79.124.17.100; Authentication-Results: mx.google.com; dkim=neutral (body hash did not verify) header.i=@gmail.com header.s=20161025 header.b=oc+GhEAr; spf=pass (google.com: domain of ffmpeg-devel-bounces@ffmpeg.org designates 79.124.17.100 as permitted sender) smtp.mailfrom=ffmpeg-devel-bounces@ffmpeg.org; dmarc=fail (p=NONE sp=QUARANTINE dis=NONE) header.from=gmail.com Received: from [127.0.1.1] (localhost [127.0.0.1]) by ffbox0-bg.mplayerhq.hu (Postfix) with ESMTP id AD25668A42C; Mon, 16 Aug 2021 12:46:01 +0300 (EEST) X-Original-To: ffmpeg-devel@ffmpeg.org Delivered-To: ffmpeg-devel@ffmpeg.org Received: from mail-lf1-f49.google.com (mail-lf1-f49.google.com [209.85.167.49]) by ffbox0-bg.mplayerhq.hu (Postfix) with ESMTPS id A3FAB689BC9 for ; Mon, 16 Aug 2021 12:45:55 +0300 (EEST) Received: by mail-lf1-f49.google.com with SMTP id i9so12430783lfg.10 for ; Mon, 16 Aug 2021 02:45:55 -0700 (PDT) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=gmail.com; s=20161025; h=from:to:cc:subject:date:message-id:mime-version :content-transfer-encoding; bh=l0jbFd16MpiqLc+Si7Jp+Mj5AThrZ+ehoEHftFz1v0o=; b=oc+GhEArA+OcWjUkGLIwach+QrPhGC5boxr2AGS/pYjresTSvjEvmZRfB4Op3ZBMKz TqW6iExqUhUrxe0IjBJmtmYWADXCn3UFL4+ZvpXofoPDUTyU8yqOKk3qCs/oAZ5B2CoG 5QehopDqktNfklnoHK6Ozva1axjP3mRWD0NxBSOxPIxwLSA0FL4onfIMlBL5WbMYSnpl sgEn7Ngh3N2uGEPpWFdp9V7K0H5l1M6R7sM90BRIj0U1H4W84V1fSFGaATAUlyxbV5cY HwAWwB0vIBRn3fG07tpN2D4GYVYmzLhf0hnVwr2a29RWyKyaRndMtIPLsgWdtCBTMVDj b26A== X-Google-DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20161025; h=x-gm-message-state:from:to:cc:subject:date:message-id:mime-version :content-transfer-encoding; bh=l0jbFd16MpiqLc+Si7Jp+Mj5AThrZ+ehoEHftFz1v0o=; b=dTijfDlpDnXjtYajT4KU9MC2CfGO9u+g1vLBG6v64zT6D6NVeiA9ZUXspAl1/PVmPH ik6x1EhwfaR+biACdk7iNjbPLV7ImX4aM40HpU4Le468WOEPN8P4WK26zCW67xEGu7sv 8wEIrTpzg0A01l81ScqWn48YMYIXhufDiFgUcWYDNjkba9Wj7xyKETNFgOaEkWbTwaHj kcSMuiFxnh230Cfz6UTHqqvLygH0RQICiv60heQ77vt0bpYN8Mi4c3UuSWAfJKwX6t7H 6EjSG/8OgvruYuqCdPYswSl9RTS+ZGgb5VzRmjoWB7D68GkNRlO7Drai9UdCnDVk1aYg gVjw== X-Gm-Message-State: AOAM532vujlX5ED/lSJ9Ys7RoMgTrO7mpzV6Xi7/UushPyWQvcKJwwEs GUVlGjChtwIb4nq3V8oNM/6KoxBvmpkLlw== X-Received: by 2002:a05:6512:5ce:: with SMTP id o14mr296782lfo.252.1629107153850; Mon, 16 Aug 2021 02:45:53 -0700 (PDT) Received: from localhost.localdomain ([109.195.102.12]) by smtp.gmail.com with ESMTPSA id v16sm205995lfq.87.2021.08.16.02.45.53 (version=TLS1_3 cipher=TLS_AES_256_GCM_SHA384 bits=256/256); Mon, 16 Aug 2021 02:45:53 -0700 (PDT) From: Mikhail Nitenko To: ffmpeg-devel@ffmpeg.org Date: Mon, 16 Aug 2021 14:45:44 +0500 Message-Id: <20210816094545.448283-1-mnitenko@gmail.com> X-Mailer: git-send-email 2.32.0 MIME-Version: 1.0 Subject: [FFmpeg-devel] [PATCH 1/2] lavc/aarch64: move transpose_4x8H to neon.S X-BeenThere: ffmpeg-devel@ffmpeg.org X-Mailman-Version: 2.1.29 Precedence: list List-Id: FFmpeg development discussions and patches List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Reply-To: FFmpeg development discussions and patches Cc: Mikhail Nitenko Errors-To: ffmpeg-devel-bounces@ffmpeg.org Sender: "ffmpeg-devel" X-TUID: rO5r/66iJlIh transpose_4x8H was declared in vp9lpf_16bpp_neon, however this macro is not unique to vp9 and could be used elsewhere. Signed-off-by: Mikhail Nitenko --- libavcodec/aarch64/neon.S | 13 +++++++++++++ libavcodec/aarch64/vp9lpf_16bpp_neon.S | 12 ------------ 2 files changed, 13 insertions(+), 12 deletions(-) diff --git a/libavcodec/aarch64/neon.S b/libavcodec/aarch64/neon.S index 0fddbecae3..1ad32c359d 100644 --- a/libavcodec/aarch64/neon.S +++ b/libavcodec/aarch64/neon.S @@ -109,12 +109,25 @@ trn2 \r5\().4H, \r0\().4H, \r1\().4H trn1 \r6\().4H, \r2\().4H, \r3\().4H trn2 \r7\().4H, \r2\().4H, \r3\().4H + trn1 \r0\().2S, \r4\().2S, \r6\().2S trn2 \r2\().2S, \r4\().2S, \r6\().2S trn1 \r1\().2S, \r5\().2S, \r7\().2S trn2 \r3\().2S, \r5\().2S, \r7\().2S .endm +.macro transpose_4x8H r0, r1, r2, r3, t4, t5, t6, t7 + trn1 \t4\().8H, \r0\().8H, \r1\().8H + trn2 \t5\().8H, \r0\().8H, \r1\().8H + trn1 \t6\().8H, \r2\().8H, \r3\().8H + trn2 \t7\().8H, \r2\().8H, \r3\().8H + + trn1 \r0\().4S, \t4\().4S, \t6\().4S + trn2 \r2\().4S, \t4\().4S, \t6\().4S + trn1 \r1\().4S, \t5\().4S, \t7\().4S + trn2 \r3\().4S, \t5\().4S, \t7\().4S +.endm + .macro transpose_8x8H r0, r1, r2, r3, r4, r5, r6, r7, r8, r9 trn1 \r8\().8H, \r0\().8H, \r1\().8H trn2 \r9\().8H, \r0\().8H, \r1\().8H diff --git a/libavcodec/aarch64/vp9lpf_16bpp_neon.S b/libavcodec/aarch64/vp9lpf_16bpp_neon.S index 9075f3d406..9869614a29 100644 --- a/libavcodec/aarch64/vp9lpf_16bpp_neon.S +++ b/libavcodec/aarch64/vp9lpf_16bpp_neon.S @@ -22,18 +22,6 @@ #include "neon.S" -.macro transpose_4x8H r0, r1, r2, r3, t4, t5, t6, t7 - trn1 \t4\().8h, \r0\().8h, \r1\().8h - trn2 \t5\().8h, \r0\().8h, \r1\().8h - trn1 \t6\().8h, \r2\().8h, \r3\().8h - trn2 \t7\().8h, \r2\().8h, \r3\().8h - - trn1 \r0\().4s, \t4\().4s, \t6\().4s - trn2 \r2\().4s, \t4\().4s, \t6\().4s - trn1 \r1\().4s, \t5\().4s, \t7\().4s - trn2 \r3\().4s, \t5\().4s, \t7\().4s -.endm - // The input to and output from this macro is in the registers v16-v31, // and v0-v7 are used as scratch registers. // p7 = v16 .. p3 = v20, p0 = v23, q0 = v24, q3 = v27, q7 = v31