From mboxrd@z Thu Jan 1 00:00:00 1970 Return-Path: Received: from ffbox0-bg.mplayerhq.hu (ffbox0-bg.ffmpeg.org [79.124.17.100]) by master.gitmailbox.com (Postfix) with ESMTP id C108E44FB5 for ; Thu, 15 Dec 2022 10:51:25 +0000 (UTC) Received: from [127.0.1.1] (localhost [127.0.0.1]) by ffbox0-bg.mplayerhq.hu (Postfix) with ESMTP id F405168BD45; Thu, 15 Dec 2022 12:51:21 +0200 (EET) Received: from mail-wr1-f43.google.com (mail-wr1-f43.google.com [209.85.221.43]) by ffbox0-bg.mplayerhq.hu (Postfix) with ESMTPS id C889068BA61 for ; Thu, 15 Dec 2022 12:51:14 +0200 (EET) Received: by mail-wr1-f43.google.com with SMTP id h12so2584524wrv.10 for ; Thu, 15 Dec 2022 02:51:14 -0800 (PST) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=obe-tv.20210112.gappssmtp.com; s=20210112; h=content-transfer-encoding:mime-version:message-id:date:subject:to :from:from:to:cc:subject:date:message-id:reply-to; bh=/PEwgkX/578fNVw2vaAn4BOMru9w83vX1hqP9jomb4o=; b=QFQwAxLCTk5qaUoAeE+cGJysLjmJ8UGKdjXzsqEmXkBPnr/c3u1YscjZhRCvzYZ9v/ MhJgwq7fH7nxMoBzET0ONiz46Izo+8Wja//kZYJe9YRFNZAkYlD2lF3TtqQwX/Zdhf/7 waM5bMTxc2nzlUQXKpSXwd5PKbiyEge30tZFxxM5tl3LwlB/Y7aE6s2DHdhJq69+IrWD 9WpMzZC0exrtDt5SbTmPjaI8NK+4nwOGTIH5IDrW3yblG76728zulp6dx3M2jkR1vDUV ghxmsNReLYk/GPu82M8oUpjjz6kuamZi5WERCX2HK94HXOnZAPiYqhBoI/UCBGFSbrC+ KIpg== X-Google-DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20210112; h=content-transfer-encoding:mime-version:message-id:date:subject:to :from:x-gm-message-state:from:to:cc:subject:date:message-id:reply-to; bh=/PEwgkX/578fNVw2vaAn4BOMru9w83vX1hqP9jomb4o=; b=BUl15LaMXC1MUL0MJX/8dG8q/f5FlPAnwOxYMShdR1Mx0hpj0QhW0+7ynMKG9fUqHe LXrXfPPu37ePj3zvbMzLItu9llrq3KkbbjOuiTxsM4HkmIaIZgYXDdMSpgdawnRIw/c3 BN4rMDwq7WwXXWtqfAZkla7rT9h5bluhxw8AhyQMIAcn+WR7V7ojvznnrvyiQWbbQu9Z Y2UOSp5Hp3H0g3/ifvEqkwo1BTTx6+7RRGmc34/kFkHT/jGuhcN/2cxrPCLbFoTBRGv+ 2EbNZ93SMcCsoYuyD6S1N2mrQ5RK7fPYFRRqxoAmOF8+llGYPTmAyXnLKYIdKA4WD/jA y6fA== X-Gm-Message-State: ANoB5pkci+VRulG9UDJrjE3KplyaMKlUUe/se8PuT86NdEpzZAwRGNse dR+QQJ+YZN6M7M/Fm9dns+TOWS6bDOIRfyD1 X-Google-Smtp-Source: AA0mqf73/pEaPjFhjvas8l+W25nNHW2jnCdD5goW3xxCX1qEusDsE7X7Kby9wAPgC+GtEypPiXfZMQ== X-Received: by 2002:a5d:4012:0:b0:242:5878:2923 with SMTP id n18-20020a5d4012000000b0024258782923mr16868391wrp.10.1671101474118; Thu, 15 Dec 2022 02:51:14 -0800 (PST) Received: from Dana.systemlords.lan (d51A44418.access.telenet.be. [81.164.68.24]) by smtp.gmail.com with ESMTPSA id x12-20020a5d650c000000b002415dd45320sm5421977wru.112.2022.12.15.02.51.13 for (version=TLS1_3 cipher=TLS_AES_256_GCM_SHA384 bits=256/256); Thu, 15 Dec 2022 02:51:13 -0800 (PST) From: James Darnley To: ffmpeg-devel@ffmpeg.org Date: Thu, 15 Dec 2022 11:49:03 +0100 Message-Id: <20221215104904.3264109-1-jdarnley@obe.tv> X-Mailer: git-send-email 2.38.0 MIME-Version: 1.0 Subject: [FFmpeg-devel] [PATCH 1/2] avcodec/x86/v210: add some comments to the improved avx2 function X-BeenThere: ffmpeg-devel@ffmpeg.org X-Mailman-Version: 2.1.29 Precedence: list List-Id: FFmpeg development discussions and patches List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Reply-To: FFmpeg development discussions and patches Content-Type: text/plain; charset="us-ascii" Content-Transfer-Encoding: 7bit Errors-To: ffmpeg-devel-bounces@ffmpeg.org Sender: "ffmpeg-devel" Archived-At: List-Archive: List-Post: --- libavcodec/x86/v210.asm | 12 ++++++------ 1 file changed, 6 insertions(+), 6 deletions(-) diff --git a/libavcodec/x86/v210.asm b/libavcodec/x86/v210.asm index 3b9e0761df..600a4ddc5f 100644 --- a/libavcodec/x86/v210.asm +++ b/libavcodec/x86/v210.asm @@ -65,18 +65,18 @@ cglobal v210_planar_unpack_%1, 5, 5, 6 + 2 * cpuflag(avx2), src, y, u, v, w mova m0, [srcq] %endif - pmullw m1, m0, m3 - pslld m0, 12 - psrlw m1, 6 ; yB yA u5 v4 y8 y7 v3 u3 y5 y4 u2 v1 y2 y1 v0 u0 - psrld m0, 22 ; 00 v5 00 y9 00 u4 00 y6 00 v2 00 y3 00 u1 00 y0 + pmullw m1, m0, m3 ; shifts the 1st and 3rd sample of each dword into the high 10 bits of each word + pslld m0, 12 ; shifts the 2nd sample of each dword into the high 10 bits of each dword + psrlw m1, 6 ; shifts the 1st and 3rd samples back into the low 10 bits + psrld m0, 22 ; shifts the 2nd sample back into the low 10 bits of each dword %if cpuflag(avx2) - vpblendd m2, m1, m0, 0x55 ; yB yA 00 y9 y8 y7 00 y6 y5 y4 00 y3 y2 y1 00 y0 + vpblendd m2, m1, m0, 0x55 ; merge the odd dwords from m0 and even from m1 ; yB yA 00 y9 y8 y7 00 y6 y5 y4 00 y3 y2 y1 00 y0 pshufb m2, m4 ; 00 00 yB yA y9 y8 y7 y6 00 00 y5 y4 y3 y2 y1 y0 vpermd m2, m6, m2 ; 00 00 00 00 yB yA y9 y8 y7 y6 y5 y4 y3 y2 y1 y0 movu [yq+2*wq], m2 - vpblendd m1, m1, m0, 0xaa ; 00 v5 u5 v4 00 u4 v3 u3 00 v2 u2 v1 00 u1 v0 u0 + vpblendd m1, m1, m0, 0xaa ; merge the even dwords from m0 and odd from m1 ; 00 v5 u5 v4 00 u4 v3 u3 00 v2 u2 v1 00 u1 v0 u0 pshufb m1, m5 ; 00 v5 v4 v3 00 u5 u4 u3 00 v2 v1 v0 00 u2 u1 u0 vpermq m1, m1, 0xd8 ; 00 v5 v4 v3 00 v2 v1 v0 00 u5 u4 u3 00 u2 u1 u0 pshufb m1, m7 ; 00 00 v5 v4 v3 v2 v1 v0 00 00 u5 u4 u3 u2 u1 u0 -- 2.38.0 _______________________________________________ ffmpeg-devel mailing list ffmpeg-devel@ffmpeg.org https://ffmpeg.org/mailman/listinfo/ffmpeg-devel To unsubscribe, visit link above, or email ffmpeg-devel-request@ffmpeg.org with subject "unsubscribe".