From mboxrd@z Thu Jan 1 00:00:00 1970 Return-Path: Received: from ffbox0-bg.mplayerhq.hu (ffbox0-bg.ffmpeg.org [79.124.17.100]) by master.gitmailbox.com (Postfix) with ESMTP id 2140C4034C for ; Mon, 20 Dec 2021 14:45:57 +0000 (UTC) Received: from [127.0.1.1] (localhost [127.0.0.1]) by ffbox0-bg.mplayerhq.hu (Postfix) with ESMTP id 8E65868AE77; Mon, 20 Dec 2021 16:45:55 +0200 (EET) Received: from mail-lf1-f74.google.com (mail-lf1-f74.google.com [209.85.167.74]) by ffbox0-bg.mplayerhq.hu (Postfix) with ESMTPS id 9E1FC68AB71 for ; Mon, 20 Dec 2021 16:45:49 +0200 (EET) Received: by mail-lf1-f74.google.com with SMTP id cf27-20020a056512281b00b004259e7fce67so1033206lfb.0 for ; Mon, 20 Dec 2021 06:45:49 -0800 (PST) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=google.com; s=20210112; h=date:in-reply-to:message-id:mime-version:references:subject:from:to :cc; bh=2X2EepDowYoA+E+PZCbzwYE6q9udLErWYvTIa8HFSRI=; b=iskcmjm4e/YovR80hHlHiSMX4m86uBEu6fZjyscneRNecZVkGda5j1CyRwMnJ6uOoi LI0pDARejBSav0xjv2OGs4cq2d108xLr6g64halhoEbE5d5TR14DM3Vdt2SXnhrsDY3H n6oQ2F1uzhLGZiuN0dVJUmUyTVYGvlR6yDnoE/HF5pLgt8US183K6anRasNJ/rH3wySp tRv6M7t9HbWw7Cp5SVTIK3L0nuTfIeW04hOQj+ODzLw2uOqP7Ceq7wKNxz/VeQ+ZnydI Y1kORa1T9wjSYX/0xYPnitBFFTCu/yCJ3iMq7cxgEza85Fil3abzSCtRHkauUj/BVUmv ldGQ== X-Google-DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20210112; h=x-gm-message-state:date:in-reply-to:message-id:mime-version :references:subject:from:to:cc; bh=2X2EepDowYoA+E+PZCbzwYE6q9udLErWYvTIa8HFSRI=; b=za5RRl/TdpFCy6i8+Emy6UhJhyxfkxB6qrnqT2Dw5pR51iCv3xP3fIfTaJkM/nesdC J6MjRtDXhQ0hpKoafDrE5fk73++IXrRKNVoPuULj9gQWFBPr6ndKED/GF/BN71oNjX8O LCU0kBIEf7awny/3EtkXvKlxjYScnouejc7A6/vh2ciX4gqNQmsrey8MVxtgRAlA41/J PZJIA/toFbVMBPvWi67fFhatLmkOylsgAVp6vYsr91VYyaY5yiWgBr3wvjReGS1uD3iZ 4c4Vw+UhsXiqwv63BCWbPTU40GDQbYrKd+7WuFNiUW0or5Rdqs93MqBERvX5HayNL5oD oQrQ== X-Gm-Message-State: AOAM533R32/52Gnsq585EL8jI1RU5+X1yTl105Bqgu9e0aQ+GtlOtyGa ZvGCTTmblN7AkeYvLriWI51bE2ZO0qm55osF2bXGgkf0ozfJ/fYKrHMEmnnXrglnhOT8EnKqM50 nUlV+SoGWwAkkd6H0Tfsp0lzDAi/aA9RtC9GUySQtR6KhvvYFKuTtniQY/kQ3RCrF7sX66gg= X-Google-Smtp-Source: ABdhPJxMYALdKLkyrlXUoBVDEVymZVJvEMB0MhbqEZsP4/oLpROA+sW4ccpJ5nXXDdZ52mcN+Cv4CZMO2Kk7FrE= X-Received: from alankelly0.zrh.corp.google.com ([2a00:79e0:61:301:922d:7ddd:85f5:5a25]) (user=alankelly job=sendgmr) by 2002:a05:6512:2289:: with SMTP id f9mr15712665lfu.619.1640011548660; Mon, 20 Dec 2021 06:45:48 -0800 (PST) Date: Mon, 20 Dec 2021 15:45:45 +0100 In-Reply-To: <166a93ba-7f15-f473-5889-1a0a879a75a4@gmail.com> Message-Id: <20211220144545.739340-1-alankelly@google.com> Mime-Version: 1.0 References: <166a93ba-7f15-f473-5889-1a0a879a75a4@gmail.com> X-Mailer: git-send-email 2.34.1.173.g76aa8bc2d0-goog From: Alan Kelly To: ffmpeg-devel@ffmpeg.org Subject: [FFmpeg-devel] [PATCH 2/2] libswscale: Test AV_CPU_FLAG_SLOW_GATHER for hscale functions. X-BeenThere: ffmpeg-devel@ffmpeg.org X-Mailman-Version: 2.1.29 Precedence: list List-Id: FFmpeg development discussions and patches List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Reply-To: FFmpeg development discussions and patches Cc: Alan Kelly Content-Type: text/plain; charset="us-ascii" Content-Transfer-Encoding: 7bit Errors-To: ffmpeg-devel-bounces@ffmpeg.org Sender: "ffmpeg-devel" Archived-At: List-Archive: List-Post: This is instead of EXTERNAL_AVX2_FAST so that the avx2 hscale functions are only used where they are faster. --- Whoops! Corrects check so that this flag is only enabled where fast avx2 and fast gathers are available. libswscale/utils.c | 2 +- libswscale/x86/swscale.c | 2 +- tests/checkasm/sw_scale.c | 2 +- 3 files changed, 3 insertions(+), 3 deletions(-) diff --git a/libswscale/utils.c b/libswscale/utils.c index d4a72d3ce1..7158384f0b 100644 --- a/libswscale/utils.c +++ b/libswscale/utils.c @@ -282,7 +282,7 @@ void ff_shuffle_filter_coefficients(SwsContext *c, int *filterPos, int filterSiz #if ARCH_X86_64 int i, j, k, l; int cpu_flags = av_get_cpu_flags(); - if (EXTERNAL_AVX2_FAST(cpu_flags)){ + if (EXTERNAL_AVX2_FAST(cpu_flags) && !(cpu_flags & AV_CPU_FLAG_SLOW_GATHER)) { if ((c->srcBpc == 8) && (c->dstBpc <= 14)){ if (dstW % 16 == 0){ if (filter != NULL){ diff --git a/libswscale/x86/swscale.c b/libswscale/x86/swscale.c index c49a05c37b..ffc7691c12 100644 --- a/libswscale/x86/swscale.c +++ b/libswscale/x86/swscale.c @@ -578,7 +578,7 @@ switch(c->dstBpc){ \ break; \ } - if (EXTERNAL_AVX2_FAST(cpu_flags)) { + if (EXTERNAL_AVX2_FAST(cpu_flags) && !(cpu_flags & AV_CPU_FLAG_SLOW_GATHER)) { if ((c->srcBpc == 8) && (c->dstBpc <= 14)) { if (c->chrDstW % 16 == 0) ASSIGN_AVX2_SCALE_FUNC(c->hcScale, c->hChrFilterSize); diff --git a/tests/checkasm/sw_scale.c b/tests/checkasm/sw_scale.c index f4912e6c2c..3c0a083b42 100644 --- a/tests/checkasm/sw_scale.c +++ b/tests/checkasm/sw_scale.c @@ -217,7 +217,7 @@ static void check_hscale(void) } ff_sws_init_scale(ctx); memcpy(filterAvx2, filter, sizeof(uint16_t) * (SRC_PIXELS * MAX_FILTER_WIDTH + MAX_FILTER_WIDTH)); - if (cpu_flags & AV_CPU_FLAG_AVX2) + if ((cpu_flags & AV_CPU_FLAG_AVX2) && !(cpu_flags & AV_CPU_FLAG_SLOW_GATHER)) ff_shuffle_filter_coefficients(ctx, filterPosAvx, width, filterAvx2, SRC_PIXELS); if (check_func(ctx->hcScale, "hscale_%d_to_%d_width%d", ctx->srcBpc, ctx->dstBpc + 1, width)) { -- 2.34.1.173.g76aa8bc2d0-goog _______________________________________________ ffmpeg-devel mailing list ffmpeg-devel@ffmpeg.org https://ffmpeg.org/mailman/listinfo/ffmpeg-devel To unsubscribe, visit link above, or email ffmpeg-devel-request@ffmpeg.org with subject "unsubscribe".