From nobody Wed Nov 19 00:14:28 2025 Delivered-To: importer@patchew.org Authentication-Results: mx.zohomail.com; spf=pass (zohomail.com: domain of gnu.org designates 209.51.188.17 as permitted sender) smtp.mailfrom=qemu-devel-bounces+importer=patchew.org@nongnu.org ARC-Seal: i=1; a=rsa-sha256; t=1613144560; cv=none; d=zohomail.com; s=zohoarc; b=d2EJHdkiUp47922+N3pxhAj6f8eZ8DWMvXApL8v6dGSofIR+D3tDXVoNPG0VFfVrQN/Bq+SH1yVcmA5kwBh/K7YbwKjS2tZYtUePD8FN8eYkcV7lKIAx0GZT/E2RQUbYAyKE7umuPF2gsXa1QDIr5p/oIbBhfxFEKh7exukVytI= ARC-Message-Signature: i=1; a=rsa-sha256; c=relaxed/relaxed; d=zohomail.com; s=zohoarc; t=1613144560; h=Cc:Date:From:In-Reply-To:List-Subscribe:List-Post:List-Id:List-Archive:List-Help:List-Unsubscribe:Message-ID:References:Sender:Subject:To; bh=lwkgVKE/IzTNvQRZfcpITrUh5zNobm+J8qvuVNL33l8=; b=BlwczWnemFdBsubWtOZAsHsgukuSf1kjzf49U6VOQh8zQpeyAzkUfMNnRPXMoUFP3kGB6+U1xc7jA7enBHDX9F76vfDxEi0+i10VyC1ZAB1uGsweV6gerWuTHsMah2cHiu5CQcnk4G+oya19c6ZF22KiMCB4+vNLsDcmTxnd+p0= ARC-Authentication-Results: i=1; mx.zohomail.com; spf=pass (zohomail.com: domain of gnu.org designates 209.51.188.17 as permitted sender) smtp.mailfrom=qemu-devel-bounces+importer=patchew.org@nongnu.org Return-Path: Received: from lists.gnu.org (lists.gnu.org [209.51.188.17]) by mx.zohomail.com with SMTPS id 1613144559987753.6440342178215; Fri, 12 Feb 2021 07:42:39 -0800 (PST) Received: from localhost ([::1]:51382 helo=lists1p.gnu.org) by lists.gnu.org with esmtp (Exim 4.90_1) (envelope-from ) id 1lAaac-0000M9-SF for importer@patchew.org; Fri, 12 Feb 2021 10:42:38 -0500 Received: from eggs.gnu.org ([2001:470:142:3::10]:50054) by lists.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.90_1) (envelope-from ) id 1lAaXG-0005wb-Ej; Fri, 12 Feb 2021 10:39:10 -0500 Received: from smtp2200-217.mail.aliyun.com ([121.197.200.217]:44531) by eggs.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.90_1) (envelope-from ) id 1lAaX4-0002WQ-IM; Fri, 12 Feb 2021 10:39:10 -0500 Received: from localhost.localdomain(mailfrom:zhiwei_liu@c-sky.com fp:SMTPD_---.JYH7gZk_1613144331) by smtp.aliyun-inc.com(10.147.44.145); Fri, 12 Feb 2021 23:38:52 +0800 X-Alimail-AntiSpam: AC=CONTINUE; BC=0.07436308|-1; CH=green; DM=|CONTINUE|false|; DS=CONTINUE|ham_system_inform|0.552192-0.00466644-0.443142; FP=0|0|0|0|0|-1|-1|-1; HT=ay29a033018047193; MF=zhiwei_liu@c-sky.com; NM=1; PH=DS; RN=6; RT=6; SR=0; TI=SMTPD_---.JYH7gZk_1613144331; From: LIU Zhiwei To: qemu-devel@nongnu.org Subject: [PATCH 17/38] target/riscv: Signed MSW 32x16 Multiply and Add Instructions Date: Fri, 12 Feb 2021 23:02:35 +0800 Message-Id: <20210212150256.885-18-zhiwei_liu@c-sky.com> X-Mailer: git-send-email 2.17.1 In-Reply-To: <20210212150256.885-1-zhiwei_liu@c-sky.com> References: <20210212150256.885-1-zhiwei_liu@c-sky.com> Received-SPF: pass (zohomail.com: domain of gnu.org designates 209.51.188.17 as permitted sender) client-ip=209.51.188.17; envelope-from=qemu-devel-bounces+importer=patchew.org@nongnu.org; helo=lists.gnu.org; Received-SPF: none client-ip=121.197.200.217; envelope-from=zhiwei_liu@c-sky.com; helo=smtp2200-217.mail.aliyun.com X-Spam_score_int: -18 X-Spam_score: -1.9 X-Spam_bar: - X-Spam_report: (-1.9 / 5.0 requ) BAYES_00=-1.9, SPF_HELO_NONE=0.001, SPF_NONE=0.001, UNPARSEABLE_RELAY=0.001 autolearn=ham autolearn_force=no X-Spam_action: no action X-BeenThere: qemu-devel@nongnu.org X-Mailman-Version: 2.1.23 Precedence: list List-Id: List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Cc: richard.henderson@linaro.org, LIU Zhiwei , qemu-riscv@nongnu.org, palmer@dabbelt.com, alistair23@gmail.com Errors-To: qemu-devel-bounces+importer=patchew.org@nongnu.org Sender: "Qemu-devel" Content-Transfer-Encoding: quoted-printable MIME-Version: 1.0 Content-Type: text/plain; charset="utf-8" Signed-off-by: LIU Zhiwei Acked-by: Alistair Francis --- target/riscv/helper.h | 17 ++ target/riscv/insn32.decode | 17 ++ target/riscv/insn_trans/trans_rvp.c.inc | 18 ++ target/riscv/packed_helper.c | 208 ++++++++++++++++++++++++ 4 files changed, 260 insertions(+) diff --git a/target/riscv/helper.h b/target/riscv/helper.h index 0bd21c8514..25aa07a7ff 100644 --- a/target/riscv/helper.h +++ b/target/riscv/helper.h @@ -1277,3 +1277,20 @@ DEF_HELPER_4(kmmsb, tl, env, tl, tl, tl) DEF_HELPER_4(kmmsb_u, tl, env, tl, tl, tl) DEF_HELPER_3(kwmmul, tl, env, tl, tl) DEF_HELPER_3(kwmmul_u, tl, env, tl, tl) + +DEF_HELPER_3(smmwb, tl, env, tl, tl) +DEF_HELPER_3(smmwb_u, tl, env, tl, tl) +DEF_HELPER_3(smmwt, tl, env, tl, tl) +DEF_HELPER_3(smmwt_u, tl, env, tl, tl) +DEF_HELPER_4(kmmawb, tl, env, tl, tl, tl) +DEF_HELPER_4(kmmawb_u, tl, env, tl, tl, tl) +DEF_HELPER_4(kmmawt, tl, env, tl, tl, tl) +DEF_HELPER_4(kmmawt_u, tl, env, tl, tl, tl) +DEF_HELPER_3(kmmwb2, tl, env, tl, tl) +DEF_HELPER_3(kmmwb2_u, tl, env, tl, tl) +DEF_HELPER_3(kmmwt2, tl, env, tl, tl) +DEF_HELPER_3(kmmwt2_u, tl, env, tl, tl) +DEF_HELPER_4(kmmawb2, tl, env, tl, tl, tl) +DEF_HELPER_4(kmmawb2_u, tl, env, tl, tl, tl) +DEF_HELPER_4(kmmawt2, tl, env, tl, tl, tl) +DEF_HELPER_4(kmmawt2_u, tl, env, tl, tl, tl) diff --git a/target/riscv/insn32.decode b/target/riscv/insn32.decode index e0be2790dc..6e63bab2d9 100644 --- a/target/riscv/insn32.decode +++ b/target/riscv/insn32.decode @@ -745,3 +745,20 @@ kmmsb 0100001 ..... ..... 001 ..... 1111111 @r kmmsb_u 0101001 ..... ..... 001 ..... 1111111 @r kwmmul 0110001 ..... ..... 001 ..... 1111111 @r kwmmul_u 0111001 ..... ..... 001 ..... 1111111 @r + +smmwb 0100010 ..... ..... 001 ..... 1111111 @r +smmwb_u 0101010 ..... ..... 001 ..... 1111111 @r +smmwt 0110010 ..... ..... 001 ..... 1111111 @r +smmwt_u 0111010 ..... ..... 001 ..... 1111111 @r +kmmawb 0100011 ..... ..... 001 ..... 1111111 @r +kmmawb_u 0101011 ..... ..... 001 ..... 1111111 @r +kmmawt 0110011 ..... ..... 001 ..... 1111111 @r +kmmawt_u 0111011 ..... ..... 001 ..... 1111111 @r +kmmwb2 1000111 ..... ..... 001 ..... 1111111 @r +kmmwb2_u 1001111 ..... ..... 001 ..... 1111111 @r +kmmwt2 1010111 ..... ..... 001 ..... 1111111 @r +kmmwt2_u 1011111 ..... ..... 001 ..... 1111111 @r +kmmawb2 1100111 ..... ..... 001 ..... 1111111 @r +kmmawb2_u 1101111 ..... ..... 001 ..... 1111111 @r +kmmawt2 1110111 ..... ..... 001 ..... 1111111 @r +kmmawt2_u 1111111 ..... ..... 001 ..... 1111111 @r diff --git a/target/riscv/insn_trans/trans_rvp.c.inc b/target/riscv/insn_tr= ans/trans_rvp.c.inc index fbc9c0b57b..e708ae7a6a 100644 --- a/target/riscv/insn_trans/trans_rvp.c.inc +++ b/target/riscv/insn_trans/trans_rvp.c.inc @@ -564,3 +564,21 @@ GEN_RVP_R_ACC_OOL(kmmsb); GEN_RVP_R_ACC_OOL(kmmsb_u); GEN_RVP_R_OOL(kwmmul); GEN_RVP_R_OOL(kwmmul_u); + +/* Most Significant Word "32x16" Multiply & Add Instructions */ +GEN_RVP_R_OOL(smmwb); +GEN_RVP_R_OOL(smmwb_u); +GEN_RVP_R_OOL(smmwt); +GEN_RVP_R_OOL(smmwt_u); +GEN_RVP_R_ACC_OOL(kmmawb); +GEN_RVP_R_ACC_OOL(kmmawb_u); +GEN_RVP_R_ACC_OOL(kmmawt); +GEN_RVP_R_ACC_OOL(kmmawt_u); +GEN_RVP_R_OOL(kmmwb2); +GEN_RVP_R_OOL(kmmwb2_u); +GEN_RVP_R_OOL(kmmwt2); +GEN_RVP_R_OOL(kmmwt2_u); +GEN_RVP_R_ACC_OOL(kmmawb2); +GEN_RVP_R_ACC_OOL(kmmawb2_u); +GEN_RVP_R_ACC_OOL(kmmawt2); +GEN_RVP_R_ACC_OOL(kmmawt2_u); diff --git a/target/riscv/packed_helper.c b/target/riscv/packed_helper.c index c1322d2fac..ea3c9f6dd8 100644 --- a/target/riscv/packed_helper.c +++ b/target/riscv/packed_helper.c @@ -1477,3 +1477,211 @@ static inline void do_kwmmul_u(CPURISCVState *env, = void *vd, void *va, } =20 RVPR(kwmmul_u, 1, 4); + +/* Most Significant Word "32x16" Multiply & Add Instructions */ +static inline void do_smmwb(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + d[H4(i)] =3D (int64_t)a[H4(i)] * b[H2(2 * i)] >> 16; +} + +RVPR(smmwb, 1, 4); + +static inline void do_smmwb_u(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + d[H4(i)] =3D ((int64_t)a[H4(i)] * b[H2(2 * i)] + (1ull << 15)) >> 16; +} + +RVPR(smmwb_u, 1, 4); + +static inline void do_smmwt(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + d[H4(i)] =3D (int64_t)a[H4(i)] * b[H2(2 * i + 1)] >> 16; +} + +RVPR(smmwt, 1, 4); + +static inline void do_smmwt_u(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + d[H4(i)] =3D ((int64_t)a[H4(i)] * b[H2(2 * i + 1)] + (1ull << 15)) >> = 16; +} + +RVPR(smmwt_u, 1, 4); + +static inline void do_kmmawb(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc; + int16_t *b =3D vb; + d[H4(i)] =3D sadd32(env, 0, (int64_t)a[H4(i)] * b[H2(2 * i)] >> 16, c[= H4(i)]); +} + +RVPR_ACC(kmmawb, 1, 4); + +static inline void do_kmmawb_u(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc; + int16_t *b =3D vb; + d[H4(i)] =3D sadd32(env, 0, ((int64_t)a[H4(i)] * b[H2(2 * i)] + + (1ull << 15)) >> 16, c[H4(i)]); +} + +RVPR_ACC(kmmawb_u, 1, 4); + +static inline void do_kmmawt(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc; + int16_t *b =3D vb; + d[H4(i)] =3D sadd32(env, 0, (int64_t)a[H4(i)] * b[H2(2 * i + 1)] >> 16, + c[H4(i)]); +} + +RVPR_ACC(kmmawt, 1, 4); + +static inline void do_kmmawt_u(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc; + int16_t *b =3D vb; + d[H4(i)] =3D sadd32(env, 0, ((int64_t)a[H4(i)] * b[H2(2 * i + 1)] + + (1ull << 15)) >> 16, c[H4(i)]); +} + +RVPR_ACC(kmmawt_u, 1, 4); + +static inline void do_kmmwb2(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + d[H4(i)] =3D INT32_MAX; + } else { + d[H4(i)] =3D (int64_t)a[H4(i)] * b[H2(2 * i)] >> 15; + } +} + +RVPR(kmmwb2, 1, 4); + +static inline void do_kmmwb2_u(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + d[H4(i)] =3D INT32_MAX; + } else { + d[H4(i)] =3D ((int64_t)a[H4(i)] * b[H2(2 * i)] + (1ull << 14)) >> = 15; + } +} + +RVPR(kmmwb2_u, 1, 4); + +static inline void do_kmmwt2(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i + 1)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + d[H4(i)] =3D INT32_MAX; + } else { + d[H4(i)] =3D (int64_t)a[H4(i)] * b[H2(2 * i + 1)] >> 15; + } +} + +RVPR(kmmwt2, 1, 4); + +static inline void do_kmmwt2_u(CPURISCVState *env, void *vd, void *va, + void *vb, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i + 1)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + d[H4(i)] =3D INT32_MAX; + } else { + d[H4(i)] =3D ((int64_t)a[H4(i)] * b[H2(2 * i + 1)] + (1ull << 14))= >> 15; + } +} + +RVPR(kmmwt2_u, 1, 4); + +static inline void do_kmmawb2(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc, result; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + result =3D INT32_MAX; + } else { + result =3D (int64_t)a[H4(i)] * b[H2(2 * i)] >> 15; + } + d[H4(i)] =3D sadd32(env, 0, result, c[H4(i)]); +} + +RVPR_ACC(kmmawb2, 1, 4); + +static inline void do_kmmawb2_u(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc, result; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + result =3D INT32_MAX; + } else { + result =3D ((int64_t)a[H4(i)] * b[H2(2 * i)] + (1ull << 14)) >> 15; + } + d[H4(i)] =3D sadd32(env, 0, result, c[H4(i)]); +} + +RVPR_ACC(kmmawb2_u, 1, 4); + +static inline void do_kmmawt2(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc, result; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i + 1)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + result =3D INT32_MAX; + } else { + result =3D (int64_t)a[H4(i)] * b[H2(2 * i + 1)] >> 15; + } + d[H4(i)] =3D sadd32(env, 0, result, c[H4(i)]); +} + +RVPR_ACC(kmmawt2, 1, 4); + +static inline void do_kmmawt2_u(CPURISCVState *env, void *vd, void *va, + void *vb, void *vc, uint8_t i) +{ + int32_t *d =3D vd, *a =3D va, *c =3D vc, result; + int16_t *b =3D vb; + if (a[H4(i)] =3D=3D INT32_MIN && b[H2(2 * i + 1)] =3D=3D INT16_MIN) { + env->vxsat =3D 0x1; + result =3D INT32_MAX; + } else { + result =3D ((int64_t)a[H4(i)] * b[H2(2 * i + 1)] + (1ull << 14)) >= > 15; + } + d[H4(i)] =3D sadd32(env, 0, result, c[H4(i)]); +} + +RVPR_ACC(kmmawt2_u, 1, 4); --=20 2.17.1