From patchwork Fri Feb 7 14:44:30 2025 Content-Type: text/plain; charset="utf-8" MIME-Version: 1.0 Content-Transfer-Encoding: 7bit X-Patchwork-Submitter: Shiju Jose X-Patchwork-Id: 13965175 Return-Path: X-Spam-Checker-Version: SpamAssassin 3.4.0 (2014-02-07) on aws-us-west-2-korg-lkml-1.web.codeaurora.org Received: from kanga.kvack.org (kanga.kvack.org [205.233.56.17]) by smtp.lore.kernel.org (Postfix) with ESMTP id 505A3C02194 for ; Fri, 7 Feb 2025 14:48:05 +0000 (UTC) Received: by kanga.kvack.org (Postfix) id C83B328000B; Fri, 7 Feb 2025 09:48:04 -0500 (EST) Received: by kanga.kvack.org (Postfix, from userid 40) id C328E280001; Fri, 7 Feb 2025 09:48:04 -0500 (EST) X-Delivered-To: int-list-linux-mm@kvack.org Received: by kanga.kvack.org (Postfix, from userid 63042) id AD38628000B; Fri, 7 Feb 2025 09:48:04 -0500 (EST) X-Delivered-To: linux-mm@kvack.org Received: from relay.hostedemail.com (smtprelay0010.hostedemail.com [216.40.44.10]) by kanga.kvack.org (Postfix) with ESMTP id 8A7D8280001 for ; Fri, 7 Feb 2025 09:48:04 -0500 (EST) Received: from smtpin01.hostedemail.com (a10.router.float.18 [10.200.18.1]) by unirelay04.hostedemail.com (Postfix) with ESMTP id A48491A1C84 for ; Fri, 7 Feb 2025 14:45:34 +0000 (UTC) X-FDA: 83093422230.01.333719C Received: from frasgout.his.huawei.com (frasgout.his.huawei.com [185.176.79.56]) by imf20.hostedemail.com (Postfix) with ESMTP id C982A1C0018 for ; Fri, 7 Feb 2025 14:45:30 +0000 (UTC) Authentication-Results: imf20.hostedemail.com; dkim=none; spf=pass (imf20.hostedemail.com: domain of shiju.jose@huawei.com designates 185.176.79.56 as permitted sender) smtp.mailfrom=shiju.jose@huawei.com; dmarc=pass (policy=quarantine) header.from=huawei.com ARC-Message-Signature: i=1; a=rsa-sha256; c=relaxed/relaxed; d=hostedemail.com; s=arc-20220608; t=1738939531; h=from:from:sender:reply-to:subject:subject:date:date: message-id:message-id:to:to:cc:cc:mime-version:mime-version: content-type:content-type: content-transfer-encoding:content-transfer-encoding: in-reply-to:in-reply-to:references:references; bh=UHasrmCt/piE1mYLQudU+Tt2bXryFZ671hqlkw9vyMk=; b=amxgu0E734IVcyY8V1h8bvnO3JiXyMfyjjzxI/EoSacsWkywHs9UOZQdTwCXZ45hxJavIy RkOoMYEr78W+6RN8CizhS5MUIN5LXLL8TOlmqZoo/ObfBBwXm21VKv9XS22utPVFbO6WjO 7yzwgJYql3QxNN7g1Z6RbBRhwSIwM5g= ARC-Authentication-Results: i=1; imf20.hostedemail.com; dkim=none; spf=pass (imf20.hostedemail.com: domain of shiju.jose@huawei.com designates 185.176.79.56 as permitted sender) smtp.mailfrom=shiju.jose@huawei.com; dmarc=pass (policy=quarantine) header.from=huawei.com ARC-Seal: i=1; s=arc-20220608; d=hostedemail.com; t=1738939531; a=rsa-sha256; cv=none; b=gHAkv/AcKQq0izz9H0jF7CNMfuGX30R270DZTEwOzwGjGk3yYi50pl2kX/+zQbZ0tvkSzt fs6pPNbrl1P4gXGJ9is1fMhZTbKtE9oo4k6ITs0gDXIC3/cLsuKKwgncXR59Ug+hNk4+jT JcBze2AlMa/vDYwKuL1/hpZ2UCJJQp0= Received: from mail.maildlp.com (unknown [172.18.186.216]) by frasgout.his.huawei.com (SkyGuard) with ESMTP id 4YqGvY2nxNz6HJZ8; Fri, 7 Feb 2025 22:44:25 +0800 (CST) Received: from frapeml500007.china.huawei.com (unknown [7.182.85.172]) by mail.maildlp.com (Postfix) with ESMTPS id A43AD1408F9; Fri, 7 Feb 2025 22:45:26 +0800 (CST) Received: from P_UKIT01-A7bmah.china.huawei.com (10.126.173.5) by frapeml500007.china.huawei.com (7.182.85.172) with Microsoft SMTP Server (version=TLS1_2, cipher=TLS_ECDHE_RSA_WITH_AES_256_GCM_SHA384) id 15.1.2507.39; Fri, 7 Feb 2025 15:45:23 +0100 From: To: , , , , CC: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Subject: [PATCH v19 01/15] EDAC: Add support for EDAC device features control Date: Fri, 7 Feb 2025 14:44:30 +0000 Message-ID: <20250207144445.1879-2-shiju.jose@huawei.com> X-Mailer: git-send-email 2.43.0.windows.1 In-Reply-To: <20250207144445.1879-1-shiju.jose@huawei.com> References: <20250207144445.1879-1-shiju.jose@huawei.com> MIME-Version: 1.0 X-Originating-IP: [10.126.173.5] X-ClientProxiedBy: lhrpeml500001.china.huawei.com (7.191.163.213) To frapeml500007.china.huawei.com (7.182.85.172) X-Rspam-User: X-Rspamd-Queue-Id: C982A1C0018 X-Stat-Signature: 8e416oiwtcnnf11wzipbxcrqnpegsiew X-Rspamd-Server: rspam03 X-HE-Tag: 1738939530-737902 X-HE-Meta: U2FsdGVkX19CAHxafn7BymdkfT7hQovEBqnV3iJo7BPb5xYbuREDIuxFj97rx7WjqAgTn1Meu1Gw5m71MkC8lE7W/X9J1eYb9TuxyLmEPWGt+yIvdwnk2HqSBOkSUB2/s/F4Y55gVuntpwq8VNUyk7qlBSsOeE0XKZts15bBrLh8vOTFbLxL7ea27Mjcd/WwzQ3LTl3p47FONnjB9+t/5NLbgyArBuoTKI8hFHzRqgW+fs0G4XJIs55977RIktlns4NejQBOrMmLAbD9Jn0tZzYOEXaMvlR2FqXAR8YmEJNIagYbpaYZSXcYU/Tv2L5oidv4RaWluQua3qgrarnWs1uR0cbs5nGVeTyEYwywP8x3lstxxXmDdvOqGXQjGEl24mmggOk78poQsIVsm7TabLlGL2F2THxehFdDx0ej3YobIPO+rnHUTUE3KREHu+2/93uNxChlE0y2dQwTjgDdIAYCb0oc83WGuFt4ptvGyE01fp1QnVN4ObmGr0PmqjYX7HZoRBLOfmeCzJ7PfkImX6dsjbv4OEWvUpxKLfE/UckKwMnIWK/MH0j+Nbfn5XGjnbOdCkY6RN0rgJ0uRjcThwgR5QxCWtu0oNFB6wfL90Y73oKyGyK5J6f+qdie0E3ewf+whXi5ieHGZg3AaikR51WVDSgZti3VdFlzlbQXMx37vIdZYFKn4MsxqWCO1T+HPYyRFu816Qv5qd2Xbue55IXggzDGyzv/b0/r1bFMufQGJ3svnD27ZvtHjfRc0zic3GSzB2SVoDGsghu6/qEMA7c9CbmnKiu0K5kOP1iGgaBO8svzqoCb8SmHJztEdv5y7VUQzq/zg8hJV7DbXCtSAD5mJAi0YJBQKnik9N70BS7BHIFbL1f5FThaVdaM3Al3p3WVvY2erYpn0xuVdVgzyHUVlAKF19do2wqkg5Jh8KsmPVE/GL++qpifWf/bdBAZ0j31aDWr2LEGvJrFUKM 60ZIJvYW LwLDCtAYB07PD3DMLdD41N6LBomoqrgo6f4kOcISxc/XzYKyKS4Z313NTkADcXref5C4uxWrzmDNKoClFWLwVaaB8OCfd0IR0R5vhdaLElvwO5JBW9bTV3ht9BSHYbhbXGuEqpomRAuywSowOoqcWnrt7I5P3UtAEdrgS5Or2X1sK5NC04fzn9zQWPkXgA8yJJmdiA1VgW7e5Q7VT9Wpfn4vHgPV53MtZ1ivxZzYHHUNj6kHAxttmj2CEfeK6x+tMP0JPiKUMR7h9o4Y= X-Bogosity: Ham, tests=bogofilter, spamicity=0.000000, version=1.2.4 Sender: owner-linux-mm@kvack.org Precedence: bulk X-Loop: owner-majordomo@kvack.org List-ID: List-Subscribe: List-Unsubscribe: From: Shiju Jose Add generic EDAC device feature controls supporting the registration of RAS features available in the system. The driver exposes control attributes for these features to userspace in /sys/bus/edac/devices/// Co-developed-by: Jonathan Cameron Signed-off-by: Jonathan Cameron Tested-by: Daniel Ferguson Signed-off-by: Shiju Jose --- Documentation/edac/features.rst | 94 +++++++++++++++++++++++++++++ Documentation/edac/index.rst | 10 ++++ drivers/edac/edac_device.c | 102 ++++++++++++++++++++++++++++++++ include/linux/edac.h | 26 ++++++++ 4 files changed, 232 insertions(+) create mode 100644 Documentation/edac/features.rst create mode 100644 Documentation/edac/index.rst diff --git a/Documentation/edac/features.rst b/Documentation/edac/features.rst new file mode 100644 index 000000000000..6b0fdc6f5d6e --- /dev/null +++ b/Documentation/edac/features.rst @@ -0,0 +1,94 @@ +.. SPDX-License-Identifier: GPL-2.0 OR GFDL-1.2-no-invariants-or-later + +============================================ +Augmenting EDAC for controlling RAS features +============================================ + +Copyright (c) 2024-2025 HiSilicon Limited. + +:Author: Shiju Jose +:License: The GNU Free Documentation License, Version 1.2 without + Invariant Sections, Front-Cover Texts nor Back-Cover Texts. + (dual licensed under the GPL v2) + +- Written for: 6.15 + +Introduction +------------ +The expansion of EDAC for controlling RAS features and exposing features +control attributes to userspace via sysfs. Some Examples: + +1. Scrub control + +2. Error Check Scrub (ECS) control + +3. ACPI RAS2 features + +4. Post Package Repair (PPR) control + +5. Memory Sparing Repair control etc. + +High level design is illustrated in the following diagram:: + + +-----------------------------------------------+ + | Userspace - Rasdaemon | + | +-------------+ | + | | RAS CXL mem | +---------------+ | + | |error handler|---->| | | + | +-------------+ | RAS dynamic | | + | +-------------+ | scrub, memory | | + | | RAS memory |---->| repair control| | + | |error handler| +----|----------+ | + | +-------------+ | | + +--------------------------|--------------------+ + | + | + +-------------------------------|------------------------------+ + | Kernel EDAC extension for | controlling RAS Features | + |+------------------------------|----------------------------+ | + || EDAC Core Sysfs EDAC| Bus | | + || +--------------------------|---------------------------+| | + || |/sys/bus/edac/devices//scrubX/ | | EDAC device || | + || |/sys/bus/edac/devices//ecsX/ |<->| EDAC MC || | + || |/sys/bus/edac/devices//repairX | | EDAC sysfs || | + || +---------------------------|--------------------------+| | + || EDAC|Bus | | + || | | | + || +----------+ Get feature | Get feature | | + || | | desc +---------|------+ desc +----------+ | | + || |EDAC scrub|<-----| EDAC device | | | | | + || +----------+ | driver- RAS |----->| EDAC mem | | | + || +----------+ | feature control| | repair | | | + || | |<-----| | +----------+ | | + || |EDAC ECS | +---------|------+ | | + || +----------+ Register RAS|features | | + || ______________________|_____________ | | + |+---------|---------------|------------------|--------------+ | + | +-------|----+ +-------|-------+ +----|----------+ | + | | | | CXL mem driver| | Client driver | | + | | ACPI RAS2 | | scrub, ECS, | | memory repair | | + | | driver | | sparing, PPR | | features | | + | +-----|------+ +-------|-------+ +------|--------+ | + | | | | | + +--------|-----------------|--------------------|--------------+ + | | | + +--------|-----------------|--------------------|--------------+ + | +---|-----------------|--------------------|-------+ | + | | | | + | | Platform HW and Firmware | | + | +--------------------------------------------------+ | + +--------------------------------------------------------------+ + + +1. EDAC Features components - Create feature specific descriptors. + For example, EDAC scrub, EDAC ECS, EDAC memory repair in the above + diagram. + +2. EDAC device driver for controlling RAS Features - Get feature's attribute + descriptors from EDAC RAS feature component and registers device's RAS + features with EDAC bus and exposes the features control attributes via + the sysfs EDAC bus. For example, /sys/bus/edac/devices//X/ + +3. RAS dynamic feature controller - Userspace sample modules in rasdaemon for + dynamic scrub/repair control to issue scrubbing/repair when excess number + of corrected memory errors are reported in a short span of time. diff --git a/Documentation/edac/index.rst b/Documentation/edac/index.rst new file mode 100644 index 000000000000..de4a3aa452cb --- /dev/null +++ b/Documentation/edac/index.rst @@ -0,0 +1,10 @@ +.. SPDX-License-Identifier: GPL-2.0 OR GFDL-1.2-no-invariants-or-later + +============== +EDAC Subsystem +============== + +.. toctree:: + :maxdepth: 1 + + features diff --git a/drivers/edac/edac_device.c b/drivers/edac/edac_device.c index 621dc2a5d034..142a661ff543 100644 --- a/drivers/edac/edac_device.c +++ b/drivers/edac/edac_device.c @@ -570,3 +570,105 @@ void edac_device_handle_ue_count(struct edac_device_ctl_info *edac_dev, block ? block->name : "N/A", count, msg); } EXPORT_SYMBOL_GPL(edac_device_handle_ue_count); + +static void edac_dev_release(struct device *dev) +{ + struct edac_dev_feat_ctx *ctx = container_of(dev, struct edac_dev_feat_ctx, dev); + + kfree(ctx->dev.groups); + kfree(ctx); +} + +const struct device_type edac_dev_type = { + .name = "edac_dev", + .release = edac_dev_release, +}; + +static void edac_dev_unreg(void *data) +{ + device_unregister(data); +} + +/** + * edac_dev_register - register device for RAS features with EDAC + * @parent: parent device. + * @name: name for the folder in the /sys/bus/edac/devices/, + * which is derived from the parent device. + * For eg. /sys/bus/edac/devices/cxl_mem0/ + * @private: parent driver's data to store in the context if any. + * @num_features: number of RAS features to register. + * @ras_features: list of RAS features to register. + * + * Return: + * * %0 - Success. + * * %-EINVAL - Invalid parameters passed. + * * %-ENOMEM - Dynamic memory allocation failed. + * + */ +int edac_dev_register(struct device *parent, char *name, + void *private, int num_features, + const struct edac_dev_feature *ras_features) +{ + const struct attribute_group **ras_attr_groups; + struct edac_dev_feat_ctx *ctx; + int attr_gcnt = 0; + int ret, feat; + + if (!parent || !name || !num_features || !ras_features) + return -EINVAL; + + /* Double parse to make space for attributes */ + for (feat = 0; feat < num_features; feat++) { + switch (ras_features[feat].ft_type) { + /* Add feature specific code */ + default: + return -EINVAL; + } + } + + ctx = kzalloc(sizeof(*ctx), GFP_KERNEL); + if (!ctx) + return -ENOMEM; + + ras_attr_groups = kcalloc(attr_gcnt + 1, sizeof(*ras_attr_groups), GFP_KERNEL); + if (!ras_attr_groups) { + ret = -ENOMEM; + goto ctx_free; + } + + attr_gcnt = 0; + for (feat = 0; feat < num_features; feat++, ras_features++) { + switch (ras_features->ft_type) { + /* Add feature specific code */ + default: + ret = -EINVAL; + goto groups_free; + } + } + + ctx->dev.parent = parent; + ctx->dev.bus = edac_get_sysfs_subsys(); + ctx->dev.type = &edac_dev_type; + ctx->dev.groups = ras_attr_groups; + ctx->private = private; + dev_set_drvdata(&ctx->dev, ctx); + + ret = dev_set_name(&ctx->dev, name); + if (ret) + goto groups_free; + + ret = device_register(&ctx->dev); + if (ret) { + put_device(&ctx->dev); + return ret; + } + + return devm_add_action_or_reset(parent, edac_dev_unreg, &ctx->dev); + +groups_free: + kfree(ras_attr_groups); +ctx_free: + kfree(ctx); + return ret; +} +EXPORT_SYMBOL_GPL(edac_dev_register); diff --git a/include/linux/edac.h b/include/linux/edac.h index b4ee8961e623..8c4b6ca2a994 100644 --- a/include/linux/edac.h +++ b/include/linux/edac.h @@ -661,4 +661,30 @@ static inline struct dimm_info *edac_get_dimm(struct mem_ctl_info *mci, return mci->dimms[index]; } + +/* RAS feature type */ +enum edac_dev_feat { + RAS_FEAT_MAX +}; + +/* EDAC device feature information structure */ +struct edac_dev_data { + u8 instance; + void *private; +}; + +struct edac_dev_feat_ctx { + struct device dev; + void *private; +}; + +struct edac_dev_feature { + enum edac_dev_feat ft_type; + u8 instance; + void *ctx; +}; + +int edac_dev_register(struct device *parent, char *dev_name, + void *parent_pvt_data, int num_features, + const struct edac_dev_feature *ras_features); #endif /* _LINUX_EDAC_H_ */