SHAP 值:全局与局部特征重要性
学习者将计算梯度提升模型的 SHAP 值,绘制蜂群图和条形汇总图,并向非技术利益相关者解释单次预测结果。
SHAP 值:全局与局部特征重要性 是 CoddyKit 上的免费 Machine Learning Academy 课时。 这是第 1 节课,共 4 节。 你可以在下方免费阅读本课时的完整内容 — 然后在浏览器中使用内置代码编辑器和全天候 AI 导师进行实践。 这是 Machine Learning Academy 学习路径的一部分,你的进度在网页和 CoddyKit 应用中同步。 Machine Learning Academy 课程共包含 4 节课。
本课时的部分内容尚未翻译,以英文显示。
Why Model Explainability Matters
Explainability is the ability to understand why a model made a specific prediction. In high-stakes domains like lending, healthcare, and hiring, regulators and users demand explanations — not just accurate predictions. SHAP (SHapley Additive exPlanations) provides a mathematically principled framework rooted in cooperative game theory to deliver these explanations for any model.
Shapley Values: The Game Theory Origin
SHAP values borrow from Shapley values in cooperative game theory, where players (features) collaborate to produce an outcome (prediction). Each feature receives a fair share of credit by averaging its marginal contribution across all possible feature orderings. This makes SHAP the only additive attribution method satisfying the axioms of efficiency, symmetry, dummy, and additivity.
Installing and Importing SHAP
The shap library supports tree models, neural networks, and any black-box model. Install it with pip install shap and import it alongside your trained model. SHAP's explainers are model-type-aware: TreeExplainer for gradient boosting and random forests gives exact values in O(T·D²) time, far faster than the naive exponential-time Shapley computation.
import shap
import xgboost as xgb
from sklearn.datasets import load_breast_cancer
from sklearn.model_selection import train_test_split
data = load_breast_cancer()
X_train, X_test, y_train, y_test = train_test_split(
data.data, data.target, test_size=0.2, random_state=42
)
model = xgb.XGBClassifier(n_estimators=100, use_label_encoder=False, eval_metric='logloss')
model.fit(X_train, y_train)
explainer = shap.TreeExplainer(model)Computing SHAP Values for the Test Set
Call explainer.shap_values(X_test) to produce a matrix where each row is a sample and each column is a feature. The SHAP value for feature j in sample i represents the contribution of feature j to pushing the prediction away from the expected base value. Positive values push toward the positive class; negative values push toward the negative class.
shap_values = explainer.shap_values(X_test)
print('SHAP values shape:', shap_values.shape) # (n_samples, n_features)
print('Base value (expected prediction):', explainer.expected_value)
print('First sample SHAP values:', shap_values[0])Global Importance: Bar Plot
Global feature importance summarises which features matter most across all predictions. The SHAP summary_plot in bar mode shows the mean absolute SHAP value per feature, ranking them from most to least important. This replaces the naive built-in feature importance that only counts split counts, which is biased toward high-cardinality features.
import matplotlib.pyplot as plt
# Bar plot: mean |SHAP| per feature
shap.summary_plot(shap_values, X_test,
feature_names=data.feature_names,
plot_type='bar')
plt.tight_layout()
plt.savefig('shap_bar.png', dpi=150)Global Importance: Beeswarm Plot
The beeswarm plot (default summary_plot) is richer than a bar chart: each dot represents one sample, coloured by feature value (red = high, blue = low). The x-axis shows the SHAP value, so you can see not only which features matter but also in which direction a high or low feature value pushes predictions. This reveals nonlinear and interaction effects at a glance.
shap.summary_plot(shap_values, X_test,
feature_names=data.feature_names)
# Dots to the right = positive contribution to predicted class
# Red dots far right = high feature value strongly increases predictionLocal Explanation: Force Plot
A force plot explains a single prediction. It shows the base value on the left and the final prediction on the right, with features as arrows that push the output higher (red) or lower (blue). The width of each arrow is proportional to the feature's SHAP value. This is the explanation you would show a loan officer asking 'why was this application denied?'
# Explain the first test sample
i = 0
shap.force_plot(
explainer.expected_value,
shap_values[i],
X_test[i],
feature_names=data.feature_names,
matplotlib=True
)Local Explanation: Waterfall Plot
The waterfall plot is a cleaner alternative to the force plot for a single sample. It stacks SHAP contributions vertically from the base value, showing each feature's contribution as a bar segment. Positive contributions are red and push toward the top; negative contributions are blue and pull down. The final stack total equals the model's raw output for that sample.
import shap
explanation = shap.Explanation(
values=shap_values[0],
base_values=explainer.expected_value,
data=X_test[0],
feature_names=list(data.feature_names)
)
shap.waterfall_plot(explanation)Dependence Plot: Feature Interactions
A SHAP dependence plot shows how a single feature's SHAP value changes as its raw value changes, coloured by a second feature to reveal interactions. For example, plotting 'worst radius' coloured by 'mean texture' reveals whether the effect of radius depends on texture. This goes beyond ordinary partial-dependence plots by accounting for all feature interactions naturally.
shap.dependence_plot(
'worst radius', # feature to plot on x-axis
shap_values,
X_test,
feature_names=list(data.feature_names),
interaction_index='mean texture' # colour by this feature
)SHAP with Any Model: KernelExplainer
When the model is a black box (SVM, neural network, any sklearn estimator), use shap.KernelExplainer, which approximates Shapley values by sampling coalitions and fitting a weighted linear model locally. It is model-agnostic but slower than TreeExplainer. Provide a background dataset summary (e.g., K-Means centroids) to speed up computation on large datasets.
from sklearn.svm import SVC
from sklearn.pipeline import Pipeline
from sklearn.preprocessing import StandardScaler
import shap
import numpy as np
pipeline = Pipeline([('scaler', StandardScaler()), ('svm', SVC(probability=True))])
pipeline.fit(X_train, y_train)
# Use 50 background samples for speed
background = shap.kmeans(X_train, 50)
explainer_k = shap.KernelExplainer(pipeline.predict_proba, background)
shap_vals_k = explainer_k.shap_values(X_test[:10]) # Explain 10 samplesValidation: SHAP Values Sum to Prediction
A key property of SHAP is efficiency: the sum of all SHAP values for a sample plus the base value must equal the model's raw output. Verifying this sanity check confirms the explainer is working correctly. Any discrepancy indicates a mismatch between the explainer type and the model, or incorrect background data.
import numpy as np
# For tree models, verify SHAP values sum to log-odds output
base = explainer.expected_value
for i in range(5):
shap_sum = shap_values[i].sum() + base
raw_pred = model.predict(X_test[i:i+1], output_margin=True)[0]
print(f'Sample {i}: SHAP sum={shap_sum:.4f}, model output={raw_pred:.4f}, match={abs(shap_sum-raw_pred)<1e-4}')Quick Check
Test your understanding of Machine Learning with Python concepts from this lesson.
Lesson Recap
In this lesson you learned: SHAP values quantify each feature's contribution to a prediction using Shapley values from game theory, global summaries (bar and beeswarm plots) reveal overall feature importance and direction, and local explanations (force and waterfall plots) justify individual predictions. Next up we explore LIME as an alternative model-agnostic explanation approach.
用 AI 导师学习 Python — 免费
在浏览器中编写并运行真实代码,获得全天候 AI 导师的即时帮助,并在网页或应用中继续学习。
- 课程
- 30
- 课程
- 120
常见问题解答
「SHAP 值:全局与局部特征重要性」课时是免费的吗?
是的 — 「SHAP 值:全局与局部特征重要性」的完整文本可在网页上免费阅读。要进行交互式练习(内置代码编辑器和全天候 AI 导师)并解锁 Machine Learning Academy 课程的其余内容,请升级到 CoddyKit PRO。 Machine Learning Academy 课程共包含 4 节课。
「SHAP 值:全局与局部特征重要性」这节课中我会学到什么?
学习者将计算梯度提升模型的 SHAP 值,绘制蜂群图和条形汇总图,并向非技术利益相关者解释单次预测结果。 你通过在浏览器中直接运行的动手代码来练习 Machine Learning Academy,全天候 AI 导师会在你学习这节课的过程中回答你的问题。
学习 Machine Learning Academy 需要有经验吗?
无需任何先前经验。CoddyKit 上的 Machine Learning Academy 课程适合初学者到高级学习者,你可以从这里开始或从头开始,按照自己的节奏学习。 这是第 1 节课,共 4 节。
「SHAP 值:全局与局部特征重要性」课时需要多长时间?
大多数 CoddyKit 课程大约需要 5–10 分钟。每节课都很精短且互动,所以你能稳步进步,并在网页和应用中从离开的地方继续。
我能在这节 Machine Learning Academy 课中编写并运行代码吗?
能。每节 Machine Learning Academy 课都包含内置代码编辑器,你可以在浏览器中直接编写并运行真实代码,并获得即时 AI 反馈 — 无需本地设置。
此课程中的所有课时
- SHAP 值:全局与局部特征重要性
- LIME:局部可解释的模型无关解释
- 公平性指标:人口统计均等与机会均等
- 偏差缓解策略:预处理、处理中与后处理