Tag: module-m2
-
【AI 核心深度 M2-020】解释双重下降(double descent)现象。(Explain the Double Descent Phenomenon in Modern Machine Learning)深度数理推导与工程落地解析
模型复杂度超过插值阈值后,测试误差先升后降;与经典 U 型曲线不同。
-
【AI 核心深度 M2-021】什么是模型选择中的’选择性偏差’?如何避免。(Define Selection Bias in Model Selection and How to Prevent Leakage in Predictive Pipelines)深度数理推导与工程落地解析
在大量模型/超参中挑验证集最好的,会高估真实性能;需用独立测试集或嵌套 CV。
-
【AI 核心深度 M2-022】解释决策树的分裂准则(信息增益 / 基尼不纯度)。(Explain Decision Tree Splitting Criteria: Information Gain, Gain Ratio, and Gini Impurity)深度数理推导与工程落地解析
选择使子节点不纯度下降最多的特征;ID3 用信息增益,CART 用基尼。
-
【AI 核心深度 M2-023】决策树如何防止过拟合?列出主要手段。(Detail Pre-Pruning and Post-Pruning Strategies to Prevent Overfitting in Decision Trees)深度数理推导与工程落地解析
预剪枝(最大深度/最小样本/最小增益)+ 后剪枝(代价复杂度)+ 集成。
-
【AI 核心深度 M2-024】决策树为什么不需要特征缩放?它的归纳偏置是什么。(Explain Why Decision Trees are Invariant to Monotonic Transformations and Their Inductive Bias)深度数理推导与工程落地解析
分裂基于阈值比较,对单调变换不敏感;归纳偏置是轴平行分割,难以建模线性/旋转关系。
-
【AI 核心深度 M2-026】如何处理缺失值与类别特征?对比常见做法。(Compare Industrial Strategies for Handling Missing Values and Categorical Features in Tree Models)深度数理推导与工程落地解析
缺失值可用代理分裂(CART)、按缺失率分流、或缺失指示变量;类别特征可用 one-hot、目标编码、或原生类别分裂(LightGBM/CatBoost)。
-
【AI 核心深度 M2-001】写出线性回归的闭式解,并说明它成立的前提。(Formulate the Closed-Form Normal Equation for Linear Regression and Its Necessary Preconditions)深度数理推导与工程落地解析
最小二乘解 w=(XᵀX)⁻¹Xᵀy,前提是 XᵀX 可逆(无完全共线性)。