· Xiaojing Yang · Machine Learning
训练集、验证集和测试集
如何划分训练集、验证集和测试集,才能让模型评估保持诚实。
如何划分训练集、验证集和测试集,才能让模型评估保持诚实。
Transformers 如何组合 self-attention、feed-forward layers、residuals 和位置信息。
统计不是公式集合,而是帮助我们理解 AI 实验不确定性、证据强度和模型评估可信度的思维工具。
离散语言如何变成向量空间,以及句向量为什么对检索和评估重要。
Domain-specific machine translation is not only a modeling problem. It is a data, terminology, evaluation, and risk problem.
A short map of the blog: foundations, research applications, bilingual notes, and how this site complements my portfolio.