跳到正文
原文
arXiv cs.AI· Emile Anand, Abdullah Ateyeh, Archer Wang, Marin Solja\v{c}i\'c·· 4 小时前AI 评分37

SMat-Attention:结构化长上下文序列建模

SMat-Attention: Structured Long-Context Sequence Modeling

AI 导读

SMat-Attention 用行支持集 VC 维为 d 的结构化因果掩码,把 softmax 注意力与线性注意力统一为可调路由,d=1 时退化为标准因果掩码。

来源:arXiv cs.AI · arxiv.org