<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>序列建模 on Answer</title>
    <link>https://answer.freetools.me/tags/%E5%BA%8F%E5%88%97%E5%BB%BA%E6%A8%A1/</link>
    <description>Recent content in 序列建模 on Answer</description>
    <generator>Hugo -- 0.152.2</generator>
    <language>zh-cn</language>
    <lastBuildDate>Thu, 12 Mar 2026 23:08:50 +0800</lastBuildDate>
    <atom:link href="https://answer.freetools.me/tags/%E5%BA%8F%E5%88%97%E5%BB%BA%E6%A8%A1/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>RNN为什么无法记住超过二十步的信息：从梯度消失到现代序列模型的四十年技术突围</title>
      <link>https://answer.freetools.me/rnn%E4%B8%BA%E4%BB%80%E4%B9%88%E6%97%A0%E6%B3%95%E8%AE%B0%E4%BD%8F%E8%B6%85%E8%BF%87%E4%BA%8C%E5%8D%81%E6%AD%A5%E7%9A%84%E4%BF%A1%E6%81%AF%E4%BB%8E%E6%A2%AF%E5%BA%A6%E6%B6%88%E5%A4%B1%E5%88%B0%E7%8E%B0%E4%BB%A3%E5%BA%8F%E5%88%97%E6%A8%A1%E5%9E%8B%E7%9A%84%E5%9B%9B%E5%8D%81%E5%B9%B4%E6%8A%80%E6%9C%AF%E7%AA%81%E5%9B%B4/</link>
      <pubDate>Thu, 12 Mar 2026 23:08:50 +0800</pubDate>
      <guid>https://answer.freetools.me/rnn%E4%B8%BA%E4%BB%80%E4%B9%88%E6%97%A0%E6%B3%95%E8%AE%B0%E4%BD%8F%E8%B6%85%E8%BF%87%E4%BA%8C%E5%8D%81%E6%AD%A5%E7%9A%84%E4%BF%A1%E6%81%AF%E4%BB%8E%E6%A2%AF%E5%BA%A6%E6%B6%88%E5%A4%B1%E5%88%B0%E7%8E%B0%E4%BB%A3%E5%BA%8F%E5%88%97%E6%A8%A1%E5%9E%8B%E7%9A%84%E5%9B%9B%E5%8D%81%E5%B9%B4%E6%8A%80%E6%9C%AF%E7%AA%81%E5%9B%B4/</guid>
      <description>深入解析循环神经网络梯度消失问题的数学本质，从Hochreiter 1991年的开创性发现到LSTM的常数误差转盘机制，揭示为什么这个困扰深度学习三十年的问题催生了从门控循环单元到Transformer的完整技术演进。</description>
    </item>
    <item>
      <title>LSTM长短期记忆网络：为什么这个门控机制统治了序列建模二十年</title>
      <link>https://answer.freetools.me/lstm%E9%95%BF%E7%9F%AD%E6%9C%9F%E8%AE%B0%E5%BF%86%E7%BD%91%E7%BB%9C%E4%B8%BA%E4%BB%80%E4%B9%88%E8%BF%99%E4%B8%AA%E9%97%A8%E6%8E%A7%E6%9C%BA%E5%88%B6%E7%BB%9F%E6%B2%BB%E4%BA%86%E5%BA%8F%E5%88%97%E5%BB%BA%E6%A8%A1%E4%BA%8C%E5%8D%81%E5%B9%B4/</link>
      <pubDate>Thu, 12 Mar 2026 12:33:22 +0800</pubDate>
      <guid>https://answer.freetools.me/lstm%E9%95%BF%E7%9F%AD%E6%9C%9F%E8%AE%B0%E5%BF%86%E7%BD%91%E7%BB%9C%E4%B8%BA%E4%BB%80%E4%B9%88%E8%BF%99%E4%B8%AA%E9%97%A8%E6%8E%A7%E6%9C%BA%E5%88%B6%E7%BB%9F%E6%B2%BB%E4%BA%86%E5%BA%8F%E5%88%97%E5%BB%BA%E6%A8%A1%E4%BA%8C%E5%8D%81%E5%B9%B4/</guid>
      <description>深入解析LSTM的核心原理、数学推导、梯度流机制，以及与GRU和Transformer的对比分析，理解为什么LSTM能够解决RNN的梯度消失问题，以及在什么场景下LSTM仍然优于Transformer。</description>
    </item>
    <item>
      <title>测试时训练：当模型在推理阶段继续学习会发生什么</title>
      <link>https://answer.freetools.me/%E6%B5%8B%E8%AF%95%E6%97%B6%E8%AE%AD%E7%BB%83%E5%BD%93%E6%A8%A1%E5%9E%8B%E5%9C%A8%E6%8E%A8%E7%90%86%E9%98%B6%E6%AE%B5%E7%BB%A7%E7%BB%AD%E5%AD%A6%E4%B9%A0%E4%BC%9A%E5%8F%91%E7%94%9F%E4%BB%80%E4%B9%88/</link>
      <pubDate>Mon, 09 Mar 2026 06:26:43 +0800</pubDate>
      <guid>https://answer.freetools.me/%E6%B5%8B%E8%AF%95%E6%97%B6%E8%AE%AD%E7%BB%83%E5%BD%93%E6%A8%A1%E5%9E%8B%E5%9C%A8%E6%8E%A8%E7%90%86%E9%98%B6%E6%AE%B5%E7%BB%A7%E7%BB%AD%E5%AD%A6%E4%B9%A0%E4%BC%9A%E5%8F%91%E7%94%9F%E4%BB%80%E4%B9%88/</guid>
      <description>深入解析Test-Time Training (TTT) 的核心原理与技术演进。从TTT层的隐藏状态即模型设计，到TTT-E2E的长上下文突破，再到TTT-Discover的科学发现能力，全面探讨测试时训练如何打破传统训练与推理的边界，让AI模型在推理过程中持续进化。</description>
    </item>
    <item>
      <title>从Transformer的二次复杂度困境到Mamba的线性突围：状态空间模型如何重塑序列建模</title>
      <link>https://answer.freetools.me/%E4%BB%8Etransformer%E7%9A%84%E4%BA%8C%E6%AC%A1%E5%A4%8D%E6%9D%82%E5%BA%A6%E5%9B%B0%E5%A2%83%E5%88%B0mamba%E7%9A%84%E7%BA%BF%E6%80%A7%E7%AA%81%E5%9B%B4%E7%8A%B6%E6%80%81%E7%A9%BA%E9%97%B4%E6%A8%A1%E5%9E%8B%E5%A6%82%E4%BD%95%E9%87%8D%E5%A1%91%E5%BA%8F%E5%88%97%E5%BB%BA%E6%A8%A1/</link>
      <pubDate>Mon, 09 Mar 2026 06:07:08 +0800</pubDate>
      <guid>https://answer.freetools.me/%E4%BB%8Etransformer%E7%9A%84%E4%BA%8C%E6%AC%A1%E5%A4%8D%E6%9D%82%E5%BA%A6%E5%9B%B0%E5%A2%83%E5%88%B0mamba%E7%9A%84%E7%BA%BF%E6%80%A7%E7%AA%81%E5%9B%B4%E7%8A%B6%E6%80%81%E7%A9%BA%E9%97%B4%E6%A8%A1%E5%9E%8B%E5%A6%82%E4%BD%95%E9%87%8D%E5%A1%91%E5%BA%8F%E5%88%97%E5%BB%BA%E6%A8%A1/</guid>
      <description>深入解析Mamba状态空间模型如何突破Transformer的O(n²)复杂度瓶颈，从S4模型到选择性SSM的数学原理，以及线性时间序列建模的技术演进。</description>
    </item>
  </channel>
</rss>
