登录
首页 >  Golang >  Go教程

Golang实现RSS订阅功能教程

时间:2026-01-22 14:04:35 215浏览 收藏

哈喽!今天心血来潮给大家带来了《Golang构建RSS订阅功能教程》,想必大家应该对Golang都不陌生吧,那么阅读本文就都不会很困难,以下内容主要涉及到,若是你正在学习Golang,千万别错过这篇文章~希望能帮助到你!

Go 的 encoding/xml 解析 RSS 经常失败,根本原因是 RSS(尤其是 Atom 混合源)含 CDATA、命名空间、HTML 实体及不规范换行;标准库不自动解码实体、不忽略命名空间冲突,需预处理字符串、用 xml.CharData 类型、显式声明命名空间字段。

如何使用Golang构建基础RSS订阅功能_Golang数据抓取与展示方法

为什么 Go 的 encoding/xml 解析 RSS 经常失败

直接用 xml.Unmarshal 解析 RSS 时,常见报错如 XML syntax error on line X: invalid character entity 或字段全为空,根本原因是 RSS(尤其是 Atom 混合源)常含 CDATA、命名空间(xmlns)、HTML 实体( )及不规范换行。Go 标准库不自动解码 HTML 实体,也不忽略命名空间冲突。

  • 必须预处理 XML 字符串:用 strings.ReplaceAll 替换   等实体为普通空格,或用 html.UnescapeString 全量解码
  • RSS 2.0 推荐用无命名空间结构体;若遇 atom: 前缀,结构体字段需加 xml:"atom:title,attr" 显式声明
  • 这类可能含 HTML 的字段,定义为 xml.CharData 类型而非 string,避免截断

如何用 net/http 安全抓取 RSS 并防超时/重定向失控

直接 http.Get(url) 抓 RSS 源极易卡死或跳转到登录页,尤其面对 Cloudflare 防护或反爬中间页时。

  • 显式设置 http.Client:超时控制必须设 Timeout: 10 * time.Second,且启用 CheckRedirect 限制跳转次数(建议 ≤3)
  • 添加基础请求头:User-Agent 设为常见浏览器值(如 "Mozilla/5.0 (X11; Linux x86_64) AppleWebKit/537.36"),否则部分站点返回 403
  • 检查响应状态码和 Content-Type 头,仅当含 application/rss+xmlapplication/atom+xml 才继续解析,避免把 HTML 错当 RSS

goquery 补救非标准 RSS(如网页内嵌 RSS 链接或伪 RSS)

很多“RSS 源”实际是普通 HTML 页面里藏了 ,或直接把文章列表渲染成 HTML——这时不能硬套 XML 解析。

  • 先用 goquery.NewDocument 加载 HTML,提取真实 RSS 地址:
    doc.Find("link[rel=alternate][type='application/rss+xml']").Attr("href")
  • 若目标无标准 RSS,直接用 goquery 提取文章列表:
    doc.Find("article h2 a, .post-title a").Each(func(i int, s *goquery.Selection) {
        title := s.Text()
        link, _ := s.Attr("href")
    })
  • 注意:goquery 不处理相对 URL,需用 url.Join 拼接基础地址

展示层避免模板注入与 XSS 的关键处理

RSS 的 </code> 和 <code><description></code> 常含未过滤 HTML,直接塞进 HTML 模板会触发 XSS。</p><ul><li>服务端渲染时,用 <code>template.HTMLEscapeString</code> 处理所有字段,而非依赖前端 JS 转义</li><li>若需保留部分格式(如 <code><p></code>、<code><strong></code>),改用 <code>bluemonday</code> 库白名单过滤:<pre>policy := bluemonday.UGCPolicy() cleanDesc := policy.Sanitize(rssItem.Description)</pre></li><li>时间字段别用原始 <code>pubDate</code> 字符串,统一转为 <code>time.Time</code> 后再格式化(<code>item.PubDate.Format("Jan 2, 2006")</code>),避免时区混乱</li></ul><p>RSS 的难点不在抓取本身,而在应对现实世界中五花八门的“伪标准”——命名空间混用、HTML 实体乱飞、重定向陷阱、以及发布者随手写的 malformed XML。真正稳定的实现,永远是组合策略:HTTP 层控超时+UA,XML 层做预清洗+容错结构体,HTML 层兜底用 goquery,展示层强制转义。</p><p>以上就是本文的全部内容了,是否有顺利帮助你解决问题?若是能给你带来学习上的帮助,请大家多多支持golang学习网!更多关于Golang的相关知识,也可关注golang学习网公众号。</p> <div style="margin:16px auto;width:100%;max-width:720px;box-sizing:border-box;padding:16px;border:1px solid #e5e7eb;border-radius:12px;background:#fff;box-shadow:0 6px 24px rgba(16,24,40,0.08);text-align:center;overflow:hidden;"> <a onclick="showThirdParty('flex')" style="display:inline-flex;width:100%;max-width:100%;justify-content:center;align-items:center;gap:8px;padding:12px 18px;border-radius:10px;background:#2d8cf0;color:#fff;text-decoration:none;font-weight:600;box-sizing:border-box;overflow-wrap:anywhere;word-break:break-word;"> 前往漫画官网入口并下载 ➜ </a> </div> <div id="third-party-overlay" style="position:fixed;left:0;top:0;width:100%;height:100%;display:none;justify-content:center;align-items:center;background:rgba(0,0,0,0.4);z-index:9999;"> <div style="background:#FFF3CD;border:1px solid #FFEEBA;padding:16px;border-radius:6px;box-sizing:border-box;max-width:480px;width:90%;text-align:center;"> <div style="font-size:14px;color:#856404;margin-bottom:12px;">您即将跳转至第三方网站,请注意保护好个人信息和财产安全!</div> <a href="https://comicdow.pdlcomic.top/1273%2F%E5%9B%A7%E6%AC%A1%E5%85%83.apk" target="_blank" rel="nofollow noopener noreferrer" style="color:#2d8cf0;text-decoration:none;" onclick="showThirdParty('none');">继续访问</a> </div> </div> <script> function showThirdParty(mode){ var el = document.getElementById('third-party-overlay'); if (!el) return; el.style.display = (mode === 'none' ? 'none' : 'flex'); } </script> </div> <div class="labsList"> </div> </div> <!-- 最新阅读 --> <div class="contBoxNor"> <div class="contTit"> <div class="tit">相关阅读</div> <a href="/articlelist.html" class="more">更多></a> </div> <ul class="latestReadList"> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  3年前  |   <a href="/articletag/56_new_0_1.html" class="aLightGray" title="map">map</a> · <a href="/articletag/993_new_0_1.html" class="aLightGray" title="实践">实践</a> · <a href="/articletag/994_new_0_1.html" class="aLightGray" title="实现原理">实现原理</a> · <a href="/special/3_new_0_1.html" target="_blank" class="aLightGray" title="golang">golang</a> </div> <div class="tit lineOverflow"><a href="/article/10762.html" title="Golangmap实践及实现原理解析" class="aBlack">Golangmap实践及实现原理解析</a></div> <div class="opt"> <span><i class="view"></i>505</span> <span class="collectBtn user_collection" data-id="10762" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  2年前  |   <a href="/special/3_new_0_1.html" target="_blank" class="aLightGray" title="golang">golang</a> <a href="javascript:;" class="aLightGray" title="Go">Go</a> <a href="javascript:;" class="aLightGray" title="编程语言选择">编程语言选择</a> <a href="javascript:;" class="aLightGray" title="区别解析">区别解析</a> </div> <div class="tit lineOverflow"><a href="/article/80612.html" title="go和golang的区别解析:帮你选择合适的编程语言" class="aBlack">go和golang的区别解析:帮你选择合适的编程语言</a></div> <div class="opt"> <span><i class="view"></i>503</span> <span class="collectBtn user_collection" data-id="80612" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  3年前  |   <a href="/articletag/1668_new_0_1.html" class="aLightGray" title="try">try</a> · <a href="/articletag/1669_new_0_1.html" class="aLightGray" title="catch">catch</a> · <a href="/special/3_new_0_1.html" target="_blank" class="aLightGray" title="golang">golang</a> </div> <div class="tit lineOverflow"><a href="/article/11451.html" title="试了下Golang实现try catch的方法" class="aBlack">试了下Golang实现try catch的方法</a></div> <div class="opt"> <span><i class="view"></i>502</span> <span class="collectBtn user_collection" data-id="11451" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  2年前  |   <a href="javascript:;" class="aLightGray" title="并发 (concurrency)">并发 (concurrency)</a> <a href="javascript:;" class="aLightGray" title="Go语言 (Go language)">Go语言 (Go language)</a> <a href="javascript:;" class="aLightGray" title="服务器架构 (Server Architecture)">服务器架构 (Server Architecture)</a> </div> <div class="tit lineOverflow"><a href="/article/53565.html" title="如何在go语言中实现高并发的服务器架构" class="aBlack">如何在go语言中实现高并发的服务器架构</a></div> <div class="opt"> <span><i class="view"></i>502</span> <span class="collectBtn user_collection" data-id="53565" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  2年前  |   <a href="javascript:;" class="aLightGray" title="工作效率">工作效率</a> <a href="javascript:;" class="aLightGray" title="Go语言">Go语言</a> <a href="javascript:;" class="aLightGray" title="项目开发">项目开发</a> </div> <div class="tit lineOverflow"><a href="/article/72902.html" title="提升工作效率的Go语言项目开发经验分享" class="aBlack">提升工作效率的Go语言项目开发经验分享</a></div> <div class="opt"> <span><i class="view"></i>502</span> <span class="collectBtn user_collection" data-id="72902" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> </ul> </div> <!-- 最新阅读 --> <div class="contBoxNor"> <div class="contTit"> <div class="tit">最新阅读</div> <a href="/articlelist.html" class="more">更多></a> </div> <ul class="latestReadList"> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  24秒前  |   </div> <div class="tit lineOverflow"><a href="/article/467662.html" title="Go中空字符串Gob编码解码方法" class="aBlack">Go中空字符串Gob编码解码方法</a></div> <div class="opt"> <span><i class="view"></i>381</span> <span class="collectBtn user_collection" data-id="467662" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  11分钟前  |   </div> <div class="tit lineOverflow"><a href="/article/467648.html" title="Go反射获取类型信息全解析" class="aBlack">Go反射获取类型信息全解析</a></div> <div class="opt"> <span><i class="view"></i>180</span> <span class="collectBtn user_collection" data-id="467648" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  29分钟前  |   </div> <div class="tit lineOverflow"><a href="/article/467625.html" title="Golang多GOPATH设置与管理技巧" class="aBlack">Golang多GOPATH设置与管理技巧</a></div> <div class="opt"> <span><i class="view"></i>231</span> <span class="collectBtn user_collection" data-id="467625" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  34分钟前  |   </div> <div class="tit lineOverflow"><a href="/article/467617.html" title="Golang单例线程安全实现解析" class="aBlack">Golang单例线程安全实现解析</a></div> <div class="opt"> <span><i class="view"></i>289</span> <span class="collectBtn user_collection" data-id="467617" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  36分钟前  |   </div> <div class="tit lineOverflow"><a href="/article/467614.html" title="Golang高效合并文件技巧分享" class="aBlack">Golang高效合并文件技巧分享</a></div> <div class="opt"> <span><i class="view"></i>134</span> <span class="collectBtn user_collection" data-id="467614" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  58分钟前  |   <a href="/special/3_new_0_1.html" target="_blank" class="aLightGray" title="golang">golang</a> <a href="javascript:;" class="aLightGray" title="JSON">JSON</a> </div> <div class="tit lineOverflow"><a href="/article/467588.html" title="Golang处理JSON请求响应实战教程" class="aBlack">Golang处理JSON请求响应实战教程</a></div> <div class="opt"> <span><i class="view"></i>227</span> <span class="collectBtn user_collection" data-id="467588" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  1小时前  |   </div> <div class="tit lineOverflow"><a href="/article/467565.html" title="IOTA多常量同赋值怎么算?图解详解" class="aBlack">IOTA多常量同赋值怎么算?图解详解</a></div> <div class="opt"> <span><i class="view"></i>101</span> <span class="collectBtn user_collection" data-id="467565" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  1小时前  |   </div> <div class="tit lineOverflow"><a href="/article/467552.html" title="Golang反射操作结构体切片详解" class="aBlack">Golang反射操作结构体切片详解</a></div> <div class="opt"> <span><i class="view"></i>111</span> <span class="collectBtn user_collection" data-id="467552" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  1小时前  |   </div> <div class="tit lineOverflow"><a href="/article/467549.html" title="Go语言实现文件上传方法解析" class="aBlack">Go语言实现文件上传方法解析</a></div> <div class="opt"> <span><i class="view"></i>246</span> <span class="collectBtn user_collection" data-id="467549" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  1小时前  |   </div> <div class="tit lineOverflow"><a href="/article/467548.html" title="Golang错误处理与error类型使用详解" class="aBlack">Golang错误处理与error类型使用详解</a></div> <div class="opt"> <span><i class="view"></i>241</span> <span class="collectBtn user_collection" data-id="467548" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  1小时前  |   </div> <div class="tit lineOverflow"><a href="/article/467544.html" title="Golang容器安全策略全解析" class="aBlack">Golang容器安全策略全解析</a></div> <div class="opt"> <span><i class="view"></i>364</span> <span class="collectBtn user_collection" data-id="467544" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> <li> <div class="info"> <a href="/articlelist/25_new_0_1.html" class="aLightGray" title="Golang">Golang</a> · <a href="/articlelist/44_new_0_1.html" class="aLightGray" title="Go教程">Go教程</a>   |  1小时前  |   </div> <div class="tit lineOverflow"><a href="/article/467537.html" title="GolangWaitGroup管理多协程实战教程" class="aBlack">GolangWaitGroup管理多协程实战教程</a></div> <div class="opt"> <span><i class="view"></i>331</span> <span class="collectBtn user_collection" data-id="467537" data-type="article" title="收藏"><i class="collect"></i>收藏</span> </div> </li> </ul> </div> <!-- 课程推荐 --> <div class="contBoxNor"> <div class="contTit"> <div class="tit">课程推荐</div> <a href="/courselist.html" class="more">更多></a> </div> <ul class="classRecomList"> <li> <a href="/course/9.html" title="前端进阶之JavaScript设计模式" class="img_box"> <img src="/uploads/20221222/52fd0f23a454c71029c2c72d206ed815.jpg" onerror="this.onerror='';this.src='/assets/images/moren/morentu.png'" alt="前端进阶之JavaScript设计模式"> </a> <dl> <dt class="lineOverflow"> 前端进阶之JavaScript设计模式 </dt> <dd class="cont1 lineOverflow">设计模式是开发人员在软件开发过程中面临一般问题时的解决方案,代表了最佳的实践。本课程的主打内容包括JS常见设计模式以及具体应用场景,打造一站式知识长龙服务,适合有JS基础的同学学习。</dd> <dd class="cont2"> <a href="/course/9.html" title="前端进阶之JavaScript设计模式" class="toStudy">立即学习</a> <span>543次学习</span> </dd> </dl> </li> <li> <a href="/course/2.html" title="GO语言核心编程课程" class="img_box"> <img src="/uploads/20221221/634ad7404159bfefc6a54a564d437b5f.png" onerror="this.onerror='';this.src='/assets/images/moren/morentu.png'" alt="GO语言核心编程课程"> </a> <dl> <dt class="lineOverflow"> GO语言核心编程课程 </dt> <dd class="cont1 lineOverflow">本课程采用真实案例,全面具体可落地,从理论到实践,一步一步将GO核心编程技术、编程思想、底层实现融会贯通,使学习者贴近时代脉搏,做IT互联网时代的弄潮儿。</dd> <dd class="cont2"> <a href="/course/2.html" title="GO语言核心编程课程" class="toStudy">立即学习</a> <span>516次学习</span> </dd> </dl> </li> <li> <a href="/course/74.html" title="简单聊聊mysql8与网络通信" class="img_box"> <img src="/uploads/20240103/bad35fe14edbd214bee16f88343ac57c.png" onerror="this.onerror='';this.src='/assets/images/moren/morentu.png'" alt="简单聊聊mysql8与网络通信"> </a> <dl> <dt class="lineOverflow"> 简单聊聊mysql8与网络通信 </dt> <dd class="cont1 lineOverflow">如有问题加微信:Le-studyg;在课程中,我们将首先介绍MySQL8的新特性,包括性能优化、安全增强、新数据类型等,帮助学生快速熟悉MySQL8的最新功能。接着,我们将深入解析MySQL的网络通信机制,包括协议、连接管理、数据传输等,让</dd> <dd class="cont2"> <a href="/course/74.html" title="简单聊聊mysql8与网络通信" class="toStudy">立即学习</a> <span>500次学习</span> </dd> </dl> </li> <li> <a href="/course/57.html" title="JavaScript正则表达式基础与实战" class="img_box"> <img src="/uploads/20221226/bbe4083bb3cb0dd135fb02c31c3785fb.jpg" onerror="this.onerror='';this.src='/assets/images/moren/morentu.png'" alt="JavaScript正则表达式基础与实战"> </a> <dl> <dt class="lineOverflow"> JavaScript正则表达式基础与实战 </dt> <dd class="cont1 lineOverflow">在任何一门编程语言中,正则表达式,都是一项重要的知识,它提供了高效的字符串匹配与捕获机制,可以极大的简化程序设计。</dd> <dd class="cont2"> <a href="/course/57.html" title="JavaScript正则表达式基础与实战" class="toStudy">立即学习</a> <span>487次学习</span> </dd> </dl> </li> <li> <a href="/course/28.html" title="从零制作响应式网站—Grid布局" class="img_box"> <img src="/uploads/20221223/ac110f88206daeab6c0cf38ebf5fe9ed.jpg" onerror="this.onerror='';this.src='/assets/images/moren/morentu.png'" alt="从零制作响应式网站—Grid布局"> </a> <dl> <dt class="lineOverflow"> 从零制作响应式网站—Grid布局 </dt> <dd class="cont1 lineOverflow">本系列教程将展示从零制作一个假想的网络科技公司官网,分为导航,轮播,关于我们,成功案例,服务流程,团队介绍,数据部分,公司动态,底部信息等内容区块。网站整体采用CSSGrid布局,支持响应式,有流畅过渡和展现动画。</dd> <dd class="cont2"> <a href="/course/28.html" title="从零制作响应式网站—Grid布局" class="toStudy">立即学习</a> <span>485次学习</span> </dd> </dl> </li> </ul> </div> </div> <!-- footer --> <link href="https://fonts.googleapis.com/icon?family=Material+Icons" rel="stylesheet"> <div class="footer"> <ul> <li ><a href="/" class="aLightGray"><em class="material-icons">home</em><span>首页</span></a></li> <li class="curr"><a href="/articlelist.html" class="aLightGray"><em class="material-icons">menu_book</em><span>阅读</span></a></li> <li ><a href="/courselist.html" class="aLightGray"><em class="material-icons">school</em><span>课程</span></a></li> <li ><a href="/ai.html" class="aLightGray"><em class="material-icons">smart_toy</em><span>AI助手</span></a></li> <li ><a href="/user.html" class="aLightGray"><em class="material-icons">person</em><span>我的</span></a></li> </ul> </div> <script src="/assets/js/require.js" data-main="/assets/js/require-frontend.js?v=1671101972"></script> <script> var _hmt = _hmt || []; (function() { var hm = document.createElement("script"); hm.src = "https://hm.baidu.com/hm.js?3dc5666f6478c7bf39cd5c91e597423d"; var s = document.getElementsByTagName("script")[0]; s.parentNode.insertBefore(hm, s); })(); </script> </body> </html>