程序员阿江(Relakkes)
|
5a27ad089c
|
Merge pull request #462 from FloRainRJY/xiaohongshu_comment_number_restrict
All_platform_comments_restrict
|
2024-10-24 15:31:13 +08:00 |
liugangdao
|
78c09c4ae1
|
fix:翻页时search id不变
|
2024-10-24 13:57:27 +08:00 |
unknown
|
7e53c4acfc
|
All_platform_comments_restrict
|
2024-10-23 16:32:02 +08:00 |
unknown
|
19269c66fd
|
xiaohongshu_comment_number_restrict
|
2024-10-22 20:33:10 +08:00 |
Relakkes
|
03e393949a
|
fix: xhs帖子详情问题更新
|
2024-10-20 00:59:08 +08:00 |
Relakkes
|
9fe3e47b0f
|
chore: 增加代码学习声明,严格禁止非法、禁止商业、不当用途
|
2024-10-20 00:43:25 +08:00 |
Relakkes
|
6dd3420743
|
fix: #423
|
2024-09-02 22:57:28 +08:00 |
Relakkes
|
d6fb255bdf
|
fix: xhs note detail error
|
2024-09-02 21:45:12 +08:00 |
Relakkes
|
65699aa1cb
|
feat: xhs支持获取评论的点赞数量
|
2024-08-24 06:07:33 +08:00 |
Relakkes
|
ab7d8142af
|
feat: weibo支持指定创作者主页
|
2024-08-24 05:52:11 +08:00 |
Relakkes
|
c70bd9e071
|
feat: 增加搜索词来源渠道
|
2024-08-23 08:29:24 +08:00 |
Relakkes
|
f371675d47
|
fix: xhs指定笔记ID获取方式增加解析html方式,原来的由于xsec_token导致失效
|
2024-08-11 22:37:10 +08:00 |
Relakkes
|
7229d29123
|
feat: xhs update
|
2024-08-04 14:54:03 +08:00 |
Wenbo Lu
|
9fb8f42f69
|
fetch_creator_notes_detail, id ->note_id
|
2024-07-26 21:55:31 +08:00 |
Relakkes
|
573ca9a659
|
feat: xhs笔记详情更新
|
2024-07-25 00:44:46 +08:00 |
AuYeung
|
1fd7827e36
|
When the query has no content, terminate the loop early
|
2024-07-18 20:44:40 +08:00 |
helloteemo
|
b95dc2c125
|
fix: 小红书下载新版本使用>3.10特性, 降低使用版本
|
2024-07-12 09:50:03 +08:00 |
helloteemo
|
6545a15ff3
|
feature: 支持小红书图片、视频下载
|
2024-07-11 22:56:30 +08:00 |
helloteemo
|
e71690a985
|
fix: 解决小红书图片水印问题
|
2024-07-11 17:39:48 +08:00 |
Relakkes
|
d3eeccbaac
|
feat: logger record current search page
|
2024-06-24 22:24:51 +08:00 |
nelzomal
|
eace7d1750
|
improve base config reading command line arg logic
|
2024-06-09 18:51:36 +08:00 |
Relakkes
|
4bba1447f8
|
feat: cache impl done
|
2024-06-02 19:57:13 +08:00 |
Relakkes
|
e64df93edd
|
feat: 由于xhs和dy现在检测playwright二维码登录了,大概率会出现滑块或者手机验证,增加登录态检测时间为5min,预留足够的时间手动过验证码。
|
2024-05-15 23:23:30 +08:00 |
Henry He
|
a2dca888ac
|
fix: 修复 f-string 双引号问题
|
2024-04-26 11:02:55 +08:00 |
bigfa
|
99d53cd945
|
fix:兼容网页端上传的图片原图获取
|
2024-04-24 14:16:35 +08:00 |
Relakkes
|
87eb8aa6a7
|
fix: #230
|
2024-04-13 20:18:04 +08:00 |
程序员阿江-Relakkes
|
a341dc2aff
|
Merge pull request #229 from Tianci-King/main
feat(core): 新增控制爬虫参数起始页面的页数start_page;perf(argparse): 向命令行解析器添加程序参数…
|
2024-04-13 13:37:35 +08:00 |
leantli
|
ad01dfba95
|
feat: 轻量化支持爬取小红书二级评论
|
2024-04-12 17:32:20 +08:00 |
Tianci-King
|
1115b0d90c
|
feat(core): 新增控制爬虫 参数起始页面的页数start_page;perf(argparse): 向命令行解析器添加程序参数起始页面页数和关键字
|
2024-04-12 00:52:47 +08:00 |
leantli
|
81a9946afd
|
feat: 支持爬取小红书二级评论
|
2024-04-11 17:16:13 +08:00 |
Relakkes
|
8f02da73ad
|
fix: #219
docs: update README.md
|
2024-04-08 00:19:50 +08:00 |
leantli
|
68a60faa7f
|
chore: 简化判断方式
|
2024-04-04 00:11:22 +08:00 |
leantli
|
133f978477
|
fix: 修复爬取视频/帖子最大数设置值较低导致不爬取的问题
|
2024-04-03 12:18:23 +08:00 |
Relakkes
|
e950e0d6e3
|
feat: add abstract api client to all platform
|
2024-03-30 21:27:25 +08:00 |
Relakkes
|
67ec49498a
|
refactor: rename xhs to xiaohongshu
|
2024-03-30 21:17:33 +08:00 |
Relakkes
|
96309dcfee
|
fix: 小红书创作者功能数据获取优化
|
2024-03-17 14:50:10 +08:00 |
Relakkes
|
59cd9f67a0
|
feat: 支持评论模式是否开启爬取选项
|
2024-03-16 11:52:42 +08:00 |
Relakkes
|
41fee4ff4f
|
feat:小红书支持获取评论中的图片链接 #145
|
2024-03-07 22:30:44 +08:00 |
jayeeliu@gmail.com
|
61ba8c5cc7
|
feat: 小红书支持通过博主ID采集笔记和评论,小红书type=search时支持配置按哪种排序方式获取笔记数据,小红书笔记增加视频地址和标签字段
|
2024-03-02 01:49:42 +08:00 |
Relakkes
|
e0f9a487e4
|
refactor: 代码优化
|
2024-01-16 00:40:07 +08:00 |
Relakkes
|
894dabcf63
|
refactor: 数据存储重构,分离不同类型的存储实现
|
2024-01-14 22:06:31 +08:00 |
Relakkes
|
e31aebbdfb
|
fix: 修复代理Bug
|
2024-01-13 15:50:02 +08:00 |
Relakkes
|
fe073801f8
|
feat: 小红书增加获取图片地址的方法
|
2024-01-04 00:11:05 +08:00 |
Relakkes
|
aba9f14f50
|
refactor: 规范日志打印
feat: B站指定视频ID爬取(bvid)
|
2023-12-23 01:04:08 +08:00 |
Relakkes
|
273c9a316b
|
fix: 修复日志打印时参数格式错误
|
2023-12-22 23:10:44 +08:00 |
Relakkes
|
2e032d09dc
|
refactor: xhs add log
|
2023-12-19 23:05:12 +08:00 |
Relakkes
|
f85fd97d25
|
fix: #10
|
2023-12-16 00:12:20 +08:00 |
peanutsplash
|
f17a85305e
|
添加功能:(哔哩哔哩,快手,小红书)每个视频/帖子抓取评论最大条数限制,评论关键词筛选
|
2023-12-13 23:53:12 +08:00 |
Relakkes
|
97d7a0c38b
|
feat: Bilibili comment done
|
2023-12-09 21:10:01 +08:00 |
Relakkes
|
1cec23f73d
|
feat: 代理IP功能 Done
|
2023-12-08 00:10:04 +08:00 |