Commit Graph

244 Commits

Author SHA1 Message Date
Relakkes 569202af78 feat: update user-agent list 2024-03-17 01:03:56 +08:00
Relakkes 59cd9f67a0 feat: 支持评论模式是否开启爬取选项 2024-03-16 11:52:42 +08:00
Relakkes 2d12ecb930 docs: 增加捐赠名单 2024-03-14 22:06:18 +08:00
Relakkes b48e4d297d docs: 将python3命令修改为python 2024-03-12 23:09:45 +08:00
Relakkes 9e2d1396b8 docs: 增加视频入门教程链接 2024-03-10 03:49:58 +08:00
Relakkes 9ca0a2b717 docs: update README.md 2024-03-09 15:37:51 +08:00
Relakkes 41fee4ff4f feat:小红书支持获取评论中的图片链接 #145 2024-03-07 22:30:44 +08:00
Relakkes 861019022a docs: 增加捐赠者名单 2024-03-05 22:45:32 +08:00
Relakkes 7606c1d993 docs: 增加捐赠者名单 2024-03-05 00:03:51 +08:00
Relakkes 149b6bcdc8 fix: 修复抖音关键词搜索为中文的情况下,有bug 2024-03-03 19:36:36 +08:00
relakkes 7002061d70
Merge pull request #142 from jayeeliu/main
小红书支持通过博主ID采集笔记和评论
2024-03-02 23:30:31 +08:00
jayeeliu@gmail.com 2b1ad18e18 Revert "feat: add aerich for migrate db"
This reverts commit 5f7cd715db.
2024-03-02 23:05:34 +08:00
jayeeliu@gmail.com 5f7cd715db feat: add aerich for migrate db 2024-03-02 15:59:20 +08:00
jayeeliu 9a5460ccca
Merge branch 'NanmiCoder:main' into main 2024-03-02 01:51:02 +08:00
jayeeliu@gmail.com 61ba8c5cc7 feat: 小红书支持通过博主ID采集笔记和评论,小红书type=search时支持配置按哪种排序方式获取笔记数据,小红书笔记增加视频地址和标签字段 2024-03-02 01:49:42 +08:00
Relakkes 384c8f9f7e fix: issue #140 2024-02-26 23:47:02 +08:00
relakkes 67d2b7cff8
Merge pull request #136 from jayeeliu/main
增加sqlite配置示例,解决用sqlite保存数据时,抓取结束不退出的问题
2024-02-22 21:32:19 +08:00
ccgo c09f9fef68 fix: 增加db.close(),解决抓取命令执行完不退出的问题 2024-02-22 00:11:41 +08:00
ccgo 4f5f83d3fa feat: 增加sqlite配置示例 2024-02-22 00:07:40 +08:00
Relakkes 03ca445221 docs: update README.md 2024-02-18 13:06:12 +08:00
Relakkes 192e72588d fix: #123 2024-01-26 23:42:26 +08:00
relakkes f1d3bc6e28
Merge pull request #124 from aa65535/patch-1
修复小红书评论重复插入
2024-01-25 14:33:09 +08:00
Jian Chang 79c0f3bd68
修复小红书评论重复插入 2024-01-25 13:01:04 +08:00
relakkes eb89a6ada8
docs: 添加赞赏者信息 2024-01-21 16:42:37 +08:00
Relakkes e940a41033 refactor: 移除评论中指定数量和过滤特定关键词的逻辑 2024-01-17 23:02:05 +08:00
Relakkes e0f9a487e4 refactor: 代码优化 2024-01-16 00:40:07 +08:00
relakkes e490123fcd
Merge pull request #112 from NanmiCoder/feature/stored_in_json
refactor: 重构数据存储代码,新增jSON
2024-01-14 22:45:00 +08:00
Relakkes 4dfa0d3fbf feat: 数据保存支持JSON格式 2024-01-14 22:40:01 +08:00
Relakkes 894dabcf63 refactor: 数据存储重构,分离不同类型的存储实现 2024-01-14 22:06:31 +08:00
Relakkes e31aebbdfb fix: 修复代理Bug 2024-01-13 15:50:02 +08:00
Relakkes 0240de7324 docs: 添加捐赠信息 2024-01-10 23:59:14 +08:00
Relakkes f85bd0c2b9 docs: 添加捐赠信息 2024-01-07 18:33:08 +08:00
Relakkes 4de14ad6a8 fix: 修复微博PC端登录后COOKIE在手机端无法使用的bug 2024-01-06 19:18:07 +08:00
Relakkes fe073801f8 feat: 小红书增加获取图片地址的方法 2024-01-04 00:11:05 +08:00
Relakkes 80ecc53cd3 fix: #107 2024-01-03 23:08:30 +08:00
Relakkes 38d6f10bf0 feat: 微博二维码登录done 2023-12-30 18:54:21 +08:00
Relakkes 27a2041929 doc: update README.md 2023-12-30 15:16:33 +08:00
Relakkes eee81622ac feat: 微博支持评论 & 指定帖子 2023-12-25 00:02:11 +08:00
Relakkes b1441ab4ae feat: 微博帖子支持保存到数据库中 2023-12-24 18:19:26 +08:00
Relakkes c5b64fdbf5 feat: 微博爬虫帖子搜索完成 2023-12-24 17:57:48 +08:00
Relakkes 9785abba5d docs: update README.md 2023-12-23 01:41:19 +08:00
Relakkes aba9f14f50 refactor: 规范日志打印
feat: B站指定视频ID爬取(bvid)
2023-12-23 01:04:08 +08:00
Relakkes 273c9a316b fix: 修复日志打印时参数格式错误 2023-12-22 23:10:44 +08:00
Relakkes bced7f6802 doc: update README.md 2023-12-22 23:04:08 +08:00
Relakkes 2e032d09dc refactor: xhs add log 2023-12-19 23:05:12 +08:00
Relakkes e2aaf0e5c1 doc: update README.md 2023-12-17 18:57:43 +08:00
Relakkes f85fd97d25 fix: #10 2023-12-16 00:12:20 +08:00
relakkes 40cbd9f4d7
Merge pull request #96 from PeanutSplash/main
fix-bug:修复抖音评论筛选
2023-12-14 00:57:17 +08:00
peanutsplash 5c6a636352 fix-bug:修复抖音评论筛选 2023-12-14 00:55:06 +08:00
relakkes c4d68f868a
Merge pull request #94 from PeanutSplash/main
添加功能:(哔哩哔哩,快手,小红书)每个视频/帖子抓取评论最大条数限制,评论关键词筛选
2023-12-14 00:01:44 +08:00