jayeeliu
|
9a5460ccca
|
Merge branch 'NanmiCoder:main' into main
|
2024-03-02 01:51:02 +08:00 |
jayeeliu@gmail.com
|
61ba8c5cc7
|
feat: 小红书支持通过博主ID采集笔记和评论,小红书type=search时支持配置按哪种排序方式获取笔记数据,小红书笔记增加视频地址和标签字段
|
2024-03-02 01:49:42 +08:00 |
Relakkes
|
384c8f9f7e
|
fix: issue #140
|
2024-02-26 23:47:02 +08:00 |
relakkes
|
67d2b7cff8
|
Merge pull request #136 from jayeeliu/main
增加sqlite配置示例,解决用sqlite保存数据时,抓取结束不退出的问题
|
2024-02-22 21:32:19 +08:00 |
ccgo
|
c09f9fef68
|
fix: 增加db.close(),解决抓取命令执行完不退出的问题
|
2024-02-22 00:11:41 +08:00 |
ccgo
|
4f5f83d3fa
|
feat: 增加sqlite配置示例
|
2024-02-22 00:07:40 +08:00 |
Relakkes
|
03ca445221
|
docs: update README.md
|
2024-02-18 13:06:12 +08:00 |
Relakkes
|
192e72588d
|
fix: #123
|
2024-01-26 23:42:26 +08:00 |
relakkes
|
f1d3bc6e28
|
Merge pull request #124 from aa65535/patch-1
修复小红书评论重复插入
|
2024-01-25 14:33:09 +08:00 |
Jian Chang
|
79c0f3bd68
|
修复小红书评论重复插入
|
2024-01-25 13:01:04 +08:00 |
relakkes
|
eb89a6ada8
|
docs: 添加赞赏者信息
|
2024-01-21 16:42:37 +08:00 |
Relakkes
|
e940a41033
|
refactor: 移除评论中指定数量和过滤特定关键词的逻辑
|
2024-01-17 23:02:05 +08:00 |
Relakkes
|
e0f9a487e4
|
refactor: 代码优化
|
2024-01-16 00:40:07 +08:00 |
relakkes
|
e490123fcd
|
Merge pull request #112 from NanmiCoder/feature/stored_in_json
refactor: 重构数据存储代码,新增jSON
|
2024-01-14 22:45:00 +08:00 |
Relakkes
|
4dfa0d3fbf
|
feat: 数据保存支持JSON格式
|
2024-01-14 22:40:01 +08:00 |
Relakkes
|
894dabcf63
|
refactor: 数据存储重构,分离不同类型的存储实现
|
2024-01-14 22:06:31 +08:00 |
Relakkes
|
e31aebbdfb
|
fix: 修复代理Bug
|
2024-01-13 15:50:02 +08:00 |
Relakkes
|
0240de7324
|
docs: 添加捐赠信息
|
2024-01-10 23:59:14 +08:00 |
Relakkes
|
f85bd0c2b9
|
docs: 添加捐赠信息
|
2024-01-07 18:33:08 +08:00 |
Relakkes
|
4de14ad6a8
|
fix: 修复微博PC端登录后COOKIE在手机端无法使用的bug
|
2024-01-06 19:18:07 +08:00 |
Relakkes
|
fe073801f8
|
feat: 小红书增加获取图片地址的方法
|
2024-01-04 00:11:05 +08:00 |
Relakkes
|
80ecc53cd3
|
fix: #107
|
2024-01-03 23:08:30 +08:00 |
Relakkes
|
38d6f10bf0
|
feat: 微博二维码登录done
|
2023-12-30 18:54:21 +08:00 |
Relakkes
|
27a2041929
|
doc: update README.md
|
2023-12-30 15:16:33 +08:00 |
Relakkes
|
eee81622ac
|
feat: 微博支持评论 & 指定帖子
|
2023-12-25 00:02:11 +08:00 |
Relakkes
|
b1441ab4ae
|
feat: 微博帖子支持保存到数据库中
|
2023-12-24 18:19:26 +08:00 |
Relakkes
|
c5b64fdbf5
|
feat: 微博爬虫帖子搜索完成
|
2023-12-24 17:57:48 +08:00 |
Relakkes
|
9785abba5d
|
docs: update README.md
|
2023-12-23 01:41:19 +08:00 |
Relakkes
|
aba9f14f50
|
refactor: 规范日志打印
feat: B站指定视频ID爬取(bvid)
|
2023-12-23 01:04:08 +08:00 |
Relakkes
|
273c9a316b
|
fix: 修复日志打印时参数格式错误
|
2023-12-22 23:10:44 +08:00 |
Relakkes
|
bced7f6802
|
doc: update README.md
|
2023-12-22 23:04:08 +08:00 |
Relakkes
|
2e032d09dc
|
refactor: xhs add log
|
2023-12-19 23:05:12 +08:00 |
Relakkes
|
e2aaf0e5c1
|
doc: update README.md
|
2023-12-17 18:57:43 +08:00 |
Relakkes
|
f85fd97d25
|
fix: #10
|
2023-12-16 00:12:20 +08:00 |
relakkes
|
40cbd9f4d7
|
Merge pull request #96 from PeanutSplash/main
fix-bug:修复抖音评论筛选
|
2023-12-14 00:57:17 +08:00 |
peanutsplash
|
5c6a636352
|
fix-bug:修复抖音评论筛选
|
2023-12-14 00:55:06 +08:00 |
relakkes
|
c4d68f868a
|
Merge pull request #94 from PeanutSplash/main
添加功能:(哔哩哔哩,快手,小红书)每个视频/帖子抓取评论最大条数限制,评论关键词筛选
|
2023-12-14 00:01:44 +08:00 |
peanutsplash
|
f17a85305e
|
添加功能:(哔哩哔哩,快手,小红书)每个视频/帖子抓取评论最大条数限制,评论关键词筛选
|
2023-12-13 23:53:12 +08:00 |
Relakkes
|
5c42076ff8
|
doc: update README.md
|
2023-12-09 23:50:33 +08:00 |
Relakkes
|
05016305e8
|
doc: update README.md
|
2023-12-09 23:49:19 +08:00 |
Relakkes
|
97d7a0c38b
|
feat: Bilibili comment done
|
2023-12-09 21:10:01 +08:00 |
Relakkes
|
a136ea2739
|
refactor: delete account_config.py
|
2023-12-09 13:56:18 +08:00 |
Relakkes
|
2d0f12f058
|
doc: fixed pydantic version
|
2023-12-09 13:31:53 +08:00 |
Relakkes
|
fd640f2008
|
doc: update README.md
|
2023-12-09 00:30:31 +08:00 |
Relakkes
|
d37d5ccd12
|
doc: add SPONSORED BY text
|
2023-12-09 00:28:29 +08:00 |
Relakkes
|
385662a7b6
|
doc: 新增代理IP使用文档
|
2023-12-09 00:06:57 +08:00 |
Relakkes
|
a40a91d7f7
|
doc: update README.md
|
2023-12-08 00:27:59 +08:00 |
Relakkes
|
72a68fbf1e
|
doc:将文档拆出去
|
2023-12-08 00:25:06 +08:00 |
Relakkes
|
1cec23f73d
|
feat: 代理IP功能 Done
|
2023-12-08 00:10:04 +08:00 |
Relakkes
|
c530bd4219
|
feat: 代理IP缓存到redis中
|
2023-12-06 23:49:56 +08:00 |