Relakkes
|
62534d7ee2
|
fix: 移出快手 client 多余的代码
|
2023-11-26 22:11:06 +08:00 |
Relakkes
|
609c3f0673
|
fix: 修复抖音数据库保存错误
|
2023-11-26 22:06:26 +08:00 |
Relakkes
|
d59113668f
|
doc: 增加快手爬虫的描述
|
2023-11-26 21:50:08 +08:00 |
relakkes
|
9faa4f4eec
|
Merge pull request #82 from NanmiCoder/feat/kuaishou_support
快手视频、评论爬取支持
|
2023-11-26 21:44:36 +08:00 |
Relakkes
|
dfb1788141
|
feat: 快手视频评论爬取done;数据保存到DB、CSV done
|
2023-11-26 21:43:39 +08:00 |
Relakkes
|
2f8541a351
|
Merge branch 'main' into feat/kuaishou_support
|
2023-11-26 15:44:26 +08:00 |
Relakkes
|
523c5f380e
|
doc: add command usage
|
2023-11-26 15:38:38 +08:00 |
Relakkes
|
bdf36ccb09
|
feat: 快手关键词搜索存储CSV完成
|
2023-11-26 01:05:52 +08:00 |
Relakkes
|
253888980c
|
doc: 增加如何更换账号的问题
|
2023-11-25 12:32:43 +08:00 |
Relakkes
|
512192a93e
|
feat: 搜索接口调试完成
|
2023-11-25 00:02:33 +08:00 |
Relakkes
|
f08b2ceb76
|
feat: 1、命令行支持快手 2、快手playwright 代码 done
|
2023-11-24 00:04:33 +08:00 |
Relakkes
|
6908da6f70
|
fix: 修复小红书client中get请求中文bug
|
2023-11-23 23:27:35 +08:00 |
Relakkes
|
95ca606938
|
feat: 快手文件目录建立
|
2023-11-23 23:13:54 +08:00 |
Relakkes
|
3790e8041e
|
doc: update README.md
|
2023-11-20 22:07:00 +08:00 |
Relakkes
|
81bc8b51e2
|
feat: 抖音支持指定视频列表爬去
|
2023-11-18 22:07:30 +08:00 |
Relakkes
|
098923d74d
|
refactor: 优化代码-变量名
|
2023-11-18 15:53:10 +08:00 |
relakkes
|
ecf9a5e893
|
Merge pull request #74 from Akiqqqqqqq/bugfix-null-key
fix null key
|
2023-11-18 15:05:16 +08:00 |
leonardoqiuyu
|
4c85355ebd
|
Merge branch 'main' into bugfix-null-key
|
2023-11-18 14:59:03 +08:00 |
Relakkes
|
700946b28a
|
feat: 小红书增加指定帖子爬取功能
fix: 修复程序一些异常 bug
refactor: 优化部分代码逻辑
|
2023-11-18 13:38:11 +08:00 |
Relakkes
|
f24c892471
|
fix: 修复评论为空导致程序异常退出的 bug
|
2023-11-18 11:18:54 +08:00 |
Akiqqqqqqq
|
602e609115
|
fix null key
|
2023-11-14 21:11:18 +08:00 |
Relakkes
|
cd3a0bdb6a
|
doc: update README.md
|
2023-11-11 20:49:32 +08:00 |
relakkes
|
68c7139dd0
|
Merge pull request #64 from NanmiCoder/feat/article_url
feat: add article url for issue #63
|
2023-11-05 15:28:10 +08:00 |
Relakkes
|
d0ded6689c
|
feat: add article url for issue #63
|
2023-11-05 15:27:18 +08:00 |
relakkes
|
7698c41d51
|
Merge pull request #40 from yangrq1018/pr
fix: get phone from phone pool when ENABLE_IP_PROXY is False
|
2023-08-31 08:09:06 -05:00 |
xiaoniu
|
ad69e63d5f
|
fix: get phone from phone pool when ENABLE_IP_PROXY is False
|
2023-08-31 14:00:57 +08:00 |
Relakkes
|
b7fe5b730a
|
style: Change the import order in the code.
|
2023-08-16 19:56:12 +08:00 |
Relakkes
|
9177c38521
|
feat: 支持数据保存到CSV中
|
2023-08-16 19:49:41 +08:00 |
Relakkes
|
c1a3f06c7a
|
fix: issue #32
|
2023-08-16 13:58:44 +08:00 |
Relakkes
|
99812b4669
|
Merge pull request #32 from renaissancezyc/main
orm添加Pydantic数据验证
|
2023-08-14 09:32:48 +08:00 |
zhoubang
|
1556e1b8cd
|
✨ feat: orm添加Pydantic数据验证
|
2023-08-13 22:39:50 +08:00 |
Relakkes
|
3ae4c5f2f4
|
fix: issue #23
|
2023-08-03 22:26:31 +08:00 |
Relakkes
|
defd8232da
|
Merge pull request #24 from jayeeliu/main
去掉搜索结果列表中的推荐搜索词(mode_type in ['rec_query', 'hot_query'])
|
2023-08-01 13:35:00 +08:00 |
JayeeLiu
|
925f0a1bc4
|
去掉搜索结果列表中的推荐搜索词(mode_type in ['rec_query', 'hot_query'])
|
2023-08-01 10:29:00 +08:00 |
Relakkes
|
1d3c2359e0
|
fix: 修复部分变量命名语义不明确
|
2023-07-30 21:30:26 +08:00 |
Relakkes
|
bf659455bb
|
fix: issue #22
|
2023-07-30 20:43:02 +08:00 |
Relakkes
|
4ff2cf8661
|
refactor: 优化代码
|
2023-07-29 15:35:40 +08:00 |
Relakkes
|
febbb133d7
|
fix: douyin 缺少一个collected_count字段
|
2023-07-28 21:23:37 +08:00 |
Relakkes
|
03565d61c6
|
fix: issue #19
|
2023-07-25 20:22:22 +08:00 |
Relakkes
|
7d63c2f9ec
|
Merge branch 'main' of github.com:NanmiCoder/MediaCrawler
|
2023-07-24 21:00:00 +08:00 |
Relakkes
|
e75707443b
|
feat: 增加配置项支持自由选择数据是否保存到关系型数据库中
|
2023-07-24 20:59:43 +08:00 |
Relakkes
|
b94924e21b
|
Merge pull request #18 from tanpenggood-fork/typo-sam-20230723
typo(field.py): words "normal" typo
|
2023-07-23 14:30:19 +08:00 |
tanpenggood
|
ae2de2f792
|
typo(field.py): words "normal" typo
|
2023-07-23 13:24:39 +08:00 |
Nanmi
|
745e59c875
|
feat: 完善类型注释,增加 mypy 类型检测
|
2023-07-16 17:57:18 +08:00 |
Relakkes
|
e5bdc63323
|
feat: xhs增加并发控制参数
|
2023-07-15 22:25:56 +08:00 |
Relakkes
|
2398a17e21
|
refactor: 优化抖音Crawler部分代码
fix: 日志初始化错误修复
|
2023-07-15 21:30:12 +08:00 |
Relakkes
|
dad8d56ab5
|
feat: issue #14
refactor: 优化小红书crawler流程代码
|
2023-07-15 17:11:53 +08:00 |
Relakkes
|
e5f4ecd8ec
|
fix: issue #12
|
2023-07-03 19:34:37 +08:00 |
Relakkes
|
0cedd7da40
|
docs: update readme.md
|
2023-07-03 12:50:55 +08:00 |
Relakkes
|
57437719bf
|
feat: 抖音三种方式登录实现 & 抖音滑块模拟滑动实现
|
2023-07-01 23:10:47 +08:00 |