Commit Graph

372 Commits

Author SHA1 Message Date
Relakkes 5aeee93fc5 feat: B站爬虫签名实现 2023-12-03 00:30:10 +08:00
Relakkes 5c920da288 feat: 快手二维码登录 2023-12-02 18:22:55 +08:00
Relakkes 986179b9c9 feat: 增加 IP 代理的最新实现 2023-12-02 16:14:36 +08:00
Relakkes a8a4d34d2a doc: update README.md 2023-12-02 10:56:13 +08:00
Relakkes e7f68dd7e1 doc: 提供仓库功能列表表格 2023-12-01 23:58:35 +08:00
relakkes 5affc8a600
Update README.md 2023-12-01 09:55:38 +08:00
Relakkes 6d743552f2 doc: update README.md 2023-11-30 23:59:02 +08:00
Relakkes 33721e5fbd feat: 快手支持指定视频列表爬取 2023-11-27 23:07:04 +08:00
Relakkes 62534d7ee2 fix: 移出快手 client 多余的代码 2023-11-26 22:11:06 +08:00
Relakkes 609c3f0673 fix: 修复抖音数据库保存错误 2023-11-26 22:06:26 +08:00
Relakkes d59113668f doc: 增加快手爬虫的描述 2023-11-26 21:50:08 +08:00
relakkes 9faa4f4eec
Merge pull request #82 from NanmiCoder/feat/kuaishou_support
快手视频、评论爬取支持
2023-11-26 21:44:36 +08:00
Relakkes dfb1788141 feat: 快手视频评论爬取done;数据保存到DB、CSV done 2023-11-26 21:43:39 +08:00
Relakkes 2f8541a351 Merge branch 'main' into feat/kuaishou_support 2023-11-26 15:44:26 +08:00
Relakkes 523c5f380e doc: add command usage 2023-11-26 15:38:38 +08:00
Relakkes bdf36ccb09 feat: 快手关键词搜索存储CSV完成 2023-11-26 01:05:52 +08:00
Relakkes 253888980c doc: 增加如何更换账号的问题 2023-11-25 12:32:43 +08:00
Relakkes 512192a93e feat: 搜索接口调试完成 2023-11-25 00:02:33 +08:00
Relakkes f08b2ceb76 feat: 1、命令行支持快手 2、快手playwright 代码 done 2023-11-24 00:04:33 +08:00
Relakkes 6908da6f70 fix: 修复小红书client中get请求中文bug 2023-11-23 23:27:35 +08:00
Relakkes 95ca606938 feat: 快手文件目录建立 2023-11-23 23:13:54 +08:00
Relakkes 3790e8041e doc: update README.md 2023-11-20 22:07:00 +08:00
Relakkes 81bc8b51e2 feat: 抖音支持指定视频列表爬去 2023-11-18 22:07:30 +08:00
Relakkes 098923d74d refactor: 优化代码-变量名 2023-11-18 15:53:10 +08:00
relakkes ecf9a5e893
Merge pull request #74 from Akiqqqqqqq/bugfix-null-key
fix null key
2023-11-18 15:05:16 +08:00
leonardoqiuyu 4c85355ebd
Merge branch 'main' into bugfix-null-key 2023-11-18 14:59:03 +08:00
Relakkes 700946b28a feat: 小红书增加指定帖子爬取功能
fix: 修复程序一些异常 bug
refactor: 优化部分代码逻辑
2023-11-18 13:38:11 +08:00
Relakkes f24c892471 fix: 修复评论为空导致程序异常退出的 bug 2023-11-18 11:18:54 +08:00
Akiqqqqqqq 602e609115 fix null key 2023-11-14 21:11:18 +08:00
Relakkes cd3a0bdb6a doc: update README.md 2023-11-11 20:49:32 +08:00
relakkes 68c7139dd0
Merge pull request #64 from NanmiCoder/feat/article_url
feat: add article url for issue #63
2023-11-05 15:28:10 +08:00
Relakkes d0ded6689c feat: add article url for issue #63 2023-11-05 15:27:18 +08:00
relakkes 7698c41d51
Merge pull request #40 from yangrq1018/pr
fix: get phone from phone pool when ENABLE_IP_PROXY is False
2023-08-31 08:09:06 -05:00
xiaoniu ad69e63d5f fix: get phone from phone pool when ENABLE_IP_PROXY is False 2023-08-31 14:00:57 +08:00
Relakkes b7fe5b730a style: Change the import order in the code. 2023-08-16 19:56:12 +08:00
Relakkes 9177c38521 feat: 支持数据保存到CSV中 2023-08-16 19:49:41 +08:00
Relakkes c1a3f06c7a fix: issue #32 2023-08-16 13:58:44 +08:00
Relakkes 99812b4669
Merge pull request #32 from renaissancezyc/main
orm添加Pydantic数据验证
2023-08-14 09:32:48 +08:00
zhoubang 1556e1b8cd feat: orm添加Pydantic数据验证 2023-08-13 22:39:50 +08:00
Relakkes 3ae4c5f2f4 fix: issue #23 2023-08-03 22:26:31 +08:00
Relakkes defd8232da
Merge pull request #24 from jayeeliu/main
去掉搜索结果列表中的推荐搜索词(mode_type in ['rec_query', 'hot_query'])
2023-08-01 13:35:00 +08:00
JayeeLiu 925f0a1bc4 去掉搜索结果列表中的推荐搜索词(mode_type in ['rec_query', 'hot_query']) 2023-08-01 10:29:00 +08:00
Relakkes 1d3c2359e0 fix: 修复部分变量命名语义不明确 2023-07-30 21:30:26 +08:00
Relakkes bf659455bb fix: issue #22 2023-07-30 20:43:02 +08:00
Relakkes 4ff2cf8661 refactor: 优化代码 2023-07-29 15:35:40 +08:00
Relakkes febbb133d7 fix: douyin 缺少一个collected_count字段 2023-07-28 21:23:37 +08:00
Relakkes 03565d61c6 fix: issue #19 2023-07-25 20:22:22 +08:00
Relakkes 7d63c2f9ec Merge branch 'main' of github.com:NanmiCoder/MediaCrawler 2023-07-24 21:00:00 +08:00
Relakkes e75707443b feat: 增加配置项支持自由选择数据是否保存到关系型数据库中 2023-07-24 20:59:43 +08:00
Relakkes b94924e21b
Merge pull request #18 from tanpenggood-fork/typo-sam-20230723
typo(field.py): words "normal" typo
2023-07-23 14:30:19 +08:00