Mike Fährmann
1f7101d606
[archivedmoe] fix thebarchive webm URLs ( #5116 )
8 months ago
Mike Fährmann
34a4ddc399
[sankaku] add 'id-format' option ( #5073 )
8 months ago
Mike Fährmann
afd20ef42c
[kemonoparty] implement filtering duplicate revisions ( #5013 )
...
set 'revisions' to '"unique"' to have it ignore duplicate revisions
8 months ago
Mike Fährmann
c28475d325
[kemonoparty] fix deleting 'name' in orginal objects ( #5103 )
...
... when computing 'revision_hash'
regression caused by 3d68eda4
dict.copy() only creates a shallow copy
I know that and still managed to get I wrong ...
8 months ago
Mike Fährmann
beacfa7436
[bunkr] update domain to 'bunkr.sk' ( #5114 )
8 months ago
Mike Fährmann
0d3af0d35b
[tests] ignore 'ytdl' categories when import fails ( #5095 )
8 months ago
Mike Fährmann
f3ad91b44f
[bunkr] update domain ( #5088 )
8 months ago
Mike Fährmann
c7a42880ab
[wikimedia] support fandom wikis ( #1443 , #2677 , #3378 )
...
Wikis hosted on fandom.com are just wikimedia instances
and support its API.
8 months ago
blankie
df718887c2
[webtoons] fix extracting comic and episode name with commas
8 months ago
Mike Fährmann
0d367ce1b9
[tests] update extractor results
8 months ago
Mike Fährmann
9ca6117c67
[hbrowse] remove module
...
website gone
8 months ago
Mike Fährmann
375eefb886
[chevereto] remove 'pixl.li'
...
"Pixl is closing down"
"All images will be deleted January 1st."
8 months ago
Mike Fährmann
b0a441f1e3
[nitter] remove 'nitter.lacontrevoie.fr'
...
"Fermeture de Nitter / Closing down Nitter"
8 months ago
Mike Fährmann
a1c1e80f67
[giantessbooru] update domain
8 months ago
Mike Fährmann
2007cb2f59
[tests] check extractor category values
8 months ago
Mike Fährmann
93b4120e77
[gelbooru] support 'all' and empty tag ( #5076 )
8 months ago
Mike Fährmann
a416d4c3d5
[sankaku] support post URLs with alphanumeric IDs ( #5073 )
8 months ago
Mike Fährmann
ea553a1d55
[wikimedia] generalize ( #1443 )
...
- support mediawiki.org
- support mariowiki.com (#3660 )
- combine code into a single extractor
(use prefix as subcategory)
- handle non-wiki instances
- unescape titles
8 months ago
Mike Fährmann
c3c1635ef3
[wikimedia] update
...
- rewrite using BaseExtractor
- support most Wiki* domains
- update docs/supportedsites
- add tests
8 months ago
Mike Fährmann
3d68eda4ab
[kemonoparty] add 'revision_hash' metadata ( #4706 , #4727 , #5013 )
...
A SHA1 hexdigest of other relevant metadata fields like
title, content, file and attachment URLs.
This value does NOT reflect which revisions are listed on the website.
Neither does 'edited' or any other metadata field (combinations).
8 months ago
Mike Fährmann
799a8206ad
merge #5061 : [webtoons] extract more metadata
...
- author_name
- comic_name
- episode_name
- username
8 months ago
Mike Fährmann
8ffa0cd3c8
[webtoons] small optimization
...
don't extract the entire 'author_area' and
avoid creating a second 'text.extract_from()' object
8 months ago
Mike Fährmann
68196589c4
[2ch] update
...
- simplify extractor code
- more metadata
- add tests
8 months ago
Mike Fährmann
69726fc82c
[tests] skip tests requiring auth when non is provided
8 months ago
blankie
bb446b1598
[webtoons] extract more metadata
8 months ago
Mike Fährmann
355b909f46
merge #5041 : [steamgriddb] add support ( #5033 )
8 months ago
Mike Fährmann
71e2c3e5a2
merge #5037 : [hatenablog] add support ( #5036 )
8 months ago
Mike Fährmann
b97af09e03
[tests] include URL in failure report
8 months ago
Mike Fährmann
58e0665fbc
[tests] load config from external file
8 months ago
Mike Fährmann
2dcfb012ea
[patreon] download 'm3u8' manifests with ytdl
8 months ago
Mike Fährmann
2191e29e14
[nijie] fix image URL for single image posts ( #5049 )
8 months ago
Mike Fährmann
39904c9e4e
[deviantart:avatar] add 'formats' option ( #4995 )
8 months ago
Mike Fährmann
887ade30a5
[batoto] support more mirror domains ( #5042 )
8 months ago
blankie
2ccb7d3bd3
[steamgriddb] add support
8 months ago
blankie
2cfe788f93
[hatenablog] fix extractor naming errors
9 months ago
blankie
61f3b2f820
[hatenablog] add support
9 months ago
Mike Fährmann
657ed93a22
[batoto] improve v2 manga URL pattern
...
and add tests
9 months ago
Mike Fährmann
33f228756a
[mangadex] add 'list' extractor ( #5025 )
...
supports listing manga and chapters from list feed
9 months ago
Mike Fährmann
c25bdbae91
[komikcast] fix 'manga' extractor ( #5027 )
9 months ago
Mike Fährmann
8e1a2b5446
[komikcast] update domain to 'komikcast.lol' ( #5027 )
9 months ago
Mike Fährmann
a441249ea2
merge #4979 : [batoto] add 'chapter' and 'manga' extractors ( #1434 , #2111 )
9 months ago
Mike Fährmann
b11c352d66
[bato] rename to 'batoto'
...
to use the same category name as the previous bato.to site
9 months ago
Mike Fährmann
3aa24c3744
[bato] simplify and update
9 months ago
Mike Fährmann
11150a7d72
[nudecollect] remove module
9 months ago
Mike Fährmann
c158927c38
merge #5016 : [zzup] add 'gallery' extractor ( #4517 , #4604 , #4659 , #4863 )
9 months ago
Mike Fährmann
217fa7f8a1
include 'test/results' in flake8 checks
9 months ago
Mike Fährmann
e61f016465
[szurubooru] support 'snootbooru.com' ( #5023 )
9 months ago
Mike Fährmann
b4bcf40278
[weibo] fix AttributeError in 'user' extractor ( #5022 )
...
yet another bug caused by a383eca7
9 months ago
Mike Fährmann
0ab0a10d2d
[jpgfish] update domain
9 months ago
enduser420
0f30136109
[zzup] add 'gallery' extractor
9 months ago