fnord
7a4a945f13
fix/support news.vice.com
2015-07-19 11:33:02 -05:00
fnord
4a63291144
Add tests for compat_urllib_parse_unquote
2015-07-17 09:46:08 -05:00
fnord
593b77064c
Don't forget trailing '%'
2015-07-17 09:45:49 -05:00
fnord
9fefc88656
fix TestCompat test_all_present
2015-07-17 07:24:07 -05:00
fnord
a0f28f90fa
remove kebab
2015-07-17 01:50:43 -05:00
fnord
851229a01f
remove debugprint
2015-07-17 01:49:55 -05:00
fnord
c9c854cea7
replace old compat_urllib_parse_unquote with backport from python3's function
...
* required unquote_to_bytes function ported as well
(uses .decode('hex') instead of dynamically populated _hextobyte global)
* required implicit conversion to bytes and/or unicode in places due to
differing type assumptions in p3
2015-07-17 01:31:29 -05:00
fnord
45eedbe58c
Generic: use compat_urllib_parse_unquote to prevent utf8 mangling
...
of the entire page in python 2.
-requires- fixed compat_urllib_parse_unquote
example - the following will save with a mangled playlist title,
instead of the kanji for 'tsunami'. This affects all utf8encoded
urls as well
youtube-dl -f18 -o '%(playlist_title)s-%(title)s.%(ext)s' \
https://gist.githubusercontent.com/atomicdryad/fcb97465e6060fc519e1/raw/61c14c1e3a4985471dcf56c281d24d7e781a4e0e/tsunami.html
2015-07-15 15:30:47 -05:00
fnord
e37c932fca
compat_urllib_parse_unquote: crash fix: only decode valid hex
...
on python 2 the following has a { "crash_rate": "100%" } of the time
as it tries to parse '" ' as hex.
2015-07-15 15:28:50 -05:00
fnord
b4e1576aee
Brightcove extractor: support customBC.createVideo(...); method
...
found in http://www.americanbar.org/groups/family_law.html and
http://america.aljazeera.com/watch/shows/america-tonight/2015/6/exclusive-hunting-isil-with-the-pkk.html
2015-06-13 06:20:30 -05:00