summaryrefslogtreecommitdiff
Commit message (Collapse)AuthorAgeFilesLines
...
* Remove pocketnow.com (#216)Marcel Dopita2016-11-041-3/+0
| | | | New site layout breaks selectors. It works perfectly without provided rule.
* Create pjmedia.com.txtFiveFilters.org2016-10-291-0/+7
|
* Update theguardian.com.txt (#215)Jeremy Benoist2016-10-211-0/+2
|
* Create webmasters.googleblog.com (#214)Simon Alberny2016-10-181-0/+9
|
* Create abc-luxe.com.txt (#213)Zlitus2016-10-171-0/+4
|
* servethehome and lowtechmagazine (#208)TheOneValen2016-10-112-5/+9
| | | | | | * https://www.servethehome.com/ articles config * http://www.lowtechmagazine.com/ article
* Create nrc.nl.txt (#211)Jeremy Benoist2016-10-111-0/+5
|
* Create imgcert.com.txt (#210)Jeremy Benoist2016-10-101-0/+3
|
* Updated heise.de.txt to work with articles from heise.de/tp and heise.de/ct ↵Marmo2016-10-091-0/+4
| | | | | | | | | | | | (#209) * Update heise.de.txt for telepolis-articles Added single_page_link for printing versions in articles at heise.de/tp/ * Update heise.de.txt for ct-articles in articles at heise.de/ct there is no print/single-page-link, so a next_page_link is needed
* Update lemonde.fr.txtFiveFilters.org2016-10-061-2/+4
|
* Small syntax fixes (#207)Max Maischein2016-10-022-2/+2
| | | | | | * Fix: "auther" -> "author" * Fix: Add test_url: in front of test url
* Update gamasutra.com.txtFiveFilters.org2016-10-011-2/+4
|
* Update reddit.com.txt (#206)Jeremy Benoist2016-09-301-2/+3
|
* Add numerama.com (#205)Simon Alberny2016-09-281-0/+1
|
* Create ploum.net.txt (#204)Jeremy Benoist2016-09-281-0/+8
|
* Create variety.com.txt (#203)Jeremy Benoist2016-09-262-2/+15
| | | | | | * Create variety.com.txt * Update journaldugeek.com.txt
* Update theverge.com.txt (#202)Jeremy Benoist2016-09-201-0/+7
|
* Fix mistake in techcrunch.com (#199)Jeremy Benoist2016-09-141-1/+1
|
* Exclude comments from bostonglobe.com (#201)Jeremy Benoist2016-09-141-1/+4
|
* Update some websites (#200)benages2016-09-143-6/+15
| | | | | | | | | | | | | | * Add nakedsecurity.sophos.com * Update technologyreview.com * Add theintercept.com * Updated theintercept.com * Updated theweek.com * Update theweek.com
* Create outsideonline.com.txt (#198)Jeremy Benoist2016-09-071-0/+6
| | | | | | * Create outsideonline.com.txt * Remove Pinterest image
* Create thepointmag.com.txt (#197)Jeremy Benoist2016-09-041-0/+5
|
* Update slate.fr.txt (#195)Jeremy Benoist2016-09-041-0/+1
|
* Create houstonchronicle.com.txt (#196)Jeremy Benoist2016-09-041-0/+4
|
* Reflets.info updated (#194)Simon Alberny2016-08-231-3/+4
|
* LWN.net: Remove subscription notice. (#193)Lukas Anzinger2016-08-181-0/+1
|
* Create rhenus.com.txtFiveFilters.org2016-08-071-0/+5
|
* fm4.ORF.at: Use HTML5 parser. (#191)Lukas Anzinger2016-08-021-0/+1
|
* thegap.at: Add site config. (#190)Lukas Anzinger2016-07-281-0/+41
| | | | | thegap.at uses relative URLs and a base tag in the HTML code. Handling of such URLs (i.e. making them absolute) is not supported by Full-Text RSS as of 3.6 so only the first page is extracted for these versions.
* Update foreignpolicy.com.txt (#187)Jeremy Benoist2016-07-281-13/+8
|
* Fix zive.cz (#189)Marcel Dopita2016-07-272-2/+4
| | | | | | * Update zive.cz * Update mobilmania.cz
* Update futurezone.at and heise.de (#188)Lukas Anzinger2016-07-262-12/+23
| | | | | | * futurezone.at: Update site config. * heise.de: Fix extraction of author and single page link.
* Fix title for TheVerge (#186)Jeremy Benoist2016-07-241-1/+1
|
* Create drgoulu.com (#184)zertrin2016-07-191-0/+6
| | | The default filter fetches too much content. This one gets the articles right.
* Create o6asan.com (#185)zertrin2016-07-191-0/+6
| | | The default filter didn't get the content correctly.
* Configs for keyboardmag.com and lenta.ru (#182)Vladimir Shabanov2016-07-172-0/+11
| | | | | | * Added lenta.ru * Added keyboardmag.com
* Update derStandard.at and ORF.at (#183)Lukas Anzinger2016-07-173-4/+5
| | | | | | * derStandard.at: Remove native ad links at the end of an article. * ORF.at: Fix extraction heuristic for author name.
* Update smbc-comics.com.txt (#181)Adrian Chu2016-07-031-1/+1
| | | filter wasn't fetching main comic
* LWN.net: Fix full text for Weekly Editions. (#180)Lukas Anzinger2016-06-261-0/+2
|
* Add lithub.com and update brainpickings.org (#179)baxq2016-06-212-0/+6
| | | | | | | | | | | | * Update brainpickings.org.txt Remove print links at bottom of articles, and ensure that headline is grabbed correctly * Create lithub.com.txt * Update lithub.com.txt Corrected typo
* add config for arstechnica.co.uk (#176)yes2016-06-171-0/+6
|
* fix content fetching of theatlantic.com with the help of single_page_link (#171)yes2016-06-161-0/+3
|
* Create allafrica.com.txtFiveFilters.org2016-06-071-0/+3
|
* Strip out some extraneous elements from brainpickings.org (#170)baxq2016-05-311-0/+3
| | | | | | | | | | | | | | * Initial commit * Initial commit * Rename .brainpickings.org.txt to brainpickings.org.txt * Delete .theatlantic.com.txt Already exists * Update brainpickings.org.txt
* Fix pagination for JDG sites (#169)Jeremy Benoist2016-05-304-13/+17
| | | | | | | | | | * Fix pagination for JDG sites All pages have a ' next ' link to view the next content, not the next page. But contents with more than one page also got this link but it's not embedded in `single-test` class node. So we first remove that block and we found the ' next ' content, it means it's multipage article. * Fix next link xpath
* Add brainpickings.org (#167)baxq2016-05-291-0/+2
| | | | | | | | | | | | * Initial commit * Initial commit * Rename .brainpickings.org.txt to brainpickings.org.txt * Delete .theatlantic.com.txt Already exists
* Update root.cz.txt (#168)Marcel Dopita2016-05-291-0/+2
|
* Update presse-citron (#165)Bilal Elmoussaoui2016-05-201-3/+3
| | | | | | | | * add ubuntugeek.com.txt * update presse-citron * shouldn't update presse-citron
* add ubuntugeek and update omgubuntu (#164)Bilal Elmoussaoui2016-05-203-5/+15
| | | | | | | | * add ubuntugeek.com.txt * update omgubuntu.co.uk * update presse-citron
* add presse-citron.net (#163)Bilal Elmoussaoui2016-05-201-0/+9
|