xmltv/grab/uk_guardian CHANGELOG, NONE, 1.1 DISCLAIMER, NONE, 1.1 INSTALL, NONE, 1.1 TODO, NONE, 1.1
Geoff <[email protected]>
| Newsgroups | gmane.comp.tv.xmltv.cvs |
|---|---|
| Message-ID | <[email protected]> |
Update of /cvsroot/xmltv/xmltv/grab/uk_guardian
In directory sfp-cvs-1.v30.ch3.sourceforge.com:/tmp/cvs-serv15925/uk_guardian
Added Files:
CHANGELOG DISCLAIMER INSTALL TODO
Log Message:
Add new UK grabber
--- NEW FILE: CHANGELOG ---
1.010 11-Oct-2013 First public release
--- NEW FILE: TODO ---
To Do
=====
--- NEW FILE: INSTALL ---
tv_grab_uk_guardian
===================
XMLTV grabber for The Guardian website.
INSTALLATION - Linux
============
Files should be installed in the following locations:
/usr/bin/
---------
tv_grab_uk_guardian (and make sure it has execute permission)
$HOME/.xmltv/supplement/tv_grab_uk_guardian/
--------------------------------------------
tv_grab_uk_guardian.map.conf
CONFIGURATION
=============
1.
Grabber configuration consists of the usual
tv_grab_uk_guardian --configure
2.
The file 'tv_grab_uk_guardian.map.conf' has two purposes. Firstly you can map the channel ids used by the site into something more meaningful to your PVR. E.g.
map==bbc-1-south==BBC 1
will change 'bbc-1-south' to 'BBC 1' in the output XML.
Note: the lines are of the form "map=={channel id}=={my name}".
The second purpose is to likewise translate genre names. So if your PVR doesn't have a category for 'Science Fiction' but uses 'Sci-fi' instead, then you can specify
cat==Science Fiction==Sci-fi
and the output XML will have 'Sci-fi'.
IMPORTANT: the downloaded 'tv_grab_uk_guardian.map.conf' contains example lines to illustrate the format - you should edit this file to suit your own purposes!
USAGE
=====
All the normal XMLTV capabilities are included.
For extended help information run
tv_grab_uk_guardian --info
VALIDATION
==========
tv_validate_grabber may report an error similar to:
"Line 5 Invalid channel-id 'BBC 1'"
This is a bug in ValidateFile.pm (lines 201-202) which insists the channel-id adheres to RFC2838 despite the xmltv.dtd only saying "preferably" not "SHOULD".
(Having channel ids of the form "bbc1.bbc.co.uk" will be rejected by many PVRs since they require the data to match their own list.)
It may also report
"tv_sort failed on the concatenated data. Probably due to overlapping data between days."
This is a bug in ValidateGrabber.pm (lines 348-353) which insists on the data retrieved being 00:00-23:59 during its "notadditive" test:
"grabbing data for tomorrow first and then for the day after tomorrow and
concatenating them does not yield the same result as grabbing the data
for tomorrow and the day after tomorrow at once."
Both these errors can be ignored.
"Grabber validated ok."
--- NEW FILE: DISCLAIMER ---
IMPORTANT
=========
Disclaimer
----------
The Guardian website's license for these data does not allow non-personal use.
Certainly any commercial use of listings data obtained by using this grabber will breach copyright law, but if you are just using the data for your own personal use then you are probably fine.
By using this grabber you aver you are using the listings data for your own personal use only and you absolve the author(s) from any liability under copyright law or otherwise.
You should be aware that, as with any screen scraping program, the website owner may understandably be upset if they detect your scraping and takes steps to block it (such as by banning your IP address from accessing their site) and, in the worst case, the website owner may change their site to make scraping impractical thereby messing things up for everyone else.
So be respectful in your use of their site:
(i) Don't run an inordinate number of scrapes,
(ii) Allow a time interval between retrievals.
Be thankful that they are making these data available.
In short, play nice!
Thanks.
* THIS SOFTWARE IS PROVIDED BY THE COPYRIGHT HOLDERS AND CONTRIBUTORS "AS IS" AND ANY
* EXPRESS OR IMPLIED WARRANTIES, INCLUDING, BUT NOT LIMITED TO, THE IMPLIED WARRANTIES OF
* MERCHANTABILITY AND FITNESS FOR A PARTICULAR PURPOSE ARE DISCLAIMED. IN NO EVENT SHALL
* THEY BE LIABLE FOR ANY DIRECT, INDIRECT, INCIDENTAL, SPECIAL, EXEMPLARY, OR CONSEQUENTIAL
* DAMAGES (INCLUDING, BUT NOT LIMITED TO, PROCUREMENT OF SUBSTITUTE GOODS OR SERVICES;
* LOSS OF USE, DATA, OR PROFITS; OR BUSINESS INTERRUPTION) HOWEVER CAUSED AND ON ANY
* THEORY OF LIABILITY, WHETHER IN CONTRACT, STRICT LIABILITY, OR TORT (INCLUDING NEGLIGENCE
* OR OTHERWISE) ARISING IN ANY WAY OUT OF THE USE OF THIS SOFTWARE, EVEN IF ADVISED OF THE
* POSSIBILITY OF SUCH DAMAGE.
------------------------------------------------------------------------------
October Webinars: Code for Performance
Free Intel webinars can help you accelerate application performance.
Explore tips for MPI, OpenMP, advanced profiling, and more. Get the most from
the latest Intel processors and coprocessors. See abstracts and register >
http://pubads.g.doubleclick.net/gampad/clk?id=60135031&iu=/4140/ostg.clktrk