[email protected] on 01 July 2005 at 17:54 +0000 wrote:
>Screen scraping HTML is
>in my opinion not good practice.  


Agreed - but I wasn't scraping HTML - I was using the Google web
service... and I think that makes it a different issue to the case of
screenscraping...?

For example, as the content was being provided from google in a 'raw' form
by Google, presumably this carries the implication that the content will
have to be republished somehow?  Which is why I posted the query... ( the
example was constructed solely to illustrate the question, rather than to
be used in an app.)

The use of the site:bbc.co.uk etc search switch is where part of the
problem lies, I think. 

 I'm not sure I can articulate at the moment the feeling I have that there
is a difference between republishing general google results (which in a
sense is publishing google's view of the world) and republishing the
results from a single site. 

Perhaps an extreme workaround would be to say that site:whatever should be
disabled in the google search api if the request is not coming *from*
whatever? A bit like the borwser imposed security constraints on who you
accept scripts from...

But this is off-topic for the list...

tony


-
Sent via the backstage.bbc.co.uk discussion group.  To unsubscribe, please 
visit http://backstage.bbc.co.uk/archives/2005/01/mailing_list.html.

Reply via email to