Publication Date:
2006
Citation:
A lightweight architecture for RSS polling of arbitrary Web sources / S. Bossa, G. Fiumara, A. Provetti (CEUR WORKSHOP PROCEEDINGS). - In: CEUR Workshop Proceedings / [a cura di] F. De Paoli, A. Di Stefano, A. Omicini, C. Santoro. - [s.l] : CEUR-Workshop, 2006. - pp. 118-123 (( Intervento presentato al 7. convegno WOA tenutosi a Catania nel 2006.
abstract:
We describe a new Web service architecture designed to make it possible to collect data from traditional plain HTML Web sites, aggregate and serve them in more advanced formats, e.g. as RSS feeds. To locate the relevant data in the plain HTML pages, the architecture requires the insertion of some meta tags in the commented text. Hence, the extra markup remains totally transparent to users and programs. Such annotated HTML documents are then routinely pulled by our Web service, which then aggregates the data and serves them over several channels, e.g. RSS 1.0 or 2.0. Also, a REST-style Web Service allows users to submit XQuery queries to the feeds database. Finally, we discuss scalability issues w.r.t. polling frequencies.
IRIS type:
03 - Contributo in volume
List of contributors:
S. Bossa, G. Fiumara, A. Provetti
Link to information sheet:
Book title:
CEUR Workshop Proceedings