<didl:DIDL xsi:schemaLocation="urn:mpeg:mpeg21:2002:02-DIDL-NS 
			 http://standards.iso.org/ittf/PubliclyAvailableStandards/MPEG-21_schema_files/did/didmodel.xsd" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:didl="urn:mpeg:mpeg21:2002:02-DIDL-NS"><didl:Item><didl:Descriptior><didl:Statement mimeType="application/xml; charset=utf-8"><dii:Identifier xsi:schemaLocation="urn:mpeg:mpeg21:2002:01-DII-NS
		 	http://standards.iso.org/ittf/PubliclyAvailableStandards/MPEG-21_schema_files/dii/dii.xsd" xmlns:dii="urn:mpeg:mpeg21:2002:01-DII-NS">http://www2009.eprints.org/125/</dii:Identifier></didl:Statement></didl:Descriptior><didl:Descriptior><didl:Statement mimeType="application/xml; charset=utf-8"><oai_dc:dc xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xmlns:dc="http://purl.org/dc/elements/1.1/">
        <dc:title>Graph Based Crawler Seed Selection</dc:title>
        <dc:creator>Zheng, Shuyi</dc:creator>
        <dc:creator>Dmitriev, Pavel</dc:creator>
        <dc:creator>Lee Giles, C.</dc:creator>
        <dc:description>This paper identiﬁes and explores the problem of seed selection in a web-scale crawler. We argue that seed selection is not a trivial but very important problem. Selecting proper seeds can increase the number of pages a crawler will discover, and can result in a collection with more “good” and less “bad” pages. Based on the analysis of the graph structure of the web, we propose several seed selection algorithms. Effectiveness of these algorithms is proved by our experimental results on real web data. </dc:description>
        <dc:date>2009-04</dc:date>
        <dc:type>Conference or Workshop Item</dc:type>
        <dc:type>PeerReviewed</dc:type>
        <dc:format>application/pdf</dc:format>
        <dc:identifier>http://www2009.eprints.org/125/1/p1089.pdf</dc:identifier>
        <dc:identifier>Zheng, Shuyi &lt;http://www2009.eprints.org/view/author/Zheng=3AShuyi=3A=3A.html&gt; and Dmitriev, Pavel &lt;http://www2009.eprints.org/view/author/Dmitriev=3APavel=3A=3A.html&gt; and Lee Giles, C. &lt;http://www2009.eprints.org/view/author/Lee_Giles=3AC=2E=3A=3A.html&gt; (2009) Graph Based Crawler Seed Selection. In: 18th International World Wide Web Conference, April 20th-24th, 2009, Madrid, Spain.</dc:identifier>
        <dc:relation>http://www2009.eprints.org/125/</dc:relation></oai_dc:dc></didl:Statement></didl:Descriptior><didl:Component><didl:Descriptior><didl:Statement mimeType="application/xml; charset=utf-8"><dii:Identifier xsi:schemaLocation="urn:mpeg:mpeg21:2002:01-DII-NS
		 	    http://standards.iso.org/ittf/PubliclyAvailableStandards/MPEG-21_schema_files/dii/dii.xsd" xmlns:dii="urn:mpeg:mpeg21:2002:01-DII-NS">http://www2009.eprints.org/125/1/</dii:Identifier></didl:Statement></didl:Descriptior><didl:Resource ref="http://www2009.eprints.org/125/1/p1089.pdf"></didl:Resource></didl:Component></didl:Item></didl:DIDL>