creators_name: Ali Bayir, Murat creators_name: Hakki Toroslu, Ismail creators_name: Cosar, Ahmet creators_name: Fidan, Guven type: conference_item datestamp: 2009-04-06 19:08:52 lastmod: 2009-04-14 04:37:05 metadata_visibility: show title: Smart Miner: A New Framework for Mining Large Scale Web Usage Data ispublished: pub full_text_status: public pres_type: paper abstract: In this paper, we propose a novel framework called SmartMiner for web usage mining problem which uses link information for producing accurate user sessions and frequent navigation patterns. Unlike the simple session concepts in the time and navigation based approaches, where sessions are sequences of web pages requested from the server or viewed in the browser, Smart Miner sessions are set of paths traversed in the web graph that corresponds to users’ navigations among web pages. We have modeled session construction as a new graph problem and utilized a new algorithm, Smart-SRA, to solve this problem efficiently. For the pattern discovery phase, we have developed an efficient version of the Apriori-All technique which uses the structure of web graph to increase the performance. From the experiments that we have performed on both real and simulated data, we have observed that Smart-Miner produces at least 30% more accurate web usage patterns than other approaches including previous session construction methods. We have also studied the effect of having the referrer information in the web server logs to show that different versions of SmartSRA produce similar results. Our another contribution is that we have implemented distributed version of the Smart Miner framework by employing Map/Reduce Paradigm. We conclude that we can efficiently process terabytes of web server logs belonging to multiple web sites by our scalable framework. date: 2009-04 pagerange: 161-161 event_title: 18th International World Wide Web Conference event_location: Madrid, Spain event_dates: April 20th-24th, 2009 event_type: conference refereed: TRUE citation: Ali Bayir, Murat and Hakki Toroslu, Ismail and Cosar, Ahmet and Fidan, Guven (2009) Smart Miner: A New Framework for Mining Large Scale Web Usage Data. In: 18th International World Wide Web Conference, April 20th-24th, 2009, Madrid, Spain. document_url: http://www2009.eprints.org/17/1/p161.pdf