Detecting soft errors by redirection classification

Taehyung Lee, Jinil Kim, Jin Wook Kim, Sung-Ryul Kim, Kunsoo Park
2009 Proceedings of the 18th international conference on World wide web - WWW '09  
A soft error redirection is a URL redirection to a page that returns the HTTP status code 200 (OK) but has actually no relevant content to the client request. Since such redirections degrade the performance of web search engines in many ways, it is highly desirable to remove as many of them as possible. We propose a novel approach to detect soft error redirections by analyzing redirection logs collected during crawling operation. Experimental results on huge crawl data show that our measure can classify soft error redirections effectively.
doi:10.1145/1526709.1526886 dblp:conf/www/LeeKKKP09 fatcat:duc7q75b6vgcbixeln3bojwdgy