基于Jaccard的移动终端自动识别并行算法及其MapReduce实现(英文)

被引:7
作者
刘军 [1 ]
李银周 [1 ]
Felix Cuadrado [2 ]
Steve Uhlig [2 ]
雷振明 [1 ]
机构
[1] Beijing Key Laboratory of Network System Architecture and Convergence,Beijing University of Posts and Telecommunications
[2] Department of Electronic Engineering and Computer Science,Queen Mary,University of London
关键词
mobile device recognition; data mining; Jaccard coefficient measurement; distributed computing; MapReduce;
D O I
暂无
中图分类号
TN929.53 [蜂窝式移动通信系统(大哥大、移动电话手机)];
学科分类号
080402 ; 080904 ; 0810 ; 081001 ;
摘要
The ability of accurate and scalable mobile device recognition is critically important for mobile network operators and ISPs to understand their customers' behaviours and enhance their user experience.In this paper,we propose a novel method for mobile device model recognition by using statistical information derived from large amounts of mobile network traffic data.Specifically,we create a Jaccardbased coefficient measure method to identify a proper keyword representing each mobile device model from massive unstructured textual HTTP access logs.To handle the large amount of traffic data generated from large mobile networks,this method is designed as a set of parallel algorithms,and is implemented through the MapReduce framework which is a distributed parallel programming model with proven low-cost and high-efficiency features.Evaluations using real data sets show that our method can accurately recognise mobile client models while meeting the scalability and producer-independency requirements of large mobile network operators.Results show that a 91.5% accuracy rate is achieved for recognising mobile client models from 2 billion records,which is dramatically higher than existing solutions.
引用
收藏
页码:71 / 84
页数:14
相关论文
共 3 条
[1]  
Learning to match ontologies on the Semantic Web[J] . AnHai Doan,Jayant Madhavan,Robin Dhamankar,Pedro Domingos,Alon Halevy. The VLDB Journal . 2003 (4)
[2]  
Etude comparative de la distribution florale dans une portion des Alpes et des Jura .2 Jaccard,P. Bulletin de la Societe Vaudoise des Sciences Naturelles . 1901
[3]  
IMEI Al-location and Approval Guidelines .2 GSM Association Official Document . 2011