A Study on Data Mining Techniques in Social Media
Total Page:16
File Type:pdf, Size:1020Kb
Load more
										Recommended publications
									
								- 
												  A Study on Social Network Analysis Through Data Mining Techniques – a Detailed SurveyISSN (Online) 2278-1021 ISSN (Print) 2319-5940 IJARCCE International Journal of Advanced Research in Computer and Communication Engineering ICRITCSA M S Ramaiah Institute of Technology, Bangalore Vol. 5, Special Issue 2, October 2016 A Study on Social Network Analysis through Data Mining Techniques – A detailed Survey Annie Syrien1, M. Hanumanthappa2 Research Scholar, Department of Computer Science and Applications, Bangalore University, India1 Professor, Department of Computer Science and Applications, Bangalore University, India2 Abstract: The paper aims to have a detailed study on data collection, data preprocessing and various methods used in developing a useful algorithms or methodologies on social network analysis in social media. The recent trends and advancements in the big data have led many researchers to focus on social media. Web enabled devices is an another important reason for this advancement, electronic media such as tablets, mobile phones, desktops, laptops and notepads enable the users to actively participate in different social networking systems. Many research has also significantly shows the advantages and challenges that social media has posed to the research world. The principal objective of this paper is to provide an overview of social media research carried out in recent years. Keywords: data mining, social media, big data. I. INTRODUCTION Data mining is an important technique in social media in scrutiny and analysis. In any industry data mining recent years as it is used to help in extraction of lot of technique based reports become the base line for making useful information, whereas this information turns to be an decisions and formulating the decisions. This report important asset to both academia and industries.
- 
												  Geoseq2seq: Information Geometric Sequence-To-Sequence NetworksGEOSEQ2SEQ:INFORMATION GEOMETRIC SEQUENCE-TO-SEQUENCE NETWORKS Alessandro Bay Biswa Sengupta Cortexica Vision Systems Ltd. Noah’s Ark Lab (Huawei Technologies UK) London, UK Imperial College London, London, UK [email protected] ABSTRACT The Fisher information metric is an important foundation of information geometry, wherein it allows us to approximate the local geometry of a probability distribu- tion. Recurrent neural networks such as the Sequence-to-Sequence (Seq2Seq) networks that have lately been used to yield state-of-the-art performance on speech translation or image captioning have so far ignored the geometry of the latent embedding, that they iteratively learn. We propose the information geometric Seq2Seq (GeoSeq2Seq) network which abridges the gap between deep recurrent neural networks and information geometry. Specifically, the latent embedding offered by a recurrent network is encoded as a Fisher kernel of a parametric Gaus- sian Mixture Model, a formalism common in computer vision. We utilise such a network to predict the shortest routes between two nodes of a graph by learning the adjacency matrix using the GeoSeq2Seq formalism; our results show that for such a problem the probabilistic representation of the latent embedding supersedes the non-probabilistic embedding by 10-15%. 1 INTRODUCTION Information geometry situates itself in the intersection of probability theory and differential geometry, wherein it has found utility in understanding the geometry of a wide variety of probability models (Amari & Nagaoka, 2000). By virtue of Cencov’s characterisation theorem, the metric on such probability manifolds is described by a unique kernel known as the Fisher information metric. In statistics, the Fisher information is simply the variance of the score function.
- 
												  Data Mining and Soft Computing© 2002 WIT Press, Ashurst Lodge, Southampton, SO40 7AA, UK. All rights reserved. Web: www.witpress.com Email [email protected] Paper from: Data Mining III, A Zanasi, CA Brebbia, NFF Ebecken & P Melli (Editors). ISBN 1-85312-925-9 Data mining and soft computing O. Ciftcioglu Faculty of Architecture, Deljl University of Technology Delft, The Netherlands Abstract The term data mining refers to information elicitation. On the other hand, soft computing deals with information processing. If these two key properties can be combined in a constructive way, then this formation can effectively be used for knowledge discovery in large databases. Referring to this synergetic combination, the basic merits of data mining and soft computing paradigms are pointed out and novel data mining implementation coupled to a soft computing approach for knowledge discovery is presented. Knowledge modeling by machine learning together with the computer experiments is described and the effectiveness of the machine learning approach employed is demonstrated. 1 Introduction Although the concept of data mining can be defined in several ways, it is clear enough to understand that it is related to information extraction especially in large databases. The methods of data mining are essentially borrowed from exact sciences among which statistics play the essential role. By contrast, soft computing is essentially used for information processing by employing methods, which are capable to deal with imprecision and uncertainty especially needed in ill-defined problem areas. In the former case, the outcomes are exact within the error bounds estimated. In the latter, the outcomes are approximate and in some cases they may be interpreted as outcomes from an intelligent behavior.
- 
												  SOCIAL MEDIA MINING Introduction Dear Instructors/Users of These SlidesSOCIAL MEDIA MINING Introduction Dear instructors/users of these slides: Please feel free to include these slides in your own material, or modify them as you see fit. If you decide to incorporate these slides into your presentations, please include the following note: R. Zafarani, M. A. Abbasi, and H. Liu, Social Media Mining: An Introduction, Cambridge University Press, 2014. Free book and slides at http://socialmediamining.info/ or include a link to the website: http://socialmediamining.info/ Social Media Mining http://socialmediamining.info/ MeasuresIntroduction and Metrics 22 Facebook • How does Facebook use your data? Social Media Mining http://socialmediamining.info/ MeasuresIntroduction and Metrics 33 What about Amazon? Social Media Mining http://socialmediamining.info/ MeasuresIntroduction and Metrics 44 Or Twitter? Social Media Mining http://socialmediamining.info/ MeasuresIntroduction and Metrics 55 Objectives of Our Course • Understand social aspects of the Web – Social Theories + Social media + Mining – Learn to collect, clean, and represent social media data – How to measure important properties of social media and simulate social media models – Find and analyze communities in social media – Understand how information propagates in social media – Understanding friendships in social media, perform recommendations, and analyze behavior • Study or ask interesting research issues – e.g., start-up ideas / research challenges • Learn representative algorithms and tools Social Media Mining http://socialmediamining.info/ MeasuresIntroduction and Metrics 66 Social Media Social Media Mining http://socialmediamining.info/ MeasuresIntroduction and Metrics 77 Definition Social Media is the use of electronic and Internet tools for the purpose of sharing and discussing information and experiences with other human beings in more efficient ways.
- 
												  Data Mining with Cloud Computing: - an OverviewInternational Journal of Advanced Research in Computer Engineering & Technology (IJARCET) Volume 5 Issue 1, January 2016 Data Mining with Cloud Computing: - An Overview Miss. Rohini A. Dhote Dr. S. P. Deshpande Department of Computer Science & Information Technology HOD at Computer Science Technology Department, HVPM, COET Amravati, Maharashtra, India. HVPM, COET Amravati, Maharashtra, India Abstract: Data mining is a process of extracting competitive environment for data analysis. The cloud potentially useful information from data, so as to makes it possible for you to access your information improve the quality of information service. The from anywhere at any time. The cloud removes the integration of data mining techniques with Cloud need for you to be in the same physical location as computing allows the users to extract useful the hardware that stores your data. The use of cloud information from a data warehouse that reduces the computing is gaining popularity due to its mobility, costs of infrastructure and storage. Data mining has attracted a great deal of attention in the information huge availability and low cost. industry and in society as a whole in recent Years Wide availability of huge amounts of data and the imminent 2. DATA MINING need for turning such data into useful information and Data Mining involves the use of sophisticated data knowledge Market analysis, fraud detection, and and tool to discover previously unknown, valid customer retention, production control and science patterns and Relationship in large data sets. These exploration. Data mining can be viewed as a result of tools can include statistical models, mathematical the natural evolution of information technology.
- 
												  A Conceptual Framework for the Mining and Analysis of the Social Media Data 1International Journal of Database Theory and Application Vol. 10, No.10 (2017), pp. 11-34 hhtp://dx.doi.org/10.14257/ijdta.2017.10.10.02 A Conceptual Framework for the Mining and Analysis of the Social Media Data 1 Sethunya R Joseph1*, Keletso Letsholo2 and Hlomani Hlomani3 1,2,3 Computer Science Department, Botswana International University of Science and Technology, Palapye, Botswana [email protected], [email protected] [email protected] Abstract Social media data possess the characteristics of Big Data such as volume, veracity, velocity, variability and value. These characteristics make its analysis a bit more challenging than conventional data. Manual analysis approaches are unable to cope with the fast pace at which data is being generated. Processing data manually is also time consuming and requires a lot of effort as compared to using computational methods. However, computational analysis methods usually cannot capture in-depth meanings (semantics) within data. On their individual capacity, each approach is insufficient. As a solution, we propose a Conceptual Framework, which integrates both the traditional approaches and computational approaches to the mining and analysis of social media data. This allows us to leverage the strengths of traditional content analysis, with its regular meticulousness and relative understanding, whilst exploiting the extensive capacity of Big Data analytics and accuracy of computational methods. The proposed Conceptual Framework was evaluated in two stages using an example case of the political landscape of Botswana data collected from Facebook and Twitter platforms. Firstly, a user study was carried through the Inductive Content Analysis (ICA) process using the collected data.
- 
												  A Social Media Mining and Ensemble Learning Model: Application to Luxury and Fast Fashion Brandsinformation Article A Social Media Mining and Ensemble Learning Model: Application to Luxury and Fast Fashion Brands Yulin Chen Department of Mass Communication, Tamkang University, New Taipei City 25137, Taiwan; [email protected] Abstract: This research proposes a framework for the fashion brand community to explore public participation behaviors triggered by brand information and to understand the importance of key image cues and brand positioning. In addition, it reviews different participation responses (likes, comments, and shares) to build systematic image and theme modules that detail planning require- ments for community information. The sample includes luxury fashion brands (Chanel, Hermès, and Louis Vuitton) and fast fashion brands (Adidas, Nike, and Zara). Using a web crawler, a total of 21,670 posts made from 2011 to 2019 are obtained. A fashion brand image model is constructed to determine key image cues in posts by each brand. Drawing on the findings of the ensemble analysis, this research divides cues used by the six major fashion brands into two modules, image cue module and image and theme cue module, to understand participation responses in the form of likes, comments, and shares. The results of the systematic image and theme module serve as a critical reference for admins exploring the characteristics of public participation for each brand and the main factors motivating public participation. Keywords: fashion brands; luxury brands; masstige; key image cues; social media mining; ensem- ble earning Citation: Chen, Y. A Social Media Mining and Ensemble Learning Model: Application to Luxury and Fast Fashion Brands. Information 2021, 1. Introduction 12, 149.
- 
												  Data Mining Techniques for Social Media AnalysisAdvances in Engineering Research (AER), volume 142 International Conference for Phoenixes on Emerging Current Trends in Engineering and Management (PECTEAM 2018) A Survey: Data Mining Techniques for Social Media Analysis 1Elangovan D, 2Dr.Subedha V, 3Sathishkumar R, 4 Ambeth kumar V D 1Research scholar, Department of Computer Science Engineering, Sathyabama University, India 2Professor, Department of Computer Science Engineering, Panimalar Institute of Technology, Chennai, India 3Assistant Professor, 4Associate Professor, Department of Computer Science Engineering, Panimalar Engineering College, Chennai, India [email protected],[email protected], [email protected],[email protected] Abstract—Data mining is the extraction of present information databases. The overall objective of the data mining technique from high volume of data sets, it’s a modern technology. The is to extract information from a huge data set and transform it main intention of the mining is to extract the information from a into a comprehensible structure for more use. The different large no of data set and convert it into a reasonable structure for data Mining techniques are further use. The social media websites like Facebook, twitter, instagram enclosed the billions of unrefined raw data. The I. Characterization. various techniques in data mining process after analyzing the II. Classification. raw data, new information can be obtained. Since this data is III. Regression. active and unstructured, conventional data mining techniques IV. Association. may not be suitable. This survey paper mainly focuses on various V. Clustering. data mining techniques used and challenges that arise while using VI. Change Detection. it. The survey of various work done in the field of social network analysis mainly focuses on future trends in research.
- 
												  NIST Data Science Symposium ProceedingsMarch 4-5 Data Science Symposium Proceedings 2014 Abstracts for posters presented at the 2014 NIST Data Science Symposium on March 4th 2014 Version 1.1 TABLE OF CONTENTS A CONCEPTUAL FRAMEWORK FOR HEALTH DATA HARMONIZATION .................................................................... 6 LEWIS E. BERMAN, & YAIR G. RAJWAN, ............................................................................................................................. 6 ICF International & Visual Science Informatics ...................................................................................................... 6 REAL-TIME ANALYTICS FOR DATA SCIENCE ............................................................................................................ 7 HIROTAKA OGAWA ......................................................................................................................................................... 7 National Institute of Advance Industrial Science and Technology, JAPAN ............................................................. 7 UTILIZATION OF A VISUAL ANALYTICAL APPROACH TO DETECT ANOMALIES IN LARGE NETWORK TRAFFIC DATA . 7 LASSINE CHERIF, SOO-YEON JI, DONG HYUN JEONG .............................................................................................................. 7 Department of Computer Science and Information Technology, Univ. of the District of Columbia and Dept. of Computer Science, Bowie State University ...........................................................................................................
- 
												  Challenges in Bioinformatics for Statistical Data MinersThis article appeared in the October 2003 issue of the “Bulletin of the Swiss Statistical Society” (46; 10-17), and is reprinted by permission from its editor. Challenges in Bioinformatics for Statistical Data Miners Dr. Diego Kuonen Statoo Consulting, PSE-B, 1015 Lausanne 15, Switzerland [email protected] Abstract Starting with possible definitions of statistical data mining and bioinformatics, this article will give a general side-by-side view of both fields, emphasising on the statistical data mining part and its techniques, illustrate possible synergies and discuss how statistical data miners may collaborate in bioinformatics’ challenges in order to unlock the secrets of the cell. What is statistics and why is it needed? Statistics is the science of “learning from data”. It includes everything from planning for the collection of data and subsequent data management to end-of-the-line activities, such as drawing inferences from numerical facts called data and presentation of results. Statistics is concerned with one of the most basic of human needs: the need to find out more about the world and how it operates in face of variation and uncertainty. Because of the increasing use of statistics, it has become very important to understand and practise statistical thinking. Or, in the words of H. G. Wells: “Statistical thinking will one day be as necessary for efficient citizenship as the ability to read and write”. But, why is statistics needed? Knowledge is what we know. Information is the communication of knowledge. Data are known to be crude information and not knowledge by itself. The way leading from data to knowledge is as follows: from data to information (data become information when they become relevant to the decision problem); from information to facts (information becomes facts when the data can support it); and finally, from facts to knowledge (facts become knowledge when they are used in the successful completion of the decision process).
- 
												  The Theoretical Framework of Data Mining & ItsInternational Journal of Social Science & Interdisciplinary Research__________________________________ ISSN 2277- 3630 IJSSIR, Vol.2 (1), January (2013) Online available at indianresearchjournals.com THE THEORETICAL FRAMEWORK OF DATA MINING & ITS TECHNIQUES SAYYAD RASHEEDUDDIN Associate Professor – HOD CSE dept Pulla Reddy Engineering College Wargal, Medak(Dist), A.P. ABSTRACT Data mining is a process which finds useful patterns from large amount of data. The research in databases and information technology has given rise to an approach to store and manipulate this precious data for further decision making. Data mining is a process of extraction of useful information and patterns from huge data. It is also called as knowledge discovery process, knowledge mining from data, knowledge extraction or data pattern analysis. The current paper discusses the following Objectives: 1. To discuss briefly about the concept of Data Mining 2. To explain about the Data Mining Algorithms & Techniques 3. To know about the Data Mining Applications Keywords: Concept of Data mining, Data mining Techniques, Data mining algorithms, Data mining applications. Introduction to Data Mining The development of Information Technology has generated large amount of databases and huge data in various areas. The research in databases and information technology has given rise to an approach to store and manipulate this precious data for further decision making. Data mining is a process of extraction of useful information and patterns from huge data. It is also called as knowledge discovery process, knowledge mining from data, knowledge extraction or data /pattern analysis. The current paper discuss about the following Objectives: 1. To discuss briefly about the concept of Data Mining 2.
- 
												  Reinforcement Learning for Data Mining Based Evolution AlgorithmReinforcement Learning for Data Mining Based Evolution Algorithm for TSK-type Neural Fuzzy Systems Design Sheng-Fuu Lin, Yi-Chang Cheng, Jyun-Wei Chang, and Yung-Chi Hsu Department of Electrical and Control Engineering, National Chiao Tung University Hsinchu, Taiwan, R.O.C. ABSTRACT membership functions of a fuzzy controller, with the fuzzy rule This paper proposed a TSK-type neural fuzzy system set assigned in advance. Juang [5] proposed the CQGAF to (TNFS) with reinforcement learning for data mining based simultaneously design the number of fuzzy rules and free evolution algorithm (R-DMEA). In R-DMEA, the group-based parameters in a fuzzy system. In [6], the group-based symbiotic symbiotic evolution (GSE) is adopted that each group represents evolution (GSE) was proposed to solve the problem that all the the set of single fuzzy rule. The R-DMEA contains both fuzzy rules were encoded into one chromosome in the structure learning and parameter learning. In the structure traditional genetic algorithm. In the GSE, each chromosome learning, the proposed R-DMEA uses the two-step self-adaptive represents a single rule of fuzzy system. The GSE uses algorithm to decide the suitable number of rules in a TNFS. In group-based population to evaluate the fuzzy rule locally. parameter learning, the R-DMEA uses the mining-based Although GSE has a better performance when compared with selection strategy to decide the selected groups in GSE by the tradition symbiotic evolution, the combination of groups is data mining method called frequent pattern growth. Illustrative fixed. example is conducted to show the performance and applicability Although the above evolution learning algorithms [4]-[6] of R-DMEA.