Proceedings Article10.1109/ICSMC.2006.384769
A Local Density Based Spatial Clustering Algorithm with Noise
Lian Duan,Deyi Xiong,Jun Lee,Feng Guo +3 more
- 01 Oct 2006
- Vol. 5, pp 4061-4066
TL;DR: A new clustering algorithm LDBSCAN relying on a local-density-based notion of clusters is proposed to solve problems of density-based clustering in spatial database and takes the advantage of the LOF to detect the noises.
read more
Abstract: Density-based clustering algorithms are attractive for the task of class identification in spatial database. However, in many cases, very different local-density clusters exist in different regions of data space, therefore, DBSCAN [Ester, M. et al., A Density-Based Algorithm for Discovering Clusters in Large Spatial Databases with Noise. In E. Simoudis, J. Han, & U. M. Fayyad (Eds.), Proc. 2nd Int. Conf. on Knowledge Discovery and Data Mining (pp. 226-231). Portland, OR: AAAI.] using a global density parameter is not suitable. As an improvement, OPTICS [Ankerst, M. et al,(1999). OPTICS: Ordering Points To Identify the Clustering Structure. In A. Delis, C. Faloutsos, & S. Ghandeharizadeh (Eds.), Proc. ACM SIGMOD Int. Conf. on Management of Data (pp. 49-60). Philadelphia, PA: ACM.] creates an augmented ordering of the database representing its density-based clustering structure, but it only generates the clusters whose local-density exceeds some threshold instead of similar local-density clusters and doesn't produce a clustering of a data set explicitly. Furthermore the parameters required by almost all the well-known clustering algorithms are hard to determine but have a significant influence on the clustering result. In this paper, a new clustering algorithm LDBSCAN relying on a local-density-based notion of clusters is proposed to solve those problems and, what is more, it is very easy for us to pick the appropriate parameters and takes the advantage of the LOF [Breunig, M. M., et al.,(2000). LOF: Identifying Density-Based Local Outliers. In W. Chen, J. F. Naughton, & P. A. Bernstein (Eds.), Proc. ACM SIGMOD Int. Conf. on Management of Data (pp. 93-104). Dalles, TX: ACM.] to detect the noises comparing with other density-based clustering algorithms. The proposed algorithm has potential applications in business intelligence and enterprise information systems.
read more
Chat with Paper
AI Agents for this Paper
Find similar papers on Google Scholar, PubMed and Arxiv
Write a critical review of this paper
Analyze citations of this paper to find unaddressed research gaps
Citations
Customer segmentation issues and strategies for an automobile dealership with two clustering techniques
TL;DR: This study investigates the problem of customer segmentation in relation to automotive customer relationship management and presents a real case study of an automobile dealer in Taiwan, using two clustering techniques, k-means and expectation maximization, and compares their results for correctness.
43
Life-space characterization from cellular telephone collected GPS data
TL;DR: The proposed scheme generated sufficient zone-based activity information to characterize an individual’s life-space, and the speed-based sensitivity assessment suggests 75 s as an appropriate GPS observation interval for smartphone based life- space data collection.
40
Digital Nations – Smart Cities, Innovation, and Sustainability
Arpan Kumar Kar,P. Ilavarasan,M. P. Gupta,Yogesh K. Dwivedi,Matti Mäntymäki,Marijn Janssen,Antonis C. Simintiras,Salah Al-Sharhan +7 more
- 01 Jan 2017
TL;DR: The findings of the study illustrate that only three factors Social Influence, Price Value and Habit of UTAUT2 model are significantly influencing the adoption of IRCTC Connect with adjusted R-Square value 0.699.
A multi-agent-based model for a negotiation support system in electronic commerce
TL;DR: A utility model based on fuzzy constraint satisfaction problems is proposed to ensure that agents in a bilateral multi-objective negotiation in electronic commerce trading reach a solution that is fair for both negotiating parties if such a solution exists.
33
A Case Study in Big Data Analytics: Exploring Twitter Sentiment Analysis and the Weather
Richard O. Sinnott,H. Duan,Y. Sun +2 more
- 01 Jan 2016
TL;DR: In this paper, the authors explore the relationship between weather and human emotion through a cloud-based Big Data solution and provide a practical demonstration of how Big Data technologies and infrastructures can be developed and delivered where nuances and correlations between combinations of large-scale and heterogeneous data can be discovered.
31
References
•Proceedings Article
A density-based algorithm for discovering clusters a density-based algorithm for discovering clusters in large spatial databases with noise
Martin Ester,Hans-Peter Kriegel,Jörg Sander,Xiaowei Xu +3 more
- 02 Aug 1996
TL;DR: In this paper, a density-based notion of clusters is proposed to discover clusters of arbitrary shape, which can be used for class identification in large spatial databases and is shown to be more efficient than the well-known algorithm CLAR-ANS.
20.3K
•Proceedings Article
A density-based algorithm for discovering clusters in large spatial Databases with Noise
Martin Ester,Hans-Peter Kriegel,Jörg Sander,Xiaowei Xu +3 more
- 01 Jan 1996
TL;DR: DBSCAN, a new clustering algorithm relying on a density-based notion of clusters which is designed to discover clusters of arbitrary shape, is presented which requires only one input parameter and supports the user in determining an appropriate value for it.
LOF: identifying density-based local outliers
Markus M. Breunig,Hans-Peter Kriegel,Raymond T. Ng,Jörg Sander +3 more
- 16 May 2000
TL;DR: This paper contends that for many scenarios, it is more meaningful to assign to each object a degree of being an outlier, called the local outlier factor (LOF), and gives a detailed formal analysis showing that LOF enjoys many desirable properties.
7.3K