Documentation of 'jsat.clustering.dissimilarity.CentroidDissimilarity' Java class
CentroidDissimilarity
jsat.clustering.dissimilarity

Class CentroidDissimilarity

    • Constructor Detail

      • CentroidDissimilarity

        public CentroidDissimilarity()
        Creates a new CentroidDissimilarity that used the EuclideanDistance
      • CentroidDissimilarity

        public CentroidDissimilarity(DistanceMetric dm)
        Creates a new CentroidDissimilarity
        Parameters:
        dm - the distance measure to use between individual points
    • Method Detail

      • dissimilarity

        public double dissimilarity(java.util.List<DataPoint> a,
                                    java.util.List<DataPoint> b)
        Description copied from interface: ClusterDissimilarity
        Provides the notion of dissimilarity between two sets of points, that may not have the same number of points.
        Specified by:
        dissimilarity in interface ClusterDissimilarity
        Overrides:
        dissimilarity in class LanceWilliamsDissimilarity
        Parameters:
        a - the first cluster of points
        b - the second cluster of points
        Returns:
        a value >= 0 that describes the dissimilarity of the two clusters. The larger the value, the more different the two clusterings are.
      • dissimilarity

        public double dissimilarity(java.util.Set<java.lang.Integer> a,
                                    java.util.Set<java.lang.Integer> b,
                                    double[][] distanceMatrix)
        Description copied from interface: ClusterDissimilarity
        Provides the notion of dissimilarity between two sets of points, that may not have the same number of points. This is done using a matrix containing all pairwise distance computations between all points.
        Specified by:
        dissimilarity in interface ClusterDissimilarity
        Overrides:
        dissimilarity in class LanceWilliamsDissimilarity
        Parameters:
        a - the first set of indices of the original data set that are in a cluster, which map to distanceMatrix
        b - the second set of indices of the original data set that are in a cluster, which map to distanceMatrix
        distanceMatrix - the upper triangual distance matrix as created by AbstractClusterDissimilarity.createDistanceMatrix(jsat.DataSet, jsat.clustering.dissimilarity.ClusterDissimilarity)
        Returns:
        a value >= 0 that describes the dissimilarity of the two clusters. The larger the value, the more different the two clusterings are.
      • dissimilarity

        public double dissimilarity(int i,
                                    int ni,
                                    int j,
                                    int nj,
                                    double[][] distanceMatrix)
        Description copied from interface: UpdatableClusterDissimilarity
        Provides the notion of dissimilarity between two sets of points, that may not have the same number of points. This is done using a matrix containing all pairwise distance computations between all points. This distance matrix will then be updated at each iteration and merging, leaving empty space in the matrix. The updates will be done by the clustering algorithm. Implementing this interface indicates that this dissimilarity measure can be accurately computed in an updatable manner that is compatible with a Lance–Williams update.
        Specified by:
        dissimilarity in interface UpdatableClusterDissimilarity
        Overrides:
        dissimilarity in class LanceWilliamsDissimilarity
        Parameters:
        i - the index of cluster i's distance in the original data set
        ni - the number of items in the cluster represented by i
        j - the index of cluster j's distance in the original data set
        nj - the number of items in the cluster represented by j
        distanceMatrix - a distance matrix originally created by AbstractClusterDissimilarity.createDistanceMatrix(jsat.DataSet, jsat.clustering.dissimilarity.ClusterDissimilarity)
        Returns:
        a value >= 0 that describes the dissimilarity of the two clusters. The larger the value, the more different the two clusterings are.
      • dissimilarity

        public double dissimilarity(int i,
                                    int ni,
                                    int j,
                                    int nj,
                                    int k,
                                    int nk,
                                    double[][] distanceMatrix)
        Description copied from interface: UpdatableClusterDissimilarity
        Provides the notion of dissimilarity between two sets of points, that may not have the same number of points. This is done using a matrix containing all pairwise distance computations between all points. This distance matrix will then be updated at each iteration and merging, leaving empty space in the matrix. The updates will be done by the clustering algorithm. Implementing this interface indicates that this dissimilarity measure can be accurately computed in an updatable manner that is compatible with a Lance–Williams update.
        This computes the dissimilarity of the union of clusters i and j, (Ci ∪ Cj), with the cluster k. This method is used by other algorithms to perform an update of the distance matrix in an efficient manner.
        Specified by:
        dissimilarity in interface UpdatableClusterDissimilarity
        Overrides:
        dissimilarity in class LanceWilliamsDissimilarity
        Parameters:
        i - the index of cluster i's distance in the original data set
        ni - the number of items in the cluster represented by i
        j - the index of cluster j's distance in the original data set
        nj - the number of items in the cluster represented by j
        k - the index of cluster k's distance in the original data set
        nk - the number of items in the cluster represented by k a distance matrix originally created by AbstractClusterDissimilarity.createDistanceMatrix(jsat.DataSet, jsat.clustering.dissimilarity.ClusterDissimilarity)
        Returns:
        a value >= 0 that describes the dissimilarity of the union of two clusters with a third cluster. The larger the value, the more different the resulting clusterings are.

DataMelt 3.0 © DataMelt by jWork.ORG

You see the box below because you did not login.