Class/Object

com.johnsnowlabs.nlp.annotators.resolution

ChunkEntityResolverModel

Related Docs: object ChunkEntityResolverModel | package resolution

Permalink

class ChunkEntityResolverModel extends AnnotatorModel[ChunkEntityResolverModel] with ResolverParams with HasStorageModel with HasEmbeddingsProperties with HasCaseSensitiveProperties with Licensed

Contains all the parameters to transform a dataset with two Input Annotations of types TOKEN and WORD_EMBEDDINGS, coming from ChunkTokenizer and ChunkEmbeddings Annotators and return the Normalized Entity for a particular trained ontology / curated dataset.

Linear Supertypes
Licensed, HasEmbeddingsProperties, HasStorageModel, HasExcludableStorage, HasStorageReader, HasCaseSensitiveProperties, HasStorageRef, ResolverParams, AnnotatorModel[ChunkEntityResolverModel], CanBeLazy, RawAnnotator[ChunkEntityResolverModel], HasOutputAnnotationCol, HasInputAnnotationCols, HasOutputAnnotatorType, ParamsAndFeaturesWritable, HasFeatures, DefaultParamsWritable, MLWritable, Model[ChunkEntityResolverModel], Transformer, PipelineStage, Logging, Params, Serializable, Serializable, Identifiable, AnyRef, Any
Ordering
  1. Alphabetic
  2. By Inheritance
Inherited
  1. ChunkEntityResolverModel
  2. Licensed
  3. HasEmbeddingsProperties
  4. HasStorageModel
  5. HasExcludableStorage
  6. HasStorageReader
  7. HasCaseSensitiveProperties
  8. HasStorageRef
  9. ResolverParams
  10. AnnotatorModel
  11. CanBeLazy
  12. RawAnnotator
  13. HasOutputAnnotationCol
  14. HasInputAnnotationCols
  15. HasOutputAnnotatorType
  16. ParamsAndFeaturesWritable
  17. HasFeatures
  18. DefaultParamsWritable
  19. MLWritable
  20. Model
  21. Transformer
  22. PipelineStage
  23. Logging
  24. Params
  25. Serializable
  26. Serializable
  27. Identifiable
  28. AnyRef
  29. Any
  1. Hide All
  2. Show All
Visibility
  1. Public
  2. All

Instance Constructors

  1. new ChunkEntityResolverModel()

    Permalink
  2. new ChunkEntityResolverModel(uid: String)

    Permalink

Type Members

  1. type AnnotationContent = Seq[Row]

    Permalink
    Attributes
    protected
    Definition Classes
    AnnotatorModel
  2. type AnnotatorType = String

    Permalink
    Definition Classes
    HasOutputAnnotatorType

Value Members

  1. final def !=(arg0: Any): Boolean

    Permalink
    Definition Classes
    AnyRef → Any
  2. final def ##(): Int

    Permalink
    Definition Classes
    AnyRef → Any
  3. final def $[T](param: Param[T]): T

    Permalink
    Attributes
    protected
    Definition Classes
    Params
  4. def $$[T](feature: StructFeature[T]): T

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  5. def $$[K, V](feature: MapFeature[K, V]): Map[K, V]

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  6. def $$[T](feature: SetFeature[T]): Set[T]

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  7. def $$[T](feature: ArrayFeature[T]): Array[T]

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  8. final def ==(arg0: Any): Boolean

    Permalink
    Definition Classes
    AnyRef → Any
  9. def _transform(dataset: Dataset[_], recursivePipeline: Option[PipelineModel]): DataFrame

    Permalink
    Attributes
    protected
    Definition Classes
    AnnotatorModel
  10. def afterAnnotate(dataset: DataFrame): DataFrame

    Permalink
    Attributes
    protected
    Definition Classes
    AnnotatorModel
  11. val allDistancesMetadata: BooleanParam

    Permalink

    whether or not to return an all distance values in the metadata.

    whether or not to return an all distance values in the metadata. Default: False

    Definition Classes
    ResolverParams
  12. val alternatives: IntParam

    Permalink

    number of results to return in the metadata after sorting by last distance calculated

    number of results to return in the metadata after sorting by last distance calculated

    Definition Classes
    ResolverParams
  13. def annotate(annotations: Seq[Annotation]): Seq[Annotation]

    Permalink

    Resolves the ResolverLabel for the given array of TOKEN and WORD_EMBEDDINGS annotations

    Resolves the ResolverLabel for the given array of TOKEN and WORD_EMBEDDINGS annotations

    annotations

    an array of TOKEN and WORD_EMBEDDINGS Annotation objects coming from ChunkTokenizer and ChunkEmbeddings respectively

    returns

    an array of Annotation objects, with the result of the entity resolution for each chunk and the following metadata all_k_results -> Sorted ResolverLabels in the top alternatives that match the distance threshold all_k_resolutions -> Respective ResolverNormalized strings all_k_distances -> Respective distance values after aggregation all_k_wmd_distances -> Respective WMD distance values all_k_tfidf_distances -> Respective TFIDF Cosinge distance values all_k_jaccard_distances -> Respective Jaccard distance values all_k_sorensen_distances -> Respective SorensenDice distance values all_k_jaro_distances -> Respective JaroWinkler distance values all_k_levenshtein_distances -> Respective Levenshtein distance values all_k_confidences -> Respective normalized probabilities based in inverse distance values target_text -> The actual searched string resolved_text -> The top ResolverNormalized string confidence -> Top probability distance -> Top distance value sentence -> Sentence index chunk -> Chunk Index token -> Token index

    Definition Classes
    ChunkEntityResolverModel → AnnotatorModel
  14. final def asInstanceOf[T0]: T0

    Permalink
    Definition Classes
    Any
  15. val auxLabelCol: Param[String]

    Permalink

    Optional column with one extra label per document.

    Optional column with one extra label per document. This extra label will be outputted later on in an additional column

  16. val auxLabelMap: StructFeature[Map[String, String]]

    Permalink
  17. def beforeAnnotate(dataset: Dataset[_]): Dataset[_]

    Permalink

    validates the dataset before applying it further down the pipeline

    validates the dataset before applying it further down the pipeline

    Attributes
    protected
    Definition Classes
    ChunkEntityResolverModel → AnnotatorModel
  18. val caseSensitive: BooleanParam

    Permalink
    Definition Classes
    HasCaseSensitiveProperties
  19. final def checkSchema(schema: StructType, inputAnnotatorType: String): Boolean

    Permalink
    Attributes
    protected
    Definition Classes
    HasInputAnnotationCols
  20. final def clear(param: Param[_]): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    Params
  21. def clone(): AnyRef

    Permalink
    Attributes
    protected[java.lang]
    Definition Classes
    AnyRef
    Annotations
    @throws( ... )
  22. val confidenceFunction: Param[String]

    Permalink

    what function to use to calculate confidence: INVERSE or SOFTMAX

    what function to use to calculate confidence: INVERSE or SOFTMAX

    Definition Classes
    ResolverParams
  23. def copy(extra: ParamMap): ChunkEntityResolverModel

    Permalink
    Definition Classes
    RawAnnotator → Model → Transformer → PipelineStage → Params
  24. def copyValues[T <: Params](to: T, extra: ParamMap): T

    Permalink
    Attributes
    protected
    Definition Classes
    Params
  25. def createDatabaseConnection(database: Name): RocksDBConnection

    Permalink
    Definition Classes
    HasStorageRef
  26. def createReader(database: Name, connection: RocksDBConnection): WordEmbeddingsReader

    Permalink

    creates WordEmbeddingsReader, based on the DB name and connection

    creates WordEmbeddingsReader, based on the DB name and connection

    database

    Name of the desired database

    connection

    Connection to the RocksDB

    returns

    The instance of the class WordEmbeddingsReader

    Attributes
    protected
    Definition Classes
    ChunkEntityResolverModel → HasStorageReader
  27. val databases: Array[Name]

    Permalink

    This cannot hold EMBEDDINGS since otherwise ER will try to re-save and read embeddings again

    This cannot hold EMBEDDINGS since otherwise ER will try to re-save and read embeddings again

    Attributes
    protected
    Definition Classes
    ChunkEntityResolverModel → HasStorageModel
  28. final def defaultCopy[T <: Params](extra: ParamMap): T

    Permalink
    Attributes
    protected
    Definition Classes
    Params
  29. def deserializeStorage(path: String, spark: SparkSession): Unit

    Permalink
    Definition Classes
    HasStorageModel
  30. def dfAnnotate: UserDefinedFunction

    Permalink
    Attributes
    protected
    Definition Classes
    AnnotatorModel
  31. val dimension: IntParam

    Permalink
    Definition Classes
    HasEmbeddingsProperties
  32. val distanceFunction: Param[String]

    Permalink

    what distance function to use for KNN: 'EUCLIDEAN' or 'COSINE'

    what distance function to use for KNN: 'EUCLIDEAN' or 'COSINE'

    Definition Classes
    ResolverParams
  33. val distanceWeights: DoubleArrayParam

    Permalink

    distance weights to apply before pooling: [WMD, TFIDF, Jaccard, SorensenDice, JaroWinkler, Levenshtein]

    distance weights to apply before pooling: [WMD, TFIDF, Jaccard, SorensenDice, JaroWinkler, Levenshtein]

    Definition Classes
    ResolverParams
  34. val enableJaccard: BooleanParam

    Permalink

    whether or not to use Jaccard token distance.

    whether or not to use Jaccard token distance. Default: True

    Definition Classes
    ResolverParams
  35. val enableJaroWinkler: BooleanParam

    Permalink

    whether or not to use Jaro-Winkler character distance.

    whether or not to use Jaro-Winkler character distance. Default: False

    Definition Classes
    ResolverParams
  36. val enableLevenshtein: BooleanParam

    Permalink

    whether or not to use Levenshtein character distance.

    whether or not to use Levenshtein character distance. Default: False

    Definition Classes
    ResolverParams
  37. val enableSorensenDice: BooleanParam

    Permalink

    whether or not to use Sorensen-Dice token distance.

    whether or not to use Sorensen-Dice token distance. Default: False

    Definition Classes
    ResolverParams
  38. val enableTfidf: BooleanParam

    Permalink

    whether or not to use TFIDF token distance.

    whether or not to use TFIDF token distance. Default: True

    Definition Classes
    ResolverParams
  39. val enableWmd: BooleanParam

    Permalink

    whether or not to use WMD token distance.

    whether or not to use WMD token distance. Default: True

    Definition Classes
    ResolverParams
  40. final def eq(arg0: AnyRef): Boolean

    Permalink
    Definition Classes
    AnyRef
  41. def equals(arg0: Any): Boolean

    Permalink
    Definition Classes
    AnyRef → Any
  42. def explainParam(param: Param[_]): String

    Permalink
    Definition Classes
    Params
  43. def explainParams(): String

    Permalink
    Definition Classes
    Params
  44. def extraValidate(structType: StructType): Boolean

    Permalink
    Attributes
    protected
    Definition Classes
    RawAnnotator
  45. def extraValidateMsg: String

    Permalink
    Attributes
    protected
    Definition Classes
    RawAnnotator
  46. final def extractParamMap(): ParamMap

    Permalink
    Definition Classes
    Params
  47. final def extractParamMap(extra: ParamMap): ParamMap

    Permalink
    Definition Classes
    Params
  48. val extramassPenalty: DoubleParam

    Permalink

    penalty for extra words in the knowledge base match during WMD calculation

    penalty for extra words in the knowledge base match during WMD calculation

    Definition Classes
    ResolverParams
  49. val features: ArrayBuffer[Feature[_, _, _]]

    Permalink
    Definition Classes
    HasFeatures
  50. def finalize(): Unit

    Permalink
    Attributes
    protected[java.lang]
    Definition Classes
    AnyRef
    Annotations
    @throws( classOf[java.lang.Throwable] )
  51. def get[T](feature: StructFeature[T]): Option[T]

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  52. def get[K, V](feature: MapFeature[K, V]): Option[Map[K, V]]

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  53. def get[T](feature: SetFeature[T]): Option[Set[T]]

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  54. def get[T](feature: ArrayFeature[T]): Option[Array[T]]

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  55. final def get[T](param: Param[T]): Option[T]

    Permalink
    Definition Classes
    Params
  56. def getAllDistancesMetadata: Boolean

    Permalink
    Definition Classes
    ResolverParams
  57. def getAlternatives: Int

    Permalink
    Definition Classes
    ResolverParams
  58. def getAuxLabelCol(): String

    Permalink
  59. def getAuxLabelMap(): Map[String, String]

    Permalink
  60. def getCaseSensitive: Boolean

    Permalink
    Definition Classes
    HasCaseSensitiveProperties
  61. final def getClass(): Class[_]

    Permalink
    Definition Classes
    AnyRef → Any
  62. def getConfidenceFunction: String

    Permalink
    Definition Classes
    ResolverParams
  63. final def getDefault[T](param: Param[T]): Option[T]

    Permalink
    Definition Classes
    Params
  64. def getDimension: Int

    Permalink
    Definition Classes
    HasEmbeddingsProperties
  65. def getDistanceFunction: String

    Permalink
    Definition Classes
    ResolverParams
  66. def getDistanceWeights: Array[Double]

    Permalink
    Definition Classes
    ResolverParams
  67. def getEnableJaccard: Boolean

    Permalink
    Definition Classes
    ResolverParams
  68. def getEnableJaroWinkler: Boolean

    Permalink
    Definition Classes
    ResolverParams
  69. def getEnableLevenshtein: Boolean

    Permalink
    Definition Classes
    ResolverParams
  70. def getEnableSorensenDice: Boolean

    Permalink
    Definition Classes
    ResolverParams
  71. def getEnableTfidf: Boolean

    Permalink
    Definition Classes
    ResolverParams
  72. def getEnableWmd: Boolean

    Permalink
    Definition Classes
    ResolverParams
  73. def getExtramassPenalty: Double

    Permalink
    Definition Classes
    ResolverParams
  74. def getIncludeStorage: Boolean

    Permalink
    Definition Classes
    HasExcludableStorage
  75. def getInputCols: Array[String]

    Permalink
    Definition Classes
    HasInputAnnotationCols
  76. def getLazyAnnotator: Boolean

    Permalink
    Definition Classes
    CanBeLazy
  77. def getMissAsEmpty: Boolean

    Permalink
    Definition Classes
    ResolverParams
  78. def getNeighbours: Int

    Permalink
    Definition Classes
    ResolverParams
  79. final def getOrDefault[T](param: Param[T]): T

    Permalink
    Definition Classes
    Params
  80. final def getOutputCol: String

    Permalink
    Definition Classes
    HasOutputAnnotationCol
  81. def getParam(paramName: String): Param[Any]

    Permalink
    Definition Classes
    Params
  82. def getPoolingStrategy: String

    Permalink
    Definition Classes
    ResolverParams
  83. def getReader[A](database: Name): StorageReader[A]

    Permalink
    Attributes
    protected
    Definition Classes
    HasStorageReader
  84. def getReturnAllKEmbeddings(): Boolean

    Permalink
  85. def getReturnCosineDistances: Boolean

    Permalink
  86. def getSearchTree: SerializableKDTree[TreeData]

    Permalink
  87. def getStorageRef: String

    Permalink
    Definition Classes
    HasStorageRef
  88. def getTermIDF: Map[String, (Int, Double)]

    Permalink
  89. def getThreshold: Double

    Permalink
    Definition Classes
    ResolverParams
  90. def getUseAuxLabel(): Boolean

    Permalink
  91. final def hasDefault[T](param: Param[T]): Boolean

    Permalink
    Definition Classes
    Params
  92. def hasParam(paramName: String): Boolean

    Permalink
    Definition Classes
    Params
  93. def hasParent: Boolean

    Permalink
    Definition Classes
    Model
  94. def hashCode(): Int

    Permalink
    Definition Classes
    AnyRef → Any
  95. val includeStorage: BooleanParam

    Permalink
    Definition Classes
    HasExcludableStorage
  96. def initializeLogIfNecessary(isInterpreter: Boolean, silent: Boolean): Boolean

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  97. def initializeLogIfNecessary(isInterpreter: Boolean): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  98. val inputAnnotatorTypes: Array[String]

    Permalink

    Annotator reference id.

    Annotator reference id. Used to identify elements in metadata or to refer to this annotator type

    Definition Classes
    ChunkEntityResolverModel → HasInputAnnotationCols
  99. final val inputCols: StringArrayParam

    Permalink
    Attributes
    protected
    Definition Classes
    HasInputAnnotationCols
  100. final def isDefined(param: Param[_]): Boolean

    Permalink
    Definition Classes
    Params
  101. final def isInstanceOf[T0]: Boolean

    Permalink
    Definition Classes
    Any
  102. final def isSet(param: Param[_]): Boolean

    Permalink
    Definition Classes
    Params
  103. def isTraceEnabled(): Boolean

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  104. val lazyAnnotator: BooleanParam

    Permalink
    Definition Classes
    CanBeLazy
  105. def log: Logger

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  106. def logDebug(msg: ⇒ String, throwable: Throwable): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  107. def logDebug(msg: ⇒ String): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  108. def logError(msg: ⇒ String, throwable: Throwable): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  109. def logError(msg: ⇒ String): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  110. def logInfo(msg: ⇒ String, throwable: Throwable): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  111. def logInfo(msg: ⇒ String): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  112. def logName: String

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  113. def logTrace(msg: ⇒ String, throwable: Throwable): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  114. def logTrace(msg: ⇒ String): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  115. def logWarning(msg: ⇒ String, throwable: Throwable): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  116. def logWarning(msg: ⇒ String): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    Logging
  117. val missAsEmpty: BooleanParam

    Permalink

    whether or not to return an empty annotation on unmatched chunks

    whether or not to return an empty annotation on unmatched chunks

    Definition Classes
    ResolverParams
  118. def msgHelper(schema: StructType): String

    Permalink
    Attributes
    protected
    Definition Classes
    HasInputAnnotationCols
  119. final def ne(arg0: AnyRef): Boolean

    Permalink
    Definition Classes
    AnyRef
  120. val neighbours: IntParam

    Permalink

    number of neighbours to consider in the KNN query to calculate WMD

    number of neighbours to consider in the KNN query to calculate WMD

    Definition Classes
    ResolverParams
  121. final def notify(): Unit

    Permalink
    Definition Classes
    AnyRef
  122. final def notifyAll(): Unit

    Permalink
    Definition Classes
    AnyRef
  123. def onWrite(path: String, spark: SparkSession): Unit

    Permalink
    Attributes
    protected
    Definition Classes
    HasStorageModel → ParamsAndFeaturesWritable
  124. val outputAnnotatorType: AnnotatorType

    Permalink
    Definition Classes
    ChunkEntityResolverModel → HasOutputAnnotatorType
  125. final val outputCol: Param[String]

    Permalink
    Attributes
    protected
    Definition Classes
    HasOutputAnnotationCol
  126. lazy val params: Array[Param[_]]

    Permalink
    Definition Classes
    Params
  127. var parent: Estimator[ChunkEntityResolverModel]

    Permalink
    Definition Classes
    Model
  128. val poolingStrategy: Param[String]

    Permalink

    pooling strategy to aggregate distances: AVERAGE or SUM

    pooling strategy to aggregate distances: AVERAGE or SUM

    Definition Classes
    ResolverParams
  129. var readers: Map[Name, StorageReader[_]]

    Permalink
    Attributes
    protected
    Definition Classes
    HasStorageReader
  130. val returnAllKEmbeddings: BooleanParam

    Permalink
  131. val returnCosineDistances: BooleanParam

    Permalink

    Whether cosine distances should be calculated between a

    Whether cosine distances should be calculated between a

    chunk and the k_candidates result embeddings

  132. def save(path: String): Unit

    Permalink
    Definition Classes
    MLWritable
    Annotations
    @Since( "1.6.0" ) @throws( ... )
  133. def saveStorage(path: String, spark: SparkSession, withinStorage: Boolean): Unit

    Permalink
    Definition Classes
    HasStorageModel
  134. val searchTree: StructFeature[SerializableKDTree[TreeData]]

    Permalink

    Search Tree.

    Search Tree. Under the hood encapsulates SerializableKDTree. Used to perform the search

  135. def serializeStorage(path: String, spark: SparkSession): Unit

    Permalink
    Definition Classes
    HasStorageModel
  136. def set[T](feature: StructFeature[T], value: T): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  137. def set[K, V](feature: MapFeature[K, V], value: Map[K, V]): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  138. def set[T](feature: SetFeature[T], value: Set[T]): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  139. def set[T](feature: ArrayFeature[T], value: Array[T]): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  140. final def set(paramPair: ParamPair[_]): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    Params
  141. final def set(param: String, value: Any): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    Params
  142. final def set[T](param: Param[T], value: T): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    Params
  143. def setAllDistancesMetadata(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  144. def setAlternatives(a: Int): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  145. def setAuxLabelCol(c: String): ChunkEntityResolverModel.this.type

    Permalink
  146. def setAuxLabelMap(m: Map[String, String]): ChunkEntityResolverModel.this.type

    Permalink

    Optional column with one extra label per document.

    Optional column with one extra label per document. This extra label will be outputted later on in an additional column

  147. def setCaseSensitive(value: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    HasCaseSensitiveProperties
  148. def setConfidenceFunction(v: String): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  149. def setDefault[T](feature: StructFeature[T], value: () ⇒ T): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  150. def setDefault[K, V](feature: MapFeature[K, V], value: () ⇒ Map[K, V]): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  151. def setDefault[T](feature: SetFeature[T], value: () ⇒ Set[T]): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  152. def setDefault[T](feature: ArrayFeature[T], value: () ⇒ Array[T]): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    HasFeatures
  153. final def setDefault(paramPairs: ParamPair[_]*): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    Params
  154. final def setDefault[T](param: Param[T], value: T): ChunkEntityResolverModel.this.type

    Permalink
    Attributes
    protected
    Definition Classes
    Params
  155. def setDimension(value: Int): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    HasEmbeddingsProperties
  156. def setDistanceFunction(value: String): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  157. def setDistanceWeights(v: Array[Double]): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  158. def setEnableJaccard(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  159. def setEnableJaroWinkler(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  160. def setEnableLevenshtein(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  161. def setEnableSorensenDice(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  162. def setEnableTfidf(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  163. def setEnableWmd(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  164. def setExtramassPenalty(emp: Double): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  165. def setIncludeStorage(value: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    HasExcludableStorage
  166. final def setInputCols(value: String*): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    HasInputAnnotationCols
  167. final def setInputCols(value: Array[String]): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    HasInputAnnotationCols
  168. def setLazyAnnotator(value: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    CanBeLazy
  169. def setMissAsEmpty(v: Boolean): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  170. def setNeighbours(k: Int): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  171. final def setOutputCol(value: String): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    HasOutputAnnotationCol
  172. def setParent(parent: Estimator[ChunkEntityResolverModel]): ChunkEntityResolverModel

    Permalink
    Definition Classes
    Model
  173. def setPoolingStrategy(value: String): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  174. def setReturnAllKEmbeddings(b: Boolean): ChunkEntityResolverModel.this.type

    Permalink
  175. def setReturnCosineDistances(value: Boolean): ChunkEntityResolverModel.this.type

    Permalink
  176. def setSearchTree(tree: SerializableKDTree[TreeData]): ChunkEntityResolverModel.this.type

    Permalink
  177. def setStorageRef(value: String): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    HasStorageRef
  178. def setTermIDF(freqs: Map[String, (Int, Double)]): ChunkEntityResolverModel.this.type

    Permalink
  179. def setThreshold(dist: Double): ChunkEntityResolverModel.this.type

    Permalink
    Definition Classes
    ResolverParams
  180. def setUseAuxLabel(b: Boolean): ChunkEntityResolverModel.this.type

    Permalink
  181. val storageRef: Param[String]

    Permalink
    Definition Classes
    HasStorageRef
  182. final def synchronized[T0](arg0: ⇒ T0): T0

    Permalink
    Definition Classes
    AnyRef
  183. val termIDF: StructFeature[Map[String, (Int, Double)]]

    Permalink

    Inverted Document Frequency of the term.

    Inverted Document Frequency of the term. Used in the TF-IDF method

  184. val threshold: DoubleParam

    Permalink

    threshold value for the aggregated distance

    threshold value for the aggregated distance

    Definition Classes
    ResolverParams
  185. def toString(): String

    Permalink
    Definition Classes
    Identifiable → AnyRef → Any
  186. final def transform(dataset: Dataset[_]): DataFrame

    Permalink
    Definition Classes
    AnnotatorModel → Transformer
  187. def transform(dataset: Dataset[_], paramMap: ParamMap): DataFrame

    Permalink
    Definition Classes
    Transformer
    Annotations
    @Since( "2.0.0" )
  188. def transform(dataset: Dataset[_], firstParamPair: ParamPair[_], otherParamPairs: ParamPair[_]*): DataFrame

    Permalink
    Definition Classes
    Transformer
    Annotations
    @Since( "2.0.0" ) @varargs()
  189. final def transformSchema(schema: StructType): StructType

    Permalink
    Definition Classes
    RawAnnotator → PipelineStage
  190. def transformSchema(schema: StructType, logging: Boolean): StructType

    Permalink
    Attributes
    protected
    Definition Classes
    PipelineStage
    Annotations
    @DeveloperApi()
  191. val uid: String

    Permalink
    Definition Classes
    ChunkEntityResolverModel → Identifiable
  192. val useAuxLabel: BooleanParam

    Permalink
  193. def validate(schema: StructType): Boolean

    Permalink
    Attributes
    protected
    Definition Classes
    RawAnnotator
  194. def validateStorageRef(dataset: Dataset[_], inputCols: Array[String], annotatorType: String): Unit

    Permalink
    Definition Classes
    HasStorageRef
  195. final def wait(): Unit

    Permalink
    Definition Classes
    AnyRef
    Annotations
    @throws( ... )
  196. final def wait(arg0: Long, arg1: Int): Unit

    Permalink
    Definition Classes
    AnyRef
    Annotations
    @throws( ... )
  197. final def wait(arg0: Long): Unit

    Permalink
    Definition Classes
    AnyRef
    Annotations
    @throws( ... )
  198. def wrapColumnMetadata(col: Column): Column

    Permalink
    Attributes
    protected
    Definition Classes
    RawAnnotator
  199. def wrapEmbeddingsMetadata(col: Column, embeddingsDim: Int, embeddingsRef: Option[String]): Column

    Permalink
    Attributes
    protected
    Definition Classes
    HasEmbeddingsProperties
  200. def wrapSentenceEmbeddingsMetadata(col: Column, embeddingsDim: Int, embeddingsRef: Option[String]): Column

    Permalink
    Attributes
    protected
    Definition Classes
    HasEmbeddingsProperties
  201. def write: MLWriter

    Permalink
    Definition Classes
    ParamsAndFeaturesWritable → DefaultParamsWritable → MLWritable

Inherited from Licensed

Inherited from HasEmbeddingsProperties

Inherited from HasStorageModel

Inherited from HasExcludableStorage

Inherited from HasStorageReader

Inherited from HasCaseSensitiveProperties

Inherited from HasStorageRef

Inherited from ResolverParams

Inherited from AnnotatorModel[ChunkEntityResolverModel]

Inherited from CanBeLazy

Inherited from RawAnnotator[ChunkEntityResolverModel]

Inherited from HasOutputAnnotationCol

Inherited from HasInputAnnotationCols

Inherited from HasOutputAnnotatorType

Inherited from ParamsAndFeaturesWritable

Inherited from HasFeatures

Inherited from DefaultParamsWritable

Inherited from MLWritable

Inherited from Model[ChunkEntityResolverModel]

Inherited from Transformer

Inherited from PipelineStage

Inherited from Logging

Inherited from Params

Inherited from Serializable

Inherited from Serializable

Inherited from Identifiable

Inherited from AnyRef

Inherited from Any

Ungrouped