Versions Compared

Key

  • This line was added.
  • This line was removed.
  • Formatting was changed.

...

Include Page
Generic Configuration Parameters
Generic Configuration Parameters

Configuration Parameters

  • splitChars (string, optional) -
    Parameter
    summary
    List of characters which should be used to split tokens
    .
    namesplitChars
    • If not present, then tokens are split on any sequence of punctuation. 
  • dontSplitChars (string, optional) -
    Parameter
    summary
    List of characters which will NOT be used to split tokens.
    namedontSplitChars

    • This is typically used to identify exceptions (characters which are not used to split tokens) when splitChars is missing.
    • These characters are included in the produced tokens.
  • Parameter
    summaryif any character in this list occurs inside a token, * that token will be split just before that character
    namesplitBeforeChars
  • Parameter
    summaryif any character in this list occurs inside a token, * that token will be split just after that character
    namesplitAfterChars
  • Parameter
    summarytrue/false whether to split on all punctuation (default: true)
    namesplitPrefixChars
  • Parameter
    summarytrue/false whether to split on all punctuation (default: true)
    namesplitSuffixChars
  • splitFlag (string, optional) -
    Parameter
    summary
    The flag to be put on the vertex between the two tokens.
    namesplitFlag

    • If missing, defaults to ALL_PUNCTUATION.

...