Versions Compared

Key

  • This line was added.
  • This line was removed.
  • Formatting was changed.

...

  • Parameter
    summaryList of characters which should be used to split tokens
    namesplitChars
    • If not present, then tokens are split on any sequence of punctuation. 
  • Parameter
    summaryList of characters which will NOT be used to split tokens.
    namedontSplitChars

    • This is typically used to identify exceptions (characters which are not used to split tokens) when splitChars is missing.
    • These characters are included in the produced tokens.
  • Parameter
    summaryif any character in this list occurs inside a token, that token will be split just before that character
    namesplitBeforeChars
  • Parameter
    summaryif any character in this list occurs inside a token, that token will be split just after that character
    namesplitAfterChars
  • Parameter
    summarytrue/false whether to split on all punctuation (default: true)
    namesplitPrefixChars
  • Parameter
    summarytrue/false whether to split on all punctuation (default: true)
    namesplitSuffixChars
  • Parameter
    summaryThe flag to be put on the vertex between the two tokens.
    namesplitFlag

    • If missing, defaults to ALL_PUNCTUATION.


Saga_config_stagecode
boundaryFlagstext block split
stageCharacterSplitter
requiredFlagstoken
languagejs
"dontSplitChars": ".",
"splitChars":"-",
"splitFlag":"DASH_SPLIT"

Example Output

Saga_config_stagecode
boundaryFlagstext block split
stageCharacterSplitter
requiredFlagstoken
languagejs
"dontSplitChars": "."

Splits on all punctuation, except periods.

For example, the token:  "SagaToolkit-1.0" will produce the following graph:

Code Block
languagetext
Saga_graph
 V-------[SagaToolkit-1.0]-------V
^----[SagaToolkit]--V--[1.0]----^
Saga_config_stage
boundaryFlagstext block split
stageCharacterSplitter
requiredFlagstoken
Code Block
languagetext
Saga_graph
 V-----[Abe-Lincoln]-----V--[likes]--V--[the]--V-----[iPhone-*&@#*&7.0]-----V
^--[Abe]--V--[Lincoln]--^                     ^--[iPhone]--V--[7]--V--[0]--^ 
Saga_config_stagecode
boundaryFlagstext block split
stageCharacterSplitter
requiredFlagstoken
languagejs
titleWith Don't Split Param
"dontSplitChars": "."
Code Block
languagetext
Saga_graph
 V-----[Abe-Lincoln]-----V--[likes]--V--[the]--V--[iPhone-*&@#*&7.0]--V
^--[Abe]--V--[Lincoln]--^                     ^--[iPhone]--V--[7.0]--^
Saga_config_stagecode
boundaryFlagstext block split
stageCharacterSplitter
requiredFlagstoken
languagejs
titleWith Split Chars Param
"splitChars": "-#."
"dontSplitChars": "."
Code Block
languagetext
Saga_graph
 V-----[Abe-Lincoln]-----V--[likes]--V--[the]--V--------[iPhone-*&@#*&7.0]--------V
^--[Abe]--V--[Lincoln]--^                     ^--[iPhone]--V--[*&@]--V--[*&7.0]--^

...