featuretools.primitives.TotalWordLength#
- class featuretools.primitives.TotalWordLength(do_not_count='[!"#$%&\'()*+,-./:;<=>?@[\\]^_`{|}~\n\t ]')[source]#
确定总字长.
- Description:
给定字符串列表,确定每个字符串中的总字长. 单词定义为一系列不以分隔符分隔的任意字符. 如果字符串为空或`NaN`,则返回`NaN`.
- Parameters:
delimiters_regex (str) – 用于将文本分割成单词的分隔符正则表达式字符串. 默认为空白字符.
Examples
>>> x = ['This is a test file', 'This is second line', 'third line $1,000', None] >>> total_word_length = TotalWordLength() >>> total_word_length(x).tolist() [15.0, 16.0, 13.0, nan]
Methods
__init__([do_not_count])flatten_nested_input_types(input_types)将嵌套的列模式输入展平成一个列表.
generate_name(base_feature_names)generate_names(base_feature_names)get_args_string()get_arguments()get_description(input_column_descriptions[, ...])get_filepath(filename)get_function()Attributes
base_ofbase_of_excludecommutativedefault_valueDefault value this feature returns if no data found.
description_templateinput_typeswoodwork.ColumnSchema types of inputs
max_stack_depthnameName of the primitive
number_output_featuresNumber of columns in feature matrix associated with this feature
return_typeColumnSchema type of return
stack_onstack_on_excludestack_on_selfuses_calc_timeuses_full_dataframe