featuretools.primitives.MedianWordLength#
- class featuretools.primitives.MedianWordLength(delimiters_regex='[ \n\t]')[source]#
确定中位单词长度.
- Description:
给定字符串列表,确定每个字符串中单词长度的中位数.单词定义为 由分隔符分隔的任意字符序列.如果字符串为空或`NaN`,则返回`NaN`.
- Parameters:
delimiters_regex (str) – 用于将文本分割成单词的分隔符正则表达式字符串. 默认为空白字符.
Examples
>>> x = ['This is a test file', 'This is second line', 'third line $1,000', None] >>> median_word_length = MedianWordLength() >>> median_word_length(x).tolist() [4.0, 4.0, 5.0, nan]
Methods
__init__([delimiters_regex])flatten_nested_input_types(input_types)将嵌套的列模式输入展平成一个列表.
generate_name(base_feature_names)generate_names(base_feature_names)get_args_string()get_arguments()get_description(input_column_descriptions[, ...])get_filepath(filename)get_function()Attributes
base_ofbase_of_excludecommutativedefault_valueDefault value this feature returns if no data found.
description_templateinput_typeswoodwork.ColumnSchema types of inputs
max_stack_depthnameName of the primitive
number_output_featuresNumber of columns in feature matrix associated with this feature
return_typeColumnSchema type of return
stack_onstack_on_excludestack_on_selfuses_calc_timeuses_full_dataframe