featuretools.primitives.NumberOfWordsInQuotes#
- class featuretools.primitives.NumberOfWordsInQuotes(quote_type='both')[source]#
确定字符串中引号内单词的数量.
- Description:
给定一个字符串列表,确定每个字符串中引号内单词的数量.
此实现处理Unicode字符.
如果字符串缺失,返回`NaN`.
- Parameters:
quote_type (str, 可选) – 指定要匹配的引号类型.
参数"single"仅匹配单引号 (' ') –
参数"double"匹配双引号 (" ") –
参数"both"匹配任意类型引号内的单词. –
默认为"both". –
Examples
>>> x = ['"python" java prolog "Diffie-Hellman" "4.99"', "Reach me at 'user@email.com'", "'Here's an interesting example!'"]
>>> number_of_words_in_quotes = NumberOfWordsInQuotes() >>> number_of_words_in_quotes(x).tolist() [3, 1, 4]
Methods
__init__([quote_type])flatten_nested_input_types(input_types)将嵌套的列模式输入展平成一个列表.
generate_name(base_feature_names)generate_names(base_feature_names)get_args_string()get_arguments()get_description(input_column_descriptions[, ...])get_filepath(filename)get_function()Attributes
base_ofbase_of_excludecommutativedefault_valueDefault value this feature returns if no data found.
description_templateinput_typeswoodwork.ColumnSchema types of inputs
max_stack_depthnameName of the primitive
number_output_featuresNumber of columns in feature matrix associated with this feature
return_typeColumnSchema type of return
stack_onstack_on_excludestack_on_selfuses_calc_timeuses_full_dataframe