...

Function Name Description Parameters Examples

Count

Outputs the number of steam elements.

Name	Description	Default Value	Optional?
`OUTPUT_ATTRIBUTES`	The name for the output attribute.	`count`	True

['FUNCTION' = 'Count']

['FUNCTION' = 'Count', 'OUTPUT_ATTRIBUTES' = 'number_of_elements']

Sum

Outputs the sum of elements.

Name	Description	Default Value	Optional?
INPUT_ATTRIBUTES	The single string or a list of the name(s) of the attribute(s) in the input tuples. By default, all input attributes are used. This could raise an error if attributes are not numeric.	(all attributes)	True
OUTPUT_ATTRIBUTES	A single string or list of output attributes. By default, the string "Sum_" concatenated with the original input attribute name is used.	"Sum_" + intput attribute name	True

['FUNCTION' = 'Sum']

['FUNCTION' = 'Sum', 'INPUT_ATTRIBUTES' = 'value1']

['FUNCTION' = 'Sum', 'INPUT_ATTRIBUTES' = ['value1', 'value2']]

Avg Average value (mean) TODO

Min Min value TODO

Max Max value TODO

Trigger The tuple that triggers the output. TODO

Variance Calculates the variance TODO

TopK Calculates the top-K list TODO

Nest Nests the valid elements as list. TODO

Examples

Code Block

language	js
linenumbers	true

counted = AGGREGATION({AGGREGATIONS = [['FUNCTION' = 'Count']], GROUP_BY = ['publisher', 'item']}, windowed)

...

Code Block

language	js
linenumbers	true

counted = AGGREGATION({AGGREGATIONS = [['FUNCTION' = 'Count'], ['FUNCTION' = 'Sum', 'INPUT_ATTRIBUTES' = 'value1']], GROUP_BY = ['publisher', 'item']}, windowed)

Code Block

language	js
linenumbers	true

/// count the number of items for each publisher
counted = AGGREGATION({AGGREGATIONS = [['FUNCTION' = 'Count']], GROUP_BY = ['publisher', 'item']}, windowed)
/// aggregate the 100 most frequent items for each publisher to an ordered list
TopKItemsByPublisher ::= AGGREGATION({AGGREGATIONS = [
	[
		'FUNCTION' = 'TopK',
		'TOP_K' = '100',                         /// number of items
		'SCORING_ATTRIBUTES' = 'Count',          /// the attribute name that defines the order
		'INPUT_ATTRIBUTES' = 'item',             /// do not use the whole input tuple, just use the 'item' attribute for creating the output top-k set
		'MIN_SCORE' = '0',                       /// remove items that reaches a score of 0 (due to the previous aggregation these are all items that has no valid tuple)
		'UNIQUE_ATTR'='item'                     /// use 'item' as a unique attribute. that means, a new tuple with an known items id replaces the previous value. (this is some kind of element window in this operator)
	]], GROUP_BY = ['publisher']}, counted)

Changing the way this operator outputs values

...

Space shortcuts

Page tree

Versions Compared

Old Version 1

New Version 2

Key

Examples

Changing the way this operator outputs values

Space shortcuts

Page tree

Page History

Versions Compared

Old Version 1

New Version 2

Key

Examples

Changing the way this operator outputs values