HN user

kiv6

3 karma

Software developer in New Brunswick, Canada. Get in touch if you have interesting work :)

Posts0
Comments1
View on HN
No posts found.

Thanks for publishing!

I'd love to hear more details about how this is used in production. For example, if you have an anomaly that occurs only twice in a long dataset, the two anomalies would match each other and would have a low matrix profile value and be considered equally normal as a pattern that recurs thousands of times, correct?

I would normally think of an anomaly as a point in a low density region of space, but Matrix Profile seems to have a more strict definition as a point that has a large distance to its nearest neighbour - is that fair?

I'm also interested in your process for setting the parameter of the subquery length. Do you have to already know something about the expected length of an anomaly/motif, or do you sweep over multiple values?

How does this tie into alerting? Do you set a threshold on the matrix profile value that would fire an automated alert? Or is this used more as an offline tool to explore the dataset?

Minor nitpick: on the Target blog post, Prof. Keogh's name is spelled wrong (as Keough)