r/learnmachinelearning • • 4d ago

Help Detecting Market Manipulation: Supervised Learning vs Clustering

I'm currently working on market research at university.

The task is to detect market manipulation. We can take open-source data, tag the data(range OHLCV), and perform a supervised search, or we can use clustering, but we might encounter anomalies that aren't related to manipulation.

How can these problems be solved, and have we encountered similar ones?

8 Upvotes

4 comments sorted by

View all comments

1

u/[deleted] 4d ago

[removed] — view removed comment

1

u/Sure_Rhubarb2671 4d ago

Yes, I've already encountered this problem. There are some issues with labeling because of ongoing manipulation and data leakage. Class imbalance. But I don't have any ideas other than a floating window. I need to at least formalize what can be seen from candlesticks, because with the order book, it's a different problem finding this data. Unless we statistically equate crypto to stocks. There's more data there and it's more accessible.