Hello everyone!
We have a substantial queue of companies waiting for their data to be published, but nothing has been released yet. The reason is simple: the datasets we process often exceed tens of terabytes. We know publishing this kind of raw, chaotic "data mush" makes no sense — it's impossible to download or use effectively.
While we acknowledge that some groups opt to publish raw data