Da

Data-generation/crowdsourcing platform, for building open datasets

Hacker News

Data-generation/crowdsourcing platform, for building open datasets

-- Background -- https://metro.exchange Metro allows data science projects to be powered by a crowd of people who self-generate the data for it. One of its primary uses is to create open datasets collaboratively, where every contributor is able to access all of the data. I've been building it for the past few months and I want to gather some feedback! We can build a useful dataset of translations, and then we can start making new DataSources to power new datasets (labeled images, named entity recog.). -- How it works -- Data generation happens on your computer, using "DataSources". A DataSource is a community-made, open-source plugin for Metro, which generates data for you. You simply install the Metro browser extension and activate the DataSources which power the project. You'll also need to signup, which doesn't require email verification right now so it takes about 10 seconds. -- Sentence Translation Project -- I made an Open Data project for gathering sentence-level translations in 7 languages, and I would you to try it out! https://metro.exchange/projects/od-sentence-translations/ It's powered by a DataSource ( https://metro.exchange/datasources/text-translation/ ) which allows you to highlight any text, right-click, press a "translate" button, and enter your translation. You'll need to 1. sign up ( https://metro.exchange/signup ), 2. install the extension, and 3. activate the DataSource on the project page ( https://metro.exchange/projects/od-sentence-translations/ ). -- Future -- I want Metro to be able to support open-data generation of any scale and eventually be the backbone for startups powered by ethical, self-generated data because it provides access to data from any platform on the internet while giving users true autonomy over their data. Any feedback, help, or just usage of the system is really useful for trying to improve the problems that I just can't see yet. Thank you!

Share card

Actual performance

3points
Did not reach leaderboard

Launch Intel predictions

Analyze your own launch →
Indie HackersFits the IH revenue-focused audience · Missing: supports, reddit linkedin, podcasting
72%72% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
Product HuntOn track for Day 1 leaderboard · Strong signals: user, computer, new · Missing: mac, agents, macos
70%70% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
Hacker NewsStrong engagement from HN community · Strong signals: ide, io · Missing: https docs, excited, just released
57%57% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
TrustMRRLess likely to generate early MRR · Strong signals: month, users · Missing: mobile apps, ios, personal
48%48% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
AppSumoMay struggle as an AppSumo deal · Strong signals: platform, users · Missing: plus, intuitive, reviews
45%45% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
12%12% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Incorrect prediction on native model

Similar products

Ca
Cached Datasets53%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Cached Datasets

Hacker News4
A
A list of annotation tools for building datasets59%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

A list of annotation tools for building datasets

Hacker News5
Op
Open Financial Datasets55%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Open Financial Datasets

Hacker News5
Bizooy
Bizooy25%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Reputation building platform

Indie Hackerscommitment-full-time
Li
Listen to over 7000 open datasets49%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Listen to over 7000 open datasets

Hacker News5
WuaBot
WuaBot25%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Building the next generation chat bot platform

Indie Hackers
I
I made this tool for navigating pandas datasets50%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

I made this tool for navigating pandas datasets

Hacker News20
De
Dexter, an open platform for building, sharing, and deploying your hack52%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Dexter, an open platform for building, sharing, and deploying your hack

Hacker News62
Da
Data wrangling – importing 300 datasets a quarter68%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Data wrangling – importing 300 datasets a quarter

Hacker News1
Ge
Geckoboard Datasets API58%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Geckoboard Datasets API

Hacker News1