Vi

Visually manipulate and clean large datasets; local and remote

Hacker News

Visually manipulate and clean large datasets; local and remote

Hi there! I'm working on an application which allows data scientists, analysts (and really whoever) to visually manipulate & clean datasets of any size, without writing code. This stems out of my own personal experience over the past 6 years of maximal frustration trying to do Data Science and being forced to do lots of engineering to get anything done. (BTW I am a software engineer first, and yes, this still annoys me). Lots of time (up to 80%) of data scientist's time can be attributed to working with data manually – think cleaning, validation, merging, visualization, data movement, dependency installation / configuration etc. The tools are also super brittle meaning that everything is ad-hoc AND unstable. Uff. Coco Alemana is my attempt to build an IDE which abstracts away the engineering layer from Data Science – as 85%+ of data scientists come from hard science instead of software engineering. You're able to load massive datasets from formats like Parquet, CSV, and JsonLines – or remotely via Amazon Athena (and more to come soon). You can then move data around like Excel & Figma had a forbidden child. Joins can be done just by dragging one column into another frame, etc. We have a bunch of cool stuff like auto-warning identification, union consolidation, easy column value mergers and renames, re-ordering, sorting, filtering, group by and the list goes on. You can download the application and load a massive file within less than 3 minutes. It's free for everyone here. I'll keep an eye out for emails and will extend trials accordingly. I'd love to hear what you all think :) Happy to receive any type of feedback, or roast ;)

Share card

Actual performance

1points
Did not reach leaderboard

Launch Intel predictions

Analyze your own launch →
Indie HackersFits the IH revenue-focused audience · Missing: supports, reddit linkedin, podcasting
67%67% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
Hacker NewsStrong engagement from HN community · Strong signals: ide, io · Missing: https docs, excited, just released
63%63% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
TrustMRRLess likely to generate early MRR · Strong signals: personal, way · Missing: mobile apps, ios, entrepreneurs
42%42% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
Product HuntUnlikely to reach the leaderboard · Strong signals: email, visual, code · Missing: mac, agents, macos
41%41% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
AppSumoMay struggle as an AppSumo deal · Strong signals: soon · Missing: plus, platform, intuitive
35%35% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
15%15% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Incorrect prediction on native model

Similar products

Ti
Tidepool – analytics for large text datasets71%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Tidepool – analytics for large text datasets

Hacker News5
Chaos
Chaos25%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Create dirty datasets out of clean datasets

Indie Hackers1analytics
Ca
Call remote functions as if they where local (ws-remote-fn)47%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Call remote functions as if they where local (ws-remote-fn)

Hacker News1
Ma
Matrices – Explore, visualize, and share large datasets57%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Matrices – Explore, visualize, and share large datasets

Hacker News8
Ca
Cached Datasets53%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Cached Datasets

Hacker News4
Su
Sustinion - the large opinion collider (in spe)55%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Sustinion - the large opinion collider (in spe)

Hacker News1
Chaos
Chaos40%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Make clean datasets dirty

Product Hunt+4
Ma
Markov chains explained visually75%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Markov chains explained visually

Hacker News1,070
Ex
Explained Visually79%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Explained Visually

Hacker News353
Vi
Visually mine π68%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Visually mine π

Hacker News5