A

A CLI to query the unsealed court files with local LLMs

Hacker News

A CLI to query the unsealed court files with local LLMs

To succeed on Hacker News (HN), you have to completely drop the "marketing" and "YouTube hook" tone. The HN community heavily downvotes clickbait, sensationalism, and marketing fluff. They love "Show HN" posts, open-source projects, CLI tools, local LLMs, and clever technical solutions to messy data problems (like parsing poorly scanned government PDFs). Here are the best titles and the exact description (to use either as a text post or your first comment) tailored specifically for the Hacker News audience. The Hacker News Titles Choose one of these. On HN, titles should be strictly factual, descriptive, and avoid emojis. Option 1 (The Classic HN Format - Recommended): Show HN: epstein-search – A CLI to query the unsealed court files with local LLMs Option 2 (Focus on the tech pipeline): Show HN: I built a local RAG CLI to make the Epstein PDFs searchable Option 3 (Straight to the point): Show HN: epstein-search – Query the Epstein document dumps offline via CLI The Hacker News Description (First Comment or Text Body) If you submit the GitHub URL directly, immediately post this as the first comment. If you submit a text post, put this in the body. Keep the tone humble, technical, and open to feedback. Hi HN, When the Epstein court documents and flight logs were unsealed, they were released the way most legal drops are: thousands of pages of messy, poorly scanned, unsearchable PDFs. Standard Ctrl+F doesn't work well due to OCR errors, and the sheer volume makes manual parsing a nightmare. To solve this, I built epstein-search, an open-source Python CLI tool that lets you search and synthesize the documents using a Retrieval-Augmented Generation (RAG) pipeline directly in your terminal. How it works: It parses and chunks the original unsealed PDF files. You can run queries against the dataset using API-based models (OpenAI/Anthropic) if you want speed. Privacy-first: If you don't want your queries logged by a third-party API, you can point it directly to a local model (via Ollama or Llama.cpp) to run the entire search and retrieval process 100% offline. The goal was to make this data accessible to researchers and OSINT investigators without requiring them to manually read thousands of pages of court dockets or hand over their search queries to OpenAI. Repo is here: https://github.com/simulationship/epstein-search

Share card

Actual performance

2points
Did not reach leaderboard

Launch Intel predictions

Analyze your own launch →
Product HuntOn track for Day 1 leaderboard · Strong signals: model, dock, new · Missing: mac, agents, macos
89%89% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
Indie HackersFits the IH revenue-focused audience · Missing: supports, reddit linkedin, podcasting
88%88% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
Hacker NewsStrong engagement from HN community · Strong signals: llama, hacker news, pipe · Missing: https docs, excited, just released
66%66% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
AppSumoMay struggle as an AppSumo deal · Missing: plus, platform, intuitive
41%41% predicted probability of success on AppSumo, based on ML models trained on real launch data.
TrustMRRLess likely to generate early MRR · Strong signals: way · Missing: mobile apps, ios, personal
33%33% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
13%13% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Incorrect prediction on native model

Similar products

A Court of Thorns and Roses 6 by Sarah J. Maas
A Court of Thorns and Roses 6 by Sarah J. Maas31%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

A Court of Thorns and Roses 6 by Sarah J. Maas

Indie Hackerscommitment-full-time
Bi
Bing Webmaster CLI for Agents and LLMs17%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Bing Webmaster CLI for Agents and LLMs

Hacker News1
To
Tombl – Easily query .toml files from bash44%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Tombl – Easily query .toml files from bash

Hacker News45
I
I scraped local court records to find dirty cops61%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

I scraped local court records to find dirty cops

Hacker News114
fe
fenic – LLMs as dataframe operators, query meaning and structure50%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

fenic – LLMs as dataframe operators, query meaning and structure

Hacker News3
Us
Use local LLMs to organize your files40%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Use local LLMs to organize your files

Hacker News6
Mo
Monglorious – Query MongoDB with Strings33%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Monglorious – Query MongoDB with Strings

Hacker News2
Qu
Query the national alcoholic beverage retailing monopoly of Finland43%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Query the national alcoholic beverage retailing monopoly of Finland

Hacker News2
VelocityPack QueryBoost
VelocityPack QueryBoost38%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

MySQL / MariaDB Query Caching and Acceleration for cPanel

Indie Hackersemployees-10-plus
Qu
Query small data with columnq CLI40%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Query small data with columnq CLI

Hacker News2