Enfer.ai – Cheap LLM Inference Service
Enfer.ai – Cheap LLM Inference Service
Hi HN! I'm excited to share enfer.ai - a low-cost LLM inference service for those who need LLM access without the high price tag. We currently serve Mistral-Nemo and its finetunes at $0.03 per mil. input tokens and $0.07 per mil. output tokens. I’d really appreciate your feedback on our approach, performance, or anything else you’d like to see. As a thank you for checking us out, here's a small voucher HN5 for $5 worth of credits.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
LLM Inference Requirements Profiler
I built a tool for renting cheap GPUs for inference
Speeding up LLM inference 2x times (possibly)
Open-source AMDGCN kernels for optimizing LLM inference
Onera – Private LLM Inference Inside AMD SEV-SNP Enclaves
Cheap IPTV Service
MonkeyPatch – Cheap, fast and predictable LLM functions in Python
BonzAI – self-sovereign, local LLM inference in the browser
NightRun, bare metal LLM inference, no OS, boots from USB
YPerf – Monitor LLM Inference API Performance