Combining LLMs and Voice Models – Part 1
Combining LLMs and Voice Models – Part 1
This is a guide that I wrote to showcase a new batch inference feature for an OSS framework that I author (nitric.io). I know things like Podcast generation via NotebookLM and also NotebookLlama exist, but wanted to demonstrate a case where an API could be built, and subsequently orchestrated in the cloud. This is just the first part for producing audio using suno/bark via an API. I'm currently working on a part 2 that will introduce an LLM to make scripts from short prompt, which will be piped to the code introduced in Part 1. Looking for feedback on improving this, there are a few things I'd like to clean up but overall am pretty happy with the outputs it produces so far. Thanks in advance for any feedback given.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
GPTCache – Redis for LLMs
prompttest – pytest for LLMs
LoongForge-Train LLMs, VLMs, diffusion and embodied models, faster
I made all LLMs play Chess against each other (inc. o1 models)
Validating models in Django 1.8
Django-spaghetti-and-meatballs – serving up ERDs from django models
Estimation of flop counts for MXNet models
Probe PyTorch Models
Implementation of DDPM (Denoising Diffusion Probabilistic Models)
Computed fields for Backbone.Models