Comment by heystefan

10 months ago

Could one usecase be generating an audiobook with this from existing books? I wonder if I could fine-tune the "characters" that speak these lines since you said it's a single pass whole the whole convo. Wonder if that's a limitation for this kind of a usecase (where speed is not imperative).

2 comments

heystefan

toebee 10 months ago

Yes! But you would need to put together a LLM system that created scripts from the book content. There is an open source project called OpenNotebookLM (https://github.com/gabrielchua/open-notebooklm) that does something similar. If you hook the Dia model to that kind of system, it will be very possible :) Thanks for the interest!

satvikpendem 10 months ago

Another project, specifically for creating audiobooks: https://github.com/prakharsr/audiobook-creator