Skip to content

The AI Podcast Studio: generate podcasts scripts and their audio version with a team of AI workers in a Podcast Studio πŸŽ™οΈπŸ“œ

License

Notifications You must be signed in to change notification settings

leopiney/neuralnoise

Folders and files

NameName
Last commit message
Last commit date

Latest commit

Β 

History

34 Commits
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 
Β 

Repository files navigation

NeuralNoise: The AI Podcast Studio

PyPI - Downloads

NeuralNoise banner

NeuralNoise is an AI-powered podcast studio that uses multiple AI agents working together. These agents collaborate to analyze content, write scripts, and generate audio, creating high-quality podcast content with minimal human input. The team generates a script that the cast team (using a TTS tool of your choice) will then record.

Features

leopiney/neuralnoise GithubStars history


Examples

Source Type NeuralNoise Download
TikTok owner sacks intern for sabotaging AI project 🌐 Web article
example_neuralnoise_0.mp4
Link
Before you buy a domain name, first check to see if it's haunted 🌐 Web article
example_neuralnoise_1.mp4
Link
Linus Torvalds Comments On The Russian Linux Maintainers Being Delisted 🌐 Web article
example_neuralnoise_2.mp4
Link
Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation πŸ“— PDF
example_neuralnoise_3.mp4
Link
Ep17. Welcome Jensen Huang | BG2 w/ Bill Gurley & Brad Gerstner πŸ“Ί YouTube
example_neuralnoise_4.mp4
Link
Notepad++ turns 21, Apple releases M4, OpenAI Search release 🌐 Multiple web articles
example_neuralnoise_5.mp4
Link

Objective

The main objective of NeuralNoise is to create a Python package that simplifies the process of generating AI podcasts. It utilizes OpenAI for content analysis and script generation, ElevenLabs for high-quality text-to-speech conversion, and Streamlit for an intuitive user interface.

Installation

To install NeuralNoise, follow these steps:

  1. Install the package:

    pip install neuralnoise
    

    or from source:

    git clone https://github.com/leopiney/neuralnoise.git
    cd neuralnoise
    pip install .
    
  2. Set up your API keys:

    • Create a .env file in the project root

    • Add your OpenAI and ElevenLabs API keys:

      OPENAI_API_KEY=your_openai_api_key
      
      # Optional
      ELEVENLABS_API_KEY=your_elevenlabs_api_key
      

Usage

To run the NeuralNoise application first make sure that you create a configuration file you want to use. There are examples in the config folder.

Then you can run the application with:

nn generate --name <name> <url|file> [<url|file>...]

Want to edit the generated script?

The generated script and audio segments are saved in the output/<name> folder. To edit the script:

  1. Locate the JSON file in this folder containing all script segments and their text content.
  2. Make your desired changes to specific segments in the JSON file. Locate the "sections" and "segments" content in this file that you want to change, then feel free to edit the content of the segments you want to change.
  3. Run the same command as before with the same name (nn generate --name <name>) to regenerate the podcast.

The application will regenerate the podcast, preserving unmodified segments and only processing the changed ones. This approach allows for efficient editing without regenerating the entire podcast from scratch.

Roadmap

  • Better PDF and articles content extraction.
  • Add interactive ways of using NeuralNoise (Gradio/Colab/etc)
  • Add local LLM provider. More generic LLM configuration. Leverage AutoGen for this.
  • Add local TTS provider
  • Add podcast generation format options: interview, narrative, etc.
  • Add podcast generation from multiple source files
  • Add more agent roles to the studio. For example, a "Content Curator" or "Content Researcher" that uses tools to find and curate content before being analyzed. Or a "Sponsor" agent that adds segways to ads in the podcast script (Γ  la LTT).
  • Add music and sound effects options
  • Real-time podcast generation with human and AI collaboration (πŸ€”)

Contributing

Contributions to NeuralNoise are welcome! Please feel free to submit a Pull Request.

License

This project is licensed under the MIT License - see the LICENSE file for details.

Related projects

About

The AI Podcast Studio: generate podcasts scripts and their audio version with a team of AI workers in a Podcast Studio πŸŽ™οΈπŸ“œ

Topics

Resources

License

Stars

Watchers

Forks

Packages

No packages published

Languages