LogoAISecKit
  • Search
  • Collection
  • Category
  • Tag
  • Blog
  • Pricing
  • Submit
LogoAISecKit

Newsletter

Join the Community

Subscribe to our newsletter for the latest news and updates

LogoAISecKit

Curated directory of 1700+ AI tools, models, frameworks, MCP servers, and cybersecurity resources

GitHub
Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
  • Pricing
  • Submit
Company
  • About Us
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
Sponsored Resources
  1. Home
  2. Category
  3. Make-An-Audio
icon of Make-An-Audio

Make-An-Audio

PyTorch implementation of a generative model for high-fidelity audio generation from text prompts.

Visit Website
image for Make-An-Audio
Visit Website

Introduction

Make-An-Audio

Make-An-Audio is a PyTorch implementation of a conditional diffusion probabilistic model designed to generate high-fidelity audio from text prompts. This repository provides an open-source implementation along with pretrained models, enabling users to create audio samples efficiently.

Key Features:
  • Text-to-Audio Generation: Generate audio samples from textual descriptions using advanced diffusion models.
  • Pretrained Models: Access pretrained models to quickly start generating audio without extensive training.
  • Flexible Training: Users can train the model on their own datasets with provided scripts and guidelines.
  • Evaluation Metrics: Includes tools for evaluating generated audio quality using metrics like FD, FAD, IS, and KL.
Benefits:
  • High Fidelity: Produces high-quality audio outputs that are suitable for various applications.
  • Open Source: Freely available for research and development, promoting collaboration and innovation in the field of audio generation.
  • Community Support: Engage with a community of developers and researchers through GitHub for feedback and improvements.
Highlights:
  • Supports various audio generation tasks including audio inpainting and audio-to-audio transformations.
  • Comprehensive documentation and examples to help users get started quickly.
  • Acknowledges contributions from other significant projects in the field, enhancing its reliability and performance.
Back

Information

  • Publisher
    AISecKit
  • Websitegithub.com
  • Published date2025/04/28

Categories

  • AI Models
  • AI Application Platforms
  • AI Audio Tools

Tags

  • Open Source
  • Text-to-Audio
  • Multimodal AI
  • Generative AI

More Products

image of Nano Bananary
AI ModelsAI Application PlatformsAI Video Tools
Visit Website
icon of Nano Bananary

Nano Bananary

Nano Bananary is an AI batch image and video generator with 142 effects.

Text-to-VideoGenerative AI
image of Twocast
AI Application PlatformsAI Productivity ToolsAI Audio Tools
Visit Website
icon of Twocast

Twocast

AI Podcast Generator for bilingual episodes, supporting multiple languages and alternative to NotebookLLM.

Content Creation
image of ZCF
AI Application PlatformsAI Productivity ToolsAI Development Frameworks
Visit Website
icon of ZCF

ZCF

Zero-Config Code Flow for Claude code & Codex, enabling seamless integration and configuration for AI development.

Open SourceClaude