27 tools found
Stable Audio is Stability AI’s generative-audio product for creating music and sound effects from text prompts. It offers web-based generation and is built around Stability AI’s audio models, including Stable Audio 2.0.
Kits AI is a browser-based music production platform for AI voice conversion, singing voice generation, vocal removal, stem splitting, mastering, and licensed artist voice models.
Bearly AI is a desktop AI assistant and browser companion for reading, writing, summarizing, research, document analysis, and chat across multiple language models, with apps and browser extensions.
HARPA AI is a Chrome-based browser automation and AI assistant that can summarize and extract webpage content, monitor pages for changes, draft text, and run reusable commands with supported AI models.
Duck.ai is DuckDuckGo’s private AI chat service. It provides access to selected third-party language models while saying chats are anonymized and are not used to train those models.
Brave Leo is an AI assistant integrated into the Brave browser. It can summarize webpages and documents, answer questions, generate text, and use multiple language models, with privacy-focused request handling.
CAMEL-AI is an open-source framework and research community for studying and building communicative multi-agent systems. Its Python library provides agent abstractions, role-playing workflows, tools, memory, models, and task-oriented agent societies.
Skyvern is an open-source platform for automating browser-based workflows with large language models and computer vision. It can navigate websites, extract information, and complete multistep tasks through APIs or a hosted service.
Genspark Super Agent is a general-purpose AI agent from MainFunc that can research, make calls, generate documents and presentations, and complete multistep tasks using integrated models and tools.
SWE-agent is an open-source system from Princeton University's NLP group that lets language models inspect repositories, edit code, and run commands to resolve GitHub issues through a purpose-built agent-computer interface.
Continue is an open-source platform for building AI coding assistants. Its IDE extensions support chat, code completion, editing, and agent workflows, with configurable language models and development context.
Aider is an open-source, terminal-based AI pair-programming tool that edits files in local Git repositories. It connects to hosted or local language models and can automatically commit code changes.
Goose is an open-source AI agent from Block that runs locally, connects language models to developer tools through extensions, and can plan and execute multi-step software engineering and automation tasks.
A macOS transcription application from Good Snooze that converts audio files and microphone recordings into text. It supports local Whisper models, speaker recognition, batch transcription, and export formats including subtitles.
Flowise is an open-source, visual development platform for building AI workflows and agents with drag-and-drop components. It supports integrations with language models, vector databases, tools, APIs, and deployment endpoints.
Gumloop is a visual, no-code platform for building AI-powered business automations. Users connect data sources, language models, and software services in workflows that can scrape, classify, generate, and route information.
n8n is a source-available workflow automation platform for connecting apps, APIs, databases, and AI models. It supports visual workflow design, custom code, self-hosting, and a managed cloud service.
Teachable Machine is a free browser-based Google tool for training simple image, audio, and pose classification models without writing code. Users collect examples, train locally, test results, and export models for projects.
Obviously AI is a no-code predictive analytics platform for building machine-learning models from tabular business data. It supports prediction, model explanation, deployment, and integration workflows.
Hailuo AI is a generative-video platform from MiniMax that creates short videos from text prompts and still images. Its web product provides access to MiniMax video models, including the Video-01 family.
Google DeepMind’s family of generative video models creates and edits video from text, image, and video prompts. Newer versions add native audio generation and are available through Google products and Vertex AI.
Udio launched in April 2024 from ex-DeepMind researchers and raised $10M from a16z, will.i.am, and Common. It generates studio-quality vocal tracks from text prompts and quickly became Suno's closest competitor. The RIAA sued Udio in June 2024 over alleged training on copyrighted recordings. Despite the litigation, Udio has continued shipping new models, including v1.5 with full-song generation and remixing.
The GitHub of machine learning — hosting models, datasets, and AI applications.
Enterprise AI platform providing NLP models for text generation, classification, and search.
API provider for Claude models, focused on AI safety research.
API platform providing access to GPT models for developers to build AI applications.
Company behind Stable Diffusion, providing open-source generative AI models.