Categories
Rankings
New
简体中文
English
Sign in
Sign up
Submit a tool
Affiliate
For vendors
Categories
Rankings
New
Submit a tool
Affiliate
For vendors
Sign in
Sign up
Search “modal”
84 matching tools
modal
Inworld
AI Companions
Inworld AI is a platform that allows developers to create and integrate AI characters into their applications, games, and virtual worlds with tools for configuring safety, knowledge, memory, narrative control, and multimodality.
inworld.ai
★ 4.0
Jiva.ai
Other
Jiva.ai is a zero-code platform for rapid multimodal AI development, enabling users to create, evaluate, and deploy AI solutions across various data types. It offers user-friendly design assistance and advanced AutoML capabilities for optimal model performance.
jiva.ai
Free
Kiro AI
AI Agents
Kiro AI is an AI-powered IDE that transforms user prompts into specifications and structured tasks to accelerate prototype development and collaboration. It automates repetitive coding tasks with AI agents, supports multimodal input, and integrates with VS Code while ensuring security and privacy.
kiro.dev
Free
★ 2.8
Lakera Guard
Security
Lakera protects generative‑AI and LLM deployments with real‑time threat detection, sub‑50 ms latency, and safeguards against prompt injection, data leakage, and jailbreaks. It offers workforce monitoring, granular policy controls, red‑team vulnerability simulation, and multilingual multimodal support.
lakera.ai
Free
LakeSail
AI Documents
LakeSail is a Rust‑native Spark Connect engine that runs Python workloads at native speed, eliminates JVM overhead, and queries multimodal lakehouse data (PDFs, images, videos, tables) inside your AWS account with zero‑ops elastic compute and built‑in governance.
lakesail.com
Free
LLM Pricing
AI Models & Training
LLM Pricing Comparison lets developers and businesses compare token costs, context lengths, and modalities for major large‑language models. An interactive calculator estimates application expenses based on input/output token volumes, helping teams budget AI workloads accurately.
llm-price.com
Free
★ 5.0
Miniflow.ai
AI Video Generation
Miniflow.ai is a multi-modal AI platform offering 200+ tools for text, image, and video generation with a no-code workflow builder. It simplifies AI integration (GPT-4, Claude) and automation while cutting costs for content creation and data analysis.
miniflow.ai
Free
★ 5.0
Nana Banana Pro
AI Photo Editing
Nana Banana Pro uses multimodal AI to edit and generate consistent character images across poses, scenes, and styles, offering style transfer, high resolution export, batch generation, photo restoration, clothes changes and anime-to-cosplay conversion for fast asset production.
nanabanana.pro
Paid
Neurahub
AI Photo Editing
Neurahub is a multi‑modal AI platform that lets users generate and edit images, videos, and code from text prompts. It also offers real‑time crypto tracking, document creation, and community visual assets for versatile projects.
neurahub.app
★ 5.0
Nexa.ai
Developer Tools & APIs
Nexa AI offers an on‑device platform that lets developers deploy vision, audio, and text models to NPUs, GPUs, and CPUs with one line of code. The SDK supports day‑zero deployment, multimodal inference, and optimizations for mobile, automotive, and IoT devices.
nexa4ai.com
Free
Non finito
AI Models & Training
Non finito is a web‑based platform that lets researchers evaluate and compare multimodal AI models across tasks like entity tracking, reasoning, QA, visual deduction, and card counting. Users input custom prompts, view outputs side‑by‑side, and collaborate in public or private spaces.
nonfinito.xyz
Paid
Ocular AI
AI Models & Training
Ocular AI unifies multimodal data from cloud, local, and external sources into a single catalog for search, versioning, and AI‑assisted labeling with human‑in‑the‑loop. It supports RLHF, GPU training pipelines, RESTful search API, and role‑based compliance controls.
useocular.com
Free
OmniChat
AI Customer Support
Omnichat is a multimodal LLM API that enables autonomous applications by integrating various AI capabilities. It enhances automation, customer service, and workflow management with human-like reasoning for better context comprehension and decision-making.
tryomni.chat
Orga AI
Developer Tools & APIs
Orga AI delivers real‑time multimodal agents that process vision, speech, and text to provide context‑aware responses. Developers embed the API/SDK into workflows for automated support, claim assessment, and high‑volume document processing across chat, voice, and hybrid channels.
orga-ai.com
Outspeed
AI Audio & Voice
Outspeed is a platform for building real-time voice and video AI applications, offering features like speech recognition, NLP, and digital avatars for industries like customer service and education. It provides a flexible SDK for custom multimodal AI solutions, ensuring low-latency processing, compliance, and seamless integration.
outspeed.ai
Free
pangeanic.com
Developer Tools & APIs
Pangeanic is a governed multilingual AI platform that builds trustworthy, private, and compliant data pipelines for text, speech, image, and multimodal content. It offers task‑specific models, RAG, cross‑lingual search, and secure deployment on private clouds.
pangeanic.com
Free
Pongo
Developer Tools & APIs
Moondream AI is an open-source vision-language model with 1.9 billion parameters, enabling efficient image captioning, object detection, and human-like response generation across various platforms through API integration for enhanced multimodal interactions.
joinpongo.com
PopboxGPT
Plugins & Extensions
PopboxGPT is a Chrome extension that embeds GPT‑3.5‑turbo into the browser, offering Alt+J quick launch, fullscreen mode, persistent session history, and a lightweight modal for drafting, research, or brainstorming across sites for all users.
popboxgpt.com
Paid
Ray2
AI Video Generation
Ray 2 is an AI video generation tool that creates high-quality visuals from text prompts and multi-modal inputs. It offers seamless motion, high resolution up to 1080p, and fast processing for efficient video production.
ray2.ai
Paid
Resemble
AI Audio & Voice
Resemble AI is a generative‑AI platform that delivers real‑time text‑to‑speech, speech‑to‑speech, and voice‑design in 60+ languages. It embeds invisible watermarks, provides multimodal deep‑fake detection across 160 models, and offers on‑prem or cloud APIs for developers and enterprises.
resemble.ai
Paid
★ 3.9
Sam Audio
AI Audio & Voice
SAM Audio uses Meta’s Segment Anything Audio Model to isolate vocals, instruments, speech and effects from mixes via multimodal prompts (text, visual, time-span). It produces target and residual stems at original sample rates for production, post, and research.
samaudio.audio
Free
Scriptaa
Other
Scriptaa is a multimodal generative AI platform that enables content creation in text, images, and audio while supporting multilingual output. It features pre-built templates and a retrieval-augmented generation framework, ensuring high-quality content tailored to brand voices.
scriptaa.io
Paid
Seedance
AI Video Generation
Seedance AI is a multimodal platform for image, video and audio generation supporting text-to-image, image-to-image, text-to-video and image-to-video workflows, with built-in video editing/enhancement, broad model selection, batch generation and team collaboration.
seedance.ai
Paid
Segwise.ai
Analytics & BI
Segwise consolidates creative data from ad networks, DSPs, and internal sources via no‑code integrations, uses multimodal AI to tag creative elements, maps tags to performance metrics, and delivers dashboards, fatigue alerts, and automated iterations for data‑driven optimization.
segwise.ai
Previous
1
2
3
4
Next
ℹ️
Snapshot mode
· This page is served from a database snapshot exported on 2026-09-16, not live data. This notice disappears once the live API is connected.