Search “Multimodal AI model”

300 matching tools · Your query was split into the terms below

multimodalmodel
Non finitoAI Models & TrainingNon finito is a web‑based platform that lets researchers evaluate and compare multimodal AI models across tasks like entity tracking, reasoning, QA, visual deduction, and card counting. Users input custom prompts, view outputs side‑by‑side, and collaborate in public or private spaces.nonfinito.xyzOcular AIAI Models & TrainingOcular AI unifies multimodal data from cloud, local, and external sources into a single catalog for search, versioning, and AI‑assisted labeling with human‑in‑the‑loop. It supports RLHF, GPU training pipelines, RESTful search API, and role‑based compliance controls.useocular.comFreepangeanic.comDeveloper Tools & APIsPangeanic is a governed multilingual AI platform that builds trustworthy, private, and compliant data pipelines for text, speech, image, and multimodal content. It offers task‑specific models, RAG, cross‑lingual search, and secure deployment on private clouds.pangeanic.comFreePongoDeveloper Tools & APIsMoondream AI is an open-source vision-language model with 1.9 billion parameters, enabling efficient image captioning, object detection, and human-like response generation across various platforms through API integration for enhanced multimodal interactions.joinpongo.comResembleAI Audio & VoiceResemble AI is a generative‑AI platform that delivers real‑time text‑to‑speech, speech‑to‑speech, and voice‑design in 60+ languages. It embeds invisible watermarks, provides multimodal deep‑fake detection across 160 models, and offers on‑prem or cloud APIs for developers and enterprises.resemble.ai3.9Sam AudioAI Audio & VoiceSAM Audio uses Meta’s Segment Anything Audio Model to isolate vocals, instruments, speech and effects from mixes via multimodal prompts (text, visual, time-span). It produces target and residual stems at original sample rates for production, post, and research.samaudio.audioFreeScriptaaOtherScriptaa is a multimodal generative AI platform that enables content creation in text, images, and audio while supporting multilingual output. It features pre-built templates and a retrieval-augmented generation framework, ensuring high-quality content tailored to brand voices.scriptaa.ioSeedanceAI Video GenerationSeedance AI is a multimodal platform for image, video and audio generation supporting text-to-image, image-to-image, text-to-video and image-to-video workflows, with built-in video editing/enhancement, broad model selection, batch generation and team collaboration.seedance.aiSuno V5 AppAI MusicSuno V5 Music Generator facilitates unique music creation across genres, allowing users to generate tracks with vocal or instrumental support, edit segments, and share compositions. It also features multi-instrument capabilities and multimodal prompts for enhanced creativity.sunov5.appTate-A-TateDeveloper Tools & APIsTate-A-Tate is a no-code visual builder for designing, testing, and deploying multimodal AI agents with workflow skills, custom code/API tools, cross-platform deployment (web, messaging, API), model integrations, enterprise connectors, real-time translation, analytics, and monetization.tate-a-tate.comTHERAiAI Image GenerationOwnAI is an AI assistant that remembers prior conversations for personalized responses. It supports text, voice, and multimodal input, offers built‑in image generation, live web search, and tone‑adjusting tools, while giving users full data privacy control.therai.meFreeWirestockAI LegalWirestock connects creatives—photographers, videographers, illustrators, designers—with AI labs, offering freelance projects and a dashboard to track earnings and progress. It supplies ethically sourced, legally cleared multimodal datasets for model training and rapid access to fresh, high‑quality data.wirestock.io3.4Release.aiAI Models & TrainingRelease.ai deploys LLM, computer‑vision, and multimodal models with sub‑100 ms latency. It auto‑scales from zero to thousands of concurrent requests, provides enterprise‑grade security (SOC 2 Type II, private networking, end‑to‑end encryption), and offers SDKs, APIs, and real‑time monitoring.release.aiFree5.0OpenAI APIAI Models & TrainingOpenAI's API provides access to GPT-3 and GPT-4 models, which performs a wide variety of natural language tasks, and Codex, which translates natural language to code.openai.comGopherAI Models & TrainingGopher by DeepMind is a 280 billion parameter language model.deepmind.comOPTAI Models & TrainingOpen Pretrained Transformers (OPT) by Facebook is a suite of decoder-only pre-trained transformers. Announcement. OPT-175B text generation hosted by Alpa.huggingface.coLLaMAAI Models & TrainingA foundational, 65-billion-parameter large language model by Meta. #opensourceai.facebook.comLlama 2AI Models & TrainingThe next generation of Meta's open source large language model. #opensourceai.meta.comClaude 3AI Models & TrainingTalk to Claude, an AI assistant from Anthropic.claude.aiVicuna-13BAI Models & TrainingAn open-source chatbot trained by fine-tuning LLaMA on user-shared conversations collected from ShareGPT.lmsys.orgMidjourneyAI Models & TrainingMidjourney is an independent research lab exploring new mediums of thought and expanding the imaginative powers of the human species.midjourney.comFree4.0ImagenAI Models & TrainingImagen by Google is a text-to-image diffusion model with an unprecedented degree of photorealism and a deep level of language understanding.imagen.research.googleCanvaAI Models & TrainingGenerate and Edit your Pictures with the help of AIcanva.comCivitaiAI Models & TrainingCommunity-driven AI model sharing tool.civitai.com4.1
Snapshot mode · This page is served from a database snapshot exported on 2026-09-16, not live data. This notice disappears once the live API is connected.