Search “inference”

40 matching tools

inference
Privatemode AIAI Chat AssistantsPrivatemode AI is a privacy-first AI assistant and inference API that ensures user data remains encrypted at all times, even during processing. Leveraging confidential computing, it provides end-to-end encryption and a zero-trust architecture for maximum security.privatemode.aiFree5.0Raphael AIAI Image GenerationRaphael AI is a browser and API image generator that routes between multiple models (Z-Image, Flux 2, Qwen-Image, Nano Banana Pro) for scene-aware photoreal, anime, and illustration outputs, with prompt-accurate controls, editor tools, fast inference, and no data retention.raphael.app3.9RunPodAI Models & TrainingRunpod supplies on‑demand GPUs in 31 regions, offering single‑node pods, multi‑node clusters, and serverless workloads. It delivers low‑latency inference, efficient fine‑tuning, instant scaling, S3‑compatible storage, real‑time logs, and sub‑200 ms cold starts.runpod.io4.5RunwareAI Photo EditingRunware offers an API and web Playground for image, video, and audio generative inference—supporting text-to-image, image-to-image, inpainting, outpainting, ControlNet, custom model uploads, background removal, upscaling, automatic captioning, and low‑latency batch execution.runware.aiSand.aiAI Video GenerationMagi-1.1 is an autoregressive video generation model that produces temporally coherent, high-resolution sequences. It provides inference code, model weights, API access, MagiAttention architecture, frame-level control, conditional generation, and pipeline integration for research and development.sand.aiFreeSesterce CloudAI Models & TrainingCloud GPU rental platform offering on-demand VMs and bare-metal servers with A100/H100/RTX4090 and other GPUs, configurable vRAM/vCPU, persistent volumes, spot instances, and API-driven provisioning for training, inference, rendering, and HPC workloads.cloud.sesterce.comFreeSiliconFlowDeveloper Tools & APIsSiliconFlow is an AI infrastructure platform enabling high-speed inference for LLMs and multimodal applications, supporting serverless, reserved, and private-cloud deployments. It offers low-latency processing, elastic compute, and built-in monitoring for scalable, cost-efficient AI workloads.siliconflow.com5.0SnapAndSolveAI EducationSnapAndSolve uses OCR and language‑model inference to turn photographed questions into quick, accurate answers. Users capture or upload images, crop for focus, and receive concise, context‑aware responses in seconds, supporting multiple languages for students, professionals, and educators.snapandsolve.comFreeSpice.aiDeveloper Tools & APIsSpice AI is an open‑source platform that fuses SQL federation with hybrid vector, full‑text, and keyword search, enabling unified queries across databases and lakes. It supports local or hosted LLM inference, real‑time change capture, and secure sandboxed AI serving.spice.aiFreeTrooper.AIAI Models & TrainingTrooper.AI provides private EU-hosted bare-metal GPU servers for model training, fine-tuning, and inference, with one-click AI environment templates, full root SSH and NVMe storage, tested CUDA on Ubuntu 22.04, scalable hardware and pause/upgrade controls.trooper.aiUnion CloudAI Models & TrainingUnion.ai is a cloud‑native AI orchestration platform that lets data scientists and ML engineers build, test, and deploy high‑velocity, pure Python workflows. It supports dynamic branching, real‑time inference, automatic failure recovery, caching, versioning, and observability dashboards.union.ai2.5up-board.orgDeveloper Tools & APIsCompact edge platform featuring the Hailo‑8 accelerator for up to 83 TOPs. Supports USB, PCIe, Ethernet, and GPIO; runs Linux ≥ 6.18 with drivers, enabling rapid AI deployment for real‑time inference in automotive, security, and industrial inspection.up-board.orgFreezgi.aiAI AgentsZGI is an AI‑native data platform that enables rapid creation, integration, and deployment of intelligent agent applications. It automates schema inference, data import, and semantic search across 1,000+ models, with workflow design, multi‑agent orchestration, and secure connectors for enterprise use.zgi.aiFreeVast.AIDeveloper Tools & APIsVast.ai supplies on‑demand GPU instances, including NVIDIA RTX, H100, and Blackwell models, deployable in seconds. Developers can programmatically provision resources via CLI, SDK or API, and scale workloads with autoscaling, serverless inference, and dedicated InfiniBand clusters.vast.aiFree2.7ModularAI Models & TrainingModular is a unified AI inference platform providing a comprehensive stack from custom GPU kernels to cloud serving. Its key strength lies in offering a seamless experience across different hardware including NVIDIA, AMD, Trainium, TPU, Qualcomm, Intel, ARM, and Apple silicon. A key feature includesmodular.comNebius Token FactoryAI Models & TrainingNebius Token Factory is an enterprise AI infrastructure platform designed for high-throughput, low-latency inference across open-source large language models. It provides developers and organizations with dedicated inference endpoints, transparent $/token pricing, and autoscaling performance, all winebius.com
Snapshot mode · This page is served from a database snapshot exported on 2026-09-16, not live data. This notice disappears once the live API is connected.