Categories
Rankings
New
简体中文
English
Sign in
Sign up
Submit a tool
Affiliate
For vendors
Categories
Rankings
New
Submit a tool
Affiliate
For vendors
Sign in
Sign up
Search “inference”
40 matching tools
inference
Privatemode AI
AI Chat Assistants
Privatemode AI is a privacy-first AI assistant and inference API that ensures user data remains encrypted at all times, even during processing. Leveraging confidential computing, it provides end-to-end encryption and a zero-trust architecture for maximum security.
privatemode.ai
Free
★ 5.0
Raphael AI
AI Image Generation
Raphael AI is a browser and API image generator that routes between multiple models (Z-Image, Flux 2, Qwen-Image, Nano Banana Pro) for scene-aware photoreal, anime, and illustration outputs, with prompt-accurate controls, editor tools, fast inference, and no data retention.
raphael.app
Paid
★ 3.9
RunPod
AI Models & Training
Runpod supplies on‑demand GPUs in 31 regions, offering single‑node pods, multi‑node clusters, and serverless workloads. It delivers low‑latency inference, efficient fine‑tuning, instant scaling, S3‑compatible storage, real‑time logs, and sub‑200 ms cold starts.
runpod.io
Paid
★ 4.5
Runware
AI Photo Editing
Runware offers an API and web Playground for image, video, and audio generative inference—supporting text-to-image, image-to-image, inpainting, outpainting, ControlNet, custom model uploads, background removal, upscaling, automatic captioning, and low‑latency batch execution.
runware.ai
Paid
Sand.ai
AI Video Generation
Magi-1.1 is an autoregressive video generation model that produces temporally coherent, high-resolution sequences. It provides inference code, model weights, API access, MagiAttention architecture, frame-level control, conditional generation, and pipeline integration for research and development.
sand.ai
Free
Sesterce Cloud
AI Models & Training
Cloud GPU rental platform offering on-demand VMs and bare-metal servers with A100/H100/RTX4090 and other GPUs, configurable vRAM/vCPU, persistent volumes, spot instances, and API-driven provisioning for training, inference, rendering, and HPC workloads.
cloud.sesterce.com
Free
SiliconFlow
Developer Tools & APIs
SiliconFlow is an AI infrastructure platform enabling high-speed inference for LLMs and multimodal applications, supporting serverless, reserved, and private-cloud deployments. It offers low-latency processing, elastic compute, and built-in monitoring for scalable, cost-efficient AI workloads.
siliconflow.com
Paid
★ 5.0
SnapAndSolve
AI Education
SnapAndSolve uses OCR and language‑model inference to turn photographed questions into quick, accurate answers. Users capture or upload images, crop for focus, and receive concise, context‑aware responses in seconds, supporting multiple languages for students, professionals, and educators.
snapandsolve.com
Free
Spice.ai
Developer Tools & APIs
Spice AI is an open‑source platform that fuses SQL federation with hybrid vector, full‑text, and keyword search, enabling unified queries across databases and lakes. It supports local or hosted LLM inference, real‑time change capture, and secure sandboxed AI serving.
spice.ai
Free
Trooper.AI
AI Models & Training
Trooper.AI provides private EU-hosted bare-metal GPU servers for model training, fine-tuning, and inference, with one-click AI environment templates, full root SSH and NVMe storage, tested CUDA on Ubuntu 22.04, scalable hardware and pause/upgrade controls.
trooper.ai
Paid
Union Cloud
AI Models & Training
Union.ai is a cloud‑native AI orchestration platform that lets data scientists and ML engineers build, test, and deploy high‑velocity, pure Python workflows. It supports dynamic branching, real‑time inference, automatic failure recovery, caching, versioning, and observability dashboards.
union.ai
★ 2.5
up-board.org
Developer Tools & APIs
Compact edge platform featuring the Hailo‑8 accelerator for up to 83 TOPs. Supports USB, PCIe, Ethernet, and GPIO; runs Linux ≥ 6.18 with drivers, enabling rapid AI deployment for real‑time inference in automotive, security, and industrial inspection.
up-board.org
Free
zgi.ai
AI Agents
ZGI is an AI‑native data platform that enables rapid creation, integration, and deployment of intelligent agent applications. It automates schema inference, data import, and semantic search across 1,000+ models, with workflow design, multi‑agent orchestration, and secure connectors for enterprise use.
zgi.ai
Free
Vast.AI
Developer Tools & APIs
Vast.ai supplies on‑demand GPU instances, including NVIDIA RTX, H100, and Blackwell models, deployable in seconds. Developers can programmatically provision resources via CLI, SDK or API, and scale workloads with autoscaling, serverless inference, and dedicated InfiniBand clusters.
vast.ai
Free
★ 2.7
Modular
AI Models & Training
Modular is a unified AI inference platform providing a comprehensive stack from custom GPU kernels to cloud serving. Its key strength lies in offering a seamless experience across different hardware including NVIDIA, AMD, Trainium, TPU, Qualcomm, Intel, ARM, and Apple silicon. A key feature includes
modular.com
Nebius Token Factory
AI Models & Training
Nebius Token Factory is an enterprise AI infrastructure platform designed for high-throughput, low-latency inference across open-source large language models. It provides developers and organizations with dedicated inference endpoints, transparent $/token pricing, and autoscaling performance, all wi
nebius.com
Previous
1
2
Next
ℹ️
Snapshot mode
· This page is served from a database snapshot exported on 2026-09-16, not live data. This notice disappears once the live API is connected.