API & Models
Browse AI tools in this category
All Tools
29 tools

Open model hub with Inference API and Spaces deployment.

CLI to run open LLMs locally for self-hosting and developer integration.

AWS Nova multimodal foundation models for chat and content.

Ray-based model deployment and LLM inference platform.

Speech AI APIs for transcription, summarization, and LeMUR LLM.

Microsoft Azure speech — TTS, STT, translation, enterprise SDKs.

Truss-based model deployment and production inference APIs.

Ultra-fast LLM inference API on Cerebras wafer-scale chips.

Node-based Stable Diffusion workflows for custom generation pipelines.

Speech-to-text and TTS APIs with low-latency streaming.

Low-cost LLM and diffusion model inference API hosting.

Generative media model API for fast image/video/audio inference.

Fast open-model inference API with tool use and fine-tuning.

Groq LPU inference with ultra-fast Llama and open models via API.

Lightweight AI app and model deployment, Python-first.

Local LLM inference engine running GGUF models on CPU/GPU.

Desktop app to run local LLMs with OpenAI-compatible API.

Serverless GPU cloud to run AI workloads by the second.

Self-hosted open chat UI for Ollama and OpenAI-compatible APIs.

Unified API gateway to call GPT, Claude, Llama, and hundreds of models.

Vector database and RAG retrieval API for LLM apps.

Open AI coding agent with self-hosting and fine-tuning.

Cloud API platform to run image, video, and language models.

Voice cloning and emotional TTS API for games and media.

GPU cloud and serverless inference for custom models.

Self-hostable open-source AI code completion server.

Cloud inference for open models—Llama, Mixtral APIs and fine-tuning.

ByteDance Volcengine cloud TTS and speech recognition APIs.

01.AI Yi models for bilingual chat and API access.
