CIOPages
DirectoryAI & ML PlatformsLLM Infrastructure & APIs

LLM Infrastructure & APIs Vendors

Large language model hosting and inference APIs. This directory lists 35 llm infrastructure & apis vendors, compiled from public sources.

AI21 Labs

Claim

Enterprise-grade AI systems for complex, secure, and scalable workflows

Aleph Alpha

Claim

Custom AI language models for secure, compliant European enterprises

Anthropic

Claim

AI research and LLM infrastructure focused on safe, scalable AI solutions

Anyscale Endpoints

Claim

Platform to build, scale, and operate AI workloads with Ray infrastructure

AutoGen (Microsoft)

Claim

Open-source LLM infrastructure for scalable AI application development

Banana Dev

Claim

Scalable GPU infrastructure for high-throughput AI model inference

Baseten

Claim

High-performance AI model deployment and inference infrastructure

Beam Cloud

Claim

AI infrastructure platform enabling scalable LLM deployment for developers

Cerebras

Claim

Ultra-fast AI infrastructure for large language model inference and training

CrewAI

Claim

Open-source platform to build and manage collaborative AI agent crews.

Crusoe Energy

Claim

Renewable-powered AI infrastructure for scalable, high-performance LLM workloads

Dify

Claim

Build, deploy, and manage autonomous AI workflows at scale

Fireworks AI

Claim

Optimized, scalable inference platform for generative AI at enterprise scale

Flowise

Claim

Visual platform to build and deploy AI agents and workflows

Fluidstack

Claim

High-performance AI infrastructure with secure, dedicated GPU clusters.

Google Gemini

Claim

Next-generation AI systems for advanced large language model infrastructure

Groq

Claim

High-speed, cost-efficient AI inference with custom silicon technology

Haystack (deepset)

Claim

Open source AI framework for building production-ready LLM agents and applications

LM Studio

Claim

Run local and private AI models on your own hardware seamlessly

Lepton AI

Claim

Connecting developers to global GPU compute for scalable AI infrastructure

Mystic AI

Claim

Enterprise-grade LLM infrastructure for scalable AI deployments

Nebius (Yandex)

Claim

Scalable AI cloud infrastructure optimized for large-scale LLM workloads

Nvidia NIM

Claim

GPU-accelerated AI inferencing microservices for enterprise LLM deployment

Octo AI

Claim

AI infrastructure solutions powering enterprise-scale large language models

Ollama

Claim

Open source infrastructure for running and integrating large language models

OpenAI

Claim

Advanced AI and large language model infrastructure for enterprises

Perplexity AI

Claim

AI-powered large language model infrastructure for enterprise applications

Runpod

Claim

On-demand GPU infrastructure for scalable AI and machine learning workloads

SambaNova

Claim

High-performance AI inference platform with custom hardware and software stack

Semantic Kernel (Microsoft)

Claim

Lightweight open-source SDK to build AI agents with seamless LLM integration

Vast.ai

Claim

Cost-effective, scalable GPU infrastructure for AI and ML workloads

Vectara

Claim

Governed, grounded, auditable AI agents for enterprise-scale applications

Voltage Park

Claim

High-performance AI infrastructure and GPU cloud for enterprise-scale AI workloads

n8n AI

Claim

AI workflow automation platform with seamless integrations and full control

vLLM

Claim

High-throughput, memory-efficient LLM inference and serving engine

Listings are compiled by CIOPages from public sources and are presented in alphabetical order — not ranked, rated, or endorsed. Profiles may be incomplete or out of date. Search & filter this category →

More in AI & ML Platforms
Represent a technology vendor?

Claim your listing to control the record — free. Upgrade only when you want buyer leads.