LLM Infrastructure & APIs Vendors
Large language model hosting and inference APIs. This directory lists 35 llm infrastructure & apis vendors, compiled from public sources.
AI21 Labs
ClaimEnterprise-grade AI systems for complex, secure, and scalable workflows
Aleph Alpha
ClaimCustom AI language models for secure, compliant European enterprises
Anthropic
ClaimAI research and LLM infrastructure focused on safe, scalable AI solutions
Anyscale Endpoints
ClaimPlatform to build, scale, and operate AI workloads with Ray infrastructure
AutoGen (Microsoft)
ClaimOpen-source LLM infrastructure for scalable AI application development
Banana Dev
ClaimScalable GPU infrastructure for high-throughput AI model inference
Baseten
ClaimHigh-performance AI model deployment and inference infrastructure
Beam Cloud
ClaimAI infrastructure platform enabling scalable LLM deployment for developers
Cerebras
ClaimUltra-fast AI infrastructure for large language model inference and training
CrewAI
ClaimOpen-source platform to build and manage collaborative AI agent crews.
Crusoe Energy
ClaimRenewable-powered AI infrastructure for scalable, high-performance LLM workloads
Dify
ClaimBuild, deploy, and manage autonomous AI workflows at scale
Fireworks AI
ClaimOptimized, scalable inference platform for generative AI at enterprise scale
Flowise
ClaimVisual platform to build and deploy AI agents and workflows
Fluidstack
ClaimHigh-performance AI infrastructure with secure, dedicated GPU clusters.
Google Gemini
ClaimNext-generation AI systems for advanced large language model infrastructure
Groq
ClaimHigh-speed, cost-efficient AI inference with custom silicon technology
Haystack (deepset)
ClaimOpen source AI framework for building production-ready LLM agents and applications
LM Studio
ClaimRun local and private AI models on your own hardware seamlessly
Lepton AI
ClaimConnecting developers to global GPU compute for scalable AI infrastructure
Mystic AI
ClaimEnterprise-grade LLM infrastructure for scalable AI deployments
Nebius (Yandex)
ClaimScalable AI cloud infrastructure optimized for large-scale LLM workloads
Nvidia NIM
ClaimGPU-accelerated AI inferencing microservices for enterprise LLM deployment
Octo AI
ClaimAI infrastructure solutions powering enterprise-scale large language models
Ollama
ClaimOpen source infrastructure for running and integrating large language models
OpenAI
ClaimAdvanced AI and large language model infrastructure for enterprises
Perplexity AI
ClaimAI-powered large language model infrastructure for enterprise applications
Runpod
ClaimOn-demand GPU infrastructure for scalable AI and machine learning workloads
SambaNova
ClaimHigh-performance AI inference platform with custom hardware and software stack
Semantic Kernel (Microsoft)
ClaimLightweight open-source SDK to build AI agents with seamless LLM integration
Vast.ai
ClaimCost-effective, scalable GPU infrastructure for AI and ML workloads
Vectara
ClaimGoverned, grounded, auditable AI agents for enterprise-scale applications
Voltage Park
ClaimHigh-performance AI infrastructure and GPU cloud for enterprise-scale AI workloads
n8n AI
ClaimAI workflow automation platform with seamless integrations and full control
vLLM
ClaimHigh-throughput, memory-efficient LLM inference and serving engine
Listings are compiled by CIOPages from public sources and are presented in alphabetical order — not ranked, rated, or endorsed. Profiles may be incomplete or out of date. Search & filter this category →
Claim your listing to control the record — free. Upgrade only when you want buyer leads.