AI Platform Pricing: Examples & Companies

167 companies in the corpus Updated full analysis
Definition

AI Platform Pricing is Pricing for general-purpose AI platforms — model APIs, inference services, and multi-model hosting providers.

Also known as: AI API PricingFoundation Model Platform Pricing

What is it

AI platform pricing is how general-purpose AI platforms — model APIs, inference services, and multi-model hosting providers — package and charge for access to model capability. It is the pricing layer that sits beneath most of the AI software market: the frontier labs, the inference marketplaces, and the application platforms that resell or wrap model access all fall inside it.

This is the broadest product category in the corpus. 167 companies carry the ai-platforms product segment, spanning three rough strata. At the base are the frontier and open-model labs — OpenAI, Anthropic, DeepSeek, and Mistral AI — that expose per-token APIs. In the middle are inference platforms — Fireworks AI, Together AI, Groq, and Novita AI — that serve open-weight models and rent GPU capacity. At the application layer are platforms that embed models into a finished workflow — HeyGen for avatar video, Runway for generative video, Gumloop and Relevance AI for agent workflows.

Because those three strata serve fundamentally different buyers, the pricing diversity inside this one category is the widest in the corpus. The same segment contains companies billing per token, per credit, per seat, per action, and per resolution. A developer buying inference wants a public rate card and no minimum; an enterprise buying a legal-AI platform wants a per-seat quote; a prosumer buying video generation wants an all-inclusive monthly credit bundle. AI platform pricing is best understood not as one model but as a set of archetypes, each attached to a buyer profile.

The through-line is that the model itself is rarely the billing unit the customer sees. Tokens are the native cost driver of inference, but most application platforms abstract them away behind credits or seats so buyers can reason about spend without doing token math. Understanding a platform’s price means first identifying which archetype it belongs to, then reading how it translates underlying model cost into a unit its buyer will accept.

The consumer AI-platform ladder · Free → $20 → $200
Six platforms, one ladder: Free → $20 → $200 FREE $0 caps apply · PLG on-ramp PRO $20 /mo · mass-market price MAX / ULTRA $200 /mo · power-user ceiling ~10× THE PRO JUMP OpenAI · Anthropic · Cursor · Perplexity · Google · You.com — all converge here

How it works

AI platform pricing resolves into five recurring archetypes. Which one a company picks is driven almost entirely by who it sells to.

ArchetypeBilling unitTypical buyerCorpus examples
Pure-usage APITokens, requests, GPU-hoursDevelopersOpenAI API, DeepSeek, Groq, Fireworks AI
Freemium + credit subscriptionMonthly credit pool + overageProsumers, SMBHeyGen, Runway, Ideogram, Gumloop
Seat + credits hybridPer-user seat + metered poolTeamsCursor, Relevance AI, Intercom
Platform fee + usage (enterprise)Committed minimum + overageEnterpriseTogether AI, Fireworks AI enterprise
Outcome-basedPer resolved taskSupport / ops teamsIntercom Fin

Pure-usage APIs publish a rate per input and output token, metered at consumption with free credit to start and volume discounts at scale. The spread is dramatic: DeepSeek’s V4-Flash starts at $0.0028/1M tokens on a cache hit, Groq prices Llama 3.1 8B at $0.05/$0.08 in/out, and OpenAI exposes the GPT-5.x API from $0.20/1M tokens. Cached and batch modes commonly cut the headline rate by 50% — Fireworks AI and Together AI both apply a 50% batch discount and layer per-hour GPU rates (H100 around $6–$7/hr) on top of serverless token pricing.

Credit subscriptions convert model cost into a monthly pool the buyer can budget. A worked example: HeyGen runs Free $0 → Creator $29 → Pro $49 → Business $149 (+$20/seat), each tier bundling a credit allotment, with a separate $5 pay-as-you-go API wallet. Runway starts at $12/mo bundling monthly credits that meter video, image, and audio generation, plus a Max tier with 9,500 rolling credits. The credit is a deliberate abstraction — it lets the vendor change underlying model economics without renegotiating, and lets the buyer avoid token math.

Seat + credits hybrids anchor predictable revenue on a per-user fee, then meter the AI-intensive work against a pool. Cursor is the canonical case; Relevance AI pairs seat tiers (Free, Pro $19/mo, Team $234/mo) with Actions usage and no-markup pass-through Vendor Credits. Outcome-based pricing is the rarest and most aligned: Intercom charges $0.99 per Fin AI resolution on top of $29–$132/seat platform tiers, billing only when the AI actually closes a ticket.

Companies using this

The 167 companies below all carry the ai-platforms product segment. They range from frontier model labs and open-model inference platforms to application-layer generation tools and vertical enterprise platforms — sort and filter the table to compare pricing models, billing units, and free-tier availability across the category.

Patterns observed

Pure-usage token pricing is the category floor, and it keeps falling. Every frontier and open-model lab in the corpus exposes a per-token rate, and the low end is dropping fast — DeepSeek’s cache-hit rate reset the floor while Groq competes on latency rather than a headline discount. The competitive pressure is so intense that cached, batch, and priority tiers have become standard — Fireworks AI and Together AI both advertise 50%-off batch modes, which is now table stakes rather than a differentiator.

Application-layer platforms abstract tokens into credits. The higher up the stack a platform sits, the less its buyer sees raw tokens. HeyGen, Runway, and Ideogram (Free, Plus $15, Pro $42, Team $20/user) all sell credit-metered subscriptions where the credit maps to a generated output — a video, an image — not to a token count. This abstraction is what lets these platforms run mass-market $15–$49 price points while insulating margin from underlying model volatility.

Multi-track pricing on one page is increasingly common. Rather than pick a single archetype, the larger platforms run several in parallel. Mistral AI prices two distinct surfaces from one page: flat-rate Vibe assistant subscriptions ($0–$24.99/user/mo) and a pure per-token developer API from $0.10/1M. OpenAI runs consumer subscriptions (Go $8, Plus $20, Pro $100 and $200, Business $20/seat) alongside its token API. The platform that serves both a developer and a prosumer buyer increasingly maintains a distinct pricing track for each.

The $200 prosumer ceiling has hardened. A cluster of consumer-facing platforms — OpenAI (ChatGPT Pro at $200), Anthropic (Claude Max from $100), Cursor, Perplexity, Google, and You.com — added a roughly $200/month top tier above the standard ~$20 Pro plan. This tier monetizes power users who would otherwise cap against lower plan limits, and it has been stable since it emerged. Following how these plans layer usage on a subscription base is the same problem covered in the guide to usage invoicing and billing cycles.

Counterexamples & variants

Not every AI platform publishes a rate card. The archetypes above assume public pricing, but the vertical enterprise strata breaks that assumption. Harvey, the legal-AI platform, sells per seat via sales-led quotes with no public rate card and reported seat minimums — the opposite of the developer-facing transparency norm. Enterprise-knowledge and regulated-vertical platforms consistently gate pricing because their buyers procure top-down and their deals are negotiated, not self-served. If you are comparing a vertical platform against a developer API on price transparency alone, you are comparing two different go-to-market cultures, not two pricing philosophies.

Outcome-based pricing is aspirational for most of the category, not standard. Intercom’s $0.99-per-resolution Fin pricing is held up as the model to emulate, but it remains rare because it requires an attributable, discrete outcome the buyer and vendor both agree happened. Most AI platforms cannot cleanly define such an outcome — what is the “outcome” of a token API or a video generator? — so they fall back to consumption or seats. Outcome pricing is a variant available to workflow platforms with a countable end state, not a template the whole category can adopt.

Credit systems vary from fully transparent to deliberately opaque. The credit archetype hides a wide range of honesty. Relevance AI passes through Vendor Credits at no markup, telling the buyer exactly what the underlying model cost was. At the other extreme, many application platforms set an internal, undisclosed credit-to-cost ratio so the buyer cannot reverse-engineer margin. Two platforms can both say “credits” and mean opposite things about transparency — the variant matters more than the label, which is why choosing and reading the underlying metric is worth understanding in its own right (see the guide to choosing the right usage metric).

What this means for buyers vs vendors

For buyers

Identify the archetype before you compare prices. A per-token rate and a per-seat quote are not comparable numbers, and a $29 credit subscription may cost far more or far less than a metered API depending on your volume — model your own usage first. For developer APIs, use a token calculator such as the Anthropic pricing calculator or the Cursor pricing calculator to translate a headline rate into a monthly bill at your actual volume, and always check whether cached and batch modes apply to your workload — a 50% batch discount changes the math entirely. For credit-based platforms, ask what the credit maps to and whether it is passed through at cost. For vertical enterprise platforms that gate pricing, expect a negotiated seat quote and factor acquisition risk into any multi-year commitment: several corpus platforms, including OpenPipe, OpenMeter, and Tavily, were acquired inside an 18-month window, and gating risk rises after a deal closes. A rate-lock clause is the cheapest insurance available.

For vendors

Your pricing archetype should follow your buyer, not your architecture. If you sell to developers, a public per-token rate with generous free credit and clear cached/batch tiers is the price of entry — opacity reads as a red flag in this segment. If you sell to prosumers or teams, credits let you set a legible monthly price while insulating margin from model-cost swings, but you owe the buyer clarity on what a credit buys. If you sell top-down into an enterprise or regulated vertical, seats and sales-led quotes remain viable, and gated pricing is defensible where deals are genuinely negotiated. The strategic risk across all archetypes is model-cost deflation: with the token floor still falling, any platform whose margin depends on a fixed markup over inference cost should assume that markup compresses. The durable positions are abstraction (credits, outcomes) and differentiation (latency, vertical depth), not a thin resale spread. For the mechanics of metering and invoicing usage at scale, the guide to usage invoicing and billing cycles covers the operational layer.

Company Product Pricing modelBilling unitsFree tier Verified
01.AIYi open-weight models + Yi API + enterprise vertical solutionsYes2026-06-11
11xAutonomous AI digital workers — Alice (outbound SDR) and Julian (inbound phone agent)No2026-06-05
1X TechnologiesNEO home humanoid robot & EVE enterprise robotics (RaaS)No2026-06-14
Abacus.AIAI super-assistant (ChatLLM) plus an enterprise agentic AI platformNo2026-06-02
AdeptACT-1 — action-oriented AI agents that operate software via the UINo2026-06-16
Agility RoboticsDigit humanoid robot + Agility Arc cloud platform (Robots-as-a-Service)No2026-06-14
AI21 LabsJamba foundation models, Maestro orchestration & enterprise AIYes2026-06-11
Aleph AlphaPhariaAI sovereign-AI platform, specialized models & professional servicesNo2026-06-11
Ambience HealthcareEnterprise AI platform for clinical documentation and point-of-care codingNo2026-06-10
AnthropicClaude API (token-based) + Claude.ai consumer subscriptions (Free/Pro/Team/Enterprise)Yes2026-07-06
AnyscaleManaged Ray platform for distributed AI training, inference, and batch processing (RayTurbo, Anyscale Compute Units)Yes2026-05-29
ApptronikApollo general-purpose humanoid robot (RaaS + outright sale)No2026-06-14
AssemblyAISpeech-to-Text & Audio AI APIsYes2026-07-06
Athina AICollaborative AI development platform for building, testing, evaluating and monitoring LLM featuresYes2026-06-04
Autodesk (Flow Studio, formerly Wonder Dynamics)AI VFX automation platform (Flow Studio)Yes2026-06-16
Baichuan AIBaichuan & medical M-series LLM APIsYes2026-06-11
BardeenAI browser automation and workflow agentsYes2026-06-10
BasetenML inference infrastructure — dedicated GPU deployments, Model APIs, and Truss frameworkYes2026-05-29
BentoMLBentoCloud — managed model-serving & inference platformYes2026-06-15
CartesiaReal-time voice AI platform (Sonic TTS, voice cloning, voice agents)Yes2026-05-29
CerebrasWafer-scale AI inference cloud and WSE hardware systemsYes2026-05-30
Character.aiConsumer AI companion and roleplay chat platformYes2026-05-29
CognosysAutonomous AI agents (rebranded Ottogrid, acquired by Cohere)Yes2026-06-16
CohereCommand, Embed, Rerank APIsYes2026-05-29
ComposioTool-calling and integration infrastructure that connects AI agents to 1,000+ apps with managed auth and tool executionYes2026-06-10
Copy.aiGTM AI workflow platformNo2026-06-15
CovariantCovariant Brain — AI for autonomous warehouse robotic pickingNo2026-07-14
CrewAIMulti-agent orchestration framework (OSS) + CrewAI AMP enterprise platformYes2026-06-10
DailyReal-time voice and video WebRTC APIs (Video SDK + Pipecat Cloud)Yes2026-07-14
Databricks (Mosaic AI)Mosaic AI — enterprise GenAI & ML on the Data Intelligence PlatformYes2026-06-15
DeepgramUsage-based speech-to-text, text-to-speech, and voice agent APIsYes2026-05-31
DeepInfraServerless inference cloud — per-token LLM/embedding APIs, per-image and per-minute media models, per-hour on-demand GPU containers, and reserved DeepCluster GPU clustersNo2026-07-14
DeepSeekDeepSeek API (V4-Flash + V4-Pro models, 1M context) with token-based pricing and aggressive cache discountsYes2026-06-05
DescriptAI-powered audio and video editingYes2026-05-31
DifyDify Cloud + self-hosted LLM app development platformYes2026-07-14
DustEnterprise AI agent deployment platformYes2026-06-24
ElevenLabsVoice AI platform across ElevenCreative, ElevenAgents, and ElevenAPIYes2026-06-30
Essential AIEnterprise foundation models & data-workflow automationNo2026-06-11
EvenUpAI Claims Intelligence Platform for personal injury law firmsNo2026-06-16
ExaAI web search API for agents — search, contents, deep research, and monitoring endpoints billed per requestYes2026-07-14
Exscientia (now part of Recursion)AI-driven drug discovery & design platformNo2026-06-16
FalGenerative-media inference platform — serverless per-output model APIs plus dedicated GPU computeNo2026-06-01
FigureGeneral-purpose humanoid robots (Figure 03) & Helix AINo2026-06-14
Fireworks AIGenerative AI inference platform — serverless per-token, on-demand GPU, fine-tuning, batch APIYes2026-05-30
FreepikAI creative suite — image, video, audio generation plus a 200M+ stock libraryYes2026-06-05
GalileoAI observability, evaluation, and guardrails platform for agents and LLM appsYes2026-06-04
GammaAI presentations, documents and websitesYes2026-06-11
GensparkAll-in-one AI agent workspace (Super Agent, AI Slides/Sheets/Docs, image/video/audio generation) on a credit-based modelYes2026-06-02
GladiaSpeech-to-text & audio intelligence APIYes2026-06-09
GleanEnterprise AI search and knowledge (Work AI) platformNo2026-05-31
GoogleGemini API & AI StudioYes2026-07-14
GranolaAI notepad for back-to-back meetingsYes2026-06-15
GrokxAI's consumer and business AI assistantYes2026-06-16
GroqGroqCloud — LPU-based ultra-low-latency inference API for Llama, GPT-OSS, Qwen, Whisper transcription, and Orpheus text-to-speechYes2026-07-14
GumloopNo-code AI workflow and agent automation platform billed on creditsYes2026-06-30
HarveyGenerative AI platform for legal and professional-services workNo2026-05-31
HebbiaMatrix — agentic AI for institutional knowledge work and document analysisNo2026-06-15
HedraAI video, avatar, image, and audio generation platform (Hedra Studio + API)Yes2026-06-04
HeyGenAI avatar and video generation platformYes2026-05-30
HiggsfieldAI video and image generation platform with a credit-metered subscriptionYes2026-06-06
Hippocratic AISafety-focused healthcare LLM — patient-facing AI agents for non-diagnostic clinical tasksNo2026-06-10
Hugging FaceAI model hub, inference endpoints & computeYes2026-06-15
Hume AIEmpathic Voice Interface (EVI) + Octave TTS + expression-measurement APIsYes2026-06-30
HyperbolicGPU cloud marketplace & serverless AI inferenceYes2026-06-15
IdeogramText-aware AI image generation platformYes2026-06-15
ImbueReasoning-agent research lab and coding-agent tools (Sculptor)No2026-06-16
Inflection AIEnterprise foundation models (Inflection 3.0) + Pi assistantNo2026-06-11
Insilico MedicinePharma.AI generative drug-discovery platform + clinical pipelineYes2026-06-14
Ironclad AIAI-powered contract lifecycle management (CLM)No2026-06-16
Isomorphic LabsAI-first drug discovery & design (Isomorphic Drug Design Engine)No2026-06-14
Janitor AIConsumer AI character chat / roleplay platformYes2026-06-16
JasperAI marketing content platformNo2026-05-31
Jina AISearch Foundation API (Embeddings, Reranker, Reader, DeepSearch, Classifier)Yes2026-06-03
LangChainAgent orchestration frameworks + LangSmith platformYes2026-06-10
LangfuseOpen-source LLM observability, evals, and prompt managementYes2026-06-09
LegoraCollaborative AI for lawyers — review, drafting, and researchNo2026-06-06
Lightning AICloud GPU/CPU Studio compute platform for building, training, and serving AI models, billed by the second with a credit pool.Yes2026-06-02
LindyAI executive assistant (iMessage/SMS) — formerly AI agent-builder platformNo2026-06-10
LiveKitOpen-source real-time (WebRTC) communications, LiveKit Cloud & Agents frameworkYes2026-06-30
LMNTLow-latency AI text-to-speech (TTS) API with voice cloningYes2026-06-04
Luma AIDream Machine — text/image-to-video, image and audio generation (plus Genie 3D)Yes2026-06-11
MakeVisual, no-code automation (iPaaS) platform connecting 3,000+ apps and AI agentsYes2026-06-11
ManusGeneral AI agent that executes multi-step tasks autonomously in the cloudYes2026-06-02
Maven AGIEnterprise AI agent platform for customer supportNo2026-06-11
MemAI-powered personal memory workspaceYes2026-07-14
Mem0Memory layer for AI agents and applicationsYes2026-06-10
MercorAI talent marketplace + enterprise data partnerships for frontier AI labsNo2026-07-14
MidjourneyAI image and video generation via subscription with GPU-hour meteringNo2026-05-29
MiniMaxFoundation models, Hailuo video & per-token APIYes2026-06-11
MintlifyAI-native developer documentationYes2026-06-15
Mistral AIOpen and commercial LLM APIsYes2026-07-06
ModalServerless compute and GPU platform — per-second billing for Python functions, batch jobs, and model servingYes2026-07-14
Moonshot AIKimi assistant + Kimi/Moonshot open-weight LLM APIYes2026-06-11
MotionMotion AI productivity platform (Pro AI, Business AI)No2026-06-08
MultiOnAutonomous web-browsing AI agent API (wound down)No2026-06-10
Murf AIAI voice / text-to-speech platform (Murf Studio app + Murf API)Yes2026-06-01
n8nFair-code workflow automation platform for technical teams, billed by monthly workflow executionsYes2026-06-02
NomicNomic Platform (AEC agentic workflows) + Atlas data-exploration app + Nomic Embed embedding/Developer APIYes2026-06-04
NooksAI sales platform — parallel dialer, AI SDR, and coachingNo2026-06-05
Notion AIAI workspace, agents, and knowledge managementYes2026-06-15
Novita AIPay-as-you-go AI cloud: 200+ model inference APIs, on-demand GPUs, and per-second agent sandboxes under one APIYes2026-07-06
Observe.AIAgentic CX platform — contact-center AI agents, conversation intelligence & auto-QANo2026-06-09
OctoAIGenerative AI inference platform (acquired by NVIDIA, sunset Oct 2024)No2026-06-15
OpenAIChatGPT consumer subscriptions + GPT-5.x API with token-based usage billingYes2026-06-30
OpenPipeOpenPipe fine-tuning and hosted inference platform (small specialized models / RL for agents)Yes2026-06-04
Paige AIFDA-cleared AI for cancer pathology — clinical diagnostics + pharma/life-sciences foundation modelsNo2026-06-10
PebblelyAI product-photography tool that generates marketing images from a product photoNo2026-06-07
Perplexity AIAI-native answer engine with citations and multi-model searchYes2026-05-29
PhindAI developer search engine and coding assistant (shut down January 2026)Yes2026-06-08
Physical IntelligenceRobotics foundation models (Vision-Language-Action policies for robots)No2026-06-14
PiPi — personal, emotionally intelligent AI assistant (consumer app)Yes2026-06-16
PipedreamWorkflow automation and integration platform for developersYes2026-06-16
PlaygroundAI image generation and graphic-design studio with a monthly credit poolYes2026-06-04
PlayHTText-to-speech & voice cloning API (PlayAI)Yes2026-06-09
PoeMulti-model AI chat subscription (by Quora)Yes2026-06-16
PolyAIEnterprise voice AI assistants for contact centersNo2026-06-09
PoolsideAI coding foundation modelNo2026-06-16
PredibaseFine-tuning & serving platform for open-source LLMsYes2026-06-15
Rad AIGenerative AI for radiology — report drafting (Reporting/Omni), automated impressions, and follow-up management (Continuity)No2026-06-10
RecraftAI image and vector generation studio plus a per-image generation APIYes2026-07-14
RecursionAI-enabled drug discovery platform (Recursion OS) — pharma partnerships, internal pipeline & NVIDIA-powered computeNo2026-06-10
Reka AINatively multimodal models (Spark, Edge, Flash, Core) + Research & Vision APIsYes2026-06-11
Relevance AINo-code platform for building AI agents and multi-agent 'AI Workforces' for sales, marketing, and operations teams.Yes2026-07-14
ReplicateCloud platform for running, fine-tuning, and deploying AI models via REST APIYes2026-05-30
Resemble AIAI deepfake detection & watermarking + voice generation APIsNo2026-07-14
Retell AIConversational voice-agent API platformNo2026-07-14
Rev AIPay-as-you-go speech-to-text, transcription, and audio-intelligence APIsYes2026-06-04
Rewind.ai (the original Rewind AI rebranded to Limitless, acquired by Meta)AI tools aggregator (token-balance) — on the domain once home to the Rewind personal-memory appYes2026-06-15
Robin AIAI legal contract review, drafting & a legal-data APINo2026-06-06
RoboflowComputer-vision platform (dataset management, model training, deployment)Yes2026-07-14
RoxAI agent swarm for sales reps (AE copilot)Yes2026-06-05
RunPodGPU cloud marketplace — Secure Cloud and Community Cloud Pods, Serverless endpoints, and persistent storageNo2026-07-14
RunwayVideo generation and AI editingYes2026-06-24
SambaNovaSambaNova Cloud inference API & RDU AI systemsYes2026-06-15
Sana AIEnterprise AI assistant (Sana Agents) and AI learning platform (Sana Learn)Yes2026-06-15
Sanctuary AIPhoenix general-purpose humanoid robot & Carbon AI control systemNo2026-06-14
Sarvam AISovereign Indic LLM, speech & translation APIsYes2026-06-11
Shield AIHivemind autonomy software, V-BAT & X-BAT autonomous aircraftNo2026-06-14
SkydioAutonomous drones, docks & flight-autonomy software for defense, public safety & enterpriseNo2026-06-14
SpeechmaticsSpeech-to-text and text-to-speech APIs with per-hour usage pricingYes2026-07-06
Stability AIBrand Studio creative platform and open generative media modelsYes2026-06-11
Suki AIAmbient clinical AI assistant for healthcare (Suki Assistant) + embeddable Suki Platform SDK/APINo2026-06-10
SunoAI music generationYes2026-05-31
SynthesiaEnterprise AI video generationYes2026-05-31
Synthflow AINo-code AI voice-agent builderNo2026-06-24
TavilyTavily Search APIYes2026-06-03
TempusPrecision-medicine platform — genomic diagnostics, multimodal clinical data licensing & oncology AI apps (NASDAQ: TEM)No2026-06-10
Thomson Reuters (CoCounsel)CoCounsel — legal generative-AI assistant (formerly Casetext)No2026-06-16
Together AIAI Acceleration Cloud — serverless inference, dedicated endpoints, GPU clusters, Code Sandbox, fine-tuningYes2026-07-14
Trigger.devBackground jobs and workflow orchestration for developersYes2026-06-16
Twelve LabsVideo understanding foundation models (Marengo for search/embeddings, Pegasus for analysis) delivered as a usage-metered APIYes2026-06-02
TypefaceArc enterprise marketing AI platformNo2026-06-16
UdioAI music generationYes2026-06-11
UniphoreBusiness AI Cloud — enterprise conversational AI & agentic automationNo2026-06-09
VapiVoice AI infrastructure for developersNo2026-06-09
VectaraEnterprise RAG-as-a-Service and agent platform for trusted, grounded, auditable AINo2026-06-02
VellumPersonal AI assistant (ex LLM application development platform)Yes2026-06-10
Viz.aiAI-powered care coordination for time-sensitive disease — stroke, aneurysm, PE, cardiac and more (Viz Neuro/Cardio/Vascular/Pulmonary suites)No2026-06-10
Voyage AIEmbedding and reranker models (text, code, multimodal) for retrieval and RAGYes2026-06-04
WeaviateAI-native vector database (open-source core + Weaviate Cloud managed serverless, dedicated/Enterprise Cloud, BYOC)Yes2026-07-06
Weights & BiasesMLOps experiment tracking, W&B Weave LLM observability/evals, Models registry, and Serverless InferenceYes2026-07-14
WriterEnterprise agentic AI platform (Palmyra models, WRITER Agent)No2026-06-15
xAIGrok API and agentic AI stackYes2026-07-14
Yellow.aiConversational CX automation platformYes2026-06-11
You.comWeb search, contents, research, and finance-research APIs for AI systemsYes2026-06-01
ZapierWorkflow-automation (iPaaS) platform connecting 9,000+ apps, with separately-metered AI Agents and Chatbots add-onsYes2026-06-30
Zhipu AIGLM foundation models, per-token API, and GLM Coding PlanYes2026-06-11

Explore this theme in the knowledge graph

FAQ

What is an AI platform?

An AI platform is a general-purpose product that delivers AI model capability — a model API, an inference service, or a multi-model hosting layer — that other software and teams build on. In this corpus, 167 companies carry the ai-platforms product segment, making it the largest single category.

How do AI platforms price their products?

There is no single model. Frontier labs and inference platforms bill per token (OpenAI's GPT-5.x API from $0.20/1M, DeepSeek from $0.0028/1M on a V4-Flash cache hit); application-layer platforms sell credit-based subscriptions (HeyGen $29, Runway $12); and vertical or agent platforms use seats, actions, or outcomes. The billing unit follows the buyer, not the underlying technology.

How much does a frontier model API cost per million tokens?

The spread is enormous. On the low end, DeepSeek's V4-Flash starts at $0.0028/1M tokens on a cache hit and Groq's Llama 3.1 8B is $0.05/$0.08 in/out. On the high end, Anthropic's Fable 5 is $10/$50 and OpenAI's flagship API tiers run several dollars per million. Cached and batch modes routinely cut the headline rate by 50%.

What is the $200 AI platform tier?

Several consumer-facing AI platforms — OpenAI (ChatGPT Pro at $200), Anthropic (Claude Max from $100), Cursor, Perplexity, Google, and You.com — added a roughly $200/month top tier above the standard ~$20 Pro plan. It monetizes power users who cap against lower plan limits without raising the mass-market price.

Are AI platform prices published or sales-gated?

Most standard model APIs and inference platforms publish full rate cards. The exceptions cluster in vertical enterprise-knowledge platforms — Harvey, for example, sells legal AI per seat via sales-led quotes with no public rate card — and in billing-infrastructure vendors that gate their own pricing.

How should I evaluate acquisition risk when choosing an AI platform vendor?

The tooling layer of this category has consolidated: corpus platforms including OpenPipe, OpenMeter, and Tavily were acquired within an 18-month window. Post-acquisition pricing has mostly held, but gating risk rises after a deal. For a critical vendor, negotiate a rate-lock clause and check the ownership structure before a long-term commitment.

Related product categories

Back to companies