What is it
AI platform pricing is how general-purpose AI platforms — model APIs, inference services, and multi-model hosting providers — package and charge for access to model capability. It is the pricing layer that sits beneath most of the AI software market: the frontier labs, the inference marketplaces, and the application platforms that resell or wrap model access all fall inside it.
This is the broadest product category in the corpus. 167 companies carry the ai-platforms product segment, spanning three rough strata. At the base are the frontier and open-model labs — OpenAI, Anthropic, DeepSeek, and Mistral AI — that expose per-token APIs. In the middle are inference platforms — Fireworks AI, Together AI, Groq, and Novita AI — that serve open-weight models and rent GPU capacity. At the application layer are platforms that embed models into a finished workflow — HeyGen for avatar video, Runway for generative video, Gumloop and Relevance AI for agent workflows.
Because those three strata serve fundamentally different buyers, the pricing diversity inside this one category is the widest in the corpus. The same segment contains companies billing per token, per credit, per seat, per action, and per resolution. A developer buying inference wants a public rate card and no minimum; an enterprise buying a legal-AI platform wants a per-seat quote; a prosumer buying video generation wants an all-inclusive monthly credit bundle. AI platform pricing is best understood not as one model but as a set of archetypes, each attached to a buyer profile.
The through-line is that the model itself is rarely the billing unit the customer sees. Tokens are the native cost driver of inference, but most application platforms abstract them away behind credits or seats so buyers can reason about spend without doing token math. Understanding a platform’s price means first identifying which archetype it belongs to, then reading how it translates underlying model cost into a unit its buyer will accept.
How it works
AI platform pricing resolves into five recurring archetypes. Which one a company picks is driven almost entirely by who it sells to.
| Archetype | Billing unit | Typical buyer | Corpus examples |
|---|---|---|---|
| Pure-usage API | Tokens, requests, GPU-hours | Developers | OpenAI API, DeepSeek, Groq, Fireworks AI |
| Freemium + credit subscription | Monthly credit pool + overage | Prosumers, SMB | HeyGen, Runway, Ideogram, Gumloop |
| Seat + credits hybrid | Per-user seat + metered pool | Teams | Cursor, Relevance AI, Intercom |
| Platform fee + usage (enterprise) | Committed minimum + overage | Enterprise | Together AI, Fireworks AI enterprise |
| Outcome-based | Per resolved task | Support / ops teams | Intercom Fin |
Pure-usage APIs publish a rate per input and output token, metered at consumption with free credit to start and volume discounts at scale. The spread is dramatic: DeepSeek’s V4-Flash starts at $0.0028/1M tokens on a cache hit, Groq prices Llama 3.1 8B at $0.05/$0.08 in/out, and OpenAI exposes the GPT-5.x API from $0.20/1M tokens. Cached and batch modes commonly cut the headline rate by 50% — Fireworks AI and Together AI both apply a 50% batch discount and layer per-hour GPU rates (H100 around $6–$7/hr) on top of serverless token pricing.
Credit subscriptions convert model cost into a monthly pool the buyer can budget. A worked example: HeyGen runs Free $0 → Creator $29 → Pro $49 → Business $149 (+$20/seat), each tier bundling a credit allotment, with a separate $5 pay-as-you-go API wallet. Runway starts at $12/mo bundling monthly credits that meter video, image, and audio generation, plus a Max tier with 9,500 rolling credits. The credit is a deliberate abstraction — it lets the vendor change underlying model economics without renegotiating, and lets the buyer avoid token math.
Seat + credits hybrids anchor predictable revenue on a per-user fee, then meter the AI-intensive work against a pool. Cursor is the canonical case; Relevance AI pairs seat tiers (Free, Pro $19/mo, Team $234/mo) with Actions usage and no-markup pass-through Vendor Credits. Outcome-based pricing is the rarest and most aligned: Intercom charges $0.99 per Fin AI resolution on top of $29–$132/seat platform tiers, billing only when the AI actually closes a ticket.
Companies using this
The 167 companies below all carry the ai-platforms product segment. They range from frontier model labs and open-model inference platforms to application-layer generation tools and vertical enterprise platforms — sort and filter the table to compare pricing models, billing units, and free-tier availability across the category.
Patterns observed
Pure-usage token pricing is the category floor, and it keeps falling. Every frontier and open-model lab in the corpus exposes a per-token rate, and the low end is dropping fast — DeepSeek’s cache-hit rate reset the floor while Groq competes on latency rather than a headline discount. The competitive pressure is so intense that cached, batch, and priority tiers have become standard — Fireworks AI and Together AI both advertise 50%-off batch modes, which is now table stakes rather than a differentiator.
Application-layer platforms abstract tokens into credits. The higher up the stack a platform sits, the less its buyer sees raw tokens. HeyGen, Runway, and Ideogram (Free, Plus $15, Pro $42, Team $20/user) all sell credit-metered subscriptions where the credit maps to a generated output — a video, an image — not to a token count. This abstraction is what lets these platforms run mass-market $15–$49 price points while insulating margin from underlying model volatility.
Multi-track pricing on one page is increasingly common. Rather than pick a single archetype, the larger platforms run several in parallel. Mistral AI prices two distinct surfaces from one page: flat-rate Vibe assistant subscriptions ($0–$24.99/user/mo) and a pure per-token developer API from $0.10/1M. OpenAI runs consumer subscriptions (Go $8, Plus $20, Pro $100 and $200, Business $20/seat) alongside its token API. The platform that serves both a developer and a prosumer buyer increasingly maintains a distinct pricing track for each.
The $200 prosumer ceiling has hardened. A cluster of consumer-facing platforms — OpenAI (ChatGPT Pro at $200), Anthropic (Claude Max from $100), Cursor, Perplexity, Google, and You.com — added a roughly $200/month top tier above the standard ~$20 Pro plan. This tier monetizes power users who would otherwise cap against lower plan limits, and it has been stable since it emerged. Following how these plans layer usage on a subscription base is the same problem covered in the guide to usage invoicing and billing cycles.
Counterexamples & variants
Not every AI platform publishes a rate card. The archetypes above assume public pricing, but the vertical enterprise strata breaks that assumption. Harvey, the legal-AI platform, sells per seat via sales-led quotes with no public rate card and reported seat minimums — the opposite of the developer-facing transparency norm. Enterprise-knowledge and regulated-vertical platforms consistently gate pricing because their buyers procure top-down and their deals are negotiated, not self-served. If you are comparing a vertical platform against a developer API on price transparency alone, you are comparing two different go-to-market cultures, not two pricing philosophies.
Outcome-based pricing is aspirational for most of the category, not standard. Intercom’s $0.99-per-resolution Fin pricing is held up as the model to emulate, but it remains rare because it requires an attributable, discrete outcome the buyer and vendor both agree happened. Most AI platforms cannot cleanly define such an outcome — what is the “outcome” of a token API or a video generator? — so they fall back to consumption or seats. Outcome pricing is a variant available to workflow platforms with a countable end state, not a template the whole category can adopt.
Credit systems vary from fully transparent to deliberately opaque. The credit archetype hides a wide range of honesty. Relevance AI passes through Vendor Credits at no markup, telling the buyer exactly what the underlying model cost was. At the other extreme, many application platforms set an internal, undisclosed credit-to-cost ratio so the buyer cannot reverse-engineer margin. Two platforms can both say “credits” and mean opposite things about transparency — the variant matters more than the label, which is why choosing and reading the underlying metric is worth understanding in its own right (see the guide to choosing the right usage metric).
What this means for buyers vs vendors
For buyers
Identify the archetype before you compare prices. A per-token rate and a per-seat quote are not comparable numbers, and a $29 credit subscription may cost far more or far less than a metered API depending on your volume — model your own usage first. For developer APIs, use a token calculator such as the Anthropic pricing calculator or the Cursor pricing calculator to translate a headline rate into a monthly bill at your actual volume, and always check whether cached and batch modes apply to your workload — a 50% batch discount changes the math entirely. For credit-based platforms, ask what the credit maps to and whether it is passed through at cost. For vertical enterprise platforms that gate pricing, expect a negotiated seat quote and factor acquisition risk into any multi-year commitment: several corpus platforms, including OpenPipe, OpenMeter, and Tavily, were acquired inside an 18-month window, and gating risk rises after a deal closes. A rate-lock clause is the cheapest insurance available.
For vendors
Your pricing archetype should follow your buyer, not your architecture. If you sell to developers, a public per-token rate with generous free credit and clear cached/batch tiers is the price of entry — opacity reads as a red flag in this segment. If you sell to prosumers or teams, credits let you set a legible monthly price while insulating margin from model-cost swings, but you owe the buyer clarity on what a credit buys. If you sell top-down into an enterprise or regulated vertical, seats and sales-led quotes remain viable, and gated pricing is defensible where deals are genuinely negotiated. The strategic risk across all archetypes is model-cost deflation: with the token floor still falling, any platform whose margin depends on a fixed markup over inference cost should assume that markup compresses. The durable positions are abstraction (credits, outcomes) and differentiation (latency, vertical depth), not a thin resale spread. For the mechanics of metering and invoicing usage at scale, the guide to usage invoicing and billing cycles covers the operational layer.
| Company | Product | Pricing model | Billing units | Free tier | Verified |
|---|---|---|---|---|---|
| 01.AI | Yi open-weight models + Yi API + enterprise vertical solutions | Yes | 2026-06-11 | ||
| 11x | Autonomous AI digital workers — Alice (outbound SDR) and Julian (inbound phone agent) | No | 2026-06-05 | ||
| 1X Technologies | NEO home humanoid robot & EVE enterprise robotics (RaaS) | No | 2026-06-14 | ||
| Abacus.AI | AI super-assistant (ChatLLM) plus an enterprise agentic AI platform | No | 2026-06-02 | ||
| Adept | ACT-1 — action-oriented AI agents that operate software via the UI | No | 2026-06-16 | ||
| Agility Robotics | Digit humanoid robot + Agility Arc cloud platform (Robots-as-a-Service) | No | 2026-06-14 | ||
| AI21 Labs | Jamba foundation models, Maestro orchestration & enterprise AI | Yes | 2026-06-11 | ||
| Aleph Alpha | PhariaAI sovereign-AI platform, specialized models & professional services | No | 2026-06-11 | ||
| Ambience Healthcare | Enterprise AI platform for clinical documentation and point-of-care coding | No | 2026-06-10 | ||
| Anthropic | Claude API (token-based) + Claude.ai consumer subscriptions (Free/Pro/Team/Enterprise) | Yes | 2026-07-06 | ||
| Anyscale | Managed Ray platform for distributed AI training, inference, and batch processing (RayTurbo, Anyscale Compute Units) | Yes | 2026-05-29 | ||
| Apptronik | Apollo general-purpose humanoid robot (RaaS + outright sale) | No | 2026-06-14 | ||
| AssemblyAI | Speech-to-Text & Audio AI APIs | Yes | 2026-07-06 | ||
| Athina AI | Collaborative AI development platform for building, testing, evaluating and monitoring LLM features | Yes | 2026-06-04 | ||
| Autodesk (Flow Studio, formerly Wonder Dynamics) | AI VFX automation platform (Flow Studio) | Yes | 2026-06-16 | ||
| Baichuan AI | Baichuan & medical M-series LLM APIs | Yes | 2026-06-11 | ||
| Bardeen | AI browser automation and workflow agents | Yes | 2026-06-10 | ||
| Baseten | ML inference infrastructure — dedicated GPU deployments, Model APIs, and Truss framework | Yes | 2026-05-29 | ||
| BentoML | BentoCloud — managed model-serving & inference platform | Yes | 2026-06-15 | ||
| Cartesia | Real-time voice AI platform (Sonic TTS, voice cloning, voice agents) | Yes | 2026-05-29 | ||
| Cerebras | Wafer-scale AI inference cloud and WSE hardware systems | Yes | 2026-05-30 | ||
| Character.ai | Consumer AI companion and roleplay chat platform | Yes | 2026-05-29 | ||
| Cognosys | Autonomous AI agents (rebranded Ottogrid, acquired by Cohere) | Yes | 2026-06-16 | ||
| Cohere | Command, Embed, Rerank APIs | Yes | 2026-05-29 | ||
| Composio | Tool-calling and integration infrastructure that connects AI agents to 1,000+ apps with managed auth and tool execution | Yes | 2026-06-10 | ||
| Copy.ai | GTM AI workflow platform | No | 2026-06-15 | ||
| Covariant | Covariant Brain — AI for autonomous warehouse robotic picking | No | 2026-07-14 | ||
| CrewAI | Multi-agent orchestration framework (OSS) + CrewAI AMP enterprise platform | Yes | 2026-06-10 | ||
| Daily | Real-time voice and video WebRTC APIs (Video SDK + Pipecat Cloud) | Yes | 2026-07-14 | ||
| Databricks (Mosaic AI) | Mosaic AI — enterprise GenAI & ML on the Data Intelligence Platform | Yes | 2026-06-15 | ||
| Deepgram | Usage-based speech-to-text, text-to-speech, and voice agent APIs | Yes | 2026-05-31 | ||
| DeepInfra | Serverless inference cloud — per-token LLM/embedding APIs, per-image and per-minute media models, per-hour on-demand GPU containers, and reserved DeepCluster GPU clusters | No | 2026-07-14 | ||
| DeepSeek | DeepSeek API (V4-Flash + V4-Pro models, 1M context) with token-based pricing and aggressive cache discounts | Yes | 2026-06-05 | ||
| Descript | AI-powered audio and video editing | Yes | 2026-05-31 | ||
| Dify | Dify Cloud + self-hosted LLM app development platform | Yes | 2026-07-14 | ||
| Dust | Enterprise AI agent deployment platform | Yes | 2026-06-24 | ||
| ElevenLabs | Voice AI platform across ElevenCreative, ElevenAgents, and ElevenAPI | Yes | 2026-06-30 | ||
| Essential AI | Enterprise foundation models & data-workflow automation | No | 2026-06-11 | ||
| EvenUp | AI Claims Intelligence Platform for personal injury law firms | No | 2026-06-16 | ||
| Exa | AI web search API for agents — search, contents, deep research, and monitoring endpoints billed per request | Yes | 2026-07-14 | ||
| Exscientia (now part of Recursion) | AI-driven drug discovery & design platform | No | 2026-06-16 | ||
| Fal | Generative-media inference platform — serverless per-output model APIs plus dedicated GPU compute | No | 2026-06-01 | ||
| Figure | General-purpose humanoid robots (Figure 03) & Helix AI | No | 2026-06-14 | ||
| Fireworks AI | Generative AI inference platform — serverless per-token, on-demand GPU, fine-tuning, batch API | Yes | 2026-05-30 | ||
| Freepik | AI creative suite — image, video, audio generation plus a 200M+ stock library | Yes | 2026-06-05 | ||
| Galileo | AI observability, evaluation, and guardrails platform for agents and LLM apps | Yes | 2026-06-04 | ||
| Gamma | AI presentations, documents and websites | Yes | 2026-06-11 | ||
| Genspark | All-in-one AI agent workspace (Super Agent, AI Slides/Sheets/Docs, image/video/audio generation) on a credit-based model | Yes | 2026-06-02 | ||
| Gladia | Speech-to-text & audio intelligence API | Yes | 2026-06-09 | ||
| Glean | Enterprise AI search and knowledge (Work AI) platform | No | 2026-05-31 | ||
| Gemini API & AI Studio | Yes | 2026-07-14 | |||
| Granola | AI notepad for back-to-back meetings | Yes | 2026-06-15 | ||
| Grok | xAI's consumer and business AI assistant | Yes | 2026-06-16 | ||
| Groq | GroqCloud — LPU-based ultra-low-latency inference API for Llama, GPT-OSS, Qwen, Whisper transcription, and Orpheus text-to-speech | Yes | 2026-07-14 | ||
| Gumloop | No-code AI workflow and agent automation platform billed on credits | Yes | 2026-06-30 | ||
| Harvey | Generative AI platform for legal and professional-services work | No | 2026-05-31 | ||
| Hebbia | Matrix — agentic AI for institutional knowledge work and document analysis | No | 2026-06-15 | ||
| Hedra | AI video, avatar, image, and audio generation platform (Hedra Studio + API) | Yes | 2026-06-04 | ||
| HeyGen | AI avatar and video generation platform | Yes | 2026-05-30 | ||
| Higgsfield | AI video and image generation platform with a credit-metered subscription | Yes | 2026-06-06 | ||
| Hippocratic AI | Safety-focused healthcare LLM — patient-facing AI agents for non-diagnostic clinical tasks | No | 2026-06-10 | ||
| Hugging Face | AI model hub, inference endpoints & compute | Yes | 2026-06-15 | ||
| Hume AI | Empathic Voice Interface (EVI) + Octave TTS + expression-measurement APIs | Yes | 2026-06-30 | ||
| Hyperbolic | GPU cloud marketplace & serverless AI inference | Yes | 2026-06-15 | ||
| Ideogram | Text-aware AI image generation platform | Yes | 2026-06-15 | ||
| Imbue | Reasoning-agent research lab and coding-agent tools (Sculptor) | No | 2026-06-16 | ||
| Inflection AI | Enterprise foundation models (Inflection 3.0) + Pi assistant | No | 2026-06-11 | ||
| Insilico Medicine | Pharma.AI generative drug-discovery platform + clinical pipeline | Yes | 2026-06-14 | ||
| Ironclad AI | AI-powered contract lifecycle management (CLM) | No | 2026-06-16 | ||
| Isomorphic Labs | AI-first drug discovery & design (Isomorphic Drug Design Engine) | No | 2026-06-14 | ||
| Janitor AI | Consumer AI character chat / roleplay platform | Yes | 2026-06-16 | ||
| Jasper | AI marketing content platform | No | 2026-05-31 | ||
| Jina AI | Search Foundation API (Embeddings, Reranker, Reader, DeepSearch, Classifier) | Yes | 2026-06-03 | ||
| LangChain | Agent orchestration frameworks + LangSmith platform | Yes | 2026-06-10 | ||
| Langfuse | Open-source LLM observability, evals, and prompt management | Yes | 2026-06-09 | ||
| Legora | Collaborative AI for lawyers — review, drafting, and research | No | 2026-06-06 | ||
| Lightning AI | Cloud GPU/CPU Studio compute platform for building, training, and serving AI models, billed by the second with a credit pool. | Yes | 2026-06-02 | ||
| Lindy | AI executive assistant (iMessage/SMS) — formerly AI agent-builder platform | No | 2026-06-10 | ||
| LiveKit | Open-source real-time (WebRTC) communications, LiveKit Cloud & Agents framework | Yes | 2026-06-30 | ||
| LMNT | Low-latency AI text-to-speech (TTS) API with voice cloning | Yes | 2026-06-04 | ||
| Luma AI | Dream Machine — text/image-to-video, image and audio generation (plus Genie 3D) | Yes | 2026-06-11 | ||
| Make | Visual, no-code automation (iPaaS) platform connecting 3,000+ apps and AI agents | Yes | 2026-06-11 | ||
| Manus | General AI agent that executes multi-step tasks autonomously in the cloud | Yes | 2026-06-02 | ||
| Maven AGI | Enterprise AI agent platform for customer support | No | 2026-06-11 | ||
| Mem | AI-powered personal memory workspace | Yes | 2026-07-14 | ||
| Mem0 | Memory layer for AI agents and applications | Yes | 2026-06-10 | ||
| Mercor | AI talent marketplace + enterprise data partnerships for frontier AI labs | No | 2026-07-14 | ||
| Midjourney | AI image and video generation via subscription with GPU-hour metering | No | 2026-05-29 | ||
| MiniMax | Foundation models, Hailuo video & per-token API | Yes | 2026-06-11 | ||
| Mintlify | AI-native developer documentation | Yes | 2026-06-15 | ||
| Mistral AI | Open and commercial LLM APIs | Yes | 2026-07-06 | ||
| Modal | Serverless compute and GPU platform — per-second billing for Python functions, batch jobs, and model serving | Yes | 2026-07-14 | ||
| Moonshot AI | Kimi assistant + Kimi/Moonshot open-weight LLM API | Yes | 2026-06-11 | ||
| Motion | Motion AI productivity platform (Pro AI, Business AI) | No | 2026-06-08 | ||
| MultiOn | Autonomous web-browsing AI agent API (wound down) | No | 2026-06-10 | ||
| Murf AI | AI voice / text-to-speech platform (Murf Studio app + Murf API) | Yes | 2026-06-01 | ||
| n8n | Fair-code workflow automation platform for technical teams, billed by monthly workflow executions | Yes | 2026-06-02 | ||
| Nomic | Nomic Platform (AEC agentic workflows) + Atlas data-exploration app + Nomic Embed embedding/Developer API | Yes | 2026-06-04 | ||
| Nooks | AI sales platform — parallel dialer, AI SDR, and coaching | No | 2026-06-05 | ||
| Notion AI | AI workspace, agents, and knowledge management | Yes | 2026-06-15 | ||
| Novita AI | Pay-as-you-go AI cloud: 200+ model inference APIs, on-demand GPUs, and per-second agent sandboxes under one API | Yes | 2026-07-06 | ||
| Observe.AI | Agentic CX platform — contact-center AI agents, conversation intelligence & auto-QA | No | 2026-06-09 | ||
| OctoAI | Generative AI inference platform (acquired by NVIDIA, sunset Oct 2024) | No | 2026-06-15 | ||
| OpenAI | ChatGPT consumer subscriptions + GPT-5.x API with token-based usage billing | Yes | 2026-06-30 | ||
| OpenPipe | OpenPipe fine-tuning and hosted inference platform (small specialized models / RL for agents) | Yes | 2026-06-04 | ||
| Paige AI | FDA-cleared AI for cancer pathology — clinical diagnostics + pharma/life-sciences foundation models | No | 2026-06-10 | ||
| Pebblely | AI product-photography tool that generates marketing images from a product photo | No | 2026-06-07 | ||
| Perplexity AI | AI-native answer engine with citations and multi-model search | Yes | 2026-05-29 | ||
| Phind | AI developer search engine and coding assistant (shut down January 2026) | Yes | 2026-06-08 | ||
| Physical Intelligence | Robotics foundation models (Vision-Language-Action policies for robots) | No | 2026-06-14 | ||
| Pi | Pi — personal, emotionally intelligent AI assistant (consumer app) | Yes | 2026-06-16 | ||
| Pipedream | Workflow automation and integration platform for developers | Yes | 2026-06-16 | ||
| Playground | AI image generation and graphic-design studio with a monthly credit pool | Yes | 2026-06-04 | ||
| PlayHT | Text-to-speech & voice cloning API (PlayAI) | Yes | 2026-06-09 | ||
| Poe | Multi-model AI chat subscription (by Quora) | Yes | 2026-06-16 | ||
| PolyAI | Enterprise voice AI assistants for contact centers | No | 2026-06-09 | ||
| Poolside | AI coding foundation model | No | 2026-06-16 | ||
| Predibase | Fine-tuning & serving platform for open-source LLMs | Yes | 2026-06-15 | ||
| Rad AI | Generative AI for radiology — report drafting (Reporting/Omni), automated impressions, and follow-up management (Continuity) | No | 2026-06-10 | ||
| Recraft | AI image and vector generation studio plus a per-image generation API | Yes | 2026-07-14 | ||
| Recursion | AI-enabled drug discovery platform (Recursion OS) — pharma partnerships, internal pipeline & NVIDIA-powered compute | No | 2026-06-10 | ||
| Reka AI | Natively multimodal models (Spark, Edge, Flash, Core) + Research & Vision APIs | Yes | 2026-06-11 | ||
| Relevance AI | No-code platform for building AI agents and multi-agent 'AI Workforces' for sales, marketing, and operations teams. | Yes | 2026-07-14 | ||
| Replicate | Cloud platform for running, fine-tuning, and deploying AI models via REST API | Yes | 2026-05-30 | ||
| Resemble AI | AI deepfake detection & watermarking + voice generation APIs | No | 2026-07-14 | ||
| Retell AI | Conversational voice-agent API platform | No | 2026-07-14 | ||
| Rev AI | Pay-as-you-go speech-to-text, transcription, and audio-intelligence APIs | Yes | 2026-06-04 | ||
| Rewind.ai (the original Rewind AI rebranded to Limitless, acquired by Meta) | AI tools aggregator (token-balance) — on the domain once home to the Rewind personal-memory app | Yes | 2026-06-15 | ||
| Robin AI | AI legal contract review, drafting & a legal-data API | No | 2026-06-06 | ||
| Roboflow | Computer-vision platform (dataset management, model training, deployment) | Yes | 2026-07-14 | ||
| Rox | AI agent swarm for sales reps (AE copilot) | Yes | 2026-06-05 | ||
| RunPod | GPU cloud marketplace — Secure Cloud and Community Cloud Pods, Serverless endpoints, and persistent storage | No | 2026-07-14 | ||
| Runway | Video generation and AI editing | Yes | 2026-06-24 | ||
| SambaNova | SambaNova Cloud inference API & RDU AI systems | Yes | 2026-06-15 | ||
| Sana AI | Enterprise AI assistant (Sana Agents) and AI learning platform (Sana Learn) | Yes | 2026-06-15 | ||
| Sanctuary AI | Phoenix general-purpose humanoid robot & Carbon AI control system | No | 2026-06-14 | ||
| Sarvam AI | Sovereign Indic LLM, speech & translation APIs | Yes | 2026-06-11 | ||
| Shield AI | Hivemind autonomy software, V-BAT & X-BAT autonomous aircraft | No | 2026-06-14 | ||
| Skydio | Autonomous drones, docks & flight-autonomy software for defense, public safety & enterprise | No | 2026-06-14 | ||
| Speechmatics | Speech-to-text and text-to-speech APIs with per-hour usage pricing | Yes | 2026-07-06 | ||
| Stability AI | Brand Studio creative platform and open generative media models | Yes | 2026-06-11 | ||
| Suki AI | Ambient clinical AI assistant for healthcare (Suki Assistant) + embeddable Suki Platform SDK/API | No | 2026-06-10 | ||
| Suno | AI music generation | Yes | 2026-05-31 | ||
| Synthesia | Enterprise AI video generation | Yes | 2026-05-31 | ||
| Synthflow AI | No-code AI voice-agent builder | No | 2026-06-24 | ||
| Tavily | Tavily Search API | Yes | 2026-06-03 | ||
| Tempus | Precision-medicine platform — genomic diagnostics, multimodal clinical data licensing & oncology AI apps (NASDAQ: TEM) | No | 2026-06-10 | ||
| Thomson Reuters (CoCounsel) | CoCounsel — legal generative-AI assistant (formerly Casetext) | No | 2026-06-16 | ||
| Together AI | AI Acceleration Cloud — serverless inference, dedicated endpoints, GPU clusters, Code Sandbox, fine-tuning | Yes | 2026-07-14 | ||
| Trigger.dev | Background jobs and workflow orchestration for developers | Yes | 2026-06-16 | ||
| Twelve Labs | Video understanding foundation models (Marengo for search/embeddings, Pegasus for analysis) delivered as a usage-metered API | Yes | 2026-06-02 | ||
| Typeface | Arc enterprise marketing AI platform | No | 2026-06-16 | ||
| Udio | AI music generation | Yes | 2026-06-11 | ||
| Uniphore | Business AI Cloud — enterprise conversational AI & agentic automation | No | 2026-06-09 | ||
| Vapi | Voice AI infrastructure for developers | No | 2026-06-09 | ||
| Vectara | Enterprise RAG-as-a-Service and agent platform for trusted, grounded, auditable AI | No | 2026-06-02 | ||
| Vellum | Personal AI assistant (ex LLM application development platform) | Yes | 2026-06-10 | ||
| Viz.ai | AI-powered care coordination for time-sensitive disease — stroke, aneurysm, PE, cardiac and more (Viz Neuro/Cardio/Vascular/Pulmonary suites) | No | 2026-06-10 | ||
| Voyage AI | Embedding and reranker models (text, code, multimodal) for retrieval and RAG | Yes | 2026-06-04 | ||
| Weaviate | AI-native vector database (open-source core + Weaviate Cloud managed serverless, dedicated/Enterprise Cloud, BYOC) | Yes | 2026-07-06 | ||
| Weights & Biases | MLOps experiment tracking, W&B Weave LLM observability/evals, Models registry, and Serverless Inference | Yes | 2026-07-14 | ||
| Writer | Enterprise agentic AI platform (Palmyra models, WRITER Agent) | No | 2026-06-15 | ||
| xAI | Grok API and agentic AI stack | Yes | 2026-07-14 | ||
| Yellow.ai | Conversational CX automation platform | Yes | 2026-06-11 | ||
| You.com | Web search, contents, research, and finance-research APIs for AI systems | Yes | 2026-06-01 | ||
| Zapier | Workflow-automation (iPaaS) platform connecting 9,000+ apps, with separately-metered AI Agents and Chatbots add-ons | Yes | 2026-06-30 | ||
| Zhipu AI | GLM foundation models, per-token API, and GLM Coding Plan | Yes | 2026-06-11 |
Explore this theme in the knowledge graph
FAQ
What is an AI platform?
An AI platform is a general-purpose product that delivers AI model capability — a model API, an inference service, or a multi-model hosting layer — that other software and teams build on. In this corpus, 167 companies carry the ai-platforms product segment, making it the largest single category.
How do AI platforms price their products?
There is no single model. Frontier labs and inference platforms bill per token (OpenAI's GPT-5.x API from $0.20/1M, DeepSeek from $0.0028/1M on a V4-Flash cache hit); application-layer platforms sell credit-based subscriptions (HeyGen $29, Runway $12); and vertical or agent platforms use seats, actions, or outcomes. The billing unit follows the buyer, not the underlying technology.
How much does a frontier model API cost per million tokens?
The spread is enormous. On the low end, DeepSeek's V4-Flash starts at $0.0028/1M tokens on a cache hit and Groq's Llama 3.1 8B is $0.05/$0.08 in/out. On the high end, Anthropic's Fable 5 is $10/$50 and OpenAI's flagship API tiers run several dollars per million. Cached and batch modes routinely cut the headline rate by 50%.
What is the $200 AI platform tier?
Several consumer-facing AI platforms — OpenAI (ChatGPT Pro at $200), Anthropic (Claude Max from $100), Cursor, Perplexity, Google, and You.com — added a roughly $200/month top tier above the standard ~$20 Pro plan. It monetizes power users who cap against lower plan limits without raising the mass-market price.
Are AI platform prices published or sales-gated?
Most standard model APIs and inference platforms publish full rate cards. The exceptions cluster in vertical enterprise-knowledge platforms — Harvey, for example, sells legal AI per seat via sales-led quotes with no public rate card — and in billing-infrastructure vendors that gate their own pricing.
How should I evaluate acquisition risk when choosing an AI platform vendor?
The tooling layer of this category has consolidated: corpus platforms including OpenPipe, OpenMeter, and Tavily were acquired within an 18-month window. Post-acquisition pricing has mostly held, but gating risk rises after a deal. For a critical vendor, negotiate a rate-lock clause and check the ownership structure before a long-term commitment.
Related product categories
- AI Coding Product PricingPricing for products whose primary surface is AI-assisted coding — IDEs, completion engines, and review agents.
- Developer Tools PricingPricing models used by tools sold to developers — IDEs, CLIs, libraries, voice-to-code, and adjacent products.
- AI Infrastructure & Cloud PricingPricing for AI compute infrastructure — GPU clouds, serverless inference, and training platforms.
- Data Platform PricingPricing for data platforms — scraping, enrichment, search API, and knowledge-graph vendors.
- Vertical SaaS PricingPricing for vertical SaaS products — AI software purpose-built for a specific industry (legal, healthcare, sales, marketing).
- PaaS PricingPricing for platform-as-a-service products that abstract away the underlying infrastructure and bill for higher-level units.
- Customer Service Platform PricingPricing for customer service software platforms — ticketing, chat, automation, and AI agent products.
- Horizontal SaaS PricingPricing for horizontal AI SaaS — productivity and workflow products sold across industries rather than to one vertical.
- Observability Platform PricingPricing for LLM and ML observability platforms — tracing, evaluation, and monitoring of model behavior in production.
- Fintech AI PricingPricing for AI-era fintech products — billing infrastructure, accounting automation, and financial operations platforms.
- Security AI PricingPricing for AI-powered security products — covering code security, voice fraud detection, SOC automation, and threat analysis.
- LLM Observability PricingPricing for platforms purpose-built to observe, debug, and optimize LLM application behavior — logging prompts, responses, latency, and cost.