A category for the underlying compute, model serving, vector, analytics and data platforms that power AI products. Use it when evaluating where to host inference, store embeddings, fine-tune models, embed analytics or run batch and sandbox workloads at production scale.
01 · Find the fit
Start with the job
Run inference for LLMs, vision or audio models at production scale
Store and search vectors for retrieval-augmented applications
Embed analytics, dashboards or interactive visualizations into products
Fine-tune, batch-process or run sandboxes on elastic GPU compute
Route requests across multiple model providers through a single API
02 · Inspect the details
Compare the tradeoffs
Inference latency and throughput characteristics
Deployment model: managed cloud, self-hosted or hybrid
Specialization: general LLM serving, vector database, analytics or visualization
Customization options for fine-tuning, training and routing
Integration surface: SDKs, REST APIs, MCP and composable components
03 · Verify before buying
Follow the evidence
Provider-specific data retention rules for certain model classes
Cost differences between real-time, batch and async workloads
Tier-dependent feature availability and rate limits
Cold start and autoscaling behavior under spiky traffic
Lock-in risk from proprietary model stacks or routing layers
Common evaluation signals
AI infrastructureFirst-party researchedgenerative AILLM inferenceimage modelsmodel fine-tuning
Profiles in this category
Find your best fit.
Browse every catalogued ai infrastructure profile, then use the evidence and freshness details to judge how much decision context is currently available.
Tool New York, Paris, London, Amsterdam, Berlin, Dubai, Denver, Madrid, Munich, Seoul, Singapor
Dataiku is an enterprise AI platform that unifies people, orchestration, and governance to connect data, ML, LLMs, and agents as one system. It delivers measurable business outcomes via Self-Managed or SaaS deployment with compliance and audit-ready oversight.
AI infrastructureFirst-party researchedenterprise AI platformAI orchestration
PricingSee official pricing
AccessUnified enterprise AI platform for agents, models, and analytics with governance
Fireworks AI is a generative AI inference platform processing 40T+ tokens daily, offering blazing-fast serverless and on-demand deployment of open-source LLMs and image models with fine-tuning, OpenAI/Anthropic-compatible APIs, and Batch processing.
AI infrastructureFirst-party researchedgenerative AILLM inference
PricingServerless pricing is pay per token with high rate limits and postpaid billing; new users · On-Demand Deployments are billed per GPU second, with no extra charges for start-up times · Cached input tokens are priced at 50% and batch inference at 50% of serverless pricing for · 50% lower cost compared to typical Serverless (synchronous) API pricing
Groq delivers fast, low-cost AI inference through its custom LPU silicon, pioneered in 2016 as the first chip purpose-built for inference. Its GroqCloud platform serves 3M developers and enterprises via tokens-as-a-service pricing, OpenAI-compatible APIs, and global data center deployment.
AI infrastructureFirst-party researchedLPU ArchitectureGroqCloud
PricingTokens-as-a-Service on-demand pricing with per-million-token input and output rates · GPT OSS 120B (128k context) priced at $0.15 input / $0.60 output per million tokens at 500 · Llama 3.1 8B Instant (128k context) priced at $0.05 input / $0.08 output per million token · Llama 3.3 70B Versatile (128k context) priced at $0.59 input / $0.79 output per million to
AccessGroqCloud is the cloud platform/console for running inference on Groq's LPU-based stack.
Qualys is a cloud-delivered cybersecurity vendor (NASDAQ: QLYS) with 25 years of innovation, offering the Enterprise TruRisk Platform with AI agents, CNAPP, and unified vulnerability, endpoint, cloud, and compliance management for global enterprises.
AI infrastructureFirst-party researchedCybersecurityAI agents
PricingFlexible Subscription Using Qualys Units (QLU)
AccessEnterprise TruRisk Platform — a unified suite of integrated cybersecurity apps
Tool Prinsengracht 769A in Amsterdam, the Netherlands
Weaviate is an open-source, AI-native database platform unifying vector search, RAG, and memory. It offers built-in embeddings, a Query Agent, multimodal support, and over 20M downloads, serving startups through enterprises with deployment-agnostic scaling.
AI infrastructureFirst-party researchedAI databasevector database
PricingFlex plan starts at $45/mo, pay-as-you-go, no commitment, with 99.5% uptime and standard s · Premium plan starts at $400/mo, prepaid contract, with up to 99.95% uptime and enterprise · Weaviate Query Agent: free to try with 1,000 requests/month; $30 per organization monthly · Weaviate Cloud is free to start across the entire product suite
AccessWeaviate is an AI-native database platform combining vector search, RAG, and memory in one
DeepInfra is an AI inference cloud offering cost-effective, scalable, production-ready machine-learning model deployment via developer-friendly APIs. It hosts 100+ models across text, image, speech, and video categories from families including Claude, Llama, DeepSeek, and Qwen, with pay-as-you-go pricing.
AI infrastructureAI inference cloudMachine learning infrastructureServerless model deployment
PricingDeepSeek-V4-Flash text-generation model priced at $0.09/M input tokens and $0.18/M output · DeepSeek-V4-Pro text-generation model priced at $1.30/M input tokens and $2.60/M output to · Kimi-K2.7-Code text-generation model priced at $0.74/M input tokens and $3.50/M output tok · NVIDIA-Nemotron-3-Ultra-550B-A55B text-generation model priced at $0.50/M input tokens and
AccessExposes a Simple API for accessing AI models
Drop-in memory infrastructure for AI agents and apps that adds persistent context across sessions. Offers Python and Node.js SDKs, 25+ integrations, SOC-2/HIPAA compliance, and vector-plus-graph retrieval.
AI infrastructureAI memory layerMemory infrastructurePersistent context
PricingCustom pricing available with the ability to book a call and a 'Start Free' option · Free tier available
Modal is a serverless cloud platform for AI and data teams offering sub-second cold starts, instant GPU autoscaling, and pay-per-use compute for inference, training, batch processing, and sandboxes.
AI infrastructureFirst-party researchedserverless platformhigh-performance AI infrastructure
PricingPay-per-use compute billed by the CPU cycle; customers never pay for idle resources. · Nvidia H100 GPU tasks are priced at $0.001097 per second. · Starter plan is $0/month plus compute, with $30/month in free credits, up to 3 workspace s · Team plan is $250/month plus compute, with $100/month in free credits, unlimited seats, 10
RunPod is an AI infrastructure platform offering on-demand GPUs and serverless compute across 31 global regions. It supports the full AI lifecycle—experiment, train, fine-tune, deploy, and scale—with 30+ GPU SKUs and three deployment modes.
AI infrastructureGPU cloudserverless computeon-demand GPUs
PricingCompute costs up to 90% lower than traditional cloud providers · Per-second billing for H100, A100, and RTX GPUs · Reduces compute costs by as much as 90% · Runpod Serverless pricing now uses Flex and Active worker types, with Active workers offer
Amplitude is an AI Analytics Platform offering AI Agents, AI Feedback, MCP, and Agent Analytics alongside Product Analytics, Web Experimentation with A/B testing, and Session Replay for startups and enterprises.
AI infrastructureFirst-party researchedAI analytics platformProduct analytics
PricingFree analytics tools available for startups
Bugcrowd is a crowdsourced cybersecurity platform combining bug bounty, pen testing, and vulnerability disclosure programs. It offers AI-Powered Security Intelligence with CrowdMatch™ triage and the Savant AI strategy, serving enterprise and regulated industries.
AI infrastructureFirst-party researchedcrowdsourced cybersecuritybug bounty platform
PricingBug bounty rewards typically range from hundreds to thousands of dollars depending on impa
Chroma is an open-source AI search infrastructure platform offering serverless vector, full-text, regex, and metadata search. Chroma Cloud provides SOC 2 Type II compliance, CMEK, BYOC, SDKs for TypeScript/Python/Rust, and scalable multi-tenant support.
AI infrastructurevector databaseopen-sourceserverless
PricingUsage-based pricing: $2.50/GiB written, $0.33/GiB per month storage, $0.0075/TiB queried,
AccessClient SDKs available for TypeScript, Python, and Rust
DataRobot is an enterprise Unified Agent Workforce Platform offering an Agentic AI Platform spanning Agentic, Generative, and Predictive AI with Governance, Observability, and Foundation capabilities. It serves Government, Oil & Gas, Life Sciences, Financial Services, and Manufacturing with Finance and Supply Chain agents.
AI infrastructureFirst-party researchedAgentic AI PlatformUnified Agent Workforce Platform
PricingSee official pricing
AccessMain offerings include Enterprise AI Suite and an Agentic AI Platform with AI Apps & Agent
Flourish is a no-code, web-based data visualization platform launched in 2018 that lets users build interactive charts, maps, and data stories. Now part of the Canva family and operated from London, it serves newsrooms, marketers, educators, and enterprises with 50+ chart types and broad embed integrations.
AI infrastructureFirst-party researchedNo-code data visualizationInteractive charts and maps
PricingTiered plans available for individuals, teams, and enterprises, including a Free plan and
AccessWeb-based platform that adapts to any screen and works across devices
H2O.ai is an end-to-end GenAI and machine learning platform offering h2oGPTe enterprise GenAI, H2O-3 open-source ML, Danube3 openweight SLMs, and Driverless AI automated ML, with airgapped, on-premises, and FedRAMP High deployments for financial, telecom, public sector, and federal organizations.
AI infrastructureFirst-party researchedEnd-to-end GenAI platformOpen-source machine learning (H2O-3)
PricingSee official pricing
AccessCourses and certifications are available on the H2O.ai Website, YouTube, Udemy, and Course
KNIME is a free, open-source visual data analytics platform supporting ETL, predictive AI, and data-aware agent building. It connects to 300+ data sources and leading AI models, enabling enterprise-grade deployment with ISO 27001-certified security.
AI infrastructureFirst-party researchedOpen sourceVisual workflow builder
PricingPro Plan and Team Plan available · Pro Plan and Team Plan offered · A free trial of the KNIME Team plan is available for sharing workflows as data apps. · KNIME Team plan costs €99/month after a 1-month free trial.
AccessExtensions are made available through the KNIME Analytics Platform.
Tool Berlin, Germany (Product & Engineering) and San Francisco, CA (GTM)
Langfuse is an open-source AI engineering platform (by ClickHouse) for tracing, evaluating, and improving LLM applications. It offers LLM observability, prompt management, evaluation, and metrics, processing 10+ billion observations monthly.
AI infrastructureFirst-party researchedLLM observabilityprompt management
PricingFree tier available ('Start free'); managed cloud launched via 'Launch App'. · Hobby plan is Free with no credit card required, 50k units/month, 30 days data access, 2 u · Core plan is $29/month with 100k units/month included (additional $8/100k units, lower wit · Pro plan is $199/month with 3 years data access, data retention management, unlimited anno
OpenRouter is a unified AI infrastructure platform providing access to 400+ models from 70+ providers through a single OpenAI-compatible API. It handles fallbacks, billing, and security for chat, image, embedding, and transcription workloads.
AI infrastructureFirst-party researchedunified LLM interfacemodel routing
PricingCredit-based pay-as-you-go model with no subscriptions; credits (e.g., $10 or $99 top-ups · Three plans offered: Free, Pay-as-you-go, and Enterprise · Pay-as-you-go tier carries a 5.5% platform fee, with fee discounts available · Small fee charged when purchasing credits; model inference pricing is passed through from
AccessConnects to 70+ providers (4 free providers and 25+ free models available)
Pinecone is a fully managed vector database that searches billions of items for similar matches in milliseconds, with automatic indexing and sub-100ms writes. It offers dense, sparse, and full-text indexes via one API, with plans from a free Starter to a $50/month Standard and Enterprise BYOC options.
AI infrastructureFirst-party researchedvector databasesimilarity search
PricingStarter plan is free, aimed at trying out and small applications · Builder plan is a flat $20/month, for solo developers and small teams · Standard plan has a $50/month minimum usage, includes a 3-week trial with $300 credits, fo · Free to create an initial index, then pay-as-you-go when scaling.
AccessAvailable on AWS, GCP, and Microsoft cloud marketplaces
Redash is an open source data visualization and dashboarding tool that connects to SQL, NoSQL, Big Data, and API sources. It provides a powerful SQL editor, drag-and-drop dashboards, scheduled refreshes, alerts, and sharing to help companies become data driven.
AI infrastructureFirst-party researchedopen sourcedata visualization
PricingSee official pricing
AccessSupports SQL, NoSQL, Big Data and API data sources
Rows is an AI-powered all-in-one spreadsheet for teams that analyzes data from 50+ sources using plain language, with no code, SQL, or formulas required. It combines built-in integrations, automated refreshes, shareable dashboards, and tiered pricing from free to Enterprise.
AI infrastructureFirst-party researchedAI spreadsheetNo-code data analysis
PricingFree plan costs $0 · Plus plan: $8/month per user on monthly billing or $6/month per user on annual billing · Pro plan: $79/month + $8/month per user on monthly billing or $59/month + $6/month per use · Listed prices are net, exclusive of any applicable sales tax or VAT
AccessInstallable as a Progressive Web App via Chrome on desktop or Safari on mobile.
Tool 228 Park Ave S PMB 216542, New York, New York 10003-1502, US
Saturn Cloud is an enterprise AI platform providing a white-labeled control plane for GPU clouds with multi-tenant isolation, day-2 support, and integrated billing, deployable across multiple cloud providers.
AI infrastructureFirst-party researchedGPU cloudwhite-label control plane
PricingBilling appears as a line-item on the customer's cloud bill, with no separate invoicing · Offers a free starting tier ("Start for free") · Saturn Cloud offers a free starting option
AccessSupports Bare Metal Kubernetes, Slurm, and Managed Inference workloads
Semgrep is an AI-assisted application security platform offering SAST, SCA, and Secrets Detection. Its products—Code, Supply Chain, Secrets, Guardian, and Multimodal—help Fintech and SaaS & Cloud teams find and fix code issues, dependency vulnerabilities, and AI-generated code risks.
AI infrastructureFirst-party researchedSASTSCA
PricingA free trial option is offered ("Try for free") · Free trial available · A free trial is available alongside a paid demo/book-a-call option. · Offers a free trial option alongside demo bookings
AccessAppSec Platform providing SAST, SCA, and Secrets detection.
Sisense is an AI-powered embedded analytics platform on a Cloud Composable architecture, delivering pro-code, low-code, and no-code flexibility. It equips developers and app creators with Compose SDK, flexible APIs, Sisense Intelligence AI features, and predictive analytics to embed into customer products.
AI infrastructureFirst-party researchedembedded analyticsAI-powered analytics
PricingCompose SDK offers a 7-day free trial
AccessPlatform capabilities include Cloud, Composable, Embedded analytics, Connectivity, Data vi