Infrastructure AI tools
AI Infrastructure platforms provide the computing resources, storage, and networking capabilities needed to develop, train, and deploy AI models at scale. These platforms offer specialized hardware like GPUs and TPUs, along with the tools to manage and optimize AI workloads efficiently.
19 verified AI-first sites in Infrastructure.
Distributed computing platform for AI development. Features Ray-based infrastructure, scalable ML training, and deployment tools. Includes monitoring and optimization capabilities. Supports enterprise AI operations.
Pricing: Free tier. Enterprise: Custom pricing.
Autonomous cloud operations platform that scales, heals, and optimizes production workloads without human-triggered runbooks. Acts on metrics across Kubernetes, containers, VMs, and serverless. Built for self-driving cloud, not carrier RAN/OSS stacks.
Pricing: Enterprise pricing based on scale
Enterprise-grade LLM infrastructure and development platform. Features the popular Haystack framework for building production-ready AI applications, LLM orchestration, and NLP solutions. Includes model deployment, optimization tools, and comprehensive AI application development capabilities. Helps organizations build, deploy, and scale AI-powered applications with production-ready infrastructure.
Pricing: Open source. Cloud: Starting at $500/month. Enterprise: Custom pricing.
Memory layer for AI applications that enables persistent context across conversations. Helps AI agents remember user preferences and past interactions to provide more personalized experiences. Features include fact extraction, memory processing, and efficient retrieval with 26% higher accuracy and 90% lower token usage than standard approaches.
Pricing: Individual: $8.33/month (billed annually) or $14.99/month. Teams: Custom pricing with dedicated support.
Web data and extraction APIs built for AI agents. Scraping, search, and structured data endpoints so agents can pull live page content into workflows without maintaining their own crawlers.
Pricing: Free trial available; paid plans from about $40/month.
Cloud browser infrastructure built for AI agents that need reliable web interaction beyond static APIs. Provides managed browser sessions, automation runtime support, and observability for debugging multi-step agent workflows at scale. Designed for teams building production-grade agents that handle authenticated and dynamic web tasks.
Pricing: Free plan available. Developer: $20/mo. Startup: $99/mo. Scale: Custom.
Managed browser runtime for AI agents on Cloudflare's network, supporting automation through Playwright, Puppeteer, and related tooling. Enables browser session execution with usage controls, concurrency support, and production deployment patterns for web-task agents. A fit for teams that need scalable, infrastructure-level browser execution in agent workflows.
Pricing: Workers Free and Paid tiers available; usage-based browser-hour and concurrency pricing applies.
Web data platform for AI agents that combines crawling, extraction, search, and interaction endpoints in one workflow. Returns LLM-ready content formats like markdown and structured JSON while supporting dynamic pages and agent-style research tasks. Useful for teams building production retrieval and web-intelligence pipelines for autonomous systems.
Pricing: Free tier available. Paid usage plans available; see site for current limits and pricing.
Cloud runtime infrastructure built for AI agents and coding assistants that need secure, disposable execution environments. Provides sandboxed containers for running generated code, browser tasks, and tool-using workflows with fast startup and API-first control. Best fit for teams shipping agent products that require reliable code execution and isolation in production.
Pricing: Free tier available. Usage-based paid plans and enterprise options available.
Open-source development runtime platform increasingly used as infrastructure for AI coding agents and autonomous developer workflows. Offers fast, reproducible dev environments with remote execution and API control, enabling agents to run tasks in isolated workspaces at scale. Strong fit for engineering teams building AI-native coding automation and agent-powered dev tooling.
Pricing: Open-source core. Cloud and enterprise offerings available; see site for current pricing.
AI infrastructure company building on open-source SGLang for inference and Miles for large-scale RL post-training. End-to-end platform for training, fine-tuning, reinforcement learning, and production inference with managed tooling. Founded by SGLang creators; $100M seed led by Accel and Spark Capital with NVIDIA and AMD backing.
Pricing: Managed infrastructure; contact for pricing and early access.
Open-source cloud browser API for AI agents and automation. Sessions API spins up Puppeteer, Playwright, or Selenium browsers with CAPTCHA solving, proxies, and session replay. Sub-second starts; 800k+ browser hours served. Free tier with monthly credits.
Pricing: Free tier; Starter $29/mo; Developers $99/mo; usage-based credits.
Cloud browser infrastructure for AI agents, scraping, and testing. API-managed sessions with Playwright/Puppeteer CDP, anti-bot proxies, and structured extract endpoints. Scales to 1,000+ concurrent isolated browser sessions.
Pricing: Usage-based API pricing; see site.
Open-source orchestration control plane for GPU fleets across AWS, GCP, Kubernetes, and SSH bare-metal. Unified CLI for dev environments, distributed training tasks, and OpenAI-compatible inference services with vLLM or SGLang. dstack Sky hosted option with GPU marketplace.
Pricing: Open source; dstack Sky and Enterprise on site.
Agentic AIOps platform that correlates IT events, reduces alert noise, and drives incident response across enterprise infrastructure. Used by IT operations teams for automated root-cause and ticket enrichment. Not a telecom OSS/BSS product.
Pricing: Enterprise subscription by event volume and modules.
Run AI inference on Cloudflare's global edge network via Workers. Serverless access to LLMs and other models with low-latency bindings, no GPU cluster to manage. Fits agents and apps already on Workers that need inference next to users.
Pricing: Workers Free and Paid; usage-based neuron pricing-see Cloudflare docs.
Infrastructure for autonomous AI agents: sandboxes, compute, and orchestration purpose-built for agent workloads. Helps teams run and scale agentic systems without assembling general-purpose cloud pieces. Focused on production agent runtime needs.
Pricing: Platform plans; see Blaxel pricing or contact sales.
AI agent accelerator with secure Devbox sandboxes for code agents. Fast micro-VM isolation, snapshots, benchmarks (including SWE-Bench), and managed scale for thousands of parallel agent environments. Enterprise VPC and compliance options.
Pricing: Credits and enterprise plans; see Runloop for current offers.
Enterprise memory layer for AI agents using context graphs for fast, governed retrieval across long-running conversations. Sub-200ms lookups with SOC 2 controls for production agent apps. Peer to other agent-memory infrastructure used by LLM products.
Pricing: Free tier and paid plans; see Zep pricing.