Services
Curated infrastructure, data, security, and API services that power production AI agents. Not a marketplace catalog — a shortlist of tools we would actually wire into a stack.
Showing 34 of 34 services
Modal: Serverless GPU Compute
Run training, fine-tuning, and inference workloads on serverless GPUs without managing infrastructure.
Robusta: AI Agent Security Audit
External security review for autonomous agent tool permissions, prompt injection risks, and deployment guardrails.
Pinecone: Managed Vector Database
Managed vector database for RAG, agent memory, and semantic search at scale.
Firecrawl: Web-to-Markdown for Agents
Turn any website into clean, structured Markdown or structured data for agent pipelines.
Scale AI: Data Labeling and RLHF
Human-in-the-loop data labeling and RLHF for training and fine-tuning custom models.
RunPod: GPU Cloud
Rent GPU compute by the hour or second for inference, training, and fine-tuning workloads.
ElevenLabs: AI Voice
Text-to-speech, voice cloning, and real-time voice agents for interactive AI experiences.
LiteLLM: Universal LLM Gateway
One API for 100+ LLM providers with load balancing, fallbacks, rate-limiting, and spend tracking.
Browserless: Headless Browser as a Service
Managed headless browsers for web scraping, form-filling, screenshots, and agent-driven browser tasks.
Composio: Tooling Layer for Agents
Pre-built, authenticated tool integrations so agents can act across SaaS apps, APIs, and databases.
Tavily: Search API for AI Agents
Real-time web search optimized for LLMs — returns clean, sourced answers instead of raw SERPs.
OpenRouter: Unified API for 100+ LLMs
One API key, one endpoint, 100+ models. OpenRouter handles routing, fallbacks, and model availability so your agent never stops because a provider is down.
Unstructured.io: Document Parsing for RAG
Extract clean, structured text from PDFs, Word docs, images, and HTML — ready for embedding and retrieval.
Portkey: AI Gateway + Guardrails
Route, observe, and secure LLM requests with enterprise-grade guardrails and model management.
Unkey: API Key Management for Agents
Managed API key infrastructure with rate limits, RBAC, usage analytics, and automatic key rotation.
Chroma: Developer-First Vector Database
Open-source vector database designed for AI applications with embeddings, full-text search, and metadata filtering.
Mem0: Memory Layer for AI Agents
Drop-in memory infrastructure that lets agents remember users, facts, and context across sessions.
Weights & Biases Weave: Agent Evaluation & Tracing
Trace, evaluate, and iterate on LLM agents with experiment tracking, prompt versioning, and production monitoring.
n8n: Workflow Automation with Agent Nodes
Open-source workflow automation with native AI nodes for building agent-driven processes without heavy code.
Replicate: Cloud API for Open-Source Models
Run open-source image, video, audio, and language models from a single API without managing GPUs.
Supabase: Postgres + Vector Extension for Agents
Open-source Firebase alternative with managed Postgres, pgvector, auth, and realtime — a common backend for agent apps.
Browserbase: Headless Browser Infrastructure
Managed headless browser platform for agent-driven web tasks — scraping, screenshots, form filling, and session persistence at scale.
Inngest: Durable Workflows for Agents
Reliable background jobs and durable execution for long-running agent tasks with retries, scheduling, and event-driven orchestration.
LanceDB: Serverless Vector Database
Open-source, serverless vector database with columnar storage and native Python/JS SDKs. Good for RAG and agent memory with no separate server to manage.
Langfuse: Open-Source LLM Observability
Trace LLM and agent runs, score outputs, manage prompts, and debug multi-step tool loops — self-host or cloud.
Modal: Serverless GPUs for Inference and Agents
Run GPU inference, fine-tunes, and agent workers as serverless Python functions — scale to zero, pay for compute seconds.
Daytona: Secure Sandbox Environments for AI Agents
Programmable sandboxed environments where agents can execute code, run tools, and operate filesystems in isolation — without risking your production infrastructure.
Helicone: LLM Gateway, Logging, and Cost Controls
Open-source-friendly gateway for routing, logging, caching, and cost controls across OpenAI-compatible providers.
Blume: Zero-Config AI-Ready Documentation
An open-source, MIT-licensed documentation framework that generates AI-ready docs from a Markdown folder — no config, no build step, just write and ship.
Cursor Router: Request-Level Model Routing
Cursor's classifier that routes each coding request to the optimal model based on query complexity — cutting AI coding spend 30-60% by matching routine work to cheaper models.
Astrix Security: NHI Governance & AI Agent Control Plane
Discover, secure, and deploy AI agents and non-human identities at scale — the leading platform for agent identity governance, access control, and threat detection.
Crossmint: Agentic Payments Infrastructure
Give your AI agents wallets, stablecoin rails, and card payment capabilities — programmable spending limits, multi-chain support, and Visa/Mastercard integration in one API.
Ferrogate: Self-Hosted AI Gateway
Open-source Rust/Pingora AI gateway for self-hosted LLM traffic control — provider routing, virtual API keys, budgets, caching, MCP tool execution, and observability.
MCP Memory (libSQL): Persistent Agent Memory
High-performance persistent memory system for Model Context Protocol powered by libSQL — vector search, semantic knowledge storage, and relationship management for AI agents.