Models & Infrastructure Tools

LLM APIs, model hosting, inference platforms, and local runtimes for running and deploying AI models.

Models and infrastructure is the builder's category: where AI models get hosted, served, connected, and measured. If you're shipping an AI product rather than using one, these are your suppliers.

Hugging Face is the gravity well — the hub where open models, datasets, and demos live, plus hosting on top. Replicate and fal.ai sell serverless inference: call a model via API, pay per second or per generation, never touch a GPU. Modal and Baseten are deployment platforms for when you outgrow serverless and want your own endpoints with real scaling controls. Pinecone is the vector database most retrieval-augmented apps started on. LiteLLM is the open-source gateway that normalizes a hundred model APIs behind one interface. Heurist and x402 round out the infrastructure edge.

Unusually, this category also includes the scorekeepers: LMArena runs the crowdsourced model leaderboard, Artificial Analysis benchmarks price and performance across providers, and Stanford HELM and SWE-bench publish rigorous academic and coding evaluations. That mix is deliberate — choosing infrastructure without benchmarks is guessing.

The differentiators are practical: cold-start latency, GPU pricing, developer experience, and lock-in. Costs range from free (the benchmarks) to enterprise contracts. The sound pattern: prototype on serverless inference, validate with public benchmarks plus your own evals, and only then commit to dedicated deployment.

Hugging Face logo

Hugging Face

The central hub for AI models, datasets, Spaces, libraries, and open-source ML collaboration.

Models & Infrastructure
Freemium
4.8
LMArena logo

LMArena

Community-powered model leaderboard for comparing AI systems through real user battles.

Models & Infrastructure
Free
4.6
SWE-bench logo

SWE-bench

Software engineering benchmark and leaderboard for evaluating AI coding agents on real GitHub issues.

Models & Infrastructure
Free
4.6
LiteLLM logo

LiteLLM

Open-source LLM gateway for routing, logging, and cost control

Models & Infrastructure
Open source
4.5
Modal logo

Modal

Serverless AI infrastructure for running code, jobs, containers, and GPUs from Python.

Models & Infrastructure
Freemium
4.5
Baseten logo

Baseten

Production AI inference platform for deploying, optimizing, and scaling models.

Models & Infrastructure
Enterprise
4.5
Artificial Analysis logo

Artificial Analysis

Independent AI model benchmarks for intelligence, speed, pricing, context, and modalities.

Models & Infrastructure
Freemium
4.5
Pinecone logo

Pinecone

Managed vector database for semantic search, RAG, recommendations, and AI retrieval.

Models & Infrastructure
Freemium
4.5
fal.ai logo

fal.ai

Fast generative media APIs for images, video, audio, and creative model workflows.

Models & Infrastructure
Paid
4.4
Stanford HELM logo

Stanford HELM

Open framework for holistic, reproducible evaluation of language and multimodal models.

Models & Infrastructure
Open source
4.4
Replicate logo

Replicate

Run open and community AI models from a web playground or API.

Models & Infrastructure
Paid
4.4
x402 logo

x402

Open payment protocol for agentic and API-based machine-to-machine commerce.

Models & Infrastructure
Open source
3.7

112 of 13 tools