[{"data":1,"prerenderedAt":1016},["ShallowReactive",2],{"categories-init":3,"stack-litellm-self-hosted":4,"stack-res-litellm-self-hosted":1015},true,{"stack_id":5,"slug":6,"name":7,"tagline":8,"long_description":9,"key_features":10,"use_cases":17,"pros":23,"cons":28,"cover_image_url":33,"scores":34,"options":48,"additions":498,"option_groups":755,"multi_select_option_types":761,"tools_by_category":762,"related_stacks":894,"faqs":958,"pricing":974,"system_requirements":993,"experience_level":1009,"project_type":1010,"stack_type_slug":902,"stack_type_icon_url":903,"published_date":691,"last_updated_date":691,"seo_meta":1011},244,"litellm-self-hosted","LiteLLM Self-Hosted","A self-hosted gateway that gives every application one OpenAI-compatible endpoint, with virtual keys, budgets, and spend tracking you control.","LiteLLM is an AI gateway. Your applications and agents call one OpenAI-compatible address, and the gateway translates each request for the right provider, whether that is Anthropic, OpenAI, Gemini, Bedrock, Azure, or a model you run yourself. Around that routing it adds **virtual API keys with budgets and rate limits per team, user, and key**, spend tracking, fallbacks between deployments, and an admin interface. The reason to run it yourself is that every prompt and every provider credential passes through it: a self-hosted gateway keeps both on infrastructure you administer, and adds no per-token charge on top of what the providers bill.\n\nThe official Docker Compose file runs **three services**: the LiteLLM proxy on port 4000, PostgreSQL for keys, teams, budgets, and spend logs, and Prometheus for metrics. PostgreSQL is what makes keys, budgets, and spend tracking work, so it is a hard dependency. Redis is not in the file; it becomes the shared state once more than one proxy instance runs, holding rate-limit counters, router state, and a response cache. The vendor's production guidance is 1 vCPU and 4 GB of memory per proxy instance, which puts the whole bundle on an 8 GB server, and Kubernetes with the Helm chart, or the Terraform modules for the large clouds, is the route once a single server is no longer enough.\n\nThe API and the admin interface share port 4000. The Compose file publishes that port on every interface, and it also publishes PostgreSQL on 5432 and Prometheus on 9090 with a default database password, so it is a starting point, not a hardened deployment: bind or close the two extra ports, replace the password, and set the master key first. **Put a reverse proxy with TLS in front of port 4000**, because provider keys and bearer tokens cross it on every request, or install LiteLLM from a self-hosted PaaS template that brings its own proxy. A tunnel is the alternative when the clients run somewhere you do not control.\n\n**Upstream models are configuration, not code.** Each entry in the config file or the admin interface names a provider, a model, and a credential, and the gateway exposes it under whatever name your applications call. Several entries can share one name to balance load or fall back across providers, and a model you run yourself is one more entry, which is how local and hosted models end up behind one address. Tracing prompts and costs in an external tool is optional, since the bundled Prometheus covers the gateway's own metrics rather than request contents.",[11,12,13,14,15,16],"LiteLLM in Docker from the official Compose file: the proxy, PostgreSQL, and Prometheus","One OpenAI-compatible endpoint in front of more than 100 model providers","Virtual API keys with budgets, rate limits, and spend tracking per key, team, and user","Load balancing and fallbacks across several deployments of the same model","An admin interface on the same port for keys, models, and spend","Prometheus metrics in the open-source edition, plus callbacks to external tracing tools",[18,19,20,21,22],"One internal endpoint for every team, with a budget and a rate limit per key instead of shared provider keys","Switching or mixing model providers without changing application code, with fallbacks when one provider is down","Putting a model server you run and hosted APIs behind the same address","Cost allocation: spend per team, project, or user from one place","Keeping provider credentials out of application code and developer laptops",[24,25,26,27],"MIT-licensed core with virtual keys, budgets, load balancing, and Prometheus metrics, and no per-token fee","Applications keep one OpenAI-style client whatever the provider behind it","Prompts and provider credentials stay on infrastructure you operate","Standard pieces everywhere: Docker, PostgreSQL, Helm, and official Terraform modules for two large clouds",[29,30,31,32],"A service every model call depends on: the vendor's production guidance is two or more instances behind a load balancer","PostgreSQL is a hard dependency, and Redis joins it as soon as there is more than one instance","SSO beyond five users, SCIM, JWT and OIDC authentication, and audit logs need a paid Enterprise license","The official Compose file publishes the database and metrics ports with a default password, so it needs hardening before it faces the internet",null,{"popularity":35,"learning_curve":38,"flexibility":41,"performance":44,"portability":46},{"score":36,"reasoning":37},4,"The most widely used open-source LLM gateway, with tens of thousands of GitHub stars and a large number of container pulls, and a place in many teams' agent and application stacks. Templates exist on two deploy platforms and as official Terraform modules, though it is a platform-team tool more than a household name.",{"score":39,"reasoning":40},3,"The idea is simple, since an application only changes its base URL, but running the gateway is not: a config file, PostgreSQL, a master key and a salt key that cannot be rotated, and hardening of ports the Compose file leaves open. Budgets and virtual keys take reading before they behave the way a team expects.",{"score":42,"reasoning":43},5,"Any provider, local or hosted, is one more entry, and routing, fallbacks, budgets, and logging callbacks are all configuration. The deployment bends as well: Compose, a deploy platform, Helm on any cloud, or the maintained Terraform modules for AWS and GCP.",{"score":36,"reasoning":45},"The vendor reports single-digit-millisecond overhead at a thousand requests a second, and the production guide covers workers, a shared cache, and a spend-write buffer for higher loads. Every call crosses the gateway and a database write path, and one instance is a single point of failure, which keeps it off the top score.",{"score":42,"reasoning":47},"An MIT-licensed core on PostgreSQL and standard containers, speaking the OpenAI request format on both sides. Moving off it means pointing applications back at a provider's base URL and exporting keys and spend from one database, and moving between Compose, Kubernetes, and a cloud is configuration.",{"agent_framework":49,"vector_db":53,"database":57,"orm":61,"authentication":65,"analytics":69,"model_inference":73,"coding_agent":77,"llm":81,"language":85,"frontend_framework":89,"cms":93,"hosting":97,"reverse_proxy":347,"self_hosted_paas":437},{"name":33,"tools":50,"descriptions":51,"aliases":52,"see_all":33},[],{},{},{"name":33,"tools":54,"descriptions":55,"aliases":56,"see_all":33},[],{},{},{"name":33,"tools":58,"descriptions":59,"aliases":60,"see_all":33},[],{},{},{"name":33,"tools":62,"descriptions":63,"aliases":64,"see_all":33},[],{},{},{"name":33,"tools":66,"descriptions":67,"aliases":68,"see_all":33},[],{},{},{"name":33,"tools":70,"descriptions":71,"aliases":72,"see_all":33},[],{},{},{"name":33,"tools":74,"descriptions":75,"aliases":76,"see_all":33},[],{},{},{"name":33,"tools":78,"descriptions":79,"aliases":80,"see_all":33},[],{},{},{"name":33,"tools":82,"descriptions":83,"aliases":84,"see_all":33},[],{},{},{"name":33,"tools":86,"descriptions":87,"aliases":88,"see_all":33},[],{},{},{"name":33,"tools":90,"descriptions":91,"aliases":92,"see_all":33},[],{},{},{"name":33,"tools":94,"descriptions":95,"aliases":96,"see_all":33},[],{},{},{"name":98,"tools":99,"descriptions":335,"aliases":343,"see_all":344},"Hosting",[100,144,174,215,249,278,302],{"tool_id":101,"name":102,"slug":103,"tooltip_description":104,"logo_url":105,"logo_bg":106,"pricing_model":107,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":113,"subcategory":116,"categories":120,"subcategories":123,"flexibility_score":39,"performance_score":36,"portability_score":36,"is_featured":125,"tags":126,"score_reasonings":136,"published_date":142,"last_updated_date":143},52,"Hetzner","hetzner","German cloud and dedicated server provider offering VPS, bare-metal servers, and storage at prices far below the big clouds, with data centers in the EU, the US, and Singapore.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fhetzner.svg","white",{"slug":108,"display_name":109,"description":110},"paid","Paid","No meaningful free tier — a subscription or one-time purchase is required to use the tool.",2,"open",{"category_id":42,"name":114,"slug":115},"Hosting & Cloud","hosting-cloud",{"subcategory_id":117,"name":118,"slug":119},42,"VPS & Servers","vps-and-servers",[121],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},0,[124],{"subcategory_id":117,"name":118,"slug":119,"category_id":42,"is_primary":3,"display_order":122},false,[127,132],{"tag_id":128,"name":129,"slug":130,"tag_type":131},12,"Self-hostable","self-hostable","feature",{"tag_id":133,"name":134,"slug":135,"tag_type":131},21,"Multi-region","multi-region",{"learning_curve":137,"flexibility":138,"performance":139,"portability":140,"popularity":141},"Standard VPS setup; any developer comfortable with Linux can deploy in an afternoon.","Standard VPS; configure anything on Linux; no managed abstraction layer constraining choices.","Competitive hardware at low cost; dedicated and VPS servers deliver strong baseline performance.","Standard Linux VPS; highly portable with no managed-service abstraction to escape.","Popular among European developers and cost-conscious teams; strong reputation for value.","2026-05-29","2026-09-27",{"tool_id":145,"name":146,"slug":147,"tooltip_description":148,"logo_url":149,"logo_bg":106,"pricing_model":150,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":151,"subcategory":152,"categories":153,"subcategories":155,"flexibility_score":39,"performance_score":39,"portability_score":39,"is_featured":125,"tags":157,"score_reasonings":168,"published_date":142,"last_updated_date":143},169,"Hostinger","hostinger","Affordable web hosting with shared, cloud, and agency plans plus KVM VPS, on LiteSpeed servers with NVMe storage and the beginner-friendly hPanel control panel.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fhostinger.svg",{"slug":108,"display_name":109,"description":110},{"category_id":42,"name":114,"slug":115},{"subcategory_id":117,"name":118,"slug":119},[154],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[156],{"subcategory_id":117,"name":118,"slug":119,"category_id":42,"is_primary":3,"display_order":122},[158,163],{"tag_id":159,"name":160,"slug":161,"tag_type":162},40,"Web","web","platform",{"tag_id":164,"name":165,"slug":166,"tag_type":167},28,"Web Development","web-development","use_case",{"learning_curve":169,"flexibility":170,"performance":171,"popularity":172,"portability":173},"Shared hosting on Hostinger is genuinely beginner-accessible — hPanel, one-click WordPress, and the Kodee AI assistant reduce the complexity of common tasks to a few clicks. VPS plans require Linux and SSH familiarity to get full value, but the guided setup and AI assistant lower the barrier compared to raw VPS providers like Hetzner.","Shared hosting is intentionally constrained — you run within Hostinger's stack and cannot change the server environment. VPS plans give full root access and can run any Linux workload. The range from shared to VPS covers most beginner-to-intermediate deployment needs, but there is no managed Kubernetes, serverless, or edge compute option.","LiteSpeed with NVMe SSD and built-in caching delivers strong performance for shared hosting — independent benchmarks show sub-1-second load times and 99.96%+ uptime on shared plans. KVM VPS performance is competitive for the price. Not at the level of dedicated cloud providers like DigitalOcean or Hetzner for compute-intensive workloads.","Hostinger is one of the top three most-recommended budget hosts globally, with over 3 million customers and a 4.7\u002F5 Trustpilot score. It dominates beginner and bootstrapper audiences and appears consistently in 'best cheap hosting' roundups. Less visible in developer-tool and startup-stack communities compared to cloud-native platforms.","Shared hosting creates moderate lock-in via hPanel and Hostinger-specific tooling, though WordPress sites can be migrated via standard exports. VPS instances are standard KVM Linux — fully portable to any other VPS provider. No proprietary runtime or framework dependencies on the VPS side.",{"tool_id":175,"name":176,"slug":177,"tooltip_description":178,"logo_url":179,"logo_bg":106,"pricing_model":180,"learning_curve_score":111,"popularity_score":36,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":184,"subcategory":185,"categories":189,"subcategories":191,"flexibility_score":36,"performance_score":36,"portability_score":36,"is_featured":125,"tags":193,"score_reasonings":209,"published_date":142,"last_updated_date":143},166,"DigitalOcean","digitalocean","Developer-focused cloud platform offering VPS (Droplets), managed Kubernetes, App Platform PaaS, and managed databases. Known for simple pricing and excellent documentation.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fdigitalocean.svg",{"slug":181,"display_name":182,"description":183},"usage_based","Usage-Based","Pricing scales with consumption: API calls, data volume, compute time, or similar metered units.",{"category_id":42,"name":114,"slug":115},{"subcategory_id":186,"name":187,"slug":188},41,"App Hosting","app-hosting",[190],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[192],{"subcategory_id":186,"name":187,"slug":188,"category_id":42,"is_primary":3,"display_order":122},[194,195,199,203,204,208],{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"tag_id":196,"name":197,"slug":198,"tag_type":131},13,"Free Tier","free-tier",{"tag_id":200,"name":201,"slug":202,"tag_type":131},20,"Auto-scaling","auto-scaling",{"tag_id":133,"name":134,"slug":135,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},24,"Docker Compatible","docker-compatible",{"tag_id":164,"name":165,"slug":166,"tag_type":167},{"learning_curve":210,"flexibility":211,"performance":212,"popularity":213,"portability":214},"DigitalOcean is specifically designed for developers rather than enterprise ops teams. The control panel is clean, Droplets launch in under 60 seconds, and App Platform requires no infrastructure knowledge at all. Thousands of community tutorials lower the barrier for beginners.","Offers VPS, PaaS, and managed Kubernetes under one roof, covering most deployment patterns from simple static sites to complex microservices. Custom networking, firewalls, VPC peering, and a full API give teams significant control. Missing serverless functions and some niche managed services limit the ceiling slightly.","NVMe SSDs and dedicated vCPU Droplets deliver competitive single-tenant performance. The 99.99% uptime SLA holds in practice per benchmark studies. Lacks the global edge network density of Cloudflare or the ultra-low-latency primitives of AWS.","Consistently ranks in the top 5 cloud providers for developer and startup usage. Strong presence on StackOverflow, Reddit, and developer surveys. Over 600,000 active customers as of 2024. Widely used as a first production cloud by self-taught developers.","Standard Linux VMs and Docker\u002FKubernetes workloads migrate freely to any other cloud or bare metal. No proprietary runtime lock-in. Spaces is S3-compatible. App Platform lock-in is higher — migrating away requires containerizing or adapting the build pipeline.",{"tool_id":216,"name":217,"slug":218,"tooltip_description":219,"logo_url":220,"logo_bg":221,"pricing_model":222,"learning_curve_score":36,"popularity_score":42,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":223,"subcategory":224,"categories":227,"subcategories":229,"flexibility_score":42,"performance_score":42,"portability_score":39,"is_featured":125,"tags":231,"score_reasonings":243,"published_date":142,"last_updated_date":143},50,"Amazon Web Services","aws","World's largest cloud platform with 200+ services. Market leader in infrastructure as a service (IaaS) and platform as a service (PaaS).","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Faws.png","dark",{"slug":181,"display_name":182,"description":183},{"category_id":42,"name":114,"slug":115},{"subcategory_id":159,"name":225,"slug":226},"Cloud Providers","cloud-providers",[228],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[230],{"subcategory_id":159,"name":225,"slug":226,"category_id":42,"is_primary":3,"display_order":122},[232,236,237,238,239],{"tag_id":233,"name":234,"slug":235,"tag_type":131},14,"Serverless","serverless",{"tag_id":200,"name":201,"slug":202,"tag_type":131},{"tag_id":133,"name":134,"slug":135,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":240,"name":241,"slug":242,"tag_type":167},36,"CI\u002FCD","ci-cd",{"learning_curve":244,"flexibility":245,"performance":246,"portability":247,"popularity":248},"Vast service catalog; IAM, VPC, and networking alone can take months to master.","Hundreds of composable services with full IaC support; virtually unlimited architectural freedom.","Global infrastructure; EC2, Lambda, and CloudFront deliver performance at any scale.","Broad API surface and proprietary services create meaningful lock-in; migration is possible but costly.","Market-leading cloud provider; used by the majority of production deployments globally.",{"tool_id":250,"name":251,"slug":252,"tooltip_description":253,"logo_url":254,"logo_bg":221,"pricing_model":255,"learning_curve_score":39,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":256,"subcategory":257,"categories":258,"subcategories":260,"flexibility_score":42,"performance_score":42,"portability_score":39,"is_featured":125,"tags":262,"score_reasonings":272,"published_date":142,"last_updated_date":143},51,"Google Cloud Platform","gcp","Google's cloud platform with strong data, analytics, AI, and Kubernetes capabilities. The third-largest cloud provider, after AWS and Azure.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fgcp.svg",{"slug":181,"display_name":182,"description":183},{"category_id":42,"name":114,"slug":115},{"subcategory_id":159,"name":225,"slug":226},[259],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[261],{"subcategory_id":159,"name":225,"slug":226,"category_id":42,"is_primary":3,"display_order":122},[263,264,265,266,267,268],{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":233,"name":234,"slug":235,"tag_type":131},{"tag_id":200,"name":201,"slug":202,"tag_type":131},{"tag_id":133,"name":134,"slug":135,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":269,"name":270,"slug":271,"tag_type":167},25,"Machine Learning","machine-learning",{"learning_curve":273,"flexibility":274,"performance":275,"portability":276,"popularity":277},"Slightly more approachable than AWS; BigQuery and Cloud Run have excellent documentation.","Deep service catalog; Cloud Run, GKE, and BigQuery combine freely for any architecture.","BigQuery's columnar engine handles petabyte queries in seconds; Cloud Run scales fast.","Similar to AWS; proprietary managed services create meaningful switching costs.","Strong in data and AI workloads; second or third cloud by market share in most segments.",{"tool_id":279,"name":280,"slug":281,"tooltip_description":282,"logo_url":283,"logo_bg":106,"pricing_model":284,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":285,"subcategory":286,"categories":287,"subcategories":289,"flexibility_score":36,"performance_score":36,"portability_score":39,"is_featured":125,"tags":291,"score_reasonings":296,"published_date":142,"last_updated_date":143},53,"Railway","railway","Modern PaaS for deploying apps, databases, and workers. Git-based deployment with infrastructure as code approach.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Frailway.svg",{"slug":181,"display_name":182,"description":183},{"category_id":42,"name":114,"slug":115},{"subcategory_id":186,"name":187,"slug":188},[288],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[290],{"subcategory_id":186,"name":187,"slug":188,"category_id":42,"is_primary":3,"display_order":122},[292,293,294,295],{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":200,"name":201,"slug":202,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"learning_curve":297,"performance":298,"portability":299,"flexibility":300,"popularity":301},"Git-push deploys work out of the box; most apps go live with zero configuration.","Container-based deployment mirrors the underlying cloud infrastructure performance.","Docker-based deployment is reasonably portable; switching to another host is manageable.","Dockerfile-based; supports any language, framework, and service composition.","Growing popularity for simple deployments; well-regarded in indie developer communities.",{"tool_id":303,"name":304,"slug":305,"tooltip_description":306,"logo_url":307,"logo_bg":106,"pricing_model":308,"learning_curve_score":312,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":313,"subcategory":314,"categories":315,"subcategories":317,"flexibility_score":39,"performance_score":39,"portability_score":39,"is_featured":125,"tags":319,"score_reasonings":329,"published_date":142,"last_updated_date":143},168,"Render","render","Fully managed cloud platform with Git-push deployments, managed Postgres, static sites, background workers, and cron jobs, all from one dashboard without writing infrastructure code.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Frender.svg",{"slug":309,"display_name":310,"description":311},"freemium","Freemium","A free tier is available; additional features, usage limits, or managed hosting require a paid plan.",1,{"category_id":42,"name":114,"slug":115},{"subcategory_id":186,"name":187,"slug":188},[316],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[318],{"subcategory_id":186,"name":187,"slug":188,"category_id":42,"is_primary":3,"display_order":122},[320,321,322,323,327,328],{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"tag_id":200,"name":201,"slug":202,"tag_type":131},{"tag_id":324,"name":325,"slug":326,"tag_type":131},23,"Git-based","git-based",{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":164,"name":165,"slug":166,"tag_type":167},{"learning_curve":330,"flexibility":331,"performance":332,"popularity":333,"portability":334},"Render is among the most beginner-accessible deployment platforms available. No Dockerfile, no CLI, no YAML — connect a repo, and the platform builds and deploys automatically. The dashboard is clean and self-explanatory. Developers new to cloud deployment consistently cite Render as the easiest on-ramp to production hosting.","Covers the standard web app deployment pattern well: HTTP services, workers, cron, databases, and static sites. Docker is supported, and private networking is first-class. However the platform does not offer managed Kubernetes, GPU machines, or fine-grained infrastructure control. Advanced networking, multi-region active-active, and custom VM configurations are out of scope.","Flat-rate VMs with up to 32 GB RAM and 8 CPUs are available for demanding workloads, and zero-downtime deploys keep services live during updates. The free tier's 30–50 second cold start after inactivity is a significant limitation for demos and hobby projects, but paid tiers are always-on. Performance is competitive for a PaaS but not benchmark-leading.","Render is the most commonly cited Heroku replacement in developer communities as of 2026. It appears regularly in Reddit discussions, dev tooling newsletters, and startup stack comparisons. Growth has been strong since Heroku removed its free tier in 2022, though it remains smaller than Railway in developer mindshare.","Workloads on Render can be containerised and moved to any Docker-compatible host with modest effort. Managed Postgres exports standard pg_dump files. However the per-service model and render.yaml configuration are Render-specific; migrating a multi-service app requires translating cron, worker, and networking config into a new platform's conventions.",{"hetzner":336,"hostinger":337,"digitalocean":338,"aws":339,"gcp":340,"railway":341,"render":342},"The default. A server in the 8 GB class costs about €11 a month on the shared-vCPU line, which holds the proxy at its recommended 4 GB with PostgreSQL and Prometheus beside it. LiteLLM publishes no host-specific guide, so this is the general Docker route: the Compose file with a real master key and database password, the database and metrics ports closed or bound to loopback, and a proxy in front of port 4000. One server is a single point of failure for every model call your applications make, so plan a second instance once the traffic matters.","The budget entry. The KVM 2 plan carries 2 vCPUs, 8 GB of RAM, and 100 GB of NVMe at $8.99 a month on a two-year term, $14.99 at the standard rate, which fits the proxy and its two companions. Root access is full and the panel adds weekly backups, but TLS, upgrades, and closing the extra ports stay your job. Run one proxy worker here: the docs size each worker at a core and 4 GB, so 8 GB covers one worker plus PostgreSQL and Prometheus, and more workers or heavier traffic mean moving up a plan.","Per-second billing and snapshots from the panel. The 8 GB Basic Droplet with 4 vCPUs is $48 a month and holds one proxy worker with PostgreSQL and Prometheus beside it, while the 4 GB plan at $24 meets the proxy's own memory floor but leaves PostgreSQL and Prometheus almost no room, so it suits a trial only. Managed PostgreSQL can take over the database later, which also removes the password-protected database port from the server. Place the Droplet in the region nearest your applications, since the gateway adds a hop to every call.","The official infrastructure-as-code route: LiteLLM maintains a Terraform module that runs the gateway, backend, and admin interface as separate ECS Fargate services behind a load balancer, with Aurora PostgreSQL and ElastiCache Redis, and the Helm chart covers EKS. Bedrock models sit in the same account, so those calls can use IAM roles instead of stored provider keys. It costs several times a single server and removes the operations: a managed database, a shared cache, and horizontal scaling out of the box.","The same official route on Google's side: a maintained Terraform module that runs the services on Cloud Run with Cloud SQL for PostgreSQL and Memorystore, or the Helm chart on GKE. Vertex AI models are available from the same project, and Cloud Run scales the gateway with traffic. It fits when the applications calling the gateway already run in Google Cloud, so requests stay on the internal network.","A container platform with a one-click template linked from LiteLLM's own README and docs: it deploys the proxy as an always-on service, and the docs say to set PORT=4000. The template runs the proxy alone, so add Railway's PostgreSQL to the project and point DATABASE_URL at it, or keys, budgets, and spend tracking have nowhere to live, and add Redis there once a second instance runs. Railway terminates TLS itself, so the reverse proxy or PaaS choice doesn't apply. Billing is usage-based on top of the $5 a month Hobby plan, with memory at about $10 per GB a month, so a proxy holding the vendor's recommended 4 GB plus a database lands in the tens of dollars, more than a small VPS in exchange for no server to patch.","A container platform with a Deploy to Render button in LiteLLM's README: the repository's blueprint runs the proxy as a web service from the main-stable image, generates the master key, and health-checks \u002Fhealth\u002Fliveliness. It provisions no database, so add Render Postgres and set DATABASE_URL, and a Key Value instance for Redis once there is more than one proxy. TLS ends at the platform, so the reverse proxy or PaaS choice doesn't apply. The vendor's 4 GB per-instance guidance means the 2 CPU and 4 GB instance at $85 a month plus the database; the free instance sleeps after 15 minutes and its free database expires after 30 days, so both suit a demo only. Pin a version tag instead of main-stable before real use.",{},{"kind":345,"slug":115,"name":114,"href":346},"category","\u002Ftools\u002Fcategories\u002Fhosting-cloud",{"name":348,"tools":349,"descriptions":432,"aliases":436,"see_all":33},"Reverse Proxy",[350,382,408],{"tool_id":351,"name":352,"slug":353,"tooltip_description":354,"logo_url":355,"logo_bg":106,"pricing_model":356,"learning_curve_score":39,"popularity_score":36,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":357,"subcategory":360,"categories":364,"subcategories":366,"flexibility_score":36,"performance_score":36,"portability_score":42,"is_featured":125,"tags":368,"score_reasonings":376,"published_date":143,"last_updated_date":33},189,"Traefik","traefik","A cloud-native reverse proxy that auto-discovers routes from Docker container labels, updating its configuration live as containers start and stop.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Ftraefik.svg",{"slug":309,"display_name":310,"description":311},{"category_id":128,"name":358,"slug":359},"DevOps & CI\u002FCD","devops-cicd",{"subcategory_id":361,"name":362,"slug":363},57,"Web Servers & Reverse Proxies","web-servers-reverse-proxies",[365],{"category_id":128,"name":358,"slug":359,"is_primary":3,"display_order":122},[367],{"subcategory_id":361,"name":362,"slug":363,"category_id":128,"is_primary":3,"display_order":122},[369,373,374,375],{"tag_id":370,"name":371,"slug":372,"tag_type":131},11,"Open Source","open-source",{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":377,"flexibility":378,"performance":379,"popularity":380,"portability":381},"Basic Docker label routing is quick to pick up, but understanding the provider model, middleware chains, and dashboard takes more upfront investment than Caddy's flat config file, especially for anyone unfamiliar with label-driven tooling.","Native providers for Docker, Kubernetes, Swarm, Consul, and ECS plus a composable middleware system cover most container-routing scenarios, though its module ecosystem is narrower than NGINX's after two decades of third-party extensions.","Built in Go with an efficient routing engine, it performs comparably to Caddy for typical self-hosted and mid-scale container workloads.","It's the de facto reverse proxy for the Docker\u002FKubernetes self-hosting community and powers routing inside Coolify and Dokploy, with a large and actively growing GitHub following, though still behind NGINX's decades-long install base.","MIT licensed, ships as a single binary and official Docker image, and runs identically across any VPS, Kubernetes cluster, or cloud provider with no lock-in.",{"tool_id":383,"name":384,"slug":385,"tooltip_description":386,"logo_url":387,"logo_bg":221,"pricing_model":388,"learning_curve_score":312,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":391,"subcategory":392,"categories":393,"subcategories":395,"flexibility_score":39,"performance_score":36,"portability_score":42,"is_featured":125,"tags":397,"score_reasonings":402,"published_date":143,"last_updated_date":33},188,"Caddy","caddy","A modern web server and reverse proxy written in Go, best known for provisioning and renewing HTTPS certificates automatically with zero configuration.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fcaddy.svg",{"slug":389,"display_name":371,"description":390},"open_source","Source code is publicly available and free to use, modify, and distribute. No paid plans from the project itself.",{"category_id":128,"name":358,"slug":359},{"subcategory_id":361,"name":362,"slug":363},[394],{"category_id":128,"name":358,"slug":359,"is_primary":3,"display_order":122},[396],{"subcategory_id":361,"name":362,"slug":363,"category_id":128,"is_primary":3,"display_order":122},[398,399,400,401],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":403,"flexibility":404,"performance":405,"popularity":406,"portability":407},"A working reverse proxy with HTTPS is typically a few lines of Caddyfile, and automatic certificate provisioning removes the TLS setup step entirely, making it one of the easiest reverse proxies to get running for the first time.","The Caddyfile and JSON config cover most reverse-proxy and static-serving needs well, but its plugin ecosystem and edge-case module coverage are considerably smaller than NGINX's after two decades of third-party modules.","Built in Go with a modern, concurrent architecture and native HTTP\u002F3 support, it performs well for typical self-hosted and small-to-mid traffic workloads, though it hasn't accumulated NGINX's extreme-scale benchmark track record.","It's a well-known, growing choice in the self-hosted community specifically for its automatic-HTTPS pitch, but its adoption and star count remain well behind NGINX's decades-long dominance.","Apache 2.0 licensed, distributed as a single static binary with an official Docker image, and runs identically on any VPS or OS with zero vendor lock-in.",{"tool_id":409,"name":410,"slug":411,"tooltip_description":412,"logo_url":413,"logo_bg":221,"pricing_model":414,"learning_curve_score":39,"popularity_score":42,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":415,"subcategory":416,"categories":417,"subcategories":419,"flexibility_score":42,"performance_score":42,"portability_score":42,"is_featured":125,"tags":421,"score_reasonings":426,"published_date":143,"last_updated_date":33},187,"NGINX","nginx","The most widely deployed web server and reverse proxy on the internet, known for event-driven performance under high concurrency.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fnginx.svg",{"slug":309,"display_name":310,"description":311},{"category_id":128,"name":358,"slug":359},{"subcategory_id":361,"name":362,"slug":363},[418],{"category_id":128,"name":358,"slug":359,"is_primary":3,"display_order":122},[420],{"subcategory_id":361,"name":362,"slug":363,"category_id":128,"is_primary":3,"display_order":122},[422,423,424,425],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":427,"flexibility":428,"performance":429,"popularity":430,"portability":431},"Basic reverse-proxy blocks are simple to copy-paste, but real production setups (upstream health checks, TLS automation, rewrite rules, rate limiting) require understanding NGINX's directive-based config language in depth, and there's no built-in automatic HTTPS to lean on.","A vast module ecosystem and a fully declarative config language let it act as a reverse proxy, load balancer, cache, static file server, or WAF front-end, with third-party modules covering nearly any edge case.","Its event-driven, non-blocking architecture was purpose-built to solve the C10k problem, and it remains a top performer in every major reverse-proxy benchmark for throughput and memory efficiency under concurrent load.","It has been the most widely deployed web server on the public internet for years, appears in essentially every stack survey, and has the largest surrounding tutorial and community knowledge base of any web server.","The open-source core is free, packaged for every OS and as an official Docker image, and runs identically on any VPS or cloud provider with no vendor lock-in.",{"traefik":433,"caddy":434,"nginx":435},"The family default and the least work here: it discovers the proxy container from Docker labels and renews certificates itself. LiteLLM serves its API and admin interface from the same port, 4000, so one router covers both. Raise the idle and response timeouts, because a proxy that cuts a connection mid-answer breaks streamed completions and slow reasoning models. The Compose file also publishes PostgreSQL and Prometheus on every interface, so close those mappings before adding any proxy.","Automatic HTTPS from a few lines of config: one site block that forwards to port 4000 covers the API and the admin interface. Caddy flushes event streams as they arrive, which suits streamed completions, and its default timeouts are generous for long generations. Provider keys and bearer tokens cross this hostname on every request, so the TLS it manages is the main reason to put it in front of the gateway at all.","The proxy many servers already run. Two of its defaults fight a gateway: response buffering, which delays streamed tokens until proxy_buffering is turned off for this location, and the 60-second read timeout, which cuts long generations unless proxy_read_timeout is raised. Forward the Host and X-Forwarded headers so the admin interface builds correct links, and raise client_max_body_size if clients send large prompts or files.",{},{"name":438,"tools":439,"descriptions":494,"aliases":497,"see_all":33},"Self-Hosted PaaS",[440,468],{"tool_id":441,"name":442,"slug":443,"tooltip_description":444,"logo_url":445,"logo_bg":221,"pricing_model":446,"learning_curve_score":39,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":447,"subcategory":448,"categories":451,"subcategories":453,"flexibility_score":36,"performance_score":39,"portability_score":42,"is_featured":125,"tags":455,"score_reasonings":462,"published_date":143,"last_updated_date":33},183,"Coolify","coolify","Open-source, self-hostable PaaS that brings a Heroku\u002FVercel-style git-push deploy workflow to your own server.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fcoolify.svg",{"slug":389,"display_name":371,"description":390},{"category_id":42,"name":114,"slug":115},{"subcategory_id":449,"name":438,"slug":450},56,"self-hosted-paas",[452],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[454],{"subcategory_id":449,"name":438,"slug":450,"category_id":42,"is_primary":3,"display_order":122},[456,457,458,459,460,461],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":324,"name":325,"slug":326,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":463,"flexibility":464,"performance":465,"popularity":466,"portability":467},"Deploying a simple app is a git-push away once Coolify is installed, but getting there requires provisioning and SSH-connecting a server first, and troubleshooting still touches Docker and Linux server concepts that a fully managed PaaS abstracts away entirely.","Supports Nixpacks, Dockerfiles, and Docker Compose directly, plus 280+ one-click services, making it capable of hosting nearly anything that runs in a container — bounded mainly by what the underlying server's resources can handle.","There is no platform-level performance ceiling of its own; a deployment's speed and headroom are a direct function of the VPS or hardware Coolify is installed on, which can range from a budget €4\u002Fmonth instance to a large dedicated server.","A fast-growing, well-starred open-source project with strong traction in the self-hosting and indie-hacker community, though it remains a niche choice relative to mainstream managed PaaS platforms like Railway, Render, or Vercel.","Apache 2.0 licensed and built entirely on standard Docker tooling with no proprietary runtime — a Coolify-managed deployment can be moved to a different server (or off Coolify entirely, back to raw Docker Compose) with minimal friction.",{"tool_id":469,"name":470,"slug":471,"tooltip_description":472,"logo_url":473,"logo_bg":221,"pricing_model":474,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":475,"subcategory":476,"categories":477,"subcategories":479,"flexibility_score":36,"performance_score":39,"portability_score":42,"is_featured":125,"tags":481,"score_reasonings":488,"published_date":143,"last_updated_date":33},184,"Dokploy","dokploy","Open-source, self-hostable PaaS that deploys apps, Docker Compose stacks, and managed databases across one or more of your own servers.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fdokploy.svg",{"slug":389,"display_name":371,"description":390},{"category_id":42,"name":114,"slug":115},{"subcategory_id":449,"name":438,"slug":450},[478],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[480],{"subcategory_id":449,"name":438,"slug":450,"category_id":42,"is_primary":3,"display_order":122},[482,483,484,485,486,487],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":324,"name":325,"slug":326,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":489,"flexibility":490,"performance":491,"popularity":492,"portability":493},"A cleaner, single-page dashboard and a deployment-first workflow make basic app deployment quick to pick up once a server is connected, though the underlying Docker Swarm concepts still surface for multi-node scaling and advanced configuration.","Native Docker Compose support plus Dockerfile and buildpack deployment covers nearly any containerizable workload, and multi-database support extends it well beyond simple app hosting — bounded mainly by the underlying server's resources.","There is no platform-level performance ceiling of its own; deployment speed and headroom are a direct function of the VPS or hardware Dokploy is installed on.","A fast-growing open-source project with a large and rapidly increasing GitHub star count, though it is younger and currently has a smaller community than the category leader, Coolify.","Apache-2.0-licensed core built on standard Docker and Docker Swarm tooling with no proprietary runtime — a Dokploy-managed deployment can move to a different server, or off Dokploy entirely back to raw Docker Compose, with minimal friction.",{"coolify":495,"dokploy":496},"A free, self-hosted deploy platform with a LiteLLM template in its service library. The template runs the litellm-database image on the main-stable tag with PostgreSQL and Redis, generates the master key and the admin login, switches to production mode, turns telemetry off, and wires Redis in as the response cache and router state, which is closer to the vendor's production advice than the bare Compose file. It runs no Prometheus, and main-stable moves with every release, so pin a version tag before real use. Coolify's own proxy handles the domain and TLS.","Free and self-hosted, with a LiteLLM template that runs the proxy beside a PostgreSQL 16 container, generates the master key and the admin password, and serves the service on your domain with HTTPS and Traefik routing. The template uses the main-latest image tag, which the vendor advises against for production, so change it to a version tag first. It has no Redis or Prometheus, which is fine for one instance. Domains and TLS come with the platform, so no separate reverse proxy is added next to it.",{},{"llm_observability":499,"model_inference":565,"tunnel":663,"caching":718},{"name":500,"tools":501,"descriptions":560,"aliases":563,"preface":564,"see_all":33},"LLM Observability",[502,537],{"tool_id":503,"name":504,"slug":505,"tooltip_description":506,"logo_url":507,"logo_bg":221,"pricing_model":508,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":509,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":510,"subcategory":514,"categories":518,"subcategories":520,"flexibility_score":42,"performance_score":36,"portability_score":42,"is_featured":125,"tags":522,"score_reasonings":530,"published_date":536,"last_updated_date":33},296,"Langfuse","langfuse","Open-source platform for tracing, evaluating, and monitoring LLM applications and AI agents. It adds prompt management, datasets, and cost and latency analytics, and runs as a managed cloud or self-hosted.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Flangfuse.png",{"slug":309,"display_name":310,"description":311},"self_hostable",{"category_id":511,"name":512,"slug":513},18,"Observability & Monitoring","observability-monitoring",{"subcategory_id":515,"name":516,"slug":517},75,"LLM & Agent Observability","llm-agent-observability",[519],{"category_id":511,"name":512,"slug":513,"is_primary":3,"display_order":122},[521],{"subcategory_id":515,"name":516,"slug":517,"category_id":511,"is_primary":3,"display_order":122},[523,524,525,526],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":527,"name":528,"slug":529,"tag_type":167},35,"Monitoring","monitoring",{"performance":531,"learning_curve":532,"flexibility":533,"popularity":534,"portability":535},"The ClickHouse-backed analytics layer handles high trace volumes and fast dashboard queries, and SDKs send traces asynchronously so they add little latency to the application.","Framework integrations add tracing with a callback handler or a decorator, and the cloud version needs only an API key. Evaluations and datasets take more thought but are optional.","OpenTelemetry-based instrumentation, SDKs in Python and TypeScript, integrations for most agent frameworks, a public API, and a self-hostable codebase let teams wire it into nearly any stack.","Around 34K GitHub stars and the most-cited open-source option in LLM observability comparisons, well known to teams shipping LLM features but a young category that few developers outside AI work have looked at yet.","MIT-licensed core, self-hosting on any infrastructure, and OpenTelemetry instrumentation mean traces and prompts can move between the cloud and self-hosted deployments or to another OTel backend.","2026-10-01",{"tool_id":538,"name":539,"slug":540,"tooltip_description":541,"logo_url":542,"logo_bg":106,"pricing_model":543,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":544,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":545,"subcategory":546,"categories":547,"subcategories":549,"flexibility_score":36,"performance_score":36,"portability_score":111,"is_featured":125,"tags":551,"score_reasonings":554,"published_date":536,"last_updated_date":33},297,"LangSmith","langsmith","LangChain's platform for tracing, evaluating, and deploying LLM applications and agents. It works with any framework and is tightest with LangChain and LangGraph, adding datasets, prompt tooling, and managed agent hosting.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Flangsmith.png",{"slug":309,"display_name":310,"description":311},"managed_only",{"category_id":511,"name":512,"slug":513},{"subcategory_id":515,"name":516,"slug":517},[548],{"category_id":511,"name":512,"slug":513,"is_primary":3,"display_order":122},[550],{"subcategory_id":515,"name":516,"slug":517,"category_id":511,"is_primary":3,"display_order":122},[552,553],{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":527,"name":528,"slug":529,"tag_type":167},{"learning_curve":555,"flexibility":556,"performance":557,"popularity":558,"portability":559},"LangChain and LangGraph apps start tracing by setting an API key and one environment variable, and the UI is built around the concepts those frameworks already use. Evaluations and deployment add more to learn.","Covers tracing, evals, prompts, monitoring, and deployment with SDKs and OpenTelemetry for any framework, but customisation stops at what the hosted product and its API expose.","Handles production trace volumes for large LangChain customers, and its managed deployment runtime supports persistence and background runs for long agents.","The default observability choice for LangChain users and the brand most often named first in LLM tracing comparisons, but the category is still young and little known outside teams building with LLMs.","Traces and datasets live in LangChain's cloud unless an Enterprise contract pays for self-hosting, and agents on LangSmith Deployment are tied to its runtime, although OpenTelemetry ingestion eases instrumenting outside LangChain.",{"langfuse":561,"langsmith":562},"The tracing backend LiteLLM supports with a built-in callback: set the public and secret keys and the host, and every request through the gateway becomes a trace with its cost and latency, with a team's or a key's traffic routable to its own Langfuse project. A self-hosted Langfuse server has to be recent enough for the SDK the proxy uses, 3.63.0 or newer for SDK v4, so upgrade it before upgrading LiteLLM. The bundled Prometheus covers gateway metrics, not prompts, which is the gap this fills.","A second callback for teams already tracing with LangChain's platform: the proxy sends each request to a LangSmith project, set with an API key and a project name, and a custom endpoint can be pointed at with LANGSMITH_BASE_URL. It is a hosted, closed-source service, with self-hosting limited to the Enterprise plan, so prompts leave your server unless you choose that plan. The free Developer plan covers one seat and 5,000 traces a month.",{},"Add LLM observability when you want to see every model call, tool call, and token cost inside a run, so a wrong answer can be traced to the step that caused it.",{"name":566,"tools":567,"descriptions":654,"aliases":658,"preface":659,"see_all":660},"Model Inference",[568,602,629],{"tool_id":569,"name":570,"slug":571,"tooltip_description":572,"logo_url":573,"logo_bg":106,"pricing_model":574,"learning_curve_score":312,"popularity_score":42,"hosting_assignment_type":575,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":576,"subcategory":579,"categories":583,"subcategories":585,"flexibility_score":42,"performance_score":36,"portability_score":42,"is_featured":125,"tags":587,"score_reasonings":595,"published_date":143,"last_updated_date":601},210,"Ollama","ollama","The most widely used way to run open-weight LLMs locally — one command downloads and serves models like Llama, Qwen, DeepSeek, GLM, and MiniMax through an OpenAI-compatible API, with an optional paid Ollama Cloud tier for larger models than local hardware can handle.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Follama.svg",{"slug":309,"display_name":310,"description":311},"deployable",{"category_id":133,"name":577,"slug":578},"AI Infrastructure","ai-infrastructure",{"subcategory_id":580,"name":581,"slug":582},60,"AI Runtime & Serving","ai-runtime-serving",[584],{"category_id":133,"name":577,"slug":578,"is_primary":3,"display_order":122},[586],{"subcategory_id":580,"name":581,"slug":582,"category_id":133,"is_primary":3,"display_order":122},[588,589,590,591],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":592,"name":593,"slug":594,"tag_type":131},16,"AI-powered","ai-powered",{"learning_curve":596,"flexibility":597,"performance":598,"popularity":599,"portability":600},"A single install command and a single command to pull and run a model make it the lowest-friction way to try a local LLM, with no configuration required to get started.","100+ supported models, an OpenAI-compatible API for drop-in tooling reuse, and both local and cloud execution modes give it very broad applicability across workflows.","Automatic hardware tuning and support for the latest open-weight models keep it competitive for local inference, though it is not purpose-built for high-throughput production serving the way dedicated inference servers are.","The most widely used local LLM runtime by a clear margin, with 178K+ GitHub stars, 52 million monthly downloads, and 2.5 billion+ cumulative downloads.","MIT-licensed, runs on macOS, Windows, and Linux across Apple Silicon, NVIDIA, and AMD hardware, and models are entirely self-hosted with no forced cloud dependency.","2026-09-28",{"tool_id":603,"name":604,"slug":605,"tooltip_description":606,"logo_url":607,"logo_bg":221,"pricing_model":608,"learning_curve_score":36,"popularity_score":36,"hosting_assignment_type":575,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":609,"subcategory":610,"categories":611,"subcategories":613,"flexibility_score":42,"performance_score":42,"portability_score":42,"is_featured":125,"tags":615,"score_reasonings":623,"published_date":536,"last_updated_date":33},298,"vLLM","vllm","Open-source, high-throughput inference and serving engine for large language models, exposing an OpenAI-compatible API. It is the most common way to self-host open-weight models in production on GPU and accelerator clusters.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fvllm.svg",{"slug":389,"display_name":371,"description":390},{"category_id":133,"name":577,"slug":578},{"subcategory_id":580,"name":581,"slug":582},[612],{"category_id":133,"name":577,"slug":578,"is_primary":3,"display_order":122},[614],{"subcategory_id":580,"name":581,"slug":582,"category_id":133,"is_primary":3,"display_order":122},[616,620,621,622],{"tag_id":312,"name":617,"slug":618,"tag_type":619},"Python","python","technology",{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":269,"name":270,"slug":271,"tag_type":167},{"learning_curve":624,"flexibility":625,"performance":626,"popularity":627,"portability":628},"Starting a server is one command, but running it well in production means understanding GPU memory, KV-cache sizing, quantization, and multi-GPU parallelism, plus the Kubernetes layer around it.","Serves hundreds of model architectures with configurable quantization, parallelism, LoRA adapters, structured outputs, and speculative decoding, and can be embedded as a Python library or run as a server.","PagedAttention and continuous batching set the throughput bar that other open-source engines are measured against, and the V1 engine cut scheduling overhead further.","Around 93K GitHub stars and the engine underneath many hosted inference services, so anyone self-hosting models knows it, while developers who only call hosted APIs rarely touch it directly.","Apache-2.0, runs on GPUs and accelerators from several vendors, installs anywhere Python or Docker runs, and exposes a standard OpenAI-style API.",{"tool_id":630,"name":631,"slug":632,"tooltip_description":633,"logo_url":634,"logo_bg":221,"pricing_model":635,"learning_curve_score":312,"popularity_score":36,"hosting_assignment_type":544,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":636,"subcategory":637,"categories":641,"subcategories":643,"flexibility_score":39,"performance_score":42,"portability_score":111,"is_featured":125,"tags":645,"score_reasonings":648,"published_date":143,"last_updated_date":536},220,"Groq","groq","Fast LLM inference on Groq's own LPU hardware: unlike aggregators such as OpenRouter, Groq runs the compute itself, with low time-to-first-token and per-token pricing among the lowest of the major providers.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fgroq.svg",{"slug":181,"display_name":182,"description":183},{"category_id":133,"name":577,"slug":578},{"subcategory_id":638,"name":639,"slug":640},77,"AI Inference Providers","ai-inference-providers",[642],{"category_id":133,"name":577,"slug":578,"is_primary":3,"display_order":122},[644],{"subcategory_id":638,"name":639,"slug":640,"category_id":133,"is_primary":3,"display_order":122},[646,647],{"tag_id":269,"name":270,"slug":271,"tag_type":167},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"flexibility":649,"performance":650,"popularity":651,"portability":652,"learning_curve":653},"Model catalogue is curated around what runs well on LPU hardware, narrower than a general-purpose aggregator, and there's no self-hosting or custom deployment option.","Sub-100ms time-to-first-token and a purpose-built LPU architecture make it the fastest, most consistent inference provider by a real technical margin, not just marketing.","A $650M fundraise and Nvidia's ~$20B LPU licensing deal signal strong external validation, though its curated model catalogue keeps it a step behind aggregators in raw adoption breadth.","Closed, hosted-only SaaS running on proprietary hardware, no self-hosting path and no way to replicate the LPU speed advantage outside Groq's own infrastructure.","OpenAI-compatible API and a single API key get requests flowing in minutes, no infrastructure or model-serving setup required.",{"ollama":655,"vllm":656,"groq":657},"A model you run on your own machine becomes one more entry: name it ollama_chat\u002F\u003Cmodel> with the Ollama server as api_base, which defaults to port 11434. The gateway then gives it a virtual key, a budget, and a fallback to a hosted model like any other entry. Tool calling depends on the model, since LiteLLM notes that not every Ollama model supports function calls and falls back to JSON mode, so test the agent's tools first. When the gateway runs in a container, localhost points at the container, so use the host's address.","For open-weight models served to many users from a GPU server: the entry uses the hosted_vllm\u002F prefix and the vLLM server's OpenAI-compatible address as api_base, and chat, embedding, and reranking endpoints are supported. The gateway adds what the inference server lacks, which is per-key budgets, rate limits, and fallback to a hosted provider when the GPU box is down. It needs a datacenter-class GPU and someone to operate it.","A hosted inference provider that runs open-weight models on its own chips: set GROQ_API_KEY and name models groq\u002F\u003Cmodel>. Its speed makes it a natural fast tier or first fallback in the gateway's routing, because every agent step waits on a response. Billing is per token with a free rate-limited tier, and other hosted providers of open models connect the same way.",{},"Add model inference when you want an open-weight model in the mix: on your own hardware for privacy, or on a hosted inference provider for speed and low per-token prices.",{"kind":661,"slug":582,"name":581,"href":662},"subcategory","\u002Ftools\u002Fcategories\u002Fai-infrastructure\u002Fai-runtime-serving",{"name":664,"tools":665,"descriptions":713,"aliases":716,"preface":717,"see_all":33},"Tunnel",[666,692],{"tool_id":667,"name":668,"slug":669,"tooltip_description":670,"logo_url":671,"logo_bg":221,"pricing_model":672,"learning_curve_score":39,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":673,"subcategory":674,"categories":678,"subcategories":680,"flexibility_score":36,"performance_score":36,"portability_score":111,"is_featured":125,"tags":682,"score_reasonings":685,"published_date":143,"last_updated_date":691},258,"Cloudflare Tunnel","cloudflare-tunnel","Cloudflare Tunnel creates an outbound-only encrypted connection from your server to Cloudflare's network, exposing services to the internet without opening inbound ports. Named tunnels are the production mode, and no-account Quick Tunnels are for testing.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fcloudflare.svg",{"slug":309,"display_name":310,"description":311},{"category_id":128,"name":358,"slug":359},{"subcategory_id":675,"name":676,"slug":677},70,"Tunneling & Secure Access","tunneling-secure-access",[679],{"category_id":128,"name":358,"slug":359,"is_primary":3,"display_order":122},[681],{"subcategory_id":675,"name":676,"slug":677,"category_id":128,"is_primary":3,"display_order":122},[683,684],{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":686,"flexibility":687,"performance":688,"popularity":689,"portability":690},"Getting a tunnel running requires a Cloudflare account, a domain on Cloudflare DNS, and configuring ingress rules for each hostname — more setup than a single-binary tunnel tool, though well documented and a common homelab pattern once learned.","Ingress rules can route many hostnames to different local services from one tunnel, Zero Trust Access adds per-hostname authentication, and the same daemon handles HTTP alongside TCP\u002FUDP via WARP.","Traffic rides Cloudflare's global Anycast network end to end, giving it the same routing and edge performance as the rest of Cloudflare's CDN.","A standard recommendation in self-hosting and homelab communities for exposing services without port forwarding, though it competes with several other tunnel tools rather than being the default name people reach for first.","Tightly coupled to a Cloudflare account and DNS zone — migrating off Cloudflare means replacing the entire exposure mechanism, not just swapping a config value.","2026-10-06",{"tool_id":693,"name":694,"slug":694,"tooltip_description":695,"logo_url":696,"logo_bg":106,"pricing_model":697,"learning_curve_score":312,"popularity_score":36,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":698,"subcategory":699,"categories":700,"subcategories":702,"flexibility_score":39,"performance_score":39,"portability_score":36,"is_featured":125,"tags":704,"score_reasonings":707,"published_date":143,"last_updated_date":33},259,"ngrok","ngrok is a globally distributed reverse proxy that creates secure public URLs for a local server in seconds, commonly used for local dev previews, webhook testing, and exposing self-hosted services.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fngrok.svg",{"slug":309,"display_name":310,"description":311},{"category_id":128,"name":358,"slug":359},{"subcategory_id":675,"name":676,"slug":677},[701],{"category_id":128,"name":358,"slug":359,"is_primary":3,"display_order":122},[703],{"subcategory_id":675,"name":676,"slug":677,"category_id":128,"is_primary":3,"display_order":122},[705,706],{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":708,"flexibility":709,"performance":710,"popularity":711,"portability":712},"A single CLI command produces a working public URL with no account configuration, DNS setup, or domain ownership required to get started.","Covers HTTP, TCP, and TLS tunnels plus Kubernetes ingress, with edge traffic policies for auth and rate limiting, though the deepest features sit behind paid plans.","Runs on a globally distributed edge network with solid latency for typical dev-preview and webhook workloads, without the scale of the largest CDN networks.","The name most developers reach for first when they need to expose a local server — long-standing mindshare in webhook testing and dev-preview workflows specifically.","A standalone client binary that works against any local service or cloud origin with no DNS zone or vendor account tie-in beyond ngrok itself, making it easy to drop into any project.",{"cloudflare-tunnel":714,"ngrok":715},"A public HTTPS address through an outbound-only connection, for clients that run somewhere you do not control: an application on someone else's platform can call a tunnel hostname while no inbound port is open on the server. Free, with Cloudflare Access in front if the admin interface should get its own login layer. Point the tunnel at port 4000 only, never at the database or metrics ports the Compose file publishes. A proxied request that gets no answer within about 125 seconds fails with a 524, and non-Enterprise plans cannot raise that, so stream long generations. A stable hostname needs a domain on Cloudflare.","The quick-start tunnel, for trying the gateway from a laptop: one command gives a public HTTPS address and no domain is needed. The free plan includes three endpoints, 1 GB of transfer, 20,000 requests, and an interstitial page on browser visits, which long completions and a busy gateway outgrow quickly, so it suits a trial and not the address every application calls. Paid plans start at $10 a month, and sustained use belongs on pay-as-you-go from $20 a month or on Cloudflare Tunnel.",{},"Add a tunnel when you're self-hosting without a static IP or can't open inbound ports — a home server, a VPS behind restrictive network policies, or anywhere a reverse proxy alone can't reach the internet.",{"name":719,"tools":720,"descriptions":751,"aliases":753,"preface":754,"see_all":33},"Caching",[721],{"tool_id":186,"name":722,"slug":723,"tooltip_description":724,"logo_url":725,"logo_bg":221,"pricing_model":726,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":575,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":727,"subcategory":730,"categories":734,"subcategories":736,"flexibility_score":36,"performance_score":42,"portability_score":36,"is_featured":125,"tags":738,"score_reasonings":745,"published_date":142,"last_updated_date":143},"Redis","redis","In-memory data store used as cache, message broker, and database. Known for speed with sub-millisecond response times.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fredis.svg",{"slug":309,"display_name":310,"description":311},{"category_id":36,"name":728,"slug":729},"Databases","databases",{"subcategory_id":731,"name":732,"slug":733},27,"Key-Value Stores","key-value-stores",[735],{"category_id":36,"name":728,"slug":729,"is_primary":3,"display_order":122},[737],{"subcategory_id":731,"name":732,"slug":733,"category_id":36,"is_primary":3,"display_order":122},[739,740,741],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":742,"name":743,"slug":744,"tag_type":131},15,"Real-time","real-time",{"learning_curve":746,"flexibility":747,"performance":748,"portability":749,"popularity":750},"Simple key-value command set; minimal configuration required to get started.","Multiple data structures; Lua scripting; Redis Modules extend core functionality significantly.","In-memory with sub-millisecond read and write; purpose-built for high-throughput caching.","Key-value commands are standard; many Redis-compatible alternatives exist.","Widely used as a cache and message broker; often a background dependency in production stacks.",{"redis":752},"The shared state a gateway needs once it runs on more than one server. One proxy instance works without it, but with two or more, the vendor's production guide makes Redis 7.0 or newer essential: it holds the rate-limit counters, the router's view of which deployments are healthy, and a response cache that every instance reads, so a budget or a limit means the same thing whichever instance answers. A cached response to a repeated identical request skips the provider call, so it saves tokens as well as time. Around 1,000 requests a second the guide adds a Redis buffer for spend writes so PostgreSQL does not deadlock. Redis is cache and counters here, not durable state. Run it as a container next to the proxy, or use the managed services the official Terraform modules provision.",{},"Add caching when the same reads or computations repeat and you want them answered from memory: a cache holds hot data, shared counters, and session state in front of a database or an upstream API.",{"exposure":756},{"name":757,"option_types":758},"Reverse Proxy or PaaS",[759,760],"reverse_proxy","self_hosted_paas",[],{"Databases":763,"Hosting & Cloud":794,"DevOps & CI\u002FCD":807,"Observability & Monitoring":839,"AI Infrastructure":866},[764],{"tool_id":159,"name":765,"slug":766,"tooltip_description":767,"logo_url":768,"logo_bg":221,"pricing_model":769,"learning_curve_score":39,"popularity_score":42,"hosting_assignment_type":575,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":770,"subcategory":771,"categories":774,"subcategories":776,"flexibility_score":42,"performance_score":36,"portability_score":42,"is_featured":3,"tags":778,"score_reasonings":788,"published_date":142,"last_updated_date":536},"PostgreSQL","postgresql","PostgreSQL is a free, open-source object-relational database known for reliability, standards compliance, and extensibility, with rich data types, JSONB, and a large ecosystem of extensions such as PostGIS and pgvector.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fpostgresql.svg",{"slug":389,"display_name":371,"description":390},{"category_id":36,"name":728,"slug":729},{"subcategory_id":269,"name":772,"slug":773},"OLTP Databases","oltp-databases",[775],{"category_id":36,"name":728,"slug":729,"is_primary":3,"display_order":122},[777],{"subcategory_id":269,"name":772,"slug":773,"category_id":36,"is_primary":3,"display_order":122},[779,782,783,784],{"tag_id":36,"name":780,"slug":781,"tag_type":619},"SQL","sql",{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":785,"name":786,"slug":787,"tag_type":131},22,"ACID Compliant","acid-compliant",{"learning_curve":789,"flexibility":790,"performance":791,"portability":792,"popularity":793},"Standard SQL is broadly known, but CTEs, window functions, and JSONB take time to master.","Extensions like PostGIS and pgvector, custom types, and PL\u002FpgSQL cover virtually any need.","Battle-tested query planner; efficient with proper indexing and regular vacuum.","Open SQL standard with no lock-in; data and skills transfer to any SQL-compatible system.","The most popular relational database among developers per Stack Overflow 2024; rapidly growing.",[795],{"tool_id":101,"name":102,"slug":103,"tooltip_description":104,"logo_url":105,"logo_bg":106,"pricing_model":796,"learning_curve_score":111,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":797,"subcategory":798,"categories":799,"subcategories":801,"flexibility_score":39,"performance_score":36,"portability_score":36,"is_featured":125,"tags":803,"score_reasonings":806,"published_date":142,"last_updated_date":143},{"slug":108,"display_name":109,"description":110},{"category_id":42,"name":114,"slug":115},{"subcategory_id":117,"name":118,"slug":119},[800],{"category_id":42,"name":114,"slug":115,"is_primary":3,"display_order":122},[802],{"subcategory_id":117,"name":118,"slug":119,"category_id":42,"is_primary":3,"display_order":122},[804,805],{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":133,"name":134,"slug":135,"tag_type":131},{"learning_curve":137,"flexibility":138,"performance":139,"portability":140,"popularity":141},[808],{"tool_id":809,"name":810,"slug":811,"tooltip_description":812,"logo_url":813,"logo_bg":221,"pricing_model":814,"learning_curve_score":36,"popularity_score":42,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":815,"subcategory":816,"categories":820,"subcategories":822,"flexibility_score":42,"performance_score":36,"portability_score":42,"is_featured":3,"tags":824,"score_reasonings":833,"published_date":142,"last_updated_date":143},72,"Docker","docker","Container platform for packaging applications and their dependencies into portable images that run the same on a laptop, in CI, and in production, with Docker Desktop, Compose, and Docker Hub.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fdocker.svg",{"slug":309,"display_name":310,"description":311},{"category_id":128,"name":358,"slug":359},{"subcategory_id":817,"name":818,"slug":819},33,"Containerization","containerization",[821],{"category_id":128,"name":358,"slug":359,"is_primary":3,"display_order":122},[823],{"subcategory_id":817,"name":818,"slug":819,"category_id":128,"is_primary":3,"display_order":122},[825,826,827,828,829],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":205,"name":206,"slug":207,"tag_type":131},{"tag_id":240,"name":241,"slug":242,"tag_type":167},{"tag_id":830,"name":831,"slug":832,"tag_type":162},43,"Cross-platform","cross-platform",{"learning_curve":834,"performance":835,"portability":836,"flexibility":837,"popularity":838},"Container images, networking, volumes, and multi-stage builds all need deliberate learning.","Container overhead is minimal; near-native performance for most workloads.","Open standard; containers built with Docker run on any container-compatible platform.","Any runtime, any architecture; multi-stage builds and Compose profiles support complex systems.","The standard for containerization; present in virtually every modern software project.",[840],{"tool_id":841,"name":842,"slug":843,"tooltip_description":844,"logo_url":845,"logo_bg":106,"pricing_model":846,"learning_curve_score":36,"popularity_score":39,"hosting_assignment_type":33,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":847,"subcategory":848,"categories":851,"subcategories":853,"flexibility_score":42,"performance_score":39,"portability_score":42,"is_featured":125,"tags":855,"score_reasonings":860,"published_date":142,"last_updated_date":143},94,"Prometheus","prometheus","Open-source systems monitoring and alerting toolkit that collects and stores metrics as time series data.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fprometheus.svg",{"slug":389,"display_name":371,"description":390},{"category_id":511,"name":512,"slug":513},{"subcategory_id":279,"name":849,"slug":850},"Infrastructure & APM","infrastructure-apm",[852],{"category_id":511,"name":512,"slug":513,"is_primary":3,"display_order":122},[854],{"subcategory_id":279,"name":849,"slug":850,"category_id":511,"is_primary":3,"display_order":122},[856,857,858,859],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":196,"name":197,"slug":198,"tag_type":131},{"tag_id":527,"name":528,"slug":529,"tag_type":167},{"learning_curve":861,"portability":862,"flexibility":863,"performance":864,"popularity":865},"PromQL and scrape configuration require practice; service discovery adds meaningful complexity.","Open source; PromQL and metrics format widely adopted as an industry standard.","Instrumentation libraries for every language; alerting rules and recording rules compose freely.","High-cardinality label sets can degrade scrape and query performance significantly.","The standard metrics collection system in Kubernetes and cloud-native environments.",[867],{"tool_id":868,"name":869,"slug":870,"tooltip_description":871,"logo_url":872,"logo_bg":221,"pricing_model":873,"learning_curve_score":111,"popularity_score":36,"hosting_assignment_type":575,"hosting_provider_restriction":112,"hosting_target_restriction":112,"hosting_compatible_tool_ids":33,"parent_tool_id":33,"category":874,"subcategory":875,"categories":879,"subcategories":881,"flexibility_score":42,"performance_score":36,"portability_score":42,"is_featured":125,"tags":883,"score_reasonings":888,"published_date":143,"last_updated_date":691},221,"LiteLLM","litellm","The open-source AI gateway you run yourself: one OpenAI-compatible endpoint in front of more than 100 model providers, with virtual keys, budgets, and spend tracking. It is the self-hosted counterpart to OpenRouter.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Flitellm.png",{"slug":309,"display_name":310,"description":311},{"category_id":133,"name":577,"slug":578},{"subcategory_id":876,"name":877,"slug":878},63,"AI Model Aggregators","ai-model-aggregators",[880],{"category_id":133,"name":577,"slug":578,"is_primary":3,"display_order":122},[882],{"subcategory_id":876,"name":877,"slug":878,"category_id":133,"is_primary":3,"display_order":122},[884,885,886,887],{"tag_id":370,"name":371,"slug":372,"tag_type":131},{"tag_id":128,"name":129,"slug":130,"tag_type":131},{"tag_id":269,"name":270,"slug":271,"tag_type":167},{"tag_id":159,"name":160,"slug":161,"tag_type":162},{"learning_curve":889,"flexibility":890,"performance":891,"popularity":892,"portability":893},"Requires standing up and maintaining real infrastructure, PostgreSQL, Redis, the proxy server itself, meaningfully more setup than a zero-ops hosted aggregator like OpenRouter.","Full control over routing logic, guardrails, virtual keys, and spend limits across 100+ providers, the most configurable option in the aggregator genre precisely because you own the deployment.","8ms P95 latency at 1,000 RPS and a Rust core deliver genuinely fast proxy performance, though real-world throughput depends on the self-hosted infrastructure backing it.","56K+ GitHub stars and production adoption at Stripe, Google ADK, Greptile, and OpenHands make it the clear self-hosted leader in this genre.","MIT licensed, genuinely self-hostable anywhere with Docker or the provided Terraform modules, no vendor lock-in to a hosted service.",[895,912,931,946],{"stack_id":896,"slug":897,"name":898,"tagline":899,"experience_level":900,"project_type":901,"stack_type_slug":902,"stack_type_icon_url":903,"score_popularity":39,"score_learning_curve":111,"catalog_display_order":33,"published_date":33,"last_updated_date":33,"core_tool_previews":904},156,"strapi-self-hosted","Strapi Self-Hosted","Self-hosted Strapi: an open-source headless CMS with PostgreSQL, on a server you control.","beginner","website","infrastructure","https:\u002F\u002Fassets.tekyous.dev\u002Ficons\u002Fstack-types\u002Finfrastructure.svg",[905,906,907],{"tool_id":159,"slug":766,"name":765,"logo_url":768,"logo_bg":221},{"tool_id":809,"slug":811,"name":810,"logo_url":813,"logo_bg":221},{"tool_id":908,"slug":909,"name":910,"logo_url":911,"logo_bg":106},88,"strapi","Strapi","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fstrapi.svg",{"stack_id":469,"slug":913,"name":914,"tagline":915,"experience_level":916,"project_type":917,"stack_type_slug":902,"stack_type_icon_url":903,"score_popularity":36,"score_learning_curve":111,"catalog_display_order":33,"published_date":33,"last_updated_date":33,"core_tool_previews":918},"grafana-self-hosted","Grafana Self-Hosted","Self-hosted Grafana and Prometheus: metrics dashboards and alerting on a server you run.","intermediate","dashboard",[919,924,925,930],{"tool_id":920,"slug":921,"name":922,"logo_url":923,"logo_bg":106},124,"sqlite","SQLite","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fsqlite.svg",{"tool_id":809,"slug":811,"name":810,"logo_url":813,"logo_bg":221},{"tool_id":926,"slug":927,"name":928,"logo_url":929,"logo_bg":221},91,"grafana","Grafana","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fgrafana.svg",{"tool_id":841,"slug":843,"name":842,"logo_url":845,"logo_bg":106},{"stack_id":932,"slug":933,"name":934,"tagline":935,"experience_level":916,"project_type":936,"stack_type_slug":902,"stack_type_icon_url":903,"score_popularity":42,"score_learning_curve":36,"catalog_display_order":33,"published_date":33,"last_updated_date":33,"core_tool_previews":937},219,"apache-airflow-self-hosted","Airflow Self-Hosted","Self-hosted Apache Airflow: the standard data-pipeline scheduler on your infrastructure.","data_pipeline",[938,939,940,941],{"tool_id":159,"slug":766,"name":765,"logo_url":768,"logo_bg":221},{"tool_id":186,"slug":723,"name":722,"logo_url":725,"logo_bg":221},{"tool_id":809,"slug":811,"name":810,"logo_url":813,"logo_bg":221},{"tool_id":942,"slug":943,"name":944,"logo_url":945,"logo_bg":221},9,"apache-airflow","Apache Airflow","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fapache-airflow.svg",{"stack_id":947,"slug":948,"name":949,"tagline":950,"experience_level":916,"project_type":951,"stack_type_slug":902,"stack_type_icon_url":903,"score_popularity":36,"score_learning_curve":39,"catalog_display_order":33,"published_date":33,"last_updated_date":33,"core_tool_previews":952},150,"n8n-self-hosted","n8n Self-Hosted","Self-hosted n8n on your own server, with full control over the database, the host, and how it's exposed to the internet.","automation",[953,954,957],{"tool_id":159,"slug":766,"name":765,"logo_url":768,"logo_bg":221},{"tool_id":785,"slug":955,"name":955,"logo_url":956,"logo_bg":221},"n8n","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Fn8n.svg",{"tool_id":809,"slug":811,"name":810,"logo_url":813,"logo_bg":221},[959,962,965,968,971],{"question":960,"answer":961},"Is there a managed LiteLLM, or do I have to self-host it?","There is no managed LiteLLM: its own site lists an open-source edition and an Enterprise edition, and both run on infrastructure you operate. So self-hosting is the product, not a cost-saving alternative to a hosted plan. What it buys is control: prompts and provider credentials pass only through your server, you contract with each provider directly, and nothing is added per token. What it costs is operating a service that every model call depends on. The hosted counterpart for the same job is a different product, OpenRouter, compared below. The free edition is enough to start: SSO is free for up to five users, and Enterprise is priced by quote, sized to request volume rather than tokens.",{"question":963,"answer":964},"What size server does the gateway need?","LiteLLM's production guide sizes each proxy instance at 1 vCPU and 4 GB of memory, as both request and limit, and scales both with the worker count: eight workers on one instance want 8 vCPU and 32 GB. The 4 GB floor is real, because the database engine inside the proxy keeps its peak memory instead of returning it. That figure covers the proxy alone. With PostgreSQL and Prometheus on the same host, which is what the Compose file starts, an 8 GB server is the practical size for one instance. For production the guide recommends at least two instances behind a load balancer, and Redis from version 7 once there is more than one, so they share rate-limit counters and router state. Disk has no official floor; it grows with the spend logs PostgreSQL keeps.",{"question":966,"answer":967},"Where does state live, and what needs backing up?","Almost everything durable is in PostgreSQL: virtual keys, teams, budgets, spend logs, and any models added through the admin interface, whose provider credentials are stored encrypted with the salt key. That salt key is the one thing to treat like a database backup, because the docs state it cannot be rotated once models have been added, and a restored database without it holds credentials nobody can read. The config file belongs in version control. Redis, when present, is cache and counters, so losing it costs warm responses and not data. Prometheus keeps 15 days of metrics in the Compose file. Nothing here is backed up for you: pg_dump the database on a schedule and store the salt key next to the backup, not in it.",{"question":969,"answer":970},"How do upgrades and the Enterprise license work?","Releases ship as container images, and the vendor says to pin a version tag in production, because the main-stable tag the Compose file uses moves with every release. Database migrations run when the proxy starts, which is fine for one instance; with several, the docs advise setting DISABLE_SCHEMA_UPDATE on the serving pods and running migrations as a separate job so replicas do not race. The code outside the enterprise directory is MIT-licensed, and virtual keys, budgets, load balancing, and Prometheus metrics are all in it. Enterprise is a separate license sold by quote, and it also covers secret-manager integration, IP allowlists, and multi-region deployment.",{"question":972,"answer":973},"LiteLLM or OpenRouter: which gateway should I use?","They do the same job from opposite sides. OpenRouter is hosted: one account, one prepaid balance, hundreds of models behind one endpoint, and a fee on credit purchases, 5.5% on the standard plan, with nothing to run. LiteLLM is the gateway you operate: you hold your own provider accounts and keys, pay each provider directly with no markup, and get budgets, virtual keys, and routing rules of your own, but you also carry its uptime and upgrades. OpenRouter fits one developer or a small team that wants every model on one bill. LiteLLM fits when keys and spend must be governed per team, when prompts cannot go through a third party, or when a model you run yourself belongs behind the same address. The two combine: OpenRouter can be one upstream provider inside LiteLLM.",{"summary":975,"starting_cost_label":976,"has_free_tier":3,"line_items":977},"The software is free, so the fixed cost is a small server: the vendor recommends 1 vCPU and 4 GB of memory per proxy instance, and with PostgreSQL and Prometheus beside it an 8 GB server is the practical size, from about €11 a month. The real spend is model usage, which each provider bills at its own prices; LiteLLM adds no per-token fee. Enterprise features, such as SSO beyond five users, need a paid license quoted by sales.","From ~€11\u002Fmo",[978,982,985,989],{"label":979,"cost":980,"note":981},"Server (VPS or cloud)","$9–48\u002Fmo","The vendor sizes each proxy instance at 1 vCPU and 4 GB of memory. With PostgreSQL and Prometheus on the same host, an 8 GB server is the practical size: about €11 a month at Hetzner, $8.99 at Hostinger on a two-year term, and $48 for an 8 GB cloud server at DigitalOcean. Production adds a second instance, a managed database, and a shared cache.",{"label":869,"cost":983,"note":984},"Free (MIT core)","The proxy, virtual keys, budgets, load balancing, and Prometheus metrics are MIT-licensed. SSO beyond five users, SCIM, JWT and OIDC authentication, and audit logs need an Enterprise license, priced by quote and never per token.",{"label":986,"cost":987,"note":988},"Model providers","Pay per token","Every request is billed by the provider it reaches, at that provider's own prices, through your own account. The gateway's spend tracking shows the same figures per key, team, and user.",{"label":990,"cost":991,"note":992},"Exposure (optional)","Free","The reverse proxies, deploy platforms, and tunnels here are free, open-source software or free tiers, and TLS certificates come from Let's Encrypt.",{"is_official":3,"source_url":994,"items":995,"notes":1008},"https:\u002F\u002Fdocs.litellm.ai\u002Fdocs\u002Fproxy\u002Fprod",[996,999,1002,1005],{"label":997,"value":998},"CPU","1 vCPU per proxy instance, one more for each extra worker",{"label":1000,"value":1001},"RAM","4 GB per proxy instance, as both request and limit",{"label":1003,"value":1004},"Disk","No official floor; PostgreSQL grows with spend logs, and Prometheus keeps 15 days of metrics in the Compose file",{"label":1006,"value":1007},"OS","Any Linux with Docker for the Compose path; Kubernetes with the Helm chart for production","Quoted from LiteLLM's production guide: each pod gets 1 vCPU and 4 GiB of memory as both requests and limits, scaled with the worker count (a container running 8 workers needs 8 vCPU and 32 GiB). The 4 GiB floor exists because the Prisma query engine's resident memory acts as a high-water mark that is not returned to the operating system. The guide sizes the proxy only. The 8 GB server named in the stack's copy is an estimate for the proxy plus PostgreSQL and Prometheus on one host, not a vendor figure. Production guidance is at least two instances behind a load balancer, with Redis 7.0 or newer once there is more than one.","advanced","ai_agents",{"title":1012,"description":1013,"og_image":33,"canonical":1014},"LiteLLM Self-Hosted: Tools, Pricing & How to Deploy | Tekyous","A self-hosted gateway that gives every application one OpenAI-compatible endpoint, with… Compare LiteLLM Self-Hosted tools, pricing & how to deploy on Tekyous.","https:\u002F\u002Ftekyous.dev\u002Fstacks\u002Flitellm-self-hosted",[],1791281141655]