[{"data":1,"prerenderedAt":103},["ShallowReactive",2],{"categories-init":3,"subcategory-ai-runtime-serving":4,"category-meta-apis-infrastructure":64,"score-types":68,"pricing-models":90},true,{"subcategory_id":5,"name":6,"slug":7,"description":8,"category_id":9,"display_order":10,"tool_count":11,"tools":12},60,"AI Runtime & Serving","ai-runtime-serving","Inference servers, model hubs, and API gateways for running or serving LLMs — local inference engines, hosted model catalogues, and aggregators that route across multiple providers.",11,10,3,[13],{"tool_id":14,"name":15,"slug":16,"tooltip_description":17,"logo_url":18,"logo_bg":19,"pricing_model":20,"learning_curve_score":24,"popularity_score":25,"hosting_assignment_type":26,"hosting_provider_restriction":27,"hosting_target_restriction":27,"hosting_compatible_tool_ids":28,"parent_tool_id":28,"category":29,"subcategory":32,"categories":33,"subcategories":36,"flexibility_score":25,"performance_score":38,"portability_score":25,"is_featured":39,"tags":40,"score_reasonings":57,"published_date":63,"last_updated_date":28},210,"Ollama","ollama","The most widely used way to run open-weight LLMs locally — one command downloads and serves models like Llama, Qwen, DeepSeek, GLM, and MiniMax through an OpenAI-compatible API, with an optional paid Ollama Cloud tier for larger models than local hardware can handle.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Follama.svg","white",{"slug":21,"display_name":22,"description":23},"freemium","Freemium","A free tier is available; additional features, usage limits, or managed hosting require a paid plan.",1,5,"deployable","open",null,{"category_id":9,"name":30,"slug":31},"APIs & Infrastructure","apis-infrastructure",{"subcategory_id":5,"name":6,"slug":7},[34],{"category_id":9,"name":30,"slug":31,"is_primary":3,"display_order":35},0,[37],{"subcategory_id":5,"name":6,"slug":7,"category_id":9,"is_primary":3,"display_order":35},4,false,[41,45,49,53],{"tag_id":9,"name":42,"slug":43,"tag_type":44},"Open Source","open-source","feature",{"tag_id":46,"name":47,"slug":48,"tag_type":44},12,"Self-hostable","self-hostable",{"tag_id":50,"name":51,"slug":52,"tag_type":44},13,"Free Tier","free-tier",{"tag_id":54,"name":55,"slug":56,"tag_type":44},16,"AI-powered","ai-powered",{"learning_curve":58,"flexibility":59,"performance":60,"popularity":61,"portability":62},"A single install command and a single command to pull and run a model make it the lowest-friction way to try a local LLM, with no configuration required to get started.","100+ supported models, an OpenAI-compatible API for drop-in tooling reuse, and both local and cloud execution modes give it very broad applicability across workflows.","Automatic hardware tuning and support for the latest open-weight models keep it competitive for local inference, though it is not purpose-built for high-throughput production serving the way dedicated inference servers are.","The most widely used local LLM runtime by a clear margin, with 178K+ GitHub stars, 52 million monthly downloads, and 2.5 billion+ cumulative downloads.","MIT-licensed, runs on macOS, Windows, and Linux across Apple Silicon, NVIDIA, and AMD hardware, and models are entirely self-hosted with no forced cloud dependency.","2026-09-27",{"category_id":9,"name":30,"slug":31,"description":65,"icon_url":66,"display_order":9,"tool_count":67},"API gateways, CDNs, and infrastructure tools.","https:\u002F\u002Fassets.tekyous.dev\u002Ficons\u002Fcategories\u002Fapis-infrastructure.svg",18,[69,73,78,82,86],{"slug":70,"name":71,"description":72,"display_order":24},"popularity","Popularity","How widely adopted the tool is in the developer community. 1 = niche; 5 = mainstream and widely used.",{"slug":74,"name":75,"description":76,"display_order":77},"learning_curve","Learning Curve","How quickly a developer can become productive with this tool. 1 = beginner-accessible; 5 = steep, requires significant prior experience.",2,{"slug":79,"name":80,"description":81,"display_order":11},"flexibility","Flexibility","How much you can customise or extend the tool for your specific needs. 1 = highly opinionated with few escape hatches; 5 = highly flexible.",{"slug":83,"name":84,"description":85,"display_order":38},"performance","Performance","How well the tool performs its primary function. For runtimes: execution speed. For services: throughput and latency. For editors: responsiveness. 1 = slow or resource-heavy; 5 = fast and efficient.",{"slug":87,"name":88,"description":89,"display_order":25},"portability","Portability","How easy it is to migrate away from this tool once you are invested in it. Based on: data\u002Fcode exportability, skills transferability to other tools, and adherence to open standards. 1 = high lock-in; 5 = fully open, skills transfer universally.",[91,94,95,99],{"slug":92,"display_name":42,"description":93,"display_order":24},"open_source","Source code is publicly available and free to use, modify, and distribute. No paid plans from the project itself.",{"slug":21,"display_name":22,"description":23,"display_order":77},{"slug":96,"display_name":97,"description":98,"display_order":11},"paid","Paid","No meaningful free tier — a subscription or one-time purchase is required to use the tool.",{"slug":100,"display_name":101,"description":102,"display_order":38},"usage_based","Usage-Based","Pricing scales with consumption: API calls, data volume, compute time, or similar metered units.",1790518754071]