[{"data":1,"prerenderedAt":105},["ShallowReactive",2],{"categories-init":3,"subcategory-ai-runtime-serving":4,"category-meta-ai-infrastructure":65,"score-types":69,"pricing-models":92},true,{"subcategory_id":5,"name":6,"slug":7,"description":8,"category_id":9,"display_order":10,"tool_count":11,"tools":12},60,"AI Runtime & Serving","ai-runtime-serving","Engines and apps for running open-weight models on hardware you control, from a laptop to a GPU cluster, usually exposed through an OpenAI-compatible API.",21,10,4,[13],{"tool_id":14,"name":15,"slug":16,"tooltip_description":17,"logo_url":18,"logo_bg":19,"pricing_model":20,"learning_curve_score":24,"popularity_score":25,"hosting_assignment_type":26,"hosting_provider_restriction":27,"hosting_target_restriction":27,"hosting_compatible_tool_ids":28,"parent_tool_id":28,"category":29,"subcategory":32,"categories":33,"subcategories":36,"flexibility_score":25,"performance_score":11,"portability_score":25,"is_featured":38,"tags":39,"score_reasonings":57,"published_date":63,"last_updated_date":64},210,"Ollama","ollama","The most widely used way to run open-weight LLMs locally — one command downloads and serves models like Llama, Qwen, DeepSeek, GLM, and MiniMax through an OpenAI-compatible API, with an optional paid Ollama Cloud tier for larger models than local hardware can handle.","https:\u002F\u002Fassets.tekyous.dev\u002Flogos\u002Ftools\u002Follama.svg","white",{"slug":21,"display_name":22,"description":23},"freemium","Freemium","A free tier is available; additional features, usage limits, or managed hosting require a paid plan.",1,5,"deployable","open",null,{"category_id":9,"name":30,"slug":31},"AI Infrastructure","ai-infrastructure",{"subcategory_id":5,"name":6,"slug":7},[34],{"category_id":9,"name":30,"slug":31,"is_primary":3,"display_order":35},0,[37],{"subcategory_id":5,"name":6,"slug":7,"category_id":9,"is_primary":3,"display_order":35},false,[40,45,49,53],{"tag_id":41,"name":42,"slug":43,"tag_type":44},11,"Open Source","open-source","feature",{"tag_id":46,"name":47,"slug":48,"tag_type":44},12,"Self-hostable","self-hostable",{"tag_id":50,"name":51,"slug":52,"tag_type":44},13,"Free Tier","free-tier",{"tag_id":54,"name":55,"slug":56,"tag_type":44},16,"AI-powered","ai-powered",{"learning_curve":58,"flexibility":59,"performance":60,"popularity":61,"portability":62},"A single install command and a single command to pull and run a model make it the lowest-friction way to try a local LLM, with no configuration required to get started.","100+ supported models, an OpenAI-compatible API for drop-in tooling reuse, and both local and cloud execution modes give it very broad applicability across workflows.","Automatic hardware tuning and support for the latest open-weight models keep it competitive for local inference, though it is not purpose-built for high-throughput production serving the way dedicated inference servers are.","The most widely used local LLM runtime by a clear margin, with 178K+ GitHub stars, 52 million monthly downloads, and 2.5 billion+ cumulative downloads.","MIT-licensed, runs on macOS, Windows, and Linux across Apple Silicon, NVIDIA, and AMD hardware, and models are entirely self-hosted with no forced cloud dependency.","2026-09-27","2026-09-28",{"category_id":9,"name":30,"slug":31,"description":66,"icon_url":67,"display_order":9,"tool_count":68},"Model runtimes, inference providers, model gateways, chat interfaces, and sandboxes for running AI models and agents.","https:\u002F\u002Fassets.tekyous.dev\u002Ficons\u002Fcategories\u002Fai-infrastructure.svg",14,[70,74,79,84,88],{"slug":71,"name":72,"description":73,"display_order":24},"popularity","Popularity","How widely adopted the tool is in the developer community. 1 = niche; 5 = mainstream and widely used.",{"slug":75,"name":76,"description":77,"display_order":78},"learning_curve","Learning Curve","How quickly a developer can become productive with this tool. 1 = beginner-accessible; 5 = steep, requires significant prior experience.",2,{"slug":80,"name":81,"description":82,"display_order":83},"flexibility","Flexibility","How much you can customise or extend the tool for your specific needs. 1 = highly opinionated with few escape hatches; 5 = highly flexible.",3,{"slug":85,"name":86,"description":87,"display_order":11},"performance","Performance","How well the tool performs its primary function. For runtimes: execution speed. For services: throughput and latency. For editors: responsiveness. 1 = slow or resource-heavy; 5 = fast and efficient.",{"slug":89,"name":90,"description":91,"display_order":25},"portability","Portability","How easy it is to migrate away from this tool once you are invested in it. Based on: data\u002Fcode exportability, skills transferability to other tools, and adherence to open standards. 1 = high lock-in; 5 = fully open, skills transfer universally.",[93,96,97,101],{"slug":94,"display_name":42,"description":95,"display_order":24},"open_source","Source code is publicly available and free to use, modify, and distribute. No paid plans from the project itself.",{"slug":21,"display_name":22,"description":23,"display_order":78},{"slug":98,"display_name":99,"description":100,"display_order":83},"paid","Paid","No meaningful free tier — a subscription or one-time purchase is required to use the tool.",{"slug":102,"display_name":103,"description":104,"display_order":11},"usage_based","Usage-Based","Pricing scales with consumption: API calls, data volume, compute time, or similar metered units.",1790871008866]