A free command-line tool that tells you which local AI models your computer can actually run — before you spend an hour downloading one that just crashes or crawls.
#model-hosting
32 tools curated in this category — including llmfit, Crun AI, Prime Intellect
techFind on mySelectas all sites and tools related to model-hosting. This selection of 32 resources is reviewed and maintained by the community. The most popular include llmfit, Crun AI, Prime Intellect. Each tool comes with a review, tags, comparisons and alternatives to help you make the best choice.
A single API that gives developers access to 100+ AI models for video, image, audio and chat, instead of integrating each vendor separately.
Rented supercomputer power and tooling for teams training their own AI models with reinforcement learning, instead of just calling someone else's finished model through an API.
An API for making computers talk and listen in a natural-sounding voice — text-to-speech, transcription, voice cloning, and live translation — built by a team that split off from a well-known French AI lab.
A free, open-source Mac app that runs AI chatbots and agents entirely on your own computer, with nothing ever sent to the cloud.
BaseRT runs AI language models directly on your Mac's own chip instead of a data center — once it's running, there's no per-message bill because there's no server in the loop at all.
Arkor lets your AI coding assistant write the training code for a custom AI model, then handles the expensive GPU work behind the scenes — so fine-tuning a model no longer requires being a machine-learning engineer.
Open-source on-device AI runtime that runs LLMs, vision models and speech models directly on phones, laptops and wearables — no cloud round-trip required.
An AI inference API built specifically for speed — over 1,000 image, video, audio and language models served through one endpoint with sub-1-second latency and no cold starts.
A control panel for getting AI models and agents from a developer's laptop into real production use — deployment, scaling, and governance — built to run on whichever cloud a company already uses instead of locking them into one.
A single API endpoint that talks to 600+ AI models from 30+ providers, so switching from one AI model to another — or spreading requests across several for cost or reliability — doesn't mean rewriting your app's code.
A safety checker for AI systems that catches hallucinations, unsafe answers and factual mistakes before they reach a user, using its own purpose-built judge models instead of relying on a general-purpose LLM to grade itself.
A European, publicly-traded cloud built specifically for AI workloads — GPU clusters, training, inference — for companies that want serious AI compute without depending on a US hyperscaler.
A marketplace of ready-to-call AI models for generating images, video, audio and 3D content — like an app store of generative AI models you access through a single API instead of hosting each model yourself.
Pay-by-the-second GPU cloud for training and running AI models, with no contracts and fast serverless cold starts.
GPU cloud built for AI research and production, scaling from a single rented GPU up to superclusters with over 165,000 GPUs.
Enterprise cloud built specifically around NVIDIA GPUs, providing large-scale compute for training and running the biggest AI models.
Cloud inference platform for running and fine-tuning open-source language models fast, without owning any GPUs.
Serverless cloud platform that runs Python code, including AI model training and inference, on GPUs with sub-second startup and pay-per-second billing.
Managed feature store and AI lakehouse platform for building production machine-learning systems with millisecond-latency feature serving.