Llama.cpp
Portable local inference engine for running quantized language models across laptops, desktops, and servers. It is infrastructure rather than a hosted ChatGPT replacement, so you supply the model and hardware.
Setup: local runtime
Functional replacements with a real free or open-source path, described in plain terms.
Some current card descriptions mention document, retrieval, or research workflows. Compare each card’s notes about model and backend setup.
Some current card descriptions mention local or self-hosted workflows. Your device and chosen model determine what you can run.
Some current card descriptions mention a provider, API, endpoint, or backend connection. A free client does not include free model access; provider plans, quotas, and charges are separate.
These are keyword signals in catalog descriptions, not a verified feature count or feature checklist. Confirm capabilities and terms on each project’s own site.