| AnythingLLM | Free | Local LLM Runners | All-in-one local RAG and agents app |
| Arize Phoenix | Free | LLM Eval & Observability | Fully open-source, self-hostable with zero setup via `uvx` (pip/conda also supported) |
| Axolotl | Free | Fine-Tuning & Training Frameworks | Config-file-driven fine-tuning workflow (YAML) covering LoRA, QLoRA, full fine-tuning, preference tuning and RL |
| DeepSeek Models | Free | Notable Open Models | Open V3-series and R1 reasoning models |
| Design Arena | Free | Leaderboards & Benchmarks | Crowdsourced Bradley-Terry (Elo-style) ranking of AI models specifically on design/aesthetic quality, not raw capability |
| Epoch AI | Free | Leaderboards & Benchmarks | Independent, non-vendor research institute tracking AI compute, training cost and capability trends over time |
| EXAONE (LG AI Research) | Free | Notable Open Models | EXAONE 4.5: LG's first open-weight vision-language model for industrial use cases |
| Falcon (TII) | Free | Notable Open Models | Falcon H1 hybrid transformer-Mamba architecture for efficient long-context inference |
| Gemma (Google) | Free | Notable Open Models | Google's lightweight open models (Gemma 4) |
| Genesis | Free | Robotics & Embodied AI | Unified multi-physics simulation engine plus photorealistic renderer for robotics/embodied-AI training, fully open source |
| GPT-OSS (OpenAI) | Free | Notable Open Models | OpenAI's open-weight reasoning models |
| GPT4All | Free | Local LLM Runners | Private local chat with your documents |
| Hunyuan (Tencent) | Free | Notable Open Models | Open-weight large text MoE models (Hunyuan A13B/Hy3 series) with permissive commercial license |
| IBM Granite | Free | Notable Open Models | Apache 2.0 licensed enterprise-grade model family |
| Jan | Free | Local LLM Runners | Open-source, offline ChatGPT alternative |
| KoboldCpp | Free | Local LLM Runners | Ships as a single portable executable with a bundled web UI (KoboldAI Lite) |
| LeRobot (Hugging Face) | Free | Robotics & Embodied AI | Open PyTorch library for training real-world robot control policies (ACT, Diffusion Policy, VLA models) |
| Llama (Meta) | Free | Notable Open Models | Meta's flagship open-weight family (Llama 4) |
| LLaMA-Factory | Free | Fine-Tuning & Training Frameworks | Zero-code Web UI (LLaMA Board) for fine-tuning 100+ open models without writing training scripts |
| llama.cpp | Free | Local LLM Runners | Pure C/C++ inference engine with minimal dependencies, runs on CPU-only hardware |
| LM Studio | Free | Local LLM Runners | Polished desktop GUI for local models |
| LMArena (Arena.ai) | Free | Leaderboards & Benchmarks | Crowd-voted 'Battle Mode' blind head-to-head comparisons across text, image, and coding/agent models |
| MiniMax M2 (open weights) | Free | Notable Open Models | Open-weight MoE text/reasoning model line distinct from MiniMax's Hailuo video product already listed |
| Mistral Open Models (Magistral, Devstral, Voxtral) | Free | Notable Open Models | Devstral: open coding-agent model tuned for SWE-bench-style tasks |
| MLX-LM (Apple) | Free | Local LLM Runners | Purpose-built for Apple Silicon's unified memory architecture, faster than llama.cpp for many models on Mac |
| Nemotron (NVIDIA) | Free | Notable Open Models | NVIDIA's open post-trained/distilled model family built on Llama and NVIDIA's own architectures |
| NVIDIA Isaac GR00T | Free | Robotics & Embodied AI | Open, Apache-2.0-licensed vision-language-action foundation model for generalist humanoid robots (GR00T N1.7) |
| NVIDIA Parakeet | Free | Speech & Transcription Engines | Tops the Hugging Face Open ASR Leaderboard with 6.05% average WER at 0.6B params |
| Ollama | Free | Local LLM Runners | The default way to run LLMs locally |
| OLMo (Ai2) | Free | Notable Open Models | Only major model family shipping full pretraining data (Dolma), code and intermediate checkpoints, not just weights |
| Open WebUI | Free | Local LLM Runners | Self-hosted ChatGPT-style interface |
| OpenAI Whisper | Free | Speech & Transcription Engines | Open-source speech-to-text, run locally free |
| Phi (Microsoft) | Free | Notable Open Models | MIT-licensed small language models optimized for on-device and edge deployment |
| Pleias | Free | Notable Open Models | Fully open-weight, open-dataset small language models (350M-3B) explicitly built for EU AI Act compliance |
| Qwen (Alibaba) | Free | Notable Open Models | State-of-the-art open MoE models (Qwen3.x) |
| SGLang | Free | Local LLM Runners | RadixAttention for automatic prefix-cache reuse across requests |
| SWE-bench | Free | Leaderboards & Benchmarks | The standard benchmark for evaluating autonomous coding agents on real GitHub issue resolution |
| Unsloth | Free | Fine-Tuning & Training Frameworks | Fine-tunes LLMs, diffusion, TTS and embedding models 2x faster with ~70% less VRAM than standard Hugging Face training |
| vLLM | Free | Local LLM Runners | PagedAttention memory management for high-throughput multi-request GPU serving |
| WhisperX | Free | Speech & Transcription Engines | 70x real-time transcription speed via batched inference on top of Whisper |
| Artificial Analysis | Freemium | Leaderboards & Benchmarks | Independent cross-provider benchmarking of intelligence, speed and price for the same model across different hosting APIs |
| AssemblyAI | Freemium | Speech & Transcription Engines | Speech-to-text with audio intelligence |
| Baseten | Freemium | Model Hosting & Inference APIs | Basic plan is pay-as-you-go with free starter credits for new accounts |
| Braintrust | Freemium | LLM Eval & Observability | Free Starter plan ships model credits and scored evals with no card required |
| Cerebras Inference | Freemium | Model Hosting & Inference APIs | Runs open models on Cerebras Wafer-Scale Engine hardware at dramatically higher tokens/sec than GPU-based APIs |
| Cloudflare Workers AI | Freemium | Model Hosting & Inference APIs | 10,000 free 'Neurons' of inference per day on every account, resetting daily |
| Deepgram | Freemium | Speech & Transcription Engines | Fast, accurate speech-to-text API |
| ElevenLabs Scribe | Freemium | Speech & Transcription Engines | Scribe v2 Realtime transcribes in under 150ms latency |
| Fireworks AI | Freemium | Model Hosting & Inference APIs | High-performance open-model serving |
| GLM (Z.ai) | Freemium | Notable Open Models | GLM-4.5-Flash and GLM-4.7-Flash text models are free via API |
| Google AI Studio | Freemium | Model Hosting & Inference APIs | Free-tier API access to Gemini Flash-family models with generous token allowances, no credit card required to start |
| Groq | Freemium | Model Hosting & Inference APIs | Ultra-fast inference on custom LPUs |
| Hugging Face | Freemium | Model Hosting & Inference APIs | The hub for open models and datasets |
| Langfuse | Freemium | LLM Eval & Observability | Fully open-source and self-hostable for free via Docker Compose/Kubernetes, not just a hosted SaaS |
| Mistral Voxtral | Freemium | Speech & Transcription Engines | Apache 2.0 open-weight models (24B and 3B) downloadable from Hugging Face |
| Msty | Freemium | Local LLM Runners | Nexus module: unified gateway for managing model connections, credentials and usage policy across providers |
| Novita AI | Freemium | Model Hosting & Inference APIs | Single API surface for 200+ LLM/image/video/audio models plus on-demand GPU and bare-metal instances |
| OpenRouter | Freemium | Model Hosting & Inference APIs | One API for 400+ models across providers |
| Promptfoo | Freemium | LLM Eval & Observability | Open-source CLI/library for evaluating prompts, models and RAG pipelines side by side, runs in CI |
| Replicate | Freemium | Model Hosting & Inference APIs | Run open models with one API call |
| Soniox | Freemium | Speech & Transcription Engines | Real-time speech-to-text priced well under typical competitor rates |
| Speechmatics | Freemium | Speech & Transcription Engines | Free starting credit with no card required, 56+ language support |
| Together AI | Freemium | Model Hosting & Inference APIs | Fast, cheap open-model inference |
| W&B Weave | Freemium | LLM Eval & Observability | Agent-native tracing model with sessions, steps, tools and sub-agents as first-class concepts (not generic spans) |
| DeepInfra | Paid | Model Hosting & Inference APIs | Very low per-token pricing across a large catalog of open models (text, image, audio) |
| Nebius Token Factory | Paid | Model Hosting & Inference APIs | European GPU-cloud-backed inference API hosting major open models (Llama, Qwen, DeepSeek, etc.) |
| SambaNova Cloud | Paid | Model Hosting & Inference APIs | Hosts open models (DeepSeek, MiniMax, Llama) on SambaNova's own RDU chips for high-speed inference |