Curated software category

LLM models

Language and foundation models compared by capability, context, licensing, deployment and best-fit workloads.

37profiles catalogued
12evidence-ready profiles
Evidence-linked details
Search and filter instantly
A clearer way to choose

Compare what matters.

Use the profiles as a shortlist, then confirm the fit against your workflow and current provider documentation.

01 · Find the fit

Start with the job

Filter by the use case, access model and pricing approach your team actually needs.

02 · Inspect the details

Compare the tradeoffs

Review capabilities, integrations, limitations and deployment requirements side by side.

03 · Verify before buying

Follow the evidence

Check source links and freshness dates, then validate important details with the provider.

Common evaluation signals
Current documentationPricing clarityImplementation fit
Profiles in this category

Find your best fit.

Browse every catalogued llm models profile, then use the evidence and freshness details to judge how much decision context is currently available.

37 llm models profiles

Gpt Oss 20b

Researched

OpenAI's gpt-oss-20b is a 22B parameter text generation model hosted on Hugging Face in 8-bit and mxfp4 (4-bit) quantizations. It deploys via the lightweight transformers serve CLI or multiple third-party inference providers, suited for evaluation, experimentation, and moderate loads.

PricingHugging Face offers paid upgrades for user or organization accounts.
AccessHugging Face
Last checkedJuly 22, 2026
View Gpt Oss 20b Full profile

Llama 3.1 8B Instruct

Researched

Llama-3.1-8B-Instruct is an 8B-parameter open-source text generation model from Meta, hosted on Hugging Face with a December 2023 knowledge cutoff and JSON-based function calling support.

PricingDedicated Pricing and Billing documentation is provided
AccessHugging Face
Last checkedJuly 22, 2026
View Llama 3.1 8B Instruct Full profile

Llama 3.2 1B Instruct Bnb 4bit

Researched

A 1B-parameter instruction-tuned Llama 3.2 model from unsloth, quantized to 4-bit via bitsandbytes and hosted on Hugging Face. It supports text with an optional ipython environment and has a December 2023 knowledge cutoff.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 23, 2026
View Llama 3.2 1B Instruct Bnb 4bit Full profile

Qwen3 8B

Researched

Qwen3-8B is an 8B-parameter open-source text generation model in the Qwen3 family, hosted on Hugging Face and HuggingChat, updated July 26, 2025, and available through multiple inference providers, though its generated content may be inaccurate.

PricingSee official pricing
AccessCan be installed via pip
Last checkedJuly 22, 2026
View Qwen3 8B Full profile

Nemotron Cascade 2 30B A3B

Researched

NVIDIA's Nemotron-Cascade-2-30B-A3B is a 32B parameter text generation model on Hugging Face, supporting configurable reasoning with budget control, tool/function calling, and deployment across multiple inference providers with GGUF quantization.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View Nemotron Cascade 2 30B A3B Full profile

GLM 5

Researched

GLM-5.2 is a 753B-parameter Mixture of Experts text generation model from zai-org, available on Hugging Face with multiple inference providers, tool calling, and up to 1M token context windows.

Pricingzai-org/GLM-5.2 on deepinfra (cheapest): $0.93 input per 1M tokens, $3.00 output per 1M to · zai-org/GLM-5.2 on novita, together, and fireworks-ai: $1.40 input per 1M tokens, $4.40 ou · zai-org/GLM-5.2 on scaleway: $2.05 input per 1M tokens, $6.27 output per 1M tokens
AccessHugging Face
Last checkedJuly 22, 2026
View GLM 5 Full profile

Hy3 Preview

Researched

Hy3-preview is a 299B-parameter text generation model by Tencent, released as a preview on Hugging Face via DeepInfra. It supports tool calling, structured output, configurable reasoning effort, and a 262,144-token context window at $0.14/$0.58 per 1M input/output tokens.

Pricing$0.14 per 1M input tokens · $0.58 per 1M output tokens
AccessHosted on Hugging Face
Last checkedJuly 22, 2026
View Hy3 Preview Full profile

Llama 3.2 1B Instruct

Researched

Llama 3.2 1B Instruct is Meta's compact 1B-parameter text generation model hosted on Hugging Face, with a December 2023 knowledge cutoff and JSON-based function/tool calling via a Jinja chat template.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View Llama 3.2 1B Instruct Full profile

NousCoder 14B

Researched

NousCoder-14B is a 15B-parameter text generation model by NousResearch, fine-tuned from Qwen3-14B. It is hosted on the Hugging Face Models hub with 4-bit and 8-bit precision variants, an OpenVINO int4 build, and an SFT derivative.

PricingSee official pricing
AccessHugging Face Models hub
Last checkedJuly 22, 2026
View NousCoder 14B Full profile

GigaChat3.1 Audio 10B A1.8B

Researched

GigaChat3.1 Audio 10B A1.8B is an ai-sage model on Hugging Face supporting text generation and automatic speech recognition, with a chat template that includes tool rendering functions.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View GigaChat3.1 Audio 10B A1.8B Full profile

Agents A1

Researched

Agents-A1 is a 35B parameter text generation model from InternScience, built on a Mixture of Experts architecture and part of the qwen3_5_moe model family. It is hosted on Hugging Face with GGUF quantized variants and multiple inference providers.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View Agents A1 Full profile

VulnLLM R 7B

Researched

VulnLLM-R-7B is an 8B parameter open-source text generation model from Virtue-AI-HUB, based on the Qwen family and hosted on Hugging Face. It is tagged for security use cases and was updated in December 2025.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View VulnLLM R 7B Full profile

AFM 4.5B

Researched

AFM-4.5B is a 4.5 billion parameter transformer-based text generation language model by Arcee AI, released July 2025. It supports 11 languages, offers open weights on Hugging Face, and is positioned as a thoughtful conversational assistant.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View AFM 4.5B Full profile

Midm 2.0 Base Instruct

Researched

Midm-2.0-Base-Instruct is an AI-based assistant model developed by KT and published on Hugging Face by K-intelligence. It supports tool/function calling via XML-tagged JSON, outputs structured formats like JSON, SQL, and code, and has a December 2024 knowledge cutoff.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View Midm 2.0 Base Instruct Full profile

Qwen3 14B

Researched

Qwen3-14B is a member of the Qwen3 model family, hosted on Hugging Face and available for direct chat via HuggingChat as an open source conversational AI model.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View Qwen3 14B Full profile

Gemma 2 2b It Abliterated

Researched

Gemma 2 2b It Abliterated is a 3B parameter text generation model from the Gemma 2 family, hosted on Hugging Face and compatible with Hugging Face Endpoints, updated July 31, 2024.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View Gemma 2 2b It Abliterated Full profile

Llama-3.1-Nemotron-8B-UltraLong-4M-Instruct is an Nvidia-developed instruction-tuned chat model from the Llama 3.1 family with 8B parameters and a 4M token context window, hosted on Hugging Face with function/tool calling support and a December 2023 knowledge cutoff.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 23, 2026
View Llama 3.1 Nemotron 8B UltraLong 4M Instruct Full profile

Qwen3 1.7B

Researched

Qwen3-1.7B is a model in the Qwen3 family, developed by Qwen and hosted on the Hugging Face platform.

PricingSee official pricing
AccessHugging Face
Last checkedJuly 22, 2026
View Qwen3 1.7B Full profile

Qwen2.5-Max

Researched

Qwen Studio offers comprehensive functionality spanning chatbot, image and video understanding, image generation, document processing, web search integration, tool utilization, and artifacts.

PricingSee official pricing
AccessWeb
Last checkedJuly 22, 2026
View Qwen2.5-Max Full profile

Claude 3.5 Sonnet

Researched

Introducing Claude 3.5 Sonnet—our most intelligent model yet. Sonnet now outperforms competitor models and Claude 3 Opus on key evaluations, at twice the speed.

PricingSee official pricing
AccessWeb
Last checkedJuly 22, 2026
View Claude 3.5 Sonnet Full profile

DeepSeek V3

Researched

DeepSeek V3 is listed in LLM models. Its official website has been verified and the profile is queued for deeper research.

PricingSee official pricing
AccessWeb
Last checkedJuly 22, 2026
View DeepSeek V3 Full profile

Falcon 3

Researched

Falcon LLM is a generative large language model (LLM) that helps advance applications and use cases to future-proof our world.

PricingSee official pricing
AccessWeb
Last checkedJuly 22, 2026
View Falcon 3 Full profile

Gemini 1.5 Flash

Researched

Gemini 1.5 Flash is listed in LLM models. Its official website has been verified and the profile is queued for deeper research.

PricingSee official pricing
AccessWeb
Last checkedJuly 22, 2026
View Gemini 1.5 Flash Full profile