Language Models AI tools
Large Language Models (LLMs) are advanced AI systems that understand and generate human-like text. These models power a wide range of applications from conversational AI to content generation, code writing, and knowledge retrieval. They serve as the foundation for many modern AI applications.
51 verified AI-first sites in Language Models.
AI lab behind ChatGPT, Codex, and the GPT model family. GPT-6 Astra (September 2026) is its flagship for professional work, coding, and computer use, with faster, cheaper GPT-6 Sol and Luna for everyday workloads. The API also covers image, speech, and realtime models, plus open-weight gpt-oss releases.
Pricing: ChatGPT: Free and Go tiers; Plus $20/month; Pro, Business, Enterprise, and Edu plans. API: pay per token.
OpenAI's AI assistant for chat, search, writing, analysis, and agentic tasks. Voice, vision, file uploads, memory, and connectors to work apps; paid plans add ChatGPT Work and Codex with GPT-6 models. One of the most widely used AI products, with hundreds of millions of weekly users.
Pricing: Free and Go tiers. Plus: $20/month. Pro, Business, Enterprise, and Edu plans.
AI safety company and creator of the Claude model family, including Claude Fable 5, Opus, Sonnet, and Haiku. Claude Fable 5 is Anthropic's frontier Mythos-class model for long-running agentic and coding work with a 1M-token context window. Offers Claude API access, enterprise deployments, and research on responsible AI.
Pricing: Free tier. Claude Pro: $20/month. Enterprise: Custom solutions.
Anthropic's Claude assistant for chat, coding, analysis, and agentic tasks. Includes Claude Fable 5 for Pro and Max subscribers - Anthropic's most capable generally available model for complex multi-step work. Also offers Opus, Sonnet, and Haiku tiers plus API access through the Claude platform.
Pricing: Free tier. Pro: $20/month. Max and Enterprise: See site.
Paris-based AI lab shipping open-weight and commercial models, including Mistral Large 3, Medium, Small, and the Devstral and Codestral coding lines. Le Chat assistant for consumers and teams, La Plateforme API, and private or on-prem deployment for European enterprises and governments.
Pricing: Free API access available. Pay-as-you-go pricing for higher usage. Enterprise licensing available.
Meta's open-weight Llama language model family (Llama 3/4 and prior). Downloadable weights for self-hosting and fine-tuning via Hugging Face and community tooling. Distinct from Muse-Meta's newer proprietary Superintelligence Labs models. Existing Llama releases remain open; frontier investment has largely shifted to Muse.
Pricing: Open weights free under Meta Llama license; commercial terms vary by release. Hosted inference billed by third-party providers.
Enterprise AI company focused on private, secure deployment. Command generative models, Embed and Rerank for retrieval, plus Transcribe, Translate, and Parse models; North workspace and Compass search build on them. Merging with Germany's Aleph Alpha (announced September 2026, pending approval) to form a transatlantic sovereign AI company.
Pricing: Free tier for development. Enterprise: Usage-based pricing. Custom solutions available.
Google DeepMind - Alphabet's frontier AI lab behind Gemini, Gemma, Imagen, Veo, and scientific systems such as AlphaFold. Merged Google Brain and DeepMind; research spans multimodal models, agents, robotics, and safety. Public research, open models, and products that power Google Search, Gemini apps, and Cloud.
Pricing: Research and many demos free. Gemini and Cloud AI: see Google pricing.
German sovereign AI company building the PhariaAI stack and models for governments and regulated enterprises, with on-prem deployment and EU AI Act compliance built in. Signed a definitive merger agreement with Cohere in September 2026; the combined company will operate as Cohere once approved.
Pricing: Custom pricing based on usage. Enterprise solutions available.
Human-centric language models and the Pi assistant. After the 2024 Microsoft talent deal, Inflection rebuilt around relational AI-updated Pi with voice, memory, and agent tools, plus Inflection AI Labs experiments such as Pi Journeys. Also licenses models for private enterprise deployments.
Pricing: Free Pi tier; enterprise licensing-see Inflection.
Chinese lab shipping competitive open-weight LLMs. DeepSeek-V4 and reasoning lines (including R1-class) for chat, coding, and agents; MoE architectures with strong price/performance. Chat app, API, and self-hosted checkpoints-often tuned for domestic Ascend as well as NVIDIA stacks.
Pricing: Free tier and API per-token pricing; see DeepSeek.
Alibaba's Qwen model family and chat app. Qwen3.8-Max (August 2026) is a 2.4T-parameter multimodal flagship with a 1M-token context and, for the first time at Max scale, open weights. Open Qwen3.5 and 3.6 sizes run from 0.8B to 397B; API via Alibaba Cloud Model Studio.
Pricing: Custom pricing based on usage. Enterprise solutions available.
Multimodal language model platform that processes streams of video, audio, and text with models trained from scratch. Offers vision, speech, and chat capabilities with on-device deployment options.
Pricing: Contact for pricing
Open-source family of AI models designed for enhanced efficiency and functionality. Includes Nemotron 3 Nano (smallest model) and larger versions for various use cases. Features improved performance, efficiency, and open-source availability for developers and researchers. Part of Nvidia's commitment to open-source AI development, enabling broader access to advanced AI models. Particularly valuable for developers seeking efficient, open-source AI models for various applications.
Pricing: Open-source; free to use and modify.
Google's multimodal language model family (Gemini). Native text, image, audio, and video understanding with long context, API access via Google AI / Vertex, and consumer apps on gemini.google.com. Frontier alternative to GPT and Claude classes.
Pricing: Free tier available. Premium subscriptions available. API pricing varies by usage.
Frontier lab building open-weight foundation models for software engineering (Laguna family). Agentic coding models via API and on-device variants, focused on code generation, testing, and refactoring.
Pricing: Enterprise pricing. Available via Amazon Bedrock and EC2. Contact for custom deployment.
Company building ultra-efficient foundation models (LFMs) for on-device and edge deployment. LFM2.5 family includes text, vision-language, and audio models (350M-8B parameters); LFM2.5-1.2B-Thinking fits under 1GB for phones. Runs on GPUs, CPUs, and NPUs; LEAP platform for customization and deployment; Apollo app for private on-device chat. Used by Shopify (multi-year partnership), Qualcomm, AMD, Capgemini; $250M raised (Dec 2024). Free for companies under $10M revenue; startup program and enterprise solutions. Particularly valuable for edge, privacy-critical, and low-latency AI applications.
Pricing: Free for companies under $10M revenue. Startup program and enterprise: contact for pricing.
Moonshot AI's Kimi assistant and model platform, now running K3 for agentic coding and knowledge work. Long-context chat, deep research, and the Kimi Code agent, with open-weight K2-series releases and an API at platform.moonshot.ai.
Pricing: Free tier. Starter: $9/mo (10M tokens). Ultra: $49/mo (70M tokens). API: pay-as-you-go. See platform.moonshot.ai for details.
Chinese AI company building language, speech, video (Hailuo), and music models. Its M3 flagship (June 2026) targets coding and agentic work at low API prices, alongside M2.7 and the MiniMax Code assistant. API and consumer products serve users in 200+ countries.
Pricing: API usage-based pricing. Enterprise solutions available.
Chinese foundation model provider behind the GLM model series, offering large language model APIs and enterprise AI capabilities for chat, reasoning, and multilingual use cases.
Pricing: API and enterprise pricing; see site for current options.
Baidu's large language model family for conversational AI, content generation, and enterprise AI applications. Includes model and platform offerings integrated into Baidu's broader AI ecosystem.
Pricing: Platform and enterprise pricing; see Baidu/ERNIE pages for current details.
AI company building foundation models and document intelligence for enterprise. Known for the Solar family of LLMs and strong multilingual and long-context performance; offers APIs and solutions for text generation, classification, and document understanding. Particularly valuable for teams seeking non-U.S.-centric model vendors with production APIs and compliance-oriented deployments.
Pricing: API and enterprise pricing; see upstage.ai for current plans and trials.
AI lab building diffusion-based reasoning language models for production latency. Mercury family (including Mercury 2) refines full responses in parallel instead of token-by-token decoding, targeting reasoning-grade quality at 1000+ tokens/sec on NVIDIA Blackwell-class hardware. OpenAI-compatible API with tunable reasoning, 128K context, native tool use, and schema-aligned JSON. Distinct from the Allen Institute's OLMo open models and from hyperscaler frontier APIs.
Pricing: API from about $0.25/1M input and $0.75/1M output tokens; see inceptionlabs.ai.
Open language model program from the Allen Institute for AI (AI2). Releases fully open OLMo 2 and related foundation weights, training data, evaluation harnesses, and recipes-not just model checkpoints. Built for reproducible research and commercial use under open licenses; complements AI2's separate Asta science-agent products. Models span efficient to large scales with transparent pretraining and ongoing community releases on Hugging Face and GitHub.
Pricing: Open-source models and artifacts free; hosted inference via third-party providers.
Open-model research lab behind the Hermes LLM family for agentic, reasoning, and roleplay workloads. Hermes 3 and 4 open-weight releases with Nous Portal API and Hermes Agent product. Also builds Nomos and Psyche models; strong community adoption for fine-tunes and agent frameworks.
Pricing: Nous Portal subscription and API; open weights free to download.
Frontier AI lab founded by Mira Murati building customizable language-model infrastructure. Tinker API for LoRA fine-tuning of open models including Qwen, Kimi, and GPT-OSS with managed distributed training. Research preview of TML-Interaction multimodal models for real-time full-duplex collaboration.
Pricing: Tinker usage-based API; private beta and waitlist on site.
Frontier code-model lab building LTM (Long-Term Memory) models for software engineering at extreme context. LTM-2-mini supports up to 100M tokens for whole-repository reasoning; custom sequence-dimension architecture for efficient long-context inference. Training on Google Cloud GB200 supercomputers; $500M+ raised.
Pricing: Research and applied access; contact for enterprise.
Chinese foundation-model lab (阶跃星辰) behind the Step family of language, vision, audio, and video models. Step-2 trillion-parameter MoE and Step-3.7 Flash for production agent workflows with tool calling and MCP compatibility. OpenAI-compatible API at platform.stepfun.com; backed by Tencent and Shanghai government.
Pricing: API pay-as-you-go; Step Plan subscriptions; see platform.stepfun.com.
Meituan's LongCat chat and API for their open LLM family, including LongCat-2.0 and Flash-Chat. Consumer chat at longcat.chat plus OpenAI- and Anthropic-compatible endpoints for agentic coding, reasoning, and long-context workflows.
Pricing: Free chat tier; API via LongCat platform keys-see site for limits and pricing.
Meta Superintelligence Labs' proprietary Muse Spark model family (including Muse Spark 1.3). Multimodal reasoning for agentic coding, tool use, and computer use-available in Meta AI and via the Meta Model API.
Pricing: In Meta AI apps; Model API pay-as-you-go-see Meta developer docs.
Microsoft's open Phi small language model family for efficient on-device and cloud reasoning. Strong quality-per-parameter SLMs for chat, coding, and agents with Azure and open-weight distribution paths.
Pricing: Open weights and Azure-hosted options-see Microsoft Phi product pages.
IBM's Granite foundation model family for enterprise language, code, and multimodal workloads. Open and governed models designed for business AI with watsonx and open-source distribution.
Pricing: Open models and watsonx commercial options-see IBM Granite.
ByteDance Seed foundation-model lab behind Seed LLMs and related multimodal systems powering Doubao and ByteDance AI products. Research and model releases for large-scale language and agent applications.
Pricing: Product and API access via ByteDance / Doubao channels-see Seed site.
Tencent Hunyuan large language model family for chat, reasoning, and enterprise AI across Tencent Cloud. Multilingual Chinese/English LLMs with APIs and product integrations across the Tencent ecosystem.
Pricing: Tencent Cloud API and product pricing-see Hunyuan.
Meta Superintelligence Labs' open-weight Muse Glimmer 30B - Apache 2.0 multimodal model built for always-on local agents. Runs on a single consumer GPU or Mac; strong tool use, long-running tasks, and coding versus peers in its size class. Distinct from closed Muse Spark; weights on Hugging Face with Ollama, LM Studio, and cloud hosts following.
Pricing: Open weights free under Apache 2.0; self-host hardware costs; hosted inference billed by third parties.
Google DeepMind's open-weight Gemma model family (Gemma 4). Dense and MoE sizes from edge (E2B/E4B) to 31B for reasoning, coding, and multimodal work. Apache 2.0 weights; the open counterpart to Gemini for self-host and fine-tunes.
Pricing: Open weights free (Apache 2.0); hosted inference billed by Google and third parties.
Amazon's Nova foundation model family on AWS for text, multimodal, and agent workloads. Size tiers from compact to Pro-class reasoning, built for Bedrock enterprise deploy alongside third-party hosted models.
Pricing: Pay-per-token on Amazon Bedrock; see AWS Nova pricing.
OpenAI's open-weight reasoning models (gpt-oss-120b and gpt-oss-20b) under Apache 2.0. MoE checkpoints for agents and private serving; native MXFP4 so 120b fits an 80GB GPU and 20b fits ~16GB. Open counterpart to GPT APIs.
Pricing: Open weights free (Apache 2.0); self-host hardware or third-party hosts.
Shanghai AI Laboratory's InternLM (书生) open language model family. Fully open weights, data, and tooling for chat, reasoning, and agents; a major Chinese open-foundation line alongside Qwen and DeepSeek.
Pricing: Open weights free; hosted inference via third-party providers.
Elon Musk's frontier AI lab building the Grok family of language models. Large-scale training infrastructure and rapid model releases, with API and product distribution through Grok and the X platform.
Pricing: Grok free and SuperGrok tiers; API pricing on xAI.
Consumer chat interface for xAI's Grok models. Multimodal Q&A with real-time context from X, image generation, and tool use for everyday and power-user workflows.
Pricing: Free tier; SuperGrok and X Premium+ for higher limits.
Technology Innovation Institute's Falcon open LLM family from Abu Dhabi. Multilingual open-weight models widely used as a sovereign and research alternative to US closed labs.
Pricing: Open weights; see TII Falcon for licenses and downloads.
Snowflake's open enterprise LLM family optimized for SQL, coding, and instruction following. Dense-MoE style models aimed at enterprise RAG and analytics workloads.
Pricing: Open weights plus Snowflake Cortex usage; see Snowflake Arctic.
Apple's on-device foundation model framework for apps on Apple silicon. Local LLM and generation APIs powering Apple Intelligence features with privacy-preserving inference.
Pricing: Included with supported Apple platforms for developers; see Apple docs.
Inflection's personal AI companion chatbot. Conversational assistant focused on helpful, emotionally aware dialogue rather than enterprise coding agents.
Pricing: Free and paid tiers; see Pi.
Databricks' mixture-of-experts open LLM (DBRX) and Mosaic AI model serving stack. Open weights plus managed inference for enterprise data platforms.
Pricing: Open weights; Databricks Mosaic AI serving priced by usage.
Chinese foundation-model lab (百川智能) behind the Baichuan open and commercial LLM families. Models and APIs for chat, reasoning, and enterprise Chinese-language applications.
Pricing: API and enterprise plans; see Baichuan.
SenseTime's SenseNova multimodal large-model platform. Native multimodal LLMs and API services for Chinese enterprise and consumer applications across text, vision, and agents.
Pricing: Platform and API plans; see SenseNova.
Open RNN-style language model family with Transformer-like quality and linear-time inference. Linux Foundation AI project combining parallelizable training with constant memory and infinite context characteristics.
Pricing: Open weights and tooling; see RWKV docs and Hugging Face.
Moonshot AI lab behind the Kimi model family. Frontier Chinese and multilingual LLMs for long-context reasoning, coding agents, and multimodal chat, with research releases and API platforms.
Pricing: API and product plans via Moonshot / Kimi; contact Moonshot.
US lab building open-weight foundation models. The Trinity family is released for anyone to inspect, fine-tune, and self-host, with the Arcee Platform for deploying and operating open models in production. Developing Genesis-Science-1 with the Department of Energy; valued at $1B+ after its 2026 Series B.
Pricing: Open weights free; Arcee Platform and enterprise plans via sales.