#
Llm
AI tools tagged Llm.
Claude Opus 4.5
Anthropic's Opus-tier model that cut pricing by two-thirds while topping SWE-bench
Claude Opus 4.6
Anthropic's Opus update that expanded the context window to 1 million tokens
Claude Opus 4.7
An Opus 4.6 upgrade targeted at the hardest software engineering tasks
Claude Opus 5
Anthropic's newest Opus-tier model, priced the same as its predecessor
Claude Sonnet 4.6
Anthropic's default Sonnet model, built to approach Opus-class results at a fifth of the price
DeepSeek V3.2
DeepSeek's V3.1 update, built around a new sparse attention mechanism
Gemini 3 Deep Think
Google's extended-reasoning mode for the hardest math and science problems
Gemini 3 Flash
Google's Flash-tier model bringing Gemini 3 Pro reasoning to lower cost and latency
Gemini 3.1 Pro
Google's generally available successor to Gemini 3 Pro
Gemini 3.5 Flash
Google's default Flash model, running about four times faster than rival frontier models
GLM-5
Zhipu's frontier open-weight model, trained entirely on domestic Huawei hardware
GPT-5 Pro
The highest-compute variant of GPT-5, built for accuracy over speed
GPT-5.1
OpenAI's coding-focused GPT-5 refresh with configurable reasoning effort
GPT-5.2
OpenAI's three-tier GPT-5.2 release: Instant, Thinking, and Pro
GPT-5.4
The first OpenAI model to ship with native computer use built in
GPT-5.5
OpenAI's model that led the ARC-AGI-2 leaderboard at launch
Grok 4.20
xAI's multi-agent model that debates internally before answering
Inkling
The first open-weight model from Mira Murati's Thinking Machines Lab
Kimi K2.5
Moonshot's native multimodal update to K2, built around coordinating swarms of sub-agents
MiniMax M2.5
MiniMax's open-weight model extending coding skill into general office work
Amazon Nova 2 Pro
Amazon's second-generation flagship Bedrock model
Command A2
Cohere's enterprise-focused flagship language model
Llama 5
Meta's next flagship open-weight multimodal model, built for on-device and datacenter use alike
Mistral Large 4
Mistral AI's flagship reasoning and coding model
Nemotron 5
NVIDIA's open-weight model tuned for efficient inference on its own hardware
Qwen4-Max
Alibaba's flagship proprietary model for its Qwen API and cloud platform
DeepSeek Chat
AI assistant for reasoning, coding, and general questions
Qwen Studio
Multimodal AI assistant powered by Qwen models
abab6.5
MiniMax's proprietary flagship model prior to its open-weight pivot
ALBERT
Google's parameter-sharing variant of BERT designed for efficiency at scale
Amazon Nova Lite
Amazon's low-cost multimodal Bedrock model tier
Amazon Nova Micro
Amazon's fastest, lowest-cost Nova tier for simple text tasks
Amazon Nova Premier
Amazon's most capable Nova tier for complex multi-step tasks
Amazon Nova Pro
Amazon's flagship multimodal foundation model on Bedrock
Apple Foundation Model
Apple's on-device and server foundation models powering Apple Intelligence
Aquila
BAAI's first open bilingual (Chinese/English) foundation model
Aquila2
BAAI's second-generation bilingual foundation model family
Arctic
Snowflake's open-weight enterprise-focused model, optimized for SQL and RAG
Aya 23
Cohere For AI's open-weight multilingual model covering 23 languages
Baichuan 2
Baichuan's open-weight bilingual foundation model
Baichuan 3
Baichuan's proprietary flagship model succeeding the open-weight Baichuan 2
Baichuan 4
Baichuan's flagship Chinese-language foundation model
Baichuan-13B
Baichuan's open-weight bilingual model, predating the Baichuan 2/3/4 generations
BitNet b1.58
Microsoft's research model demonstrating competitive LLM performance with 1.58-bit ternary weights
BLOOM
The open, multilingual model built by the BigScience collaboration
BLOOMZ
An instruction-tuned variant of BLOOM trained to follow cross-lingual instructions
BTLM-3B
Cerebras' compact open-weight model tuned to punch above its parameter count
ByT5
Google's token-free variant of T5 that operates directly on raw bytes
c2 model
Character.AI's in-house conversational model powering its chat companions
Cerebras-GPT
Cerebras' fully open family of models trained on its wafer-scale chips
ChatGLM3
Zhipu's third-generation open-weight bilingual dialogue model
Chinchilla
DeepMind's 2022 research model that reset scaling-law assumptions for LLM training
Claude 1
Anthropic's first publicly released Claude model
Claude 2
Anthropic's second-generation Claude, ahead of the Claude 2.1 refresh
Claude 2.1
Anthropic's 2023 model that pushed context length to 200K tokens
Claude 3 Haiku
Anthropic's fastest model in the original Claude 3 family
Claude 3 Opus
Anthropic's 2024 flagship, the top tier of the original Claude 3 family
Claude 3 Sonnet
The balanced mid-tier of Anthropic's original Claude 3 lineup
Claude 3.5 Haiku
Anthropic's fast, affordable model matching the prior Claude 3 Opus on some benchmarks
Claude 3.5 Sonnet
The 2024 Claude release that popularized computer-use agent capabilities
Claude 3.7 Sonnet
Anthropic's first hybrid reasoning model, blending fast answers with extended thinking
Claude Fable 5
Anthropic's first Mythos-class model, tuned for the hardest coding benchmarks
Claude Haiku 4.5
Anthropic's fastest and cheapest current-generation model
Claude Instant
Anthropic's original fast, low-cost model tier
Claude Opus 4
Anthropic's flagship model at launch, built for sustained, long-horizon agentic work
Claude Opus 4.1
An incremental refinement of Claude Opus 4 with improved coding and agentic accuracy
Claude Opus 4.8
Anthropic's most capable Opus-tier model for deep analysis and long-horizon tasks
Claude Sonnet 4
Anthropic's mid-tier Claude 4 model, balancing coding strength with everyday cost
Claude Sonnet 4.5
Anthropic's most capable Sonnet-tier model at launch, tuned heavily for coding and computer use
Claude Sonnet 5
The first Sonnet-class Claude model to top Anthropic's flagship writing benchmarks
Code Llama
Meta's open-weight code-specialized Llama variant
Codestral
Mistral's code-generation and completion model
CodeT5+
Salesforce's open-weight encoder-decoder code model family
Command A
Cohere's flagship enterprise model, tuned for agentic and RAG workloads
Command Light
Cohere's smallest, fastest Command tier
Command R Plus
Cohere's flagship RAG- and tool-use-optimized model above Command R
Command R+
Cohere's retrieval-augmented-generation-optimized model
Command R7B
Cohere's smallest current-generation Command model
DBRX
Databricks' open-weight mixture-of-experts model
DBRX Instruct
Instruction-tuned variant of Databricks' open-weight DBRX model
DeepSeek Coder V2
DeepSeek's open-weight code-specialized mixture-of-experts model
DeepSeek LLM 67B
DeepSeek's first dense open-weight foundation model
DeepSeek R1
DeepSeek's reasoning model that matched closed frontier models on math and code
DeepSeek V3
DeepSeek's efficient mixture-of-experts model trained at a fraction of rival costs
DeepSeek V4-Pro
The top open-weight model of 2026, leading on agentic coding and graduate reasoning
DeepSeek-Math
DeepSeek's open-weight model specialized for mathematical reasoning
DeepSeek-MoE-16B
DeepSeek's early fine-grained mixture-of-experts model that informed its later MoE architectures
DeepSeek-R1-Distill-Llama-70B
A Llama-based distillation of DeepSeek-R1's reasoning traces, the largest distilled variant
DeepSeek-R1-Distill-Qwen-32B
A Qwen-based distillation of DeepSeek-R1's reasoning traces into a smaller open model
DeepSeek-V2
DeepSeek's 2024 mixture-of-experts model that undercut rivals on API price
DeepSeek-VL2
DeepSeek's open-weight vision-language mixture-of-experts model
DistilBERT
Hugging Face's distilled, 40%-smaller version of BERT that retains most of its performance
Dolly 2.0
Databricks' first fully open, commercially-usable instruction-following model
Doubao 1.5 Pro
ByteDance's flagship foundation model behind its Doubao assistant
Doubao-Seed 1.6
ByteDance's reasoning-tuned tier of its Doubao foundation model
ELECTRA
Google's more sample-efficient pretraining approach using a replaced-token-detection objective
ELMo
AI2's deep contextualized word representation model that predated the transformer era
ERNIE 4.0
Baidu's flagship model prior to the 4.5 generation
ERNIE 4.5
Baidu's flagship multimodal foundation model
EXAONE 3.0
LG's earlier EXAONE generation, preceding EXAONE 3.5
EXAONE 3.5
LG's open-weight bilingual Korean-English model family
Falcon 180B
TII's open-weight model that led open leaderboards on release
Falcon 3
TII's efficient small-model open-weight family
Falcon 40B
TII's earlier open-weight model that led leaderboards before Falcon 180B
Falcon 7B
TII's compact Falcon tier that helped popularize the original Falcon release
Falcon Mamba 7B
TII's open-weight model built on the state-space Mamba architecture
Flan-T5
Google's instruction-tuned, open-weight successor to T5
Fuyu-8B
Adept's open-weight multimodal model built for digital-agent perception
Galactica
Meta's 2022 model trained on scientific literature, pulled days after launch
Gemini 1.0 Pro
The mid-tier of Google's first Gemini generation
Gemini 1.0 Ultra
Google's first Gemini-generation flagship, launched at the start of 2024
Gemini 1.5 Flash
Google's fast, low-cost Gemini 1.5 tier, distilled from Gemini 1.5 Pro
Gemini 1.5 Pro
Google's 2024 model that introduced the 1M-token context window
Gemini 2.0 Flash
Google's low-latency 2.0-generation multimodal model
Gemini 2.0 Flash Thinking
Early experimental reasoning variant of Gemini 2.0 Flash that shows its chain of thought
Gemini 2.5 Flash
Google's low-latency, cost-efficient Gemini tier
Gemini 2.5 Pro
Google's 2025 reasoning-focused Gemini release
Gemini 3 Pro
Google's flagship multimodal model with a 1M-token context window
Gemini 3.5 Pro
Google's next-generation frontier model, announced but not yet fully released
Gemini Nano
Google's on-device Gemini tier built into Pixel and Android
Gemma
Google's original open-weight Gemma release, built from the same research as Gemini
Gemma 2
Google's prior-generation open-weight model family
Gemma 3
Google's open-weight model family built from Gemini research
GLM-4
Zhipu's 2024 general-purpose open-weight model
GLM-4-Plus
Zhipu's proprietary flagship tier alongside the open-weight GLM line
GLM-4.6
Zhipu's value-tier open-weight model for day-to-day coding
GLM-5.2
Zhipu's flagship open-weight model, shipped alongside Kimi K2.7 in mid-2026
Gopher
DeepMind's 280B-parameter research model that informed the Chinchilla scaling laws
GPT-1
OpenAI's original 2018 paper model that introduced the GPT architecture
GPT-2
OpenAI's 2019 model, once withheld from release over misuse concerns
GPT-3
The 2020 model that first showed large-scale few-shot learning was possible
GPT-3 Davinci
The largest of the original GPT-3 model sizes
GPT-3.5 Turbo
The model that powered ChatGPT's original public launch
GPT-4
The 2023 release that reset expectations for what LLMs could do
GPT-4 Turbo
Faster, cheaper successor to GPT-4 with a larger context window
GPT-4.1
OpenAI's 2025 model with a 1M-token context window, tuned for coding
GPT-4.5
OpenAI's largest pre-GPT-5 model, tuned for natural conversation
GPT-4o
OpenAI's omni model, natively multimodal across text, image, and audio
GPT-4o mini
OpenAI's small, cost-efficient multimodal model
GPT-5
OpenAI's unified reasoning and chat model family
GPT-5.6 Luna
OpenAI's fast, cost-efficient GPT-5.6 tier for high-volume, latency-sensitive tasks
GPT-5.6 Sol
OpenAI's flagship model for advanced math, science, and cybersecurity reasoning
GPT-5.6 Terra
OpenAI's balanced GPT-5.6 tier for everyday coding, reasoning, and agentic tasks
GPT-J
EleutherAI's early open-weight GPT-3-style model
GPT-Neo 2.7B
EleutherAI's early GPT-3-style open model, a predecessor to GPT-J and GPT-NeoX
GPT-NeoX-20B
EleutherAI's open-weight model, one of the largest public checkpoints of its era
GPT-OSS-120B
OpenAI's first open-weight model release since GPT-2, matching o4-mini on many reasoning benchmarks
GPT-OSS-20B
The smaller, single-GPU-friendly tier of OpenAI's open-weight model release
Granite 13B
IBM's earlier enterprise foundation model, preceding the Granite 3 generation
Granite 3.0
IBM's open-weight enterprise model family for regulated industries
Granite Code
IBM's open-weight code generation model family
Grok 3
xAI's 2025 model trained on the Colossus supercomputer
Grok 4
xAI's 2025 flagship reasoning model
Grok 4 Fast
xAI's low-latency tier of Grok 4 for high-throughput use
Grok 4.5
xAI's coding-focused frontier model
Grok-1
xAI's first model, later open-weighted under Apache 2.0
Grok-1.5
xAI's transitional model that added long-context reasoning ahead of Grok 2
Grok-2
xAI's 2024 model that added native image generation
Hermes 2
Nous Research's earlier open-weight fine-tune line
Hermes 3
Nous Research's open-weight fine-tune focused on steerability and reasoning
Hunyuan Large
Tencent's open-weight mixture-of-experts model
Hunyuan Turbo
Tencent's low-latency Hunyuan tier for production workloads
Hunyuan-A13B
Tencent's open-weight mixture-of-experts reasoning model
HyperCLOVA X
Naver's flagship model tuned for Korean language and culture
Inflection-1
Inflection's original model that first powered the Pi assistant
Inflection-2.5
Inflection's model built to power the empathetic Pi assistant
InstructGPT
OpenAI's 2022 model that introduced RLHF instruction-tuning at scale
InternLM2
Shanghai AI Lab's open-weight foundation model with strong long-context and reasoning ability
InternLM2.5
Shanghai AI Lab's updated InternLM generation with stronger tool-use and reasoning
Jamba 1.5 Large
AI21's hybrid transformer-Mamba model built for very long context
Jamba 1.5 Mini
AI21's compact hybrid transformer-Mamba model
Jamba Large 1.6
AI21's updated flagship hybrid transformer-Mamba model
Jurassic-1
AI21's first large language model, a GPT-3 era contemporary
Jurassic-2
AI21's earlier proprietary large language model line
Jurassic-2 Mid
The mid-sized tier of AI21's Jurassic-2 family, preceding the Jamba architecture switch
Kimi K1.5
Moonshot's first reasoning-tuned Kimi release
Kimi K2
Moonshot's original K2-generation open-weight base model
Kimi K2.6
Moonshot's open-weight model built for long-running agentic and tool-use tasks
Kimi K2.7 Code
Moonshot's open-weight agentic coding model, tuned for tool-use and long-running tasks
Kimi-VL
Moonshot's open-weight vision-language model
LaMDA
Google's 2021 conversational model that underpinned the original Bard
LLaMA
Meta's original 2023 open-weight release that kicked off the open LLM boom
Llama 2
Meta's 2023 open-weight release that jump-started the open LLM ecosystem
Llama 3
Meta's 2024 open-weight release that closed most of the gap to closed models
Llama 3.1 405B
Meta's largest dense open-weight model, competitive with closed frontier models
Llama 3.1 70B
The mid-sized tier of Meta's Llama 3.1 release, between the 8B and 405B models
Llama 3.1 8B
Meta's compact tier of the Llama 3.1 generation
Llama 3.2
Meta's first Llama generation with vision-capable and on-device tiers
Llama 3.3
Meta's efficient 70B open-weight model matching larger predecessors
Llama 4 Maverick
Meta's flagship open-weight multimodal mixture-of-experts model
Llama 4 Scout
Meta's smaller, faster Llama 4 tier with a 10M-token context window
Llama Guard 3
Meta's open-weight safety classifier model for moderating LLM input and output
Luminous
Aleph Alpha's European enterprise foundation model with source-attribution features
Med-PaLM
Google's first LLM tuned to answer medical exam and consumer health questions
Med-PaLM 2
Google's medical LLM that reached expert-level scores on US medical licensing exam questions
Megatron-Turing NLG
NVIDIA and Microsoft's 530B-parameter research model, among the largest dense LLMs of its time
MiniCPM
OpenBMB's efficient small model designed to run well on phones and edge devices
MiniMax M3
First open-weight model to combine frontier coding, a 1M-token context, and native multimodality
MiniMax-01
MiniMax's 2025 open-weight model with a 4M-token context window
Ministral 8B
Mistral's compact model built for on-device and edge deployment
Mistral 7B
Mistral's original compact open-weight model that outperformed larger rivals
Mistral Large 2
Mistral's 2024 flagship model, the generation preceding Mistral Large 3
Mistral Large 3
Mistral's flagship frontier-tier language model
Mistral Medium 3
Mistral's mid-tier model balancing cost and frontier-level performance
Mistral NeMo
Mistral's 12B open-weight model built with NVIDIA
Mistral Saba
Mistral's regional model tuned for Arabic and South Asian languages
Mistral Small 3
Mistral's efficient small model tier for latency-sensitive workloads
Mixtral 8x22B
Mistral's open-weight sparse mixture-of-experts model
Mixtral 8x7B
Mistral's original open-weight mixture-of-experts model that popularised sparse MoE LLMs
MPT-30B
MosaicML's open-weight commercial-use model, later folded into Databricks
mT5
Google's multilingual variant of T5, covering 101 languages
Nemotron Nano
NVIDIA's small, efficient Nemotron tier optimised for edge deployment
Nemotron-3 8B
NVIDIA's earlier compact enterprise-focused Nemotron model
Nemotron-4 340B
NVIDIA's open-weight model built to generate synthetic training data
o1
OpenAI's first reasoning model trained to think before answering
o1-mini
OpenAI's first compact reasoning model, tuned for coding and math
o1-preview
Public preview of OpenAI's first reasoning model, ahead of the full o1 release
o1-pro
Higher-compute variant of o1 for the hardest reasoning tasks
o3
OpenAI's deep multi-step reasoning model for hard technical problems
o3-mini
OpenAI's small, fast reasoning model for STEM tasks
o4-mini
OpenAI's fast, low-cost reasoning model for high-volume agentic use
OLMo 2
AI2's fully open model, released with training data and code alongside weights
OLMo 7B
AI2's first fully open model release, with open weights, data, and training code
OLMoE
AI2's fully open mixture-of-experts model
OPT-175B
Meta's 2022 open-weight model released to mirror GPT-3's scale
Orca 2
Microsoft Research's reasoning-focused fine-tune of Llama 2
PaLM
Google's 2022 540B-parameter model that set the stage for Gemini
PaLM 2
Google's 2023 large language model that powered the original Bard
Palmyra Creative
Writer's model tuned for marketing and brand-voice writing tasks
Palmyra Med
Writer's domain-tuned model for clinical and medical text tasks
Palmyra X 004
Writer's enterprise foundation model with built-in tool calling
Palmyra-Fin
Writer's finance-specialized enterprise foundation model
PanGu-Alpha
Huawei's first large-scale Chinese autoregressive language model
PanGu-Sigma
Huawei's trillion-parameter sparse mixture-of-experts model
Pharia-1
Aleph Alpha's open-weight European-sovereignty-focused model
Phi-2
Microsoft's earlier 2.7B model demonstrating outsized small-model performance
Phi-3
Microsoft's compact open-weight model family for on-device use
Phi-3-medium
The 14B-parameter tier of Microsoft's Phi-3 family, tuned for stronger reasoning
Phi-3-small
The 7B-parameter tier of Microsoft's Phi-3 small-language-model family
Phi-3.5-mini
Microsoft's updated compact model in the Phi-3 line
Phi-4
Microsoft's small reasoning-focused model that punches above its parameter count
Phi-4-reasoning
Reasoning-tuned variant of Phi-4 trained with chain-of-thought supervision
Pixtral Large
Mistral's flagship vision-language model
Pythia
EleutherAI's fully open suite of models built for interpretability research
Qwen-7B
Alibaba's original Qwen release that started the model family
Qwen-Max
Alibaba's largest proprietary Qwen tier, served via API only
Qwen1.5-110B
Alibaba's largest dense model in the Qwen1.5 generation
Qwen2-72B
Alibaba's second-generation flagship dense open-weight model
Qwen2.5-72B
Alibaba's widely-deployed dense open-weight model
Qwen2.5-Max
Alibaba's largest proprietary MoE model, positioned against GPT-4o and DeepSeek-V3
Qwen2.5-VL
Alibaba's open-weight vision-language model
Qwen3-235B-A22B
Alibaba's flagship open-weight mixture-of-experts model, the cheapest capable frontier model
Qwen3-Coder
Alibaba's open-weight agentic coding model
QwQ-32B
Alibaba's open-weight reasoning model competitive with much larger models
QwQ-32B-Preview
Alibaba's first public preview of its QwQ reasoning line, ahead of the full QwQ-32B release
RedPajama-INCITE
Together AI's fully open model trained on the reproduced RedPajama dataset
Reka Core
Reka's flagship multimodal foundation model
Reka Edge
Reka's smallest, on-device-friendly model tier
Reka Flash
Reka's efficient mid-tier multimodal model
replit-code-v1.5
Replit's open-weight code-completion model
RETRO
DeepMind's retrieval-augmented language model that matches larger models with far fewer parameters
RoBERTa
Meta's robustly-optimized retraining of BERT that improved on it across benchmarks
Samba-1
SambaNova's composition-of-experts enterprise model
Samsung Gauss
Samsung's original in-house foundation model, ahead of Gauss 2
Samsung Gauss 2
Samsung's in-house foundation model powering Galaxy AI features
SenseChat 5
SenseTime's flagship general-purpose foundation model
SenseChat-3
SenseTime's third-generation SenseChat model, preceding SenseChat-5
SenseNova 5.5
SenseTime's flagship multimodal foundation model
Skywork-13B
Kunlun Tech's open-weight bilingual foundation model
SmolLM
Hugging Face's original family of small, fully open language models
SmolLM2
Hugging Face's fully open small-model family for on-device use
Solar 10.7B
Upstage's depth-upscaled open model that preceded the Solar Pro line
Solar Pro 2
Upstage's efficient open-weight model that punches above its parameter count
Sonar Large
Perplexity's larger web-grounded model, ahead of the Sonar Pro tier
Sonar Pro
Perplexity's search-grounded answer model built on top of open-weight LLMs
Sonar Reasoning Pro
Perplexity's search-grounded reasoning model
Sonar Small
Perplexity's low-cost web-grounded search model tier
Spark 4.0
iFlytek's flagship Chinese-language foundation model
Spark 4.0 Ultra
iFlytek's top-tier Spark model with enhanced mathematical reasoning
Sparrow
DeepMind's dialogue agent research model trained to be more helpful and less harmful
StableLM 2
Stability AI's open-weight small language model
StableLM 3B
Stability AI's early compact open-weight language model, preceding StableLM 2
StableLM Zephyr
Stability AI's chat-tuned small model built on StableLM 3B
StarCoder2
The BigCode collaboration's open-weight code generation model
Step-1
StepFun's earlier flagship model that preceded Step-2
Step-2
StepFun's trillion-parameter reasoning model
T5
Google's text-to-text transfer transformer, an early unifying NLP framework
text-davinci-003
The last and most capable of OpenAI's original GPT-3.5 completion models
Titan Text
Amazon's original in-house foundation model line for Bedrock
Titan Text Express
Amazon's mid-tier Titan model for general text generation, between Lite and the base Titan Text
Turing-NLG
Microsoft's 17B-parameter model that was among the largest published language models of its time
UL2
Google's unified pretraining framework blending multiple denoising objectives
Vicuna-13B
LMSYS's influential early open chat model fine-tuned from Llama on ShareGPT conversations
Vicuna-33B
The largest Vicuna release, LMSYS's ShareGPT fine-tune of Llama
WizardLM-2
Microsoft's instruction-tuned model family built with complex, evolved synthetic training data
XGen-7B
Salesforce's open-weight long-context research model
XLNet
A permutation-based autoregressive pretraining model that outperformed BERT on many tasks
Yi-1.5-34B
01.AI's improved open-weight Yi generation
Yi-34B
01.AI's open-weight bilingual model
Yi-Coder
01.AI's open-weight code-specialized model
Yi-Large
01.AI's flagship proprietary model, competitive with GPT-4-class models on Chinese benchmarks
Yuan 2.0
Inspur's open-weight foundation model tuned for enterprise Chinese-language tasks
Zephyr 141B
Hugging Face's larger mixture-of-experts Zephyr fine-tune, distinct from the 7B release
Zephyr 7B
Hugging Face's open-weight fine-tune that popularized DPO alignment
Landbot
Landbot is a no-code chatbot builder for creating conversational flows on …
SambaNova
SambaNova is an AI hardware and software platform built for running large …
AI21 Labs
Enterprise AI platform with Jamba hybrid state-space models
ChatGPT
Advanced AI chatbot by OpenAI for conversations and assistance
Claude
AI assistant by Anthropic for helpful, harmless, and honest conversations
CoreWeave
GPU cloud infrastructure for AI workloads
Gemini
Google's AI chatbot with multimodal capabilities
Gemini Advanced
Google's most capable AI assistant with Gemini Ultra
Hugging Face
The AI community platform for models, datasets, and spaces
Kimi
Moonshot AI's long-context chatbot with 1M token window
Lambda Labs
GPU cloud and workstations for deep learning
Le Chat
Mistral AI's free conversational assistant with web search
Modal
Serverless GPU cloud for AI and data workloads
Paperspace
Cloud GPU platform for AI development and deployment
Perplexity
AI-powered search engine and chatbot for research
Poe
AI chatbot aggregator giving access to GPT-4, Claude, and more
RunPod
GPU cloud for AI training and inference
Stability AI
The company behind Stable Diffusion and open generative AI
Together AI
Fast inference cloud for open-source AI models
Vast.ai
GPU marketplace for affordable AI compute
Venice AI
Private AI chat and image generation with no data retention
01.AI
Open-source Yi large language models by 01.AI
Aleph Alpha
European sovereign AI models for enterprise and government
Nous Research
Open-source AI research and fine-tuned model community
Reka AI
Multimodal AI models for text, image, and video understanding
ChatPlayground AI
Compare the best AI models side by side in one place
Cohere
Enterprise LLMs for search, summarisation, and classification
Grok
xAI's AI assistant with real-time X/Twitter data
Microsoft Copilot
Microsoft's AI assistant across Windows, Office, and the web