#

Llm

AI tools tagged Llm.

200+ tools · handpicked & curated
// price
Claude Opus 4.5 Anthropic's Opus-tier model that cut pricing by two-thirds while topping SWE-bench
Claude Opus 4.6 Anthropic's Opus update that expanded the context window to 1 million tokens
Claude Opus 4.7 An Opus 4.6 upgrade targeted at the hardest software engineering tasks
Claude Opus 5 Anthropic's newest Opus-tier model, priced the same as its predecessor
Claude Sonnet 4.6 Anthropic's default Sonnet model, built to approach Opus-class results at a fifth of the price
DeepSeek V3.2 DeepSeek's V3.1 update, built around a new sparse attention mechanism
Gemini 3 Deep Think Google's extended-reasoning mode for the hardest math and science problems
Gemini 3 Flash Google's Flash-tier model bringing Gemini 3 Pro reasoning to lower cost and latency
Gemini 3.1 Pro Google's generally available successor to Gemini 3 Pro
Gemini 3.5 Flash Google's default Flash model, running about four times faster than rival frontier models
GLM-5 Zhipu's frontier open-weight model, trained entirely on domestic Huawei hardware
GPT-5 Pro The highest-compute variant of GPT-5, built for accuracy over speed
GPT-5.1 OpenAI's coding-focused GPT-5 refresh with configurable reasoning effort
GPT-5.2 OpenAI's three-tier GPT-5.2 release: Instant, Thinking, and Pro
GPT-5.4 The first OpenAI model to ship with native computer use built in
GPT-5.5 OpenAI's model that led the ARC-AGI-2 leaderboard at launch
Grok 4.20 xAI's multi-agent model that debates internally before answering
Inkling The first open-weight model from Mira Murati's Thinking Machines Lab
Kimi K2.5 Moonshot's native multimodal update to K2, built around coordinating swarms of sub-agents
MiniMax M2.5 MiniMax's open-weight model extending coding skill into general office work
Amazon Nova 2 Pro Amazon's second-generation flagship Bedrock model
Command A2 Cohere's enterprise-focused flagship language model
Llama 5 Meta's next flagship open-weight multimodal model, built for on-device and datacenter use alike
Mistral Large 4 Mistral AI's flagship reasoning and coding model
Nemotron 5 NVIDIA's open-weight model tuned for efficient inference on its own hardware
Qwen4-Max Alibaba's flagship proprietary model for its Qwen API and cloud platform
DeepSeek Chat AI assistant for reasoning, coding, and general questions
Qwen Studio Multimodal AI assistant powered by Qwen models
abab6.5 MiniMax's proprietary flagship model prior to its open-weight pivot
ALBERT Google's parameter-sharing variant of BERT designed for efficiency at scale
Amazon Nova Lite Amazon's low-cost multimodal Bedrock model tier
Amazon Nova Micro Amazon's fastest, lowest-cost Nova tier for simple text tasks
Amazon Nova Premier Amazon's most capable Nova tier for complex multi-step tasks
Amazon Nova Pro Amazon's flagship multimodal foundation model on Bedrock
Apple Foundation Model Apple's on-device and server foundation models powering Apple Intelligence
Aquila BAAI's first open bilingual (Chinese/English) foundation model
Aquila2 BAAI's second-generation bilingual foundation model family
Arctic Snowflake's open-weight enterprise-focused model, optimized for SQL and RAG
Aya 23 Cohere For AI's open-weight multilingual model covering 23 languages
Baichuan 2 Baichuan's open-weight bilingual foundation model
Baichuan 3 Baichuan's proprietary flagship model succeeding the open-weight Baichuan 2
Baichuan 4 Baichuan's flagship Chinese-language foundation model
Baichuan-13B Baichuan's open-weight bilingual model, predating the Baichuan 2/3/4 generations
BitNet b1.58 Microsoft's research model demonstrating competitive LLM performance with 1.58-bit ternary weights
BLOOM The open, multilingual model built by the BigScience collaboration
BLOOMZ An instruction-tuned variant of BLOOM trained to follow cross-lingual instructions
BTLM-3B Cerebras' compact open-weight model tuned to punch above its parameter count
ByT5 Google's token-free variant of T5 that operates directly on raw bytes
c2 model Character.AI's in-house conversational model powering its chat companions
Cerebras-GPT Cerebras' fully open family of models trained on its wafer-scale chips
ChatGLM3 Zhipu's third-generation open-weight bilingual dialogue model
Chinchilla DeepMind's 2022 research model that reset scaling-law assumptions for LLM training
Claude 1 Anthropic's first publicly released Claude model
Claude 2 Anthropic's second-generation Claude, ahead of the Claude 2.1 refresh
Claude 2.1 Anthropic's 2023 model that pushed context length to 200K tokens
Claude 3 Haiku Anthropic's fastest model in the original Claude 3 family
Claude 3 Opus Anthropic's 2024 flagship, the top tier of the original Claude 3 family
Claude 3 Sonnet The balanced mid-tier of Anthropic's original Claude 3 lineup
Claude 3.5 Haiku Anthropic's fast, affordable model matching the prior Claude 3 Opus on some benchmarks
Claude 3.5 Sonnet The 2024 Claude release that popularized computer-use agent capabilities
Claude 3.7 Sonnet Anthropic's first hybrid reasoning model, blending fast answers with extended thinking
Claude Fable 5 Anthropic's first Mythos-class model, tuned for the hardest coding benchmarks
Claude Haiku 4.5 Anthropic's fastest and cheapest current-generation model
Claude Instant Anthropic's original fast, low-cost model tier
Claude Opus 4 Anthropic's flagship model at launch, built for sustained, long-horizon agentic work
Claude Opus 4.1 An incremental refinement of Claude Opus 4 with improved coding and agentic accuracy
Claude Opus 4.8 Anthropic's most capable Opus-tier model for deep analysis and long-horizon tasks
Claude Sonnet 4 Anthropic's mid-tier Claude 4 model, balancing coding strength with everyday cost
Claude Sonnet 4.5 Anthropic's most capable Sonnet-tier model at launch, tuned heavily for coding and computer use
Claude Sonnet 5 The first Sonnet-class Claude model to top Anthropic's flagship writing benchmarks
Code Llama Meta's open-weight code-specialized Llama variant
Codestral Mistral's code-generation and completion model
CodeT5+ Salesforce's open-weight encoder-decoder code model family
Command A Cohere's flagship enterprise model, tuned for agentic and RAG workloads
Command Light Cohere's smallest, fastest Command tier
Command R Plus Cohere's flagship RAG- and tool-use-optimized model above Command R
Command R+ Cohere's retrieval-augmented-generation-optimized model
Command R7B Cohere's smallest current-generation Command model
DBRX Databricks' open-weight mixture-of-experts model
DBRX Instruct Instruction-tuned variant of Databricks' open-weight DBRX model
DeepSeek Coder V2 DeepSeek's open-weight code-specialized mixture-of-experts model
DeepSeek LLM 67B DeepSeek's first dense open-weight foundation model
DeepSeek R1 DeepSeek's reasoning model that matched closed frontier models on math and code
DeepSeek V3 DeepSeek's efficient mixture-of-experts model trained at a fraction of rival costs
DeepSeek V4-Pro The top open-weight model of 2026, leading on agentic coding and graduate reasoning
DeepSeek-Math DeepSeek's open-weight model specialized for mathematical reasoning
DeepSeek-MoE-16B DeepSeek's early fine-grained mixture-of-experts model that informed its later MoE architectures
DeepSeek-R1-Distill-Llama-70B A Llama-based distillation of DeepSeek-R1's reasoning traces, the largest distilled variant
DeepSeek-R1-Distill-Qwen-32B A Qwen-based distillation of DeepSeek-R1's reasoning traces into a smaller open model
DeepSeek-V2 DeepSeek's 2024 mixture-of-experts model that undercut rivals on API price
DeepSeek-VL2 DeepSeek's open-weight vision-language mixture-of-experts model
DistilBERT Hugging Face's distilled, 40%-smaller version of BERT that retains most of its performance
Dolly 2.0 Databricks' first fully open, commercially-usable instruction-following model
Doubao 1.5 Pro ByteDance's flagship foundation model behind its Doubao assistant
Doubao-Seed 1.6 ByteDance's reasoning-tuned tier of its Doubao foundation model
ELECTRA Google's more sample-efficient pretraining approach using a replaced-token-detection objective
ELMo AI2's deep contextualized word representation model that predated the transformer era
ERNIE 4.0 Baidu's flagship model prior to the 4.5 generation
ERNIE 4.5 Baidu's flagship multimodal foundation model
EXAONE 3.0 LG's earlier EXAONE generation, preceding EXAONE 3.5
EXAONE 3.5 LG's open-weight bilingual Korean-English model family
Falcon 180B TII's open-weight model that led open leaderboards on release
Falcon 3 TII's efficient small-model open-weight family
Falcon 40B TII's earlier open-weight model that led leaderboards before Falcon 180B
Falcon 7B TII's compact Falcon tier that helped popularize the original Falcon release
Falcon Mamba 7B TII's open-weight model built on the state-space Mamba architecture
Flan-T5 Google's instruction-tuned, open-weight successor to T5
Fuyu-8B Adept's open-weight multimodal model built for digital-agent perception
Galactica Meta's 2022 model trained on scientific literature, pulled days after launch
Gemini 1.0 Pro The mid-tier of Google's first Gemini generation
Gemini 1.0 Ultra Google's first Gemini-generation flagship, launched at the start of 2024
Gemini 1.5 Flash Google's fast, low-cost Gemini 1.5 tier, distilled from Gemini 1.5 Pro
Gemini 1.5 Pro Google's 2024 model that introduced the 1M-token context window
Gemini 2.0 Flash Google's low-latency 2.0-generation multimodal model
Gemini 2.0 Flash Thinking Early experimental reasoning variant of Gemini 2.0 Flash that shows its chain of thought
Gemini 2.5 Flash Google's low-latency, cost-efficient Gemini tier
Gemini 2.5 Pro Google's 2025 reasoning-focused Gemini release
Gemini 3 Pro Google's flagship multimodal model with a 1M-token context window
Gemini 3.5 Pro Google's next-generation frontier model, announced but not yet fully released
Gemini Nano Google's on-device Gemini tier built into Pixel and Android
Gemma Google's original open-weight Gemma release, built from the same research as Gemini
Gemma 2 Google's prior-generation open-weight model family
Gemma 3 Google's open-weight model family built from Gemini research
GLM-4 Zhipu's 2024 general-purpose open-weight model
GLM-4-Plus Zhipu's proprietary flagship tier alongside the open-weight GLM line
GLM-4.6 Zhipu's value-tier open-weight model for day-to-day coding
GLM-5.2 Zhipu's flagship open-weight model, shipped alongside Kimi K2.7 in mid-2026
Gopher DeepMind's 280B-parameter research model that informed the Chinchilla scaling laws
GPT-1 OpenAI's original 2018 paper model that introduced the GPT architecture
GPT-2 OpenAI's 2019 model, once withheld from release over misuse concerns
GPT-3 The 2020 model that first showed large-scale few-shot learning was possible
GPT-3 Davinci The largest of the original GPT-3 model sizes
GPT-3.5 Turbo The model that powered ChatGPT's original public launch
GPT-4 The 2023 release that reset expectations for what LLMs could do
GPT-4 Turbo Faster, cheaper successor to GPT-4 with a larger context window
GPT-4.1 OpenAI's 2025 model with a 1M-token context window, tuned for coding
GPT-4.5 OpenAI's largest pre-GPT-5 model, tuned for natural conversation
GPT-4o OpenAI's omni model, natively multimodal across text, image, and audio
GPT-4o mini OpenAI's small, cost-efficient multimodal model
GPT-5 OpenAI's unified reasoning and chat model family
GPT-5.6 Luna OpenAI's fast, cost-efficient GPT-5.6 tier for high-volume, latency-sensitive tasks
GPT-5.6 Sol OpenAI's flagship model for advanced math, science, and cybersecurity reasoning
GPT-5.6 Terra OpenAI's balanced GPT-5.6 tier for everyday coding, reasoning, and agentic tasks
GPT-J EleutherAI's early open-weight GPT-3-style model
GPT-Neo 2.7B EleutherAI's early GPT-3-style open model, a predecessor to GPT-J and GPT-NeoX
GPT-NeoX-20B EleutherAI's open-weight model, one of the largest public checkpoints of its era
GPT-OSS-120B OpenAI's first open-weight model release since GPT-2, matching o4-mini on many reasoning benchmarks
GPT-OSS-20B The smaller, single-GPU-friendly tier of OpenAI's open-weight model release
Granite 13B IBM's earlier enterprise foundation model, preceding the Granite 3 generation
Granite 3.0 IBM's open-weight enterprise model family for regulated industries
Granite Code IBM's open-weight code generation model family
Grok 3 xAI's 2025 model trained on the Colossus supercomputer
Grok 4 xAI's 2025 flagship reasoning model
Grok 4 Fast xAI's low-latency tier of Grok 4 for high-throughput use
Grok 4.5 xAI's coding-focused frontier model
Grok-1 xAI's first model, later open-weighted under Apache 2.0
Grok-1.5 xAI's transitional model that added long-context reasoning ahead of Grok 2
Grok-2 xAI's 2024 model that added native image generation
Hermes 2 Nous Research's earlier open-weight fine-tune line
Hermes 3 Nous Research's open-weight fine-tune focused on steerability and reasoning
Hunyuan Large Tencent's open-weight mixture-of-experts model
Hunyuan Turbo Tencent's low-latency Hunyuan tier for production workloads
Hunyuan-A13B Tencent's open-weight mixture-of-experts reasoning model
HyperCLOVA X Naver's flagship model tuned for Korean language and culture
Inflection-1 Inflection's original model that first powered the Pi assistant
Inflection-2.5 Inflection's model built to power the empathetic Pi assistant
InstructGPT OpenAI's 2022 model that introduced RLHF instruction-tuning at scale
InternLM2 Shanghai AI Lab's open-weight foundation model with strong long-context and reasoning ability
InternLM2.5 Shanghai AI Lab's updated InternLM generation with stronger tool-use and reasoning
Jamba 1.5 Large AI21's hybrid transformer-Mamba model built for very long context
Jamba 1.5 Mini AI21's compact hybrid transformer-Mamba model
Jamba Large 1.6 AI21's updated flagship hybrid transformer-Mamba model
Jurassic-1 AI21's first large language model, a GPT-3 era contemporary
Jurassic-2 AI21's earlier proprietary large language model line
Jurassic-2 Mid The mid-sized tier of AI21's Jurassic-2 family, preceding the Jamba architecture switch
Kimi K1.5 Moonshot's first reasoning-tuned Kimi release
Kimi K2 Moonshot's original K2-generation open-weight base model
Kimi K2.6 Moonshot's open-weight model built for long-running agentic and tool-use tasks
Kimi K2.7 Code Moonshot's open-weight agentic coding model, tuned for tool-use and long-running tasks
Kimi-VL Moonshot's open-weight vision-language model
LaMDA Google's 2021 conversational model that underpinned the original Bard
LLaMA Meta's original 2023 open-weight release that kicked off the open LLM boom
Llama 2 Meta's 2023 open-weight release that jump-started the open LLM ecosystem
Llama 3 Meta's 2024 open-weight release that closed most of the gap to closed models
Llama 3.1 405B Meta's largest dense open-weight model, competitive with closed frontier models
Llama 3.1 70B The mid-sized tier of Meta's Llama 3.1 release, between the 8B and 405B models
Llama 3.1 8B Meta's compact tier of the Llama 3.1 generation
Llama 3.2 Meta's first Llama generation with vision-capable and on-device tiers
Llama 3.3 Meta's efficient 70B open-weight model matching larger predecessors
Llama 4 Maverick Meta's flagship open-weight multimodal mixture-of-experts model
Llama 4 Scout Meta's smaller, faster Llama 4 tier with a 10M-token context window
Llama Guard 3 Meta's open-weight safety classifier model for moderating LLM input and output
Luminous Aleph Alpha's European enterprise foundation model with source-attribution features
Med-PaLM Google's first LLM tuned to answer medical exam and consumer health questions
Med-PaLM 2 Google's medical LLM that reached expert-level scores on US medical licensing exam questions
Megatron-Turing NLG NVIDIA and Microsoft's 530B-parameter research model, among the largest dense LLMs of its time
MiniCPM OpenBMB's efficient small model designed to run well on phones and edge devices
MiniMax M3 First open-weight model to combine frontier coding, a 1M-token context, and native multimodality
MiniMax-01 MiniMax's 2025 open-weight model with a 4M-token context window
Ministral 8B Mistral's compact model built for on-device and edge deployment
Mistral 7B Mistral's original compact open-weight model that outperformed larger rivals
Mistral Large 2 Mistral's 2024 flagship model, the generation preceding Mistral Large 3
Mistral Large 3 Mistral's flagship frontier-tier language model
Mistral Medium 3 Mistral's mid-tier model balancing cost and frontier-level performance
Mistral NeMo Mistral's 12B open-weight model built with NVIDIA
Mistral Saba Mistral's regional model tuned for Arabic and South Asian languages
Mistral Small 3 Mistral's efficient small model tier for latency-sensitive workloads
Mixtral 8x22B Mistral's open-weight sparse mixture-of-experts model
Mixtral 8x7B Mistral's original open-weight mixture-of-experts model that popularised sparse MoE LLMs
MPT-30B MosaicML's open-weight commercial-use model, later folded into Databricks
mT5 Google's multilingual variant of T5, covering 101 languages
Nemotron Nano NVIDIA's small, efficient Nemotron tier optimised for edge deployment
Nemotron-3 8B NVIDIA's earlier compact enterprise-focused Nemotron model
Nemotron-4 340B NVIDIA's open-weight model built to generate synthetic training data
o1 OpenAI's first reasoning model trained to think before answering
o1-mini OpenAI's first compact reasoning model, tuned for coding and math
o1-preview Public preview of OpenAI's first reasoning model, ahead of the full o1 release
o1-pro Higher-compute variant of o1 for the hardest reasoning tasks
o3 OpenAI's deep multi-step reasoning model for hard technical problems
o3-mini OpenAI's small, fast reasoning model for STEM tasks
o4-mini OpenAI's fast, low-cost reasoning model for high-volume agentic use
OLMo 2 AI2's fully open model, released with training data and code alongside weights
OLMo 7B AI2's first fully open model release, with open weights, data, and training code
OLMoE AI2's fully open mixture-of-experts model
OPT-175B Meta's 2022 open-weight model released to mirror GPT-3's scale
Orca 2 Microsoft Research's reasoning-focused fine-tune of Llama 2
PaLM Google's 2022 540B-parameter model that set the stage for Gemini
PaLM 2 Google's 2023 large language model that powered the original Bard
Palmyra Creative Writer's model tuned for marketing and brand-voice writing tasks
Palmyra Med Writer's domain-tuned model for clinical and medical text tasks
Palmyra X 004 Writer's enterprise foundation model with built-in tool calling
Palmyra-Fin Writer's finance-specialized enterprise foundation model
PanGu-Alpha Huawei's first large-scale Chinese autoregressive language model
PanGu-Sigma Huawei's trillion-parameter sparse mixture-of-experts model
Pharia-1 Aleph Alpha's open-weight European-sovereignty-focused model
Phi-2 Microsoft's earlier 2.7B model demonstrating outsized small-model performance
Phi-3 Microsoft's compact open-weight model family for on-device use
Phi-3-medium The 14B-parameter tier of Microsoft's Phi-3 family, tuned for stronger reasoning
Phi-3-small The 7B-parameter tier of Microsoft's Phi-3 small-language-model family
Phi-3.5-mini Microsoft's updated compact model in the Phi-3 line
Phi-4 Microsoft's small reasoning-focused model that punches above its parameter count
Phi-4-reasoning Reasoning-tuned variant of Phi-4 trained with chain-of-thought supervision
Pixtral Large Mistral's flagship vision-language model
Pythia EleutherAI's fully open suite of models built for interpretability research
Qwen-7B Alibaba's original Qwen release that started the model family
Qwen-Max Alibaba's largest proprietary Qwen tier, served via API only
Qwen1.5-110B Alibaba's largest dense model in the Qwen1.5 generation
Qwen2-72B Alibaba's second-generation flagship dense open-weight model
Qwen2.5-72B Alibaba's widely-deployed dense open-weight model
Qwen2.5-Max Alibaba's largest proprietary MoE model, positioned against GPT-4o and DeepSeek-V3
Qwen2.5-VL Alibaba's open-weight vision-language model
Qwen3-235B-A22B Alibaba's flagship open-weight mixture-of-experts model, the cheapest capable frontier model
Qwen3-Coder Alibaba's open-weight agentic coding model
QwQ-32B Alibaba's open-weight reasoning model competitive with much larger models
QwQ-32B-Preview Alibaba's first public preview of its QwQ reasoning line, ahead of the full QwQ-32B release
RedPajama-INCITE Together AI's fully open model trained on the reproduced RedPajama dataset
Reka Core Reka's flagship multimodal foundation model
Reka Edge Reka's smallest, on-device-friendly model tier
Reka Flash Reka's efficient mid-tier multimodal model
replit-code-v1.5 Replit's open-weight code-completion model
RETRO DeepMind's retrieval-augmented language model that matches larger models with far fewer parameters
RoBERTa Meta's robustly-optimized retraining of BERT that improved on it across benchmarks
Samba-1 SambaNova's composition-of-experts enterprise model
Samsung Gauss Samsung's original in-house foundation model, ahead of Gauss 2
Samsung Gauss 2 Samsung's in-house foundation model powering Galaxy AI features
SenseChat 5 SenseTime's flagship general-purpose foundation model
SenseChat-3 SenseTime's third-generation SenseChat model, preceding SenseChat-5
SenseNova 5.5 SenseTime's flagship multimodal foundation model
Skywork-13B Kunlun Tech's open-weight bilingual foundation model
SmolLM Hugging Face's original family of small, fully open language models
SmolLM2 Hugging Face's fully open small-model family for on-device use
Solar 10.7B Upstage's depth-upscaled open model that preceded the Solar Pro line
Solar Pro 2 Upstage's efficient open-weight model that punches above its parameter count
Sonar Large Perplexity's larger web-grounded model, ahead of the Sonar Pro tier
Sonar Pro Perplexity's search-grounded answer model built on top of open-weight LLMs
Sonar Reasoning Pro Perplexity's search-grounded reasoning model
Sonar Small Perplexity's low-cost web-grounded search model tier
Spark 4.0 iFlytek's flagship Chinese-language foundation model
Spark 4.0 Ultra iFlytek's top-tier Spark model with enhanced mathematical reasoning
Sparrow DeepMind's dialogue agent research model trained to be more helpful and less harmful
StableLM 2 Stability AI's open-weight small language model
StableLM 3B Stability AI's early compact open-weight language model, preceding StableLM 2
StableLM Zephyr Stability AI's chat-tuned small model built on StableLM 3B
StarCoder2 The BigCode collaboration's open-weight code generation model
Step-1 StepFun's earlier flagship model that preceded Step-2
Step-2 StepFun's trillion-parameter reasoning model
T5 Google's text-to-text transfer transformer, an early unifying NLP framework
text-davinci-003 The last and most capable of OpenAI's original GPT-3.5 completion models
Titan Text Amazon's original in-house foundation model line for Bedrock
Titan Text Express Amazon's mid-tier Titan model for general text generation, between Lite and the base Titan Text
Turing-NLG Microsoft's 17B-parameter model that was among the largest published language models of its time
UL2 Google's unified pretraining framework blending multiple denoising objectives
Vicuna-13B LMSYS's influential early open chat model fine-tuned from Llama on ShareGPT conversations
Vicuna-33B The largest Vicuna release, LMSYS's ShareGPT fine-tune of Llama
WizardLM-2 Microsoft's instruction-tuned model family built with complex, evolved synthetic training data
XGen-7B Salesforce's open-weight long-context research model
XLNet A permutation-based autoregressive pretraining model that outperformed BERT on many tasks
Yi-1.5-34B 01.AI's improved open-weight Yi generation
Yi-34B 01.AI's open-weight bilingual model
Yi-Coder 01.AI's open-weight code-specialized model
Yi-Large 01.AI's flagship proprietary model, competitive with GPT-4-class models on Chinese benchmarks
Yuan 2.0 Inspur's open-weight foundation model tuned for enterprise Chinese-language tasks
Zephyr 141B Hugging Face's larger mixture-of-experts Zephyr fine-tune, distinct from the 7B release
Zephyr 7B Hugging Face's open-weight fine-tune that popularized DPO alignment
Landbot Landbot is a no-code chatbot builder for creating conversational flows on …
SambaNova SambaNova is an AI hardware and software platform built for running large …
AI21 Labs Enterprise AI platform with Jamba hybrid state-space models
ChatGPT Advanced AI chatbot by OpenAI for conversations and assistance
Claude AI assistant by Anthropic for helpful, harmless, and honest conversations
CoreWeave GPU cloud infrastructure for AI workloads
Gemini Google's AI chatbot with multimodal capabilities
Gemini Advanced Google's most capable AI assistant with Gemini Ultra
Hugging Face The AI community platform for models, datasets, and spaces
Kimi Moonshot AI's long-context chatbot with 1M token window
Lambda Labs GPU cloud and workstations for deep learning
Le Chat Mistral AI's free conversational assistant with web search
Modal Serverless GPU cloud for AI and data workloads
Paperspace Cloud GPU platform for AI development and deployment
Perplexity AI-powered search engine and chatbot for research
Poe AI chatbot aggregator giving access to GPT-4, Claude, and more
RunPod GPU cloud for AI training and inference
Stability AI The company behind Stable Diffusion and open generative AI
Together AI Fast inference cloud for open-source AI models
Vast.ai GPU marketplace for affordable AI compute
Venice AI Private AI chat and image generation with no data retention
01.AI Open-source Yi large language models by 01.AI
Aleph Alpha European sovereign AI models for enterprise and government
Nous Research Open-source AI research and fine-tuned model community
Reka AI Multimodal AI models for text, image, and video understanding
ChatPlayground AI Compare the best AI models side by side in one place
Cohere Enterprise LLMs for search, summarisation, and classification
Grok xAI's AI assistant with real-time X/Twitter data
Microsoft Copilot Microsoft's AI assistant across Windows, Office, and the web