Skip to main content

LLM Models

Selecting the right Large Language Model (LLM) for your application is a critical decision that impacts performance, cost, and user experience. This guide provides a comprehensive comparison of leading LLMs to help you make an informed choice based on your specific requirements.

How to Select the Right LLM

When choosing an LLM, consider these key factors:
  1. Task Complexity: For complex reasoning, research, or creative tasks, prioritize models with high accuracy scores (8-10), even if they’re slower or more expensive. For simpler, routine tasks, models with moderate accuracy (6-8) but higher speed may be sufficient.
  2. Response Time Requirements: If your application needs real-time interactions, prioritize models with speed ratings of 8-10. Customer-facing applications generally benefit from faster models to maintain engagement.
  3. Context Needs: If your application processes long documents or requires maintaining extended conversations, select models with context window ratings of 8 or higher. Some specialized tasks might work fine with smaller context windows.
  4. Budget Constraints: Cost varies dramatically across models. Free and low-cost options (0-2 on our relative scale) can be excellent for startups or high-volume applications, while premium models (5+) might be justified for mission-critical enterprise applications where accuracy is paramount.
  5. Specific Capabilities: Some models excel at particular tasks like code generation, multimodal understanding, or multilingual support. Review the use cases to find models that specialize in your specific needs.
The ideal approach is often to start with a model that balances your primary requirements, then test alternatives to fine-tune performance. Many organizations use multiple models: premium options for complex tasks and more affordable models for routine operations.

Vendor Overview

OpenAI: Offers the most diverse range of models with industry-leading capabilities, though often at premium price points, with particular strengths in reasoning and multimodal applications. Anthropic (Claude): Focuses on highly reliable, safety-aligned models with exceptional context length capabilities, making them ideal for document analysis and complex reasoning tasks. Google: Provides models with impressive context windows and competitive pricing, with the Gemini series offering particularly strong performance in creative and analytical tasks. Perplexity: Specializes in research-oriented models with unique web search integration, offering free access to powerful research capabilities and real-time information. Other Vendors: Offer open-source and specialized models that provide strong performance at minimal or no cost, making advanced AI accessible for deployment in resource-constrained environments.

OpenAI Models

Anthropic (Claude) Models

Google Models

Perplexity Models

Open Source Models

Model Deprecation

In the LLM Engine dropdown, there’s a section labeled “Legacy Models Soon To Be Deprecated”. These are models we plan to remove soon, and we’ll automatically migrate agents using them to a recommended alternative.