Key Specifications

SpecificationJamba 1.5 LargeMistral Large 2
Vendorothermistral
Versionjamba-1-5-largelarge-2
Release Date2024-08-072024-07-24
Context Window256000 tokens128000 tokens
Input Modalitiestexttext
Output Modalitiestexttext
LicenseJamba Open Model LicenseMistral Research License
SOC2✗✓
HIPAA✗✗
GDPR✗✓
ISO 27001✗✓

Benchmark Results

BenchmarkJamba 1.5 LargeMistral Large 2Winner
ARC93.8—Jamba 1.5 Large
BBH82.881Jamba 1.5 Large
GPQA50.8—Jamba 1.5 Large
GSM8K83.593Mistral Large 2
HUMANEVAL73.792Mistral Large 2
IFEVAL70.8—Jamba 1.5 Large
MATH6071Mistral Large 2
MMLU84.384Jamba 1.5 Large
MUSR51.1—Jamba 1.5 Large
WINOGRANDE84.3—Jamba 1.5 Large

Pricing Comparison

Tier (per Mtok)Jamba 1.5 LargeMistral Large 2
Input$2$2
Output$8$6
Cache Read$0$0
Cache Write$0$0

Jamba 1.5 Large kontra Mistral Large 2

Przegląd modelu

Jamba 1.5 Large and Mistral Large 2 are both notable options in the AI model market. This page compares their benchmarks, pricing, and compliance.

Kluczowe specyfikacje

DostawcaData wydaniaOkno kontekstoweLicencja
Other / Mistral2024-08-07 / 2024-07-24256K / 128KJamba Open Model License / Mistral Research License

Wydajność benchmarków

BenchmarkJamba 1.5 LargeMistral Large 2Zwycięzca
ARC93.8—A
BBH (BIG-Bench Hard)82.881.0A
GPQA50.8—A
GSM8K (Grade School Math 8K)83.593.0B
HumanEval73.792.0B
IFEval70.8—A
MATH60.071.0B
MMLU (Massive Multitask Language Understanding)84.384.0Tie
MUSR51.1—A
WinoGrande84.3—A

Porównanie cen

WejścieWyjścieOdczyt pamięci podręcznejZapis pamięci podręcznej
— / —— / —— / —— / —

za milion tokenów — A / B

Mocne strony & Słabe strony

Jamba 1.5 Large

  • ✅ MMLU score 84.3, strong knowledge reasoning.
  • ✅ 上下文窗口 256K,支持长文本。
  • ✅ 采用 MoE 混合专家架构。
  • ⚠️ 闭源专有模型,不支持自托管。

Mistral Large 2

  • ✅ MMLU score 84.0, strong knowledge reasoning.
  • ✅ HumanEval 92.0, excellent code generation.
  • ✅ GSM8K 93.0, robust math reasoning.
  • ⚠️ 闭源专有模型,不支持自托管。

Opinia redakcji

Jamba 1.5 Large and Mistral Large 2 each have their strengths. Choose based on workload (code, long context, vision), referencing the tables above.

FAQ

Which model is better for coding tasks?

Refer to the HumanEval benchmark table; the model with a higher score is better suited for coding tasks.

Which model is cheaper?

Refer to the pricing comparison table above; the model with lower input/output prices is more cost-effective.

Which has a longer context window?

Refer to the key specifications table; the model with a larger context window is better for long documents.

Referencje

Editor's Take

See Editor's Take section.