CohereCohere

Aya Expanse 8B

現行
39.4Kダウンロード431いいねOct 2024公開日8K トークンコンテキストCC-BY-NC-4.0ライセンス8 入門品質

Aya Expanse 8B (8B parameters) requires approximately 8.3 GB of VRAM with Q4_K_M quantization. For the best balance of quality and speed, we recommend hardware with at least 10 GB of VRAM.

はじめに

— コピー&ペーストでローカル実行

Copy-paste commands to run Aya Expanse 8B on your machine.

Run

docker run --rm -it ghcr.io/ggerganov/llama.cpp:full \ --hf-repo "CohereForAI/aya-expanse-8b" \ --hf-file "aya-expanse-8b-Q4_K_M.gguf" \ -c 4096 -ngl 99

Quick specs

Parameters8B
Architecturedense
Context8K tokens
Modalitytext
Min RAM3.1 GB
Rec. RAM4.9 GB (Q4_K_M)
LicenseCC-BY-NC-4.0
FamilyAya
Chat

About this model

Aya Expanse 8B is Cohere's multilingual model supporting 23 languages with strong cross-lingual transfer. Designed for global applications requiring high-quality generation across diverse languages.

関連モデル

あなたのハードウェア

検出中...

おすすめ

最適なハードウェア

Aya Expanse 8Bのおすすめ

このモデルを実行

量子化オプション

量子化レベル別VRAM推定値

No hardware detected — fit column shows raw VRAM estimates

QuantBitsVRAMQualityFit
Q2_K
2
3.1 GB
Low
Q3_K_S
3
3.9 GB
Low
NVFP4
4
4.5 GB
Medium
Q4_K_M
4
4.9 GB
Medium
Q5_K_M
5
5.8 GB
High
Q6_K
6
6.6 GB
High
Q8_0
8
8.6 GB
Very High
F16
16
16.4 GB
Maximum

Quality benchmarks

Aya Expanse 8B benchmark scores

Benchmark verified

Reasoning

MMLU-Pro22.3%
GPQA Diamond7.0%
MATH-5008.6%
ARC Challenge

General

Chatbot Arena
IFEval63.6%

Source: community · 2025-01-01

ハードウェア互換性

全ハードウェアの適合度推定

カリキュレーターを開く

Computing compatibility...

メモリ内訳

Reference: RTX 2060 6GB

Weights4.9 GB
KV Cache2.0 GB
Runtime0.9 GB
Headroom0.6 GB

よくある質問

FAQ — Aya Expanse 8B

関連項目