CohereCohere

Aya Expanse 8B

Aktuell
39.4KDownloads431LikesOct 2024Veröffentlicht8K TokenKontextCC-BY-NC-4.0Lizenz8 EinstiegQualität

Aya Expanse 8B (8B parameters) requires approximately 8.3 GB of VRAM with Q4_K_M quantization. For the best balance of quality and speed, we recommend hardware with at least 10 GB of VRAM.

Loslegen

— kopieren & einfügen, um lokal auszuführen

Copy-paste commands to run Aya Expanse 8B on your machine.

Run

docker run --rm -it ghcr.io/ggerganov/llama.cpp:full \ --hf-repo "CohereForAI/aya-expanse-8b" \ --hf-file "aya-expanse-8b-Q4_K_M.gguf" \ -c 4096 -ngl 99

Quick specs

Parameters8B
Architecturedense
Context8K tokens
Modalitytext
Min RAM3.1 GB
Rec. RAM4.9 GB (Q4_K_M)
LicenseCC-BY-NC-4.0
FamilyAya
Chat

About this model

Aya Expanse 8B is Cohere's multilingual model supporting 23 languages with strong cross-lingual transfer. Designed for global applications requiring high-quality generation across diverse languages.

Verwandte Modelle

Deine Hardware

Erkennung...

Schnellauswahl

Beste Hardware

Top-Empfehlungen für Aya Expanse 8B

Dieses Modell ausführen

Quantisierungsoptionen

VRAM-Schätzungen nach Quantisierungsstufe

No hardware detected — fit column shows raw VRAM estimates

QuantBitsVRAMQualityFit
Q2_K
2
3.1 GB
Low
Q3_K_S
3
3.9 GB
Low
NVFP4
4
4.5 GB
Medium
Q4_K_M
4
4.9 GB
Medium
Q5_K_M
5
5.8 GB
High
Q6_K
6
6.6 GB
High
Q8_0
8
8.6 GB
Very High
F16
16
16.4 GB
Maximum

Quality benchmarks

Aya Expanse 8B benchmark scores

Benchmark verified

Reasoning

MMLU-Pro22.3%
GPQA Diamond7.0%
MATH-5008.6%
ARC Challenge

General

Chatbot Arena
IFEval63.6%

Source: community · 2025-01-01

Hardware-Kompatibilität

Eignungsschätzungen für alle Hardware

Rechner öffnen

Computing compatibility...

Speicheraufschlüsselung

Reference: RTX 2060 6GB

Weights4.9 GB
KV Cache2.0 GB
Runtime0.9 GB
Headroom0.6 GB

Häufig gestellte Fragen

FAQ — Aya Expanse 8B

Siehe auch