InternScience
Agents-A1 35B A3B
FrontierAgents-A1 35B A3B (35.099998474121094B parameters) requires approximately 23.2 GB of VRAM with Q4_K_M quantization. As a Mixture of Experts model with 3B active parameters, it uses less memory than its total parameter count suggests. For the best balance of quality and speed, we recommend hardware with at least 27 GB of VRAM.
Comece agora
— copie e cole para rodar localmenteCopy-paste commands to run Agents-A1 35B A3B on your machine.
Run
docker run --rm -it ghcr.io/ggerganov/llama.cpp:full \
--hf-repo "InternScience/Agents-A1" \
--hf-file "Agents-A1-Q4_K_M.gguf" \
-c 4096 -ngl 99Quick specs
About this model
- •35B total / ~3B active — runs single-GPU at 262K context.
- •SOTA on Seal-0 (56.4), HiPhO (46.4), FrontierScience-Olympiad (79.0), IFBench (80.6), IFEval (94.8).
- •Competitive with GPT-5.5, DeepSeek-V4-Pro and Kimi-K2.6 on agentic benchmarks.
- •Vision-capable (VLM) for GUI and document agent workflows.
- •Agent-horizon scaling: long-horizon trajectories plus heterogeneous agent abilities.
Modelos relacionados
Escolhas rápidas
Melhor hardware
Melhores opções para Agents-A1 35B A3B
Rodar este modelo
Opções de quantização
Estimativas de VRAM por nível de quantização
No hardware detected — fit column shows raw VRAM estimates
| Quant | Bits | VRAM | Quality | Fit |
|---|---|---|---|---|
Q1_0_G128 | 1.125 | 5.1 GB | Very Low | — |
Q2_0_G128 | 1.71 | 9.4 GB | Low | — |
Q2_K | 2 | 13.7 GB | Low | — |
Q3_K_S | 3 | 17.2 GB | Low | — |
NVFP4 | 4 | 19.7 GB | Medium | — |
Q4_K_M | 4 | 21.4 GB | Medium | — |
Q5_K_M | 5 | 25.3 GB | High | — |
Q6_K | 6 | 28.8 GB | High | — |
Q8_0 | 8 | 37.6 GB | Very High | — |
F16 | 16 | 72.0 GB | Maximum | — |
Quality benchmarks
Agents-A1 35B A3B benchmark scores
General
Source: official · 2026-06-22
Compatibilidade de hardware
Estimativas de compatibilidade para todo o hardware
Computing compatibility...
Detalhamento de memória
Reference: RTX 2060 6GB
Perguntas frequentes
FAQ — Agents-A1 35B A3B
Veja também