InternScience
Agents-A1 35B A3B
FronteraAgents-A1 35B A3B (35.099998474121094B parameters) requires approximately 23.2 GB of VRAM with Q4_K_M quantization. As a Mixture of Experts model with 3B active parameters, it uses less memory than its total parameter count suggests. For the best balance of quality and speed, we recommend hardware with at least 27 GB of VRAM.
Comenzar
— copia y pega para ejecutar en localCopy-paste commands to run Agents-A1 35B A3B on your machine.
Run
docker run --rm -it ghcr.io/ggerganov/llama.cpp:full \
--hf-repo "InternScience/Agents-A1" \
--hf-file "Agents-A1-Q4_K_M.gguf" \
-c 4096 -ngl 99Quick specs
About this model
- •35B total / ~3B active — runs single-GPU at 262K context.
- •SOTA on Seal-0 (56.4), HiPhO (46.4), FrontierScience-Olympiad (79.0), IFBench (80.6), IFEval (94.8).
- •Competitive with GPT-5.5, DeepSeek-V4-Pro and Kimi-K2.6 on agentic benchmarks.
- •Vision-capable (VLM) for GUI and document agent workflows.
- •Agent-horizon scaling: long-horizon trajectories plus heterogeneous agent abilities.
Modelos relacionados
Selecciones rápidas
Mejor hardware
Mejores opciones para Agents-A1 35B A3B
Ejecutar este modelo
Opciones de cuantización
Estimaciones de VRAM por nivel de cuantización
No hardware detected — fit column shows raw VRAM estimates
| Quant | Bits | VRAM | Quality | Fit |
|---|---|---|---|---|
Q1_0_G128 | 1.125 | 5.1 GB | Very Low | — |
Q2_0_G128 | 1.71 | 9.4 GB | Low | — |
Q2_K | 2 | 13.7 GB | Low | — |
Q3_K_S | 3 | 17.2 GB | Low | — |
NVFP4 | 4 | 19.7 GB | Medium | — |
Q4_K_M | 4 | 21.4 GB | Medium | — |
Q5_K_M | 5 | 25.3 GB | High | — |
Q6_K | 6 | 28.8 GB | High | — |
Q8_0 | 8 | 37.6 GB | Very High | — |
F16 | 16 | 72.0 GB | Maximum | — |
Quality benchmarks
Agents-A1 35B A3B benchmark scores
General
Source: official · 2026-06-22
Compatibilidad de hardware
Estimaciones de encaje en todo el hardware
Computing compatibility...
Desglose de memoria
Reference: RTX 2060 6GB
Preguntas frecuentes
FAQ — Agents-A1 35B A3B
Ver también