AliceAI-Foundation-80B-A3B-Base is a base language model from Yandex built with a hybrid MoE architecture. It has 80B total parameters with 3B active per token and supports a context length of 262,144 tokens. Yandex trained it from scratch without third-party model weights and released the weights under Apache 2.0. The model card reports strong Russian-language factual knowledge and competitive math and code results against larger open base models.
A solid 81.29B-parameter MoE language model from yandex. A pragmatic middle-ground choice when you need open weights without a flagship-sized footprint. Newly released, so production-readiness is still being shaken out.
Generated from this model’s benchmarks and ranking signals. Editor reviews refine it over time.
Access model weights, configuration files, and documentation.
No benchmark data available for this model yet.
See how different quantization levels affect VRAM requirements and quality for this model.
| Format | VRAM Required | Quality | |
|---|---|---|---|
| Q2_K | 7.9 GB | Low | |
| Q4_K_MRecommended | 8.5 GB | Good | |
| Q5_K_M | 8.8 GB | Very Good | |
| Q6_K | 9.2 GB | Excellent | |
| Q8_0 | 9.9 GB | Near Perfect | |
| FP16 | 12.8 GB | Full |
The top devices for this model at 4-bit, ranked by fit and speed.
| Device | Grade | Speed | VRAM |
|---|---|---|---|
| ACEMAGIC M1A Pro (i9-13900HK + ARC A770)ACEMAGIC | SS | 48.3 tok/s | 8.5 GB |
| AMD Radeon RX 7700 XTAMD | SS | 40.8 tok/s | 8.5 GB |
| AMD Radeon RX 7800 XTAMD | SS | 58.9 tok/s | 8.5 GB |
| AMD Radeon RX 9070AMD | SS | 60.4 tok/s | 8.5 GB |
| AMD Radeon RX 9070 XTAMD | SS | 60.4 tok/s | 8.5 GB |
Energy cost on Raspberry Pi 5 (8GB) (~3.2 tok/s, Q4_K_M) vs flagship API pricing.
| Source | Cost per 1M tokens |
|---|---|
Local (energy only)AliceAI-Foundation-80B-A3B-Base on Raspberry Pi 5 (8GB) · ~3.2 tok/s · 12W | $0.125 |
GPT-6 SolOpenAI · in $2.00 · out $10.00 | $4.40 |
Claude Opus 5.5Anthropic · in $4.00 · out $20.00 | $8.80 |
Gemini 3.5 FlashGoogle · in $1.50 · out $9.00 | $3.75 |
Grok 4.5xAI · in $2.00 · out $6.00 | $3.20 |
API prices blended at 70% input / 30% output.
Hardware amortisation not included. Run the full ROI calculator for payback math.
Cheapest current cloud rentals with at least 9 GB VRAM, refreshed hourly.
Advertising disclosure: we earn commissions when you shop through the links below.
| Option | Cost / GPU-hour |
|---|---|
NVIDIA GeForce RTX 2080 TiVast.ai · Spot · 11 GB VRAM | $0.04 |
NVIDIA GeForce RTX 3060Vast.ai · Spot · 12 GB VRAM | $0.05 |
NVIDIA GeForce RTX 3080 TiVast.ai · Spot · 12 GB VRAM | $0.05 |
NVIDIA GeForce RTX 3060Vast.ai · On-Demand · 12 GB VRAM | $0.05 |
NVIDIA GeForce RTX 3090Vast.ai · Spot · 24 GB VRAM | $0.08 |
Per-GPU rate across RunPod, the Vast.ai marketplace, DigitalOcean, and Vultr.
Spot tier is interruptible. Plan for restarts when comparing against on-demand prices.

Explore the Provider
Aggregate stats, leaderboard, release timeline, and benchmark coverage across every yandex model we track.

Check which of your devices can run this model and how fast.

Compare hosted per-token prices before you commit to local hardware.

Break-even math for self-hosting, renting a GPU, or paying per token.
Live status for the hosted APIs you might use instead of running this locally.