Invergent’s decision model built on Gemma 4 26B-A4B, a mixture of experts with 8 of 128 experts active per token. It reads text, structured data or images with a 262k-token context and answers Choice, Noul and Score questions in one forward pass, served fastest by the open surogate engine.
A workable 25.81B-parameter MoE decision model from Invergent. Treat the modality benchmarks above as the leading indicator of fit — composite scoring across modalities is still maturing. Newly released, so production-readiness is still being shaken out.
Generated from this model’s benchmarks and ranking signals. Editor reviews refine it over time.
Access model weights, configuration files, and documentation.
The top devices for this model at 4-bit, ranked by fit and speed.
| Device | Grade | VRAM |
|---|---|---|
| Acer Veriton GN100 AI MiniAcer | SS | 11.2 GB |
| AMD Instinct MI300XAMD | SS | 11.2 GB |
| AMD Instinct MI325XAMD | SS | 11.2 GB |
| AMD Instinct MI355XAMD | SS | 11.2 GB |
| AMD Radeon RX 7900 XTAMD | SS | 11.2 GB |
Cheapest current cloud rentals with at least 11 GB VRAM, refreshed hourly.
Advertising disclosure: we earn commissions when you shop through the links below.
| Option | Cost / GPU-hour |
|---|---|
NVIDIA GeForce RTX 3060Vast.ai · Spot · 12 GB VRAM | $0.05 |
NVIDIA GeForce RTX 3060Vast.ai · On-Demand · 12 GB VRAM | $0.05 |
NVIDIA Tesla V100 16GBVast.ai · Spot · 16 GB VRAM | $0.07 |
NVIDIA GeForce RTX 3090Vast.ai · Spot · 24 GB VRAM | $0.08 |
NVIDIA Tesla V100 16GBVast.ai · On-Demand · 16 GB VRAM | $0.08 |
Per-GPU rate across RunPod, the Vast.ai marketplace, DigitalOcean, and Vultr.
Spot tier is interruptible. Plan for restarts when comparing against on-demand prices.

Explore the Provider
Aggregate stats, leaderboard, release timeline, and benchmark coverage across every Invergent model we track.

Check which of your devices can run this model and how fast.

Compare hosted per-token prices before you commit to local hardware.

Break-even math for self-hosting, renting a GPU, or paying per token.
Live status for the hosted APIs you might use instead of running this locally.