14 ago - Gemini
KAPUALabs
Gemini 3.1 Flash Lite at a glance
Good enough on 6/52 tasks at the 90% bar. Best value on 2 tasks. Doesn't qualify on any: Financial Analysis & Trading Decisions, Content Summarization & Synthesis, Long-form Content Generation, Social
Provider Gemini Model name gemini-3.1-flash-lite
Official resources
- Model docs
- Pricing
- Google AI
How good does a model need to be? At least 90% of the best-performing model.
Cost mode: Batch if supported Sync only
Qualifies on 6 / 52 tasks (at 90% bar) Best value on 2 tasks
Cost vs quality across all tasks
0% 25% 50% 75% 100% 8 9 10 Quality score (8–10) Cost-efficiency vs cheapest (1.0 = cheapest) Engagement Reply Review — quality 9.80, this model IS the cheapest good-enough option ★ S-1 TOC Extraction — quality 8.61, 1.2x the cost of the cheapest good-enough option Language Detection — quality 9.95, 1.4x the cost of the cheapest good-enough option Engagement Triage — quality 8.38, 1.6x the cost of the cheapest good-enough option X Post Selection — quality 8.47, 1.6x the cost of the cheapest good-enough option Claim Refinement — quality 8.47, this model IS the cheapest good-enough option ★
within ~1.3× of the best-value model
- 1.3–2×
- >2×
- ★ this model is the best-value pick on that task. Top-right = best quadrant. Only tasks where this model qualifies at the 90% bar are plotted.
Per-task breakdown
Task Category Quality (% of best) Confidence Overpay Claim Refinement★ Infrastructure &
• Utility 100% RANKED cheapest Engagement Reply Review★ Relevance, Classification &
• Matching 99% RANKED cheapest S-1 TOC Extraction Structured Data &
• Fact Extraction 92% MEDIUM 1.2x Language Detection Relevance, Classification &
• Matching 99% RANKED 1.4x X Post Selection Relevance, Classification &
• Matching 97% HIGH 1.6x Engagement Triage Relevance, Classification &
• Matching 94% RANKED 1.6x
Overpay — how much more you pay by running this model instead of the best-value model that clears the quality bar on that task (marked ★). "16x" means you overpay 16× — the same output for 16× the best-value good-enough option; ★ means this model is that option (no overpayment). Confidence — how sure we are about the quality score (more judgments + more agreement = higher confidence): RANKED many independent judges scored this model's outputs and their agreement is very high (most confident) — HIGH many judges have scored it and they mostly agree (well-pinned) — MEDIUM enough judges have weighed in to publish, but they disagree more than we'd like (treat with a small grain of salt). LOW-confidence cells are hidden everywhere on the site. See the methodology for the exact thresholds.
14 ago - Codroipo
Umana
14 ago - Rovigo
Privato
14 ago - Cittadella
Food Racers
14 ago - Lucca
La Risorsa Umana