Benchmark menu

Model results

Gemini 3.5 Flash Lite

See how Gemini 3.5 Flash Lite performed, which reasoning settings were tested, how often it produced a working player program, and what those runs cost.

Gemini 3.5 Flash Lite

Provider: Google

Rank
33
Status
Complete 3/3
Rank change
Baseline
Score
34.4
Last tested
2026-08-23
Settings tested
All settings tested

Scores by reasoning setting

Available settings

MinimalLowMediumHigh

Checked on 2026-07-24 · Provider documentation

Scores from tested settings

Tested settingScoreResult statusDisplay group
Minimal40.6Tested and includedBaseline
Medium27.0Tested and includedBalanced
High35.4Tested and includedIntensive

How settings are grouped

  • Baseline: Minimal — tested and included
  • Balanced: Medium — tested and included
  • Intensive: High — tested and included

Player program results

High: 81% / 6% / 13% · n 16Medium: 88% / 0% / 12% · n 16Minimal: 56% / 38% / 6% · n 16 All settings tested

Generation cost

Average estimated cost
$0.04
Median estimated cost
$0.04
Estimated range
$0.01–$0.07
Cost data
24 combinations / 48 recorded runs
Observed total
$1.85
Price source
Provider-reported
Price list
Not recorded for this older data
Average output
13k tokens
Output range
1.5k–34k tokens
Samples
48

Estimated cost gives equal weight to each model, game, and reasoning-setting combination.

Test setup

  • Combines this model's tested reasoning settings
  • Uses the standard GameBench player-program prompt