BasedAGIBasedAGI

Model Profile

grok/grok-4.20-0309-reasoning

External Benchmark Shadowexternal_benchmark_shadowpublic
4,096 ctx

Use this page to decide where this model is a strong fit. Rankings below are benchmark-backed by use case, with explicit confidence and contributor metrics.

Identity

ID: external/grok/grok-4-20-0309-reasoning

Author: grok

Origin: external_benchmark_shadow

Arch: unknown

Benchmark Coverage

Scored use cases: 12

Avg confidence: 26.3%

Evidence points: 210

Raw rows: 400

Weighted rows: 19

Catalog Metadata

Parameters: unknown

Context window: 4096

Downloads: 0

Intelligence Profile

IQ63%EQAccuracyCreativityBased

Dimension Breakdown

IQ22 benchmarks
63.4%
EQ0 benchmarks

No eq benchmarks found

Insufficient data
Accuracy0 benchmarks

No accuracy benchmarks found

Insufficient data
Creativity0 benchmarks

No creativity benchmarks found

Insufficient data
Based0 benchmarks

No based benchmarks found

Insufficient data

1/5 dimensions scored · Last updated Apr 14, 2026

Benchmark Signals

Click through to the benchmark source behind this model profile.

Some fit rows have limited benchmark evidence.

5 of 12 scored use cases have low confidence or thin contributor coverage.

Coverage Diagnostics

actively scored

Use-Case Scores

117

Total Measurements

400

Weighted Measurements

19

Weighted Sources

13

Raw Source Coverage

vals_mmlu_pro 60vals_finance_agent 40vals_multimodal_index 32vals_medqa 28corpfin_taxeval_public 24vals_legal_bench 24

Weighted Source Coverage

vals_finance_agent 5vals_corp_fin_v2 3vals_case_law_v2 1vals_gpqa 1vals_lcb 1vals_legal_bench 1

Best Use Cases for This Model

Use CaseScore
Thesis red teaming

use_case.fin.thesis_red_team

26.6%
Accounts payable invoice extraction (text)

use_case.fin.ap_invoice_extraction

24.7%
Earnings call synthesis

use_case.fin.earnings_call_synthesis

24.0%
Transaction anomaly narrative

use_case.fin.transaction_anomaly_narrative

23.6%
KYC profile synthesis

use_case.fin.kyc_profile_synthesis

22.6%
AML alert triage

use_case.fin.aml_alert_triage

22.6%
Filings summarization (10-K/10-Q)

use_case.fin.filings_summarization

21.0%
Quant research code generation

use_case.fin.alpha_research_codegen

18.1%
Component selection assistant

use_case.eng.component_selection

17.1%
Simulation setup assistant

use_case.eng.simulation_setup_assistant

16.8%
Runbook step assistant

use_case.sre.runbook_steps

14.0%
Knowledge base Q&A (fast, no citations)

use_case.business.kb_qna_fast

14.0%