Risk Dashboard

Tracking Grok model capabilities against Frontier Artificial Intelligence Framework (FAIF) thresholds

Last Updated: 2025-12-30
Risk Overview
Number of Xai models at each risk level (Frontier Artificial Intelligence Framework (FAIF))
4
Low Risk
0
Medium Risk
0
High Risk
0
Critical Risk
Total Models Assessed:4
Near Threshold:
Grok Code Fast 1concerning-propensities
Low Risk Models:
Grok 4low
Grok 4.1low
Grok 4 Fastlow
Grok Code Fast 1low
Near Threshold

No models are currently near threshold limits.

Risk Progression Over Time
Combined risk (average across all categories) for each model
All CategoriesAbuse PotentialConcerning PropensitiesDual-Use Capabilities
Latest Updates

xAI Frontier Artificial Intelligence Framework Published

Dec 30, 2025

xAI publishes comprehensive Frontier Artificial Intelligence Framework (FAIF) defining three risk categories (Abuse Potential, Concerning Propensities, Dual-Use Capabilities) and standardized evaluation benchmarks.

Grok 4.1 Released

Nov 17, 2025

xAI releases Grok 4.1, the most capable Grok model to date. Achieves superhuman biology capabilities (87% on WMDP Bio) with improved safety training and reduced dishonesty/sycophancy compared to Grok 4.

grok-4.1System Card

Grok 4 Fast Released

Sep 19, 2025

xAI releases Grok 4 Fast, a faster variant of Grok 4 with strong cybersecurity capabilities (81.4% accuracy) and moderate dual-use performance. Maintains low abuse potential through safety mechanisms.

grok-4-fastSystem Card

Grok Code Fast 1 Released

Aug 26, 2025

xAI releases Grok Code Fast 1, a specialized model for coding tasks with lower dual-use capabilities but elevated dishonesty rate due to specialized safety training. Designed for agentic coding applications.

grok-code-fast-1System Card

Grok 4 Released

Aug 20, 2025

xAI releases Grok 4 with expert-level capabilities in biology and strong chemistry/cybersecurity performance. Model includes comprehensive safety measures including refusal policies and input filters for harmful requests.

Model Risk Comparison
Compare risk scores across different Grok models and categories
Key Insights
Critical findings from official xAI model cards and framework

Superhuman Biology Capabilities

Grok 4 and Grok 4.1 demonstrate expert-level and superhuman performance on biological threat benchmarks (WMDP Bio accuracy of 87% for Grok 4.1), exceeding human baselines. xAI has implemented comprehensive input filters for bioweapons knowledge to mitigate misuse risks.

Strong Abuse Mitigations in Place

All Grok models maintain near-zero response rates (0.00-0.02) for harmful requests through system prompt refusal policies and input filtering. Refusals remain robust against jailbreak attempts and adversarial attacks.

Specialized Model Trade-offs

Grok Code Fast 1, a specialized model for agentic coding, shows elevated dishonesty rates (71.9% on MASK benchmark). This trade-off was accepted due to the narrow use-case focus and limited general-purpose exposure as a specialized tool.

Agentic Security Risks

Grok Code Fast 1 shows elevated vulnerability to agentic abuse (17% completion rate on AgentHarm benchmark) and hijacking attacks (26.9% success rate), reflecting challenges in securing specialized agent models.

Quantitative Benchmark-Based Framework

xAI uses rigorous quantitative benchmarks (WMDP, VCT, BioLP-Bench, CyBench, MASK) for risk assessment rather than qualitative ratings. All models maintain low overall risk through enforced safety measures and continuous monitoring.

Risk Categories
xAI Frontier Artificial Intelligence Framework (FAIF) evaluation domains

Abuse Potential

Evaluates the model's vulnerability to exploitation through jailbreaks, adversarial attacks, or harmful queries, such as those involving CBRN weapons, CSAM, or self-harm, focusing on refusal rates and robustness.

Concerning Propensities

Assesses inherent model behaviors that could lead to loss of control, including deception, sycophancy, and political bias, measured via benchmarks like MASK to maintain honesty and alignment.

Dual-Use Capabilities

Examines the model's potential for misuse in high-risk domains like biology, chemistry, cybersecurity, and persuasion, using benchmarks to ensure capabilities do not enable catastrophic outcomes.

Official Sources
All data sourced from official xAI documentation

All risk assessments, scores, and threshold proximity values are extracted directly from official xAI model cards and the Frontier Artificial Intelligence Framework. Each data point is grounded in quantitative benchmark results and includes citations to source documents.