Skip to content

lab

Center for AI Standards and Innovation (CAISI)

U.S. government body within NIST that develops AI standards, guidelines, and evaluations, including safety, security, and misuse testing of models.

Known aliases

  • CAISI
  • Center for AI Standards and Innovation
  • NIST CAISI
  • US Center for AI Standards and Innovation

Relationships

No evidence-backed relationships are recorded.

Current stories

invest1 publisher

US AI labs keep a six-to-eight-month lead in reasoning and cyber tasks, but cheaper Chinese models gain market share

Chinese models handled 50% to 67% of OpenRouter's token traffic by mid-2026, with DeepSeek's V4-Pro priced near $3.96 per million output tokens. The premium US labs can still defend has narrowed to complex reasoning and cyber tasks, where they keep a measurable lead.

Reality

Evidence35
Adoption50
Hype gap+25
Incentives
Insufficient
Confidence35
build4 publishers

Anthropic says a freely downloadable model builds exploits nearly as well as its restricted tool

Anthropic says Zhipu's freely downloadable GLM-5.3 built working V8 exploits in 50 of 410 tries, against 56 for its own restricted Claude Mythos Preview. With the weights public, its safeguards come off cheaply, so a lab that restricts its own model no longer keeps the capability out of reach.

Perspective Coverage

4 publishers
Builder
Builder 41%
Operator
Operator 38%
Investor
Investor 21%

Reality

Evidence62
Adoption30
Hype gap+15
Incentives72
Confidence62
science2 publishers

Google, OpenAI and Anthropic reportedly plan their own body to set frontier AI's pre-release tests

Google, OpenAI and Anthropic are reportedly building SAFA, a government-independent body to set pre-release AI testing rules, aiming to launch in early 2027. Its members, evaluators and enforcement powers are unannounced, so buyers still have to set testing terms with each vendor themselves.

Reality

Evidence35
Adoption
Insufficient
Hype gap+25
Incentives65
Confidence40
build1 publisher

Tool permissions set the maximum harm a hijacked security agent can do

Dev.to author dharani2d argues agent security depends on who picks the next tool call, citing Excessive Agency's rise from sixth to third at OWASP. The proposed control plane keeps identity, authorization, argument checks and approvals in deterministic code outside the model.

Publishers:dev.to

Reality

Evidence45
Adoption
Insufficient
Hype gap+10
Incentives
Insufficient
Confidence40