shipfeedAI news, curated daily

16:07:26 CET
26 AUG16:07:26shipfeed
pull to refreshlast sync
Just in — 30 new
§ safety · storyline

New Platform Peers Inside AI’s Black Box

Goodfire releases Silico platform for AI interpretability with $1 million grant program for researchers.

today · · primary fetch1 sourceupdated today ·

Prompt Claude, ChatGPT, Gemini, or any other popular large language model (LLM) with a question like “What is the best film ever made?” and the response will vary, and you (and most worryingly, the people who built the LLM) have little idea exactly how it came up with that specific answer. This mysterious behavior can be useful in some situations. But—as a recent incident where OpenAI could not explain why its advanced pre-release model hacked AI company Hugging Face highlighted—it can have negative and alarming consequences too. And when frontier AI models are writing code, generating results humans could not achieve alone, and performing other important tasks across society, the need to interpret AI ‘thinking’ and outputs has never been greater.

Goodfire, an AI lab focused solely on this very problem, recently made its cutting-edge Silico platform, filled with tools to interpret the behavior of AI, generally available to the public. As part of this, the company recently announced a new grant program offering $1 million in free Silico usage for academic and nonprofit interpretability researchers. These efforts aim to democratize AI interpretability, placing techniques previously…

read full article on spectrum.ieee.org
§ sources1 publication · timeline below
  1. spectrum.ieee.orgNew Platform Peers Inside AI’s Black Boxprimary