LLM Mechanistic Interpretability

categorie

LLM Mechanistic Interpretability
Scott Alexander ci spiega l'argomento tecnologico del momento, come riuscire ad interpretare (e quindi eventualmente riuscire a modificare) il funzionamento di un modello linguistico, ovvero capire come funziona la "AI". Questo perchè un LLM "cresce", non viene "costruito", e quindi nessuno lo sa per davvero. Solo nel 2023 si è capito, per esempio, che non c'è una associazione uno a uno tra concetti e neuroni, ma una corrispondenza da molti a molti. Nel 2025 però quale sia la natura di questa relazione molti a molti è ancora del tutto incomprensibile.

Let’s Try To Learn About Mechanistic Interpretability Techniques


Add new comment

The content of this field is kept private and will not be shown publicly.

Full HTML 2

  • Web page addresses and email addresses turn into links automatically.
  • Lines and paragraphs break automatically.

Filtered HTML

  • Web page addresses and email addresses turn into links automatically.
  • Allowed HTML tags: <a href hreflang> <em> <strong> <cite> <blockquote cite> <code> <ul type> <ol start type='1 A I'> <li> <dl> <dt> <dd> <h2 id='jump-*'> <h3 id> <h4 id> <h5 id> <h6 id>
  • Lines and paragraphs break automatically.
CAPTCHA
This question is for testing whether or not you are a human visitor and to prevent automated spam submissions.