Don't miss the latest stories
Advertise Newsletter
Network
  • The Creative Finder
  • The Bazaar
  • Deals
  • Status Is Down
Community
  • Sign up / Log in
  • Discussion Forums
  • Calendar of Events
NEW

Follow

Share this

Artificial Intelligence
Future
Research
Technology
UI/UX
  • Technology
  • UI/UX
MENU
  • Advertise with us
  • Submit tip/feedback
  • Work with us
  • Subscribe to newsletter
  • Subscribe to RSS
Advertise here
Advertisement

OpenAI’s Newest Tool Explains Why Language Models Behave In Certain Ways

By Nicole Rodrigues, 10 May 2023

Subscribe to newsletter
Like us on Facebook
Photo 246538875 © Rafael Henrique | Dreamstime.com

 

When talking to a chatbot, you may sometimes come across either a genius response from it, or it might say something completely unrelated that will have you wondering where along the lines it connected those dots. OpenAI’s newest tool is pursuing a more transparent relationship with users and its language learning models (LLM) by creating a tool that can identify why these machines behave in specific ways.
 
LLMs have “neurons” that help them pick up on certain patterns, just like the human brain. When a prompt is typed into them, they can highlight keywords, and trace them back to different neurons that hold information on the subject. In the example used by OpenAI, if you ask it about the Marvel universe, it will pick up on “Marvel neurons” and present relevant (and sometimes not so relevant) information as it answers.
 
This tool can examine the tiny parts of a chatbot’s brain, and explain why it responded a certain way. For example, it looks at the words the chatbot processed, and waits for a specific part of the neurons to “light up.” Then, it sends that information to GPT-4, which generates an explanation.
 
GPT-4 checks if the explanation is correct by giving it more words to process, and sees if it can predict the same behavior. If it matches, then the explanation is accurate.
 
“Using this methodology, we can basically, for every single neuron, come up with some kind of preliminary natural language explanation for what it’s doing and also have a score for how well that explanation matches the actual behavior,” Jeff Wu, scalable alignment team leader at OpenAI, said.
 
In total, all 307,200 neurons could be explained via the tool. Such a system could one day open up a better understanding of why biases still run rampant in artifical intelligence, and how to crack down on them. However, OpenAI admits it’s still far off in this regard. That said, it is now available for others to use and refine on GitHub.
 
 
 
[via TechCrunch and SiliconANGLE, Photo 246538875 © Rafael Henrique | Dreamstime.com]

Advertisement
Advertisement
Receive interesting stories like this one in your inbox
Advertise here

More related news

Advertise here
Also check out these recent news
Web Design
Link to news page

When Your Website Goes Down, This Is the Page Customers Meet Instead

2027
Link to news page

2027 Already Has A Color Of The Year And It’s Beginning On ‘Grounded’ Territory

IKEA
Link to news page

IKEA & Xbox Press Play On Furniture & Storage Inspired By The Iconic Controller

Fashion
Link to news page

Vogue Presents ‘United Flags of Fashion’ With Top Designers For All 50 States

Coca-Cola
Link to news page

Coca-Cola Pours Fresh Life Into Its Iconic Branding With Worldwide Redesign