Solega Co. Done For Your E-Commerce solutions.
  • Home
  • E-commerce
  • Start Ups
  • Project Management
  • Artificial Intelligence
  • Investment
  • More
    • Cryptocurrency
    • Finance
    • Real Estate
    • Travel
No Result
View All Result
  • Home
  • E-commerce
  • Start Ups
  • Project Management
  • Artificial Intelligence
  • Investment
  • More
    • Cryptocurrency
    • Finance
    • Real Estate
    • Travel
No Result
View All Result
No Result
View All Result
Home Artificial Intelligence

How OpenAI stress-tests its large language models

Solega Team by Solega Team
November 22, 2024
in Artificial Intelligence
Reading Time: 3 mins read
0
How OpenAI stress-tests its large language models
0
SHARES
2
VIEWS
Share on FacebookShare on Twitter


When OpenAI examined DALL-E 3 final yr, it used an automatic course of to cowl much more variations of what customers would possibly ask for. It used GPT-4 to generate requests producing photos that might be used for misinformation or that depicted intercourse, violence, or self-harm. OpenAI then up to date DALL-E 3 in order that it might both refuse such requests or rewrite them earlier than producing a picture. Ask for a horse in ketchup now, and DALL-E is sensible to you: “It seems there are challenges in producing the picture. Would you want me to attempt a distinct request or discover one other concept?”

In concept, automated red-teaming can be utilized to cowl extra floor, however earlier methods had two main shortcomings: They have a tendency to both fixate on a slim vary of high-risk behaviors or give you a variety of low-risk ones. That’s as a result of reinforcement studying, the expertise behind these methods, wants one thing to purpose for—a reward—to work effectively. As soon as it’s received a reward, akin to discovering a high-risk habits, it’ll hold making an attempt to do the identical factor repeatedly. With out a reward, however, the outcomes are scattershot. 

“They sort of collapse into ‘We discovered a factor that works! We’ll hold giving that reply!’ or they will give a number of examples which can be actually apparent,” says Alex Beutel, one other OpenAI researcher. “How will we get examples which can be each numerous and efficient?”

An issue of two components

OpenAI’s reply, outlined within the second paper, is to separate the issue into two components. As an alternative of utilizing reinforcement studying from the beginning, it first makes use of a big language mannequin to brainstorm potential undesirable behaviors. Solely then does it direct a reinforcement-learning mannequin to determine the right way to convey these behaviors about. This offers the mannequin a variety of particular issues to purpose for. 

Beutel and his colleagues confirmed that this strategy can discover potential assaults often called oblique immediate injections, the place one other piece of software program, akin to a web site, slips a mannequin a secret instruction to make it do one thing its person hadn’t requested it to. OpenAI claims that is the primary time that automated red-teaming has been used to seek out assaults of this type. “They don’t essentially appear like flagrantly unhealthy issues,” says Beutel.

Will such testing procedures ever be sufficient? Ahmad hopes that describing the corporate’s strategy will assist folks perceive red-teaming higher and comply with its lead. “OpenAI shouldn’t be the one one doing red-teaming,” she says. Individuals who construct on OpenAI’s fashions or who use ChatGPT in new methods ought to conduct their very own testing, she says: “There are such a lot of makes use of—we’re not going to cowl each one.”

For some, that’s the entire drawback. As a result of no one is aware of precisely what giant language fashions can and can’t do, no quantity of testing can rule out undesirable or dangerous behaviors absolutely. And no community of red-teamers will ever match the number of makes use of and misuses that a whole bunch of hundreds of thousands of precise customers will suppose up. 

That’s very true when these fashions are run in new settings. Folks typically hook them as much as new sources of information that may change how they behave, says Nazneen Rajani, founder and CEO of Collinear AI, a startup that helps companies deploy third-party fashions safely. She agrees with Ahmad that downstream customers ought to have entry to instruments that allow them take a look at giant language fashions themselves. 



Source link

Tags: languagelargeModelsOpenAIstresstests
Previous Post

The Art of Inspiring Shoppers

Next Post

Experts gather in Atlanta for Supercomputing 2024, showcasing several breakthrough technologies

Next Post
Experts gather in Atlanta for Supercomputing 2024, showcasing several breakthrough technologies

Experts gather in Atlanta for Supercomputing 2024, showcasing several breakthrough technologies

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

POPULAR POSTS

  • 10 Ways To Get a Free DoorDash Gift Card

    10 Ways To Get a Free DoorDash Gift Card

    0 shares
    Share 0 Tweet 0
  • They Combed the Co-ops of Upper Manhattan With $700,000 to Spend

    0 shares
    Share 0 Tweet 0
  • Saal.AI and Cisco Systems Inc Ink MoU to Explore AI and Big Data Innovations at GITEX Global 2024

    0 shares
    Share 0 Tweet 0
  • Exxon foe Engine No. 1 to build fossil fuel plants with Chevron

    0 shares
    Share 0 Tweet 0
  • They Wanted a House in Chicago for Their Growing Family. Would $650,000 Be Enough?

    0 shares
    Share 0 Tweet 0
Solega Blog

Categories

  • Artificial Intelligence
  • Cryptocurrency
  • E-commerce
  • Finance
  • Investment
  • Project Management
  • Real Estate
  • Start Ups
  • Travel

Connect With Us

Recent Posts

How Cox Automotive Ties Employee Experience to Customer Success

How Cox Automotive Ties Employee Experience to Customer Success

June 22, 2025
7 Essential Things To Do When You Get Paid

7 Essential Things To Do When You Get Paid

June 22, 2025

© 2024 Solega, LLC. All Rights Reserved | Solega.co

No Result
View All Result
  • Home
  • E-commerce
  • Start Ups
  • Project Management
  • Artificial Intelligence
  • Investment
  • More
    • Cryptocurrency
    • Finance
    • Real Estate
    • Travel

© 2024 Solega, LLC. All Rights Reserved | Solega.co