Close Menu
Sedona.Biz – The Voice of Sedona and The Verde ValleySedona.Biz – The Voice of Sedona and The Verde Valley
    Sedona.Biz – The Voice of Sedona and The Verde ValleySedona.Biz – The Voice of Sedona and The Verde Valley
    • Home
    • Sedona
      • Steve’s Corner
      • Bear Howard Chronicles
      • Business Profiles
      • Mind and Body
      • Real Estate
      • Sedona News
    • About
    • Advertise
    • Shop
    • Sedona’s Best
    Sedona.Biz – The Voice of Sedona and The Verde ValleySedona.Biz – The Voice of Sedona and The Verde Valley
    Home » AI Safety: Why LLMs Benchmarks or Evaluations May Not Decide Alignment, AGI
    Sedona News

    AI Safety: Why LLMs Benchmarks or Evaluations May Not Decide Alignment, AGI

    July 3, 2024No Comments
    Facebook Twitter Pinterest LinkedIn Email Reddit WhatsApp
    AI Safety: Why LLMs Benchmarks or Evaluations May Not Decide Alignment, AGI
    Share
    Facebook Twitter LinkedIn Pinterest Email Reddit WhatsApp

    By David Stephen

    Advances for AI safety are not currently a problem of evaluations or benchmarks for models, since new benchmarks alone are unlikely to solve the current problems of misinformation and deepfakes—images, audios and videos. There are several present risks with AI that new evaluation methods may do little or nothing to solve.

    How is it possible to trace the AI source of some misinformation or voice cloning for deception? How can a post-guardrail AI model that produces a problematic output be penalized for its actions?

    Already there are several benchmark and evaluation rankings for LLMs. While they measure certain criteria, they do not solve some current problems, nor do they provide a sure way to determine what is or when artificial general intelligence [AGI] or artificial superintelligence [ASI] may arrive.

    There is a recent feature on WIRED, We’re Still Waiting for the Next Big Leap in AI, with the quotes, “Gauging the rate of progress in AI using conventional benchmarks like those touted by Anthropic for Claude can be misleading. AI developers are strongly incentivized to design their creations to score highly in these benchmarks, and the data used for these standardized tests can be swept into their training data. Benchmarks within the research community are riddled with data contamination, inconsistent rubrics and reporting, and unverified annotator expertise.”

    Anthropic just announced, A new initiative for developing third-party model evaluations, stating that, “We’re introducing a new initiative to fund evaluations developed by third-party organizations that can effectively measure advanced capabilities in AI models. We are interested in sourcing three key areas of evaluation development: AI Safety Level assessments; Advanced capability and safety metrics; Infrastructure, tools, and methods for developing evaluations

    How would this counter general misinformation? How would they prevent deepfake videos, audios and images, outside scrutinized areas like politics and elections? There are several AI tools in search results that make several misuses possible. How can there be a collective safety approach against some of their outputs?

    How can AGI or ASI be determined or measured with comparison to how human intelligence works? If human intelligence is based in the human mind, how does the human mind mechanize intelligence? AI already has access to lots of memory. It can make inferences about the world through language without self-experience. If it were the human mind, with access to resources on things, without experience, how does AI currently compare

    There are several detection tools for AI outputs, with varying levels of accuracy, but knowing that something came by AI, may not matter if the thing is already used to cause harm. How can outputs around certain keywords be tracked, across AI outputs indexed on search engines, using web crawling and scraping?

    If an AI model is misused, how can it begin to lose access to some of its parameters, as a consequence for its actions? There are directions that some AI safety and alignment research are going that may not be helpful for current risks—or existential risks. There are also benchmarks that are sought for AGI, without exploring the human mind.

    The threats and risks of AI exceed the safety of individual frontier models. The capabilities of AI exceed its limitation to language. Approaching answers from extended angles would make a better case for the common purpose.

     

     

    Comments are closed.

    Related Coverage

    Sedona Airport to Host Annual Wings & Wheels ‘26 On Oct. 10

    October 6, 2026

    Sedona Claims Bragging Rights in Inaugural Sister Cities Golf Exchange

    October 4, 2026

    Rotary Youth Exchange Opens Doors to International Adventure — Abroad and at Home

    October 4, 2026

    Arts & Crafts Fair Returns to Sedona Heritage Museum

    October 4, 2026

    City of Sedona has Help for Sedona Homeowners to Switch to Heat Pumps and Lower Energy Bills 

    October 4, 2026

    Howl-O-Ween Dog Costume Parade Returns to Tlaquepaque

    October 4, 2026
    BV’s Italian Kitchen Launches Lunch Menu

    Cottonwood gets another great place to have lunch at and this great new place happens to be BV’s Italian Kitchen. ‘Breakfast may start the day and dinner may get all the romance, but somewhere between the two sits the meal that too often gets treated like an afterthought. Click photo to learn more.

    Sedona Realtor
    In The Living Room Music Series

    Every other Monday, the Mary D. Fisher Theatre transforms into your living room for a FUN, intimate, interactive night of music and conversation! Enjoy LIVE music and ask the artist your questions during the concert. Epic music. Real conversations. Unforgettable Mondays. Click the photo to claim your seat!

     

    Sedona’s Backstage Pass

     

    Tune in weekly for Shondra’s behind-the-scenes conversations with the Creators, Curators, and Visionaries who are the heartbeat of Sedona’s Creativity. Spotify Click HERE. Apple Podcast Click HERE.

     

     

    Recent Comments
    • Jill Dougherty on SEDONA VERDE VALLEY: TWO IDENTITIES, ONE REGIONAL FUTURE
    • Jill Dougherty on Blessing the new moon is an ancient and modern practice
    • Derek Pfaff on SEDONA VERDE VALLEY: TWO IDENTITIES, ONE REGIONAL FUTURE
    • Jill Dougherty on SEDONA VERDE VALLEY: TWO IDENTITIES, ONE REGIONAL FUTURE
    • JB on To Kill or Not to Kill: That’s Tennessees’s Question
    • Robert Hagedorn on Blessing the new moon is an ancient and modern practice
    • Buddy Oakes on To Kill or Not to Kill: That’s Tennessees’s Question
    • Jill Dougherty on Fashion Lab Sedona Takes on Fast Fashion With a Runway Made in Sedona
    • jim kautz on Sedona home prices firmed up in August. Here is what the data actually shows
    • Craig on WILL HISTORY REMEMBER TRUMP AS ONE OF AMERICA’S GREAT PRESIDENTS?
    Don’t miss a beat – signup for our weekly newsletter
    Cactus Quill
    Categories
    Your ad could be here
    The Voice of Sedona and The Verde Valley

    News

    • Sedona News
    • Verde Valley News
    • Editorials/Opinion
    • Letter to The Editor

    Community

    • Arts and Culture
    • Mind and Body
    • Spiritual
    • Community Events
    • Sedona Restaurants

    More

    • Sedona Real Estate
    • Shop
    • Advertise
    • About
    • Contact
    • Editorial Policy

    Connect

    f
    Get the best of Sedona delivered to your inbox.
    Our Network: TheSedonan.com • SedonaBest.com
    © 2026 Sedona.Biz · Privacy Policy · Editorial Policy · Contact

    Type above and press Enter to search. Press Esc to cancel.