World Lifestyler
  • Art & Culture
    • Architecture
    • Art & Exhibitions
    • Books
    • Design
    • Film & Music
  • Competitions
    • Dining Experiences
    • Hotel Stays
    • Luxury Experiences
    • Product Giveaways
    • Reader Exclusives
    • Travel Giveaways
  • Food & Drink
    • Chefs
    • Coffee Culture
    • Food Destinations
    • Recipes
    • Restaurants
    • Wine & Spirits
  • Lifestyle
    • Design
    • Fashion
    • Health & Wellbeing
    • Homes & Property
    • Love & Romance
  • People
    • Creatives
    • Entrepreneurs
    • Icons
    • Interviews
    • Profiles
    • Rising Talent
  • Travel
    • Adventure & Experience Travel
    • City Guides
    • Destinations
    • Hotels
    • Secret Spots
    • Travel Trends
  • Art & Culture
    • Architecture
    • Art & Exhibitions
    • Books
    • Design
    • Film & Music
  • Competitions
    • Dining Experiences
    • Hotel Stays
    • Luxury Experiences
    • Product Giveaways
    • Reader Exclusives
    • Travel Giveaways
  • Food & Drink
    • Chefs
    • Coffee Culture
    • Food Destinations
    • Recipes
    • Restaurants
    • Wine & Spirits
  • Lifestyle
    • Design
    • Fashion
    • Health & Wellbeing
    • Homes & Property
    • Love & Romance
  • People
    • Creatives
    • Entrepreneurs
    • Icons
    • Interviews
    • Profiles
    • Rising Talent
  • Travel
    • Adventure & Experience Travel
    • City Guides
    • Destinations
    • Hotels
    • Secret Spots
    • Travel Trends
No Result
View All Result
WORLD LIFESTYLER
No Result
View All Result
Home Press Releases Press Releases - Lifestyle

Snorkel AI Highlights First Wave of Open Benchmarks Grants Projects

Cision PR Newswire by Cision PR Newswire
July 24, 2026
in Press Releases - Lifestyle
Reading Time: 3 mins read
0
Share on FacebookShare on Twitter

SAN FRANCISCO, July 24, 2026 /PRNewswire/ — Snorkel AI today highlighted the first group of projects supported through Open Benchmarks Grants, a $3 million commitment to support open-source datasets, benchmarks, and evaluation research.

Snorkel AI, The Frontier AI Data Lab

Launched in February 2026, Open Benchmarks Grants has received hundreds of applications from researchers, labs, and engineers working to address a growing challenge: AI systems are advancing faster than the field’s ability to rigorously measure their performance on realistic, consequential work.

“From complex environments and huge autonomy horizons to rich, sophisticated outputs, these projects tackle some of the field’s hardest evaluation challenges,” said Fred Sala, a member of the Open Benchmarks Grants steering committee and assistant professor at the University of Wisconsin–Madison. “I’m excited to see the broader research community use, validate, and build on them.”

Open Benchmarks Grants provides selected teams with funding, expert data development support, research and engineering collaboration, and platform resources. Supported projects include:

  • Frontier-Bench (formerly Terminal-Bench 3.0), developed with Laude Institute and the Harbor community, is a harder, more domain-diverse successor to Terminal-Bench 2.1 — built in the open, task by task, under continuous adversarial review.
  • Agents’ Last Exam, developed with UC Berkeley RDI and the RDI Foundation, evaluates agents on long-horizon, economically valuable professional workflows. It spans 55 sub-industries and includes more than 1,500 tasks toward a 5,000-task target, sourced and validated by more than 300 industry experts.
  • OSWorld 2.0, developed with XLANG Lab, evaluates computer-use agents on 108 long-horizon workflows across 31 self-hosted web environments and professional desktop applications.
  • Continual Learning Bench, developed with UC Berkeley SkyLab and the University of Wisconsin–Madison, measures whether agents genuinely improve across sequential, stateful tasks.
  • SlopCode Bench, developed with the University of Wisconsin–Madison, measures how code quality degrades as coding agents repeatedly modify and extend their own solutions.
  • Terminal-Bench 2.1, developed with Stanford University, Laude Institute and the Harbor community, evaluates agents on challenging work in terminal environments. The release corrected 28 tasks and introduced continuous validation.

With support from Open Benchmarks Grants, Terminal-Bench Science is also now in development, extending the Terminal-Bench framework to computational research workflows across the life, physical, earth, and mathematical sciences.

Beyond the grants program, Snorkel led the development of Senior SWE-Bench with the research teams at Princeton University and the University of Wisconsin–Madison. The benchmark evaluates coding agents on senior-level engineering work, including implementing features from realistic instructions, investigating bugs that require runtime analysis, and producing code that follows existing codebase conventions.

Open Benchmarks Grants was established with support from Hugging Face, Prime Intellect, Together AI, Factory, Harbor, and PyTorch. Applications remain open and are reviewed on a rolling basis.

Learn more and apply for a grant at benchmarks.snorkel.ai.

About Snorkel AI
Snorkel AI is the frontier AI data lab, helping teams build the data and environments behind high-performing frontier and agentic AI. We combine technology with research-driven AI data development to create datasets, benchmarks, evals, and custom solutions for real-world AI systems. Founded out of the Stanford AI Lab in 2019, Snorkel works with leading AI labs and enterprises to move from better data to better outcomes. 

media@snorkel.ai

Cision View original content to download multimedia:https://www.prnewswire.com/news-releases/snorkel-ai-highlights-first-wave-of-open-benchmarks-grants-projects-302833805.html

SOURCE Snorkel AI

Cision PR Newswire

Cision PR Newswire

Related Posts

GAC Becomes Official Automotive Partner of Melbourne City FC, Strengthening Commitment to the Australian Market

July 24, 2026

DC Climate Week 2026 Highlights the Growing Role of Distributed Energy Resources With Insights from Emily Sanford Fisher

July 24, 2026

Jury Awards $17 Million to Woman Whose Bladder Was Mistakenly Removed During Routine Surgery

July 24, 2026

NTEC Scholarship Program Surpasses $1 Million in Awards to Navajo Students

July 24, 2026

GSMA Welcomes Abuja Declaration on Meaningful Connectivity for Africa and Joins Partners to Launch ATLAS Umoja

July 24, 2026

In HelloNation, Behavioral Health Experts Cody Luke and David Spencer Explain When It Is Time to See a Therapist in Idaho Falls

July 24, 2026

Popular News

  • DC Climate Week 2026 Highlights the Growing Role of Distributed Energy Resources With Insights from Emily Sanford Fisher

    0 shares
    Share 0 Tweet 0
  • GAC Becomes Official Automotive Partner of Melbourne City FC, Strengthening Commitment to the Australian Market

    0 shares
    Share 0 Tweet 0
  • Jury Awards $17 Million to Woman Whose Bladder Was Mistakenly Removed During Routine Surgery

    0 shares
    Share 0 Tweet 0
  • Solid Joins Snowflake and Industry Leaders to Advance Open Standards for AI-Ready Semantic Context

    0 shares
    Share 0 Tweet 0
  • Timeshifter Marks Circadian Awareness Day (24/7) with “Why Timing Matters in Healthcare”

    0 shares
    Share 0 Tweet 0

About & Contact

  • About Us
  • Branding Style Guide
  • Contact Us
  • Help Centre
  • Media Kit
  • Site Map

Explore Content

  • Events
  • Newsletter
  • Press Releases
  • Topics

Legal & Privacy

  • Advertiser & Partner Policy
  • Communications & Newsletter Policy
  • Contributor Agreement
  • Copyright Policy
  • Privacy Policy
  • Prohibited Content Policy
  • Terms of Service

Tiny Media Brands

  • Silicon Valleys Journal
  • The AI Journal
  • The City Banker
  • The Wall Street Banker
  • World Lifestyler

© 2025 World Lifestyler

No Result
View All Result
  • Home

© 2025 World Lifestyler