resources

Resource library

Explore our complete library of resources including blogs, benchmarks, research papers and more.
Image for Evaluating Coding Agent Capabilities with Terminal-Bench: Snorkel’s Role in Building the Next Generation Benchmark
Blog

Evaluating Coding Agent Capabilities with Terminal-Bench: Snorkel’s Role in Building the Next Generation Benchmark

Announcing a $3M commitment to launch Open Benchmarks Grants
September 30, 2025
Image for Closing the Evaluation Gap in Agentic AI
Blog

Closing the Evaluation Gap in Agentic AI

Announcing a $3M commitment to launch Open Benchmarks Grants

February 11, 2026
Image for Benchtalks #1: Alex Shaw (Terminal-Bench, Harbor) – Building the Benchmark Factory
Blog

Benchtalks #1: Alex Shaw (Terminal-Bench, Harbor) – Building the Benchmark Factory

Announcing a $3M commitment to launch Open Benchmarks Grants
March 31, 2026
Image for Building FinQA: An Open RL Environment for Financial Reasoning Agents
Blog

Building FinQA: An Open RL Environment for Financial Reasoning Agents

Announcing a $3M commitment to launch Open Benchmarks Grants
March 30, 2026
Image for The science of rubric design
Blog

The science of rubric design

Announcing a $3M commitment to launch Open Benchmarks Grants
September 11, 2025
of
Type: All Types
Sort: Newest
Efficiently Modeling Long Sequences with Structured State Spaces
This paper introduces the Structured State Space sequence model (s4), which uses a new parameterization for the state-space model to improve long-range dependency handling both mathematically and empirically.
Research Paper
Efficiently Modeling Long Sequences with Structured State Spaces

This paper introduces the Structured State Space sequence model (s4), which uses a new parameterization for the state-space model to improve long-range dependency handling both mathematically and empirically.

Mar 29, 2022

A. Gu, et al

Learn more about Efficiently Modeling Long Sequences with Structured State Spaces
Learning from Multiple Noisy Partial Labelers
This work enables users to create partial labelers that output subsets of possible class labels would greatly expand the expressivity of programmatic weak supervision.
Research Paper
Learning from Multiple Noisy Partial Labelers

This work enables users to create partial labelers that output subsets of possible class labels would greatly expand the expressivity of programmatic weak supervision.

Mar 28, 2022

P. Yu, et al

Learn more about Learning from Multiple Noisy Partial Labelers
Snorkel AI welcomes industry leaders to the team
Blog
Snorkel AI welcomes industry leaders to the team

 

Mar 21, 2022
Learn more about Snorkel AI welcomes industry leaders to the team
Learning with imperfect labels and visual data with Anima Anandkumar
Blog
Learning with imperfect labels and visual data with Anima Anandkumar

The future of data-centric AI talk series Background Anima Anandkumar holds dual positions in academia and industry. She is a Bren professor at Caltech and the director of machine learning research at NVIDIA. Anima also has a long list of accomplishments ranging from the Alfred P. Sloan scholarship to the prestigious NSF career award and many more. She recently joined…

Mar 18, 2022
Learn more about Learning with imperfect labels and visual data with Anima Anandkumar
Weak Supervision Modeling with Fred Sala
Blog
Weak Supervision Modeling with Fred Sala

Understanding the label model. Machine learning whiteboard (MLW) open-source series Background Frederic Sala, is an assistant professor at the University of Wisconsin-Madison, and a research scientist at Snorkel AI. Previously, he was a postdoc in Chris Re’s lab at Stanford. His research focuses on data-driven systems and weak supervision. In this talk, Fred focuses on weak supervision modeling. This machine…

Mar 17, 2022
Learn more about Weak Supervision Modeling with Fred Sala
Tips for using a data-centric AI approach
Blog
Tips for using a data-centric AI approach

The future of data-centric AI talk series Background An AI system consists of two parts: the model— algorithm or some code—and data. The dominant paradigm in machine-learning researchers has been for most data scientists, including myself, to download a fixed dataset and iterate on the model. That this has become conventional is a tribute to how successful this model-centric approach…

Mar 09, 2022
Learn more about Tips for using a data-centric AI approach
Resilient enterprise AI application development
Blog
Resilient enterprise AI application development

Using a data-centric approach to capture the best of rule-based systems and ML models for enterprise AI One of the biggest challenges to making AI practical for the enterprise is keeping the AI application relevant (and therefore valuable) in the face of ever-changing input data and evolving business objectives. Practitioners typically use one of two approaches to build these AI applications:…

Mar 03, 2022
Learn more about Resilient enterprise AI application development
How AI can be used to rapidly respond to information warfare in the Russia-Ukraine conflict
Blog
How AI can be used to rapidly respond to information warfare in the Russia-Ukraine conflict

Proliferating web technology has contributed to information warfare in recent conflicts. Artificial Intelligence (AI) can play a significant role in stemming disinformation campaigns, cyber-attacks, and informing diplomacy in the rapidly evolving situation in Ukraine. Snorkel AI is dedicated to supporting the National Security community and other enterprise organizations with state-of-the-art AI technology. We see this as our responsibility in the…

Feb 28, 2022
Learn more about How AI can be used to rapidly respond to information warfare in the Russia-Ukraine conflict
Blog
How Genentech extracted information for clinical trial analytics with Snorkel Flow

Genentech, a global biotech leader and member of the Roche Group, leveraged Snorkel Flow to extract critical information from lengthy clinical trial protocol (CTP) pdf documents. They built AI applications that used NER, entity linking, text extraction, and classification models to determine inclusion/ exclusion criteria and to analyze Schedules of Assessments. Genentech’s team achieved 95-99% model accuracy by using Snorkel…

Feb 26, 2022
Learn more about How Genentech extracted information for clinical trial analytics with Snorkel Flow
1 51 52 53 65
Image
Image

Join our newsletter

For expert advice, the latest research, and exclusive events.
By submitting this form, I acknowledge I will receive email updates from Snorkel AI, and I agree to the Terms of Use and acknowledge that my information will be used in accordance with the Privacy Policy.