Applied AI

Demo: Using Snorkel Flow to train Microsoft Azure Form Recognizer models

January 5, 2023
2 min read

Last month Snorkel AI highlighted how we’re deepening our partnership with Microsoft Azure AI to help enterprises and government agencies solve their most impactful problems. One example of how Snorkel AI helps organizations leverage Azure AI services for proprietary data and custom objectives is the new Snorkel Flow integration for Microsoft Azure Form Recognizer (currently in private preview).  

Azure Form Recognizer is an AI service that provides pre-built and customizable models for analyzing forms and PDFs. In addition to pre-built models supporting standard forms like W-2s, invoices, receipts, business cards, etc., Form Recognizer includes custom training support to fine-tune its powerful neural document models and support proprietary datasets, non-standard formats, and custom objectives. 

Check out this demo to see how Snorkel Flow can be used to easily and quickly train Azure Form Recognizer custom models. In this demo, we showcase the integration with a real-world use case from real estate and construction. Companies in this space manage massive volumes of forms and contracts with high variability depending on the year or specific process being followed. Without Snorkel Flow, it would require significant manual effort to extract specific key dollar amounts (such as the original contract amount, the current contract amount, and any contract value modifications included in the document) from a large data set of real estate and construction contracts. 

The demo also shows how Snorkel Flow can be used to orchestrate document content preprocessing (including OCR, layout information, and more through Form Recognizer), kick off custom Form Recognizer training jobs directly from the UI, and auto-generate performance analyses over custom Form Recognizer models to guide your next steps.

We’re excited to continue deepening our partnership with Microsoft Azure AI to accelerate AI development for enterprises. Schedule a custom demo tailored to your use case and Azure stack with our ML experts today.

Share this article

Recommended articles

View all articles
Image
Why Frontier Agents Fail Real Engineering Work: Two Terminal-Bench 3.0 Task Deep Dives
Terminal-Bench 3.0 (formerly Frontier-Bench) recently launched, built to track what AI agents can and can’t do across real computer work. Terminal-Bench 2.1 has been saturating, with top agents reaching 84%; on Terminal-Bench 3.0, the best model, Claude Opus 5, achieves just 43.5%. Terminal-Bench 3.0 raises the bar with 74 authentic, verifiable tasks across 7 domains, designed to expose meaningful gaps
August 20, 2026
Derek Pham
,
Srikar Kodati
Image
Claude Opus 5: Performance and Error Analysis on Frontier Coding Tasks
Anthropic’s Claude Opus 5 recently debuted as the second model overall on the current Senior SWE-bench leaderboard, behind Fable 5. It also achieves the highest score of any evaluated model on the benchmark’s Bug & Performance Investigation category, reinforcing the rapid progress frontier coding models continue to make on increasingly realistic software engineering tasks. Just as notable, Opus 5 reaches
July 27, 2026
Ankit Aich
Image
Inside Frontier-Bench: two Snorkel-built tasks that frontier agents still can’t crack
Frontier-Bench launched this week – the successor to Terminal-Bench, built to track what AI agents can and can’t do across real computer work. Terminal-Bench 2.1 has been saturating, with top agents clearing 75-84%; on Frontier-Bench’s launch set of 74 tasks across 7 domains, the best mode Opus 5 achieving 43.3%. Snorkel AI contributed as a task author and data partner,
July 23, 2026
Derek Pham
,
Srikar Kodati
Image
Image

Join our newsletter

For expert advice, the latest research, and exclusive events.
By submitting this form, I acknowledge I will receive email updates from Snorkel AI, and I agree to the Terms of Use and acknowledge that my information will be used in accordance with the Privacy Policy.