Snorkel - fast processing of training data
The training focuses on the practical use of the Snorkel framework for efficient processing and tagging of training data. Participants will learn techniques for programming tagging functions and methods for combining and verifying tagging. The program covers both theoretical fundamentals and practical applications in real projects. The classes are implemented in the form of workshops using real data sets.
Issues
-
Snorkel Architecture
-
Meaningful functions
-
Combining markings
-
Quality verification
-
Automation of marking
-
Integration with ML pipelines
-
Preparation of training data
-
Monitoring of assay quality
-
Strategies for improving accuracy
-
Use cases
-
Export and use of markings
-
Update models
Benefits
- Upon completion of the training, the participant will be able to effectively use Snorkel to quickly process and tag training data
- He/she will gain the ability to design and implement tagging functions for different types of data
- He/she will be able to evaluate and improve the quality of automated determinations
- Will master techniques for combining determinations from different sources
- Will learn to integrate Snorkel with existing machine learning pipelines
- Will gain hands-on experience in automating the data determination process
- Will learn methods for monitoring and improving assay quality
- Will be able to apply best practices in training data preparation
Who is this training for?
Prerequisites
- Knowledge of the basics of machine learning
- Experience in Python programming
- Basic knowledge of data preparation
- Understand supervised learning processes
Training program
Architecture of the framework
- Marking functions and their types
- Programming functions meaning
- Data management
- Advanced labeling techniques
Combining markings
- Quality verification
- Strategies for improving accuracy
Process automation
- Integration with ML pipelines
- Preparing data for training
- Export of markings
Quality monitoring
- Update models
- Practical applications
- Text marking
- Image classification
Sequence analysis
- Use cases
Delivery Methods
Online
- Convenience of participating from anywhere
- Interactive live sessions with trainer
- Materials available for 30 days
- No travel costs
On-site
- Direct contact with trainer and group
- Intensive hands-on workshops
- Networking with other participants
- Full focus on learning
Frequently asked questions
What are the prerequisites for this training?
For Snorkel - fast processing of training data we recommend: Knowledge of the basics of machine learning; Experience in Python programming; Basic knowledge of data preparation.
What is the format and duration of this training?
The training lasts 1 day and is available in online and on-site format. Sessions run from 9:00 AM to 4:00 PM. We can also customize the schedule to fit your team's needs.
Who is this training designed for?
This training is designed for: Data scientists working with large data sets; Machine learning engineers; Specialists in data preparation.
Request a quote
Funding Options
Check funding options for your company
Development Services Database
Up to 80% funding for SMEs from EU funds
Check availabilityNational Training Fund
Up to 100% funding for employers
Learn moreTrusted by
We train teams at Poland's largest companies
Interested in this training?
Contact us - we'll prepare an offer tailored to your organization's needs.