Welcome back to part three of our three part series on sharding and parallelism. In this episode we’ll put everything together, covering the training loop, data loading, checkpointing, and a complete, practical example with a Transformer block.
Resources:
Learn more → https://goo.gle/learning-jax
Subscribe to Google for Developers → https://goo.gle/developers
Speaker: Robert Crowe
Resources:
Learn more → https://goo.gle/learning-jax
Subscribe to Google for Developers → https://goo.gle/developers
Speaker: Robert Crowe
- Category
- Project
- Tags
- Google, developers, pr_pr: AI DevRel (fka Core ML);
Sign in or sign up to post comments.
Be the first to comment




