MindSpore Transformers

Introduction

  • Quick Start
  • Overall Structure
  • Model Support Library

Installation

  • Installation Guide

Training Guide

  • Training Guide

Function Features

  • Function Overview
  • Starting Tasks
  • Configuration File Description
  • Logs
  • Datasets
  • Hyperparameters and Optimizers for Training
  • Distributed Parallel Training
  • Training Graphics Memory Optimization
  • Weight Saving and Loading
  • Resumable Training
  • Training Metric Monitoring and Profiling
  • Other Training Features
  • Static Graph Features

Environment Variables

  • Environment Variables

Contribution Guide

  • MindSpore Transformers Contribution Guidelines
  • Modelers Contribution Guidelines

FAQ

  • Model-Related FAQ
  • Feature-Related FAQ

Static Graph Implementation (Deprecated)

  • Overall Structure
  • Full-process Guide to Large Models
    • Training Guide
    • Pretraining
    • Supervised Fine-Tuning (SFT)
    • Inference
    • Service Deployment
    • Evaluation
  • Features
  • Advanced Development
  • Excellent Practice
  • Environment Variable Descriptions
MindSpore Transformers
  • »
  • Full-process Guide to Large Models
  • View page source

Full-process Guide to Large Models

  • Training Guide
  • Pretraining
  • Supervised Fine-Tuning (SFT)
  • Inference
  • Service Deployment
  • Evaluation
Previous Next

© Copyright MindSpore.

Built with Sphinx using a theme provided by Read the Docs.