Thank you for sending your enquiry! One of our team members will contact you shortly.
Thank you for sending your booking! One of our team members will contact you shortly.
Course Outline
Introduction to Cambricon and MLU Architecture
- Overview of Cambricon’s AI chip portfolio
- MLU architecture and instruction pipeline
- Supported model types and use cases
Installing the Development Toolchain
- Installing BANGPy and Neuware SDK
- Environment setup for Python and C++
- Model compatibility and preprocessing
Model Development with BANGPy
- Tensor structure and shape management
- Computation graph construction
- Custom operation support in BANGPy
Deploying with Neuware Runtime
- Converting and loading models
- Execution and inference control
- Edge and datacenter deployment practices
Performance Optimization
- Memory mapping and layer tuning
- Execution tracing and profiling
- Common bottlenecks and fixes
Integrating MLU into Applications
- Using Neuware APIs for application integration
- Streaming and multi-model support
- Hybrid CPU-MLU inference scenarios
End-to-End Project and Use Case
- Lab: Deploying a vision or NLP model
- Edge inference with BANGPy integration
- Testing accuracy and throughput
Summary and Next Steps
Requirements
- A solid grasp of machine learning model architectures
- Proficiency in Python and/or C++
- A working knowledge of model deployment strategies and acceleration principles
Target Audience
- Developers specializing in embedded AI solutions
- ML engineers focused on edge or data center deployments
- Professionals developing within the Chinese AI infrastructure ecosystem
21 Hours
Testimonials (1)
That we can cover advance topic and work with real-life example