Deploy and Scale AI Models with Cloud Run
About this course
AI inference is the process of using a trained machine learning model to make predictions on new, unseen data by applying learned patterns. This course is designed for developers, data scientists, and ML engineers interested in quickly deploying AI inference services on Cloud Run. It is useful for those familiar with cloud-based serverless application deployment solutions, but who may not have experience with running AI inference using Google Cloud serverless products. The course includes examples that deploys a model for AI inference with GPUs and integrates gen AI apps with data storage services.
Price shown by Coursera — confirm on their site.
Enroll on CourseraYou'll be redirected to Coursera to complete enrollment.
- Listed & compared by CourseAsk
- English · All Levels
More courses like this
1020 Exam Style Practice Questions DY0-001 CompTIA DataAI
Udemy · Certificate
Coursera
R Programming: Data Analysis and Modeling
Coursera · MOOC / Non-credit
Coursera
Azure ML: Explore & Configure the Machine Learning Workspace
Coursera · Certificate
Coursera
Transformer Architectures and Multimodal Models
Coursera · MOOC / Non-credit
More courses from Coursera
Coursera
TCP/IP and Internet
Birla Institute of Technology & Science, Pilani · MOOC / Non-credit
Coursera
Agile Project Management
University of Colorado Boulder · Master's Degree
Coursera
Conservation and Sustainable Development
University of Michigan · MOOC / Non-credit
Coursera
Extra-Galactic Astronomy
University of Cambridge · MOOC / Non-credit