When you enroll in this course, you'll also be enrolled in this Specialization.
Learn new concepts from industry experts
Gain a foundational understanding of a subject or tool
Develop job-relevant skills with hands-on projects
Earn a shareable career certificate
There are 2 modules in this course
"Unlock Multimodal Search" is an intermediate, hands-on course for developers and ML engineers ready to build the next generation of AI-powered search. Text-only search is no longer enough; this 90-minute course will teach you how to create applications that can search across different data types, such as finding text from an image. Using the powerful open-source vector database Weaviate, you will move from theory to a functioning demonstration. This course requires basic skills in Docker, APIs, Python, and the command line (CLI). Familiarity with vector databases. Docker Desktop must be installed.
This course is focused on execution. You will learn to configure a Weaviate schema to handle both image and text embeddings for a single object, ingest multimodal data, and perform powerful cross-modal queries. Through a final, hands-on project that mirrors a real-world job task, you will not only build an image-to-text search demo but also learn how to measure its accuracy with precision metrics. By the end, you'll be equipped to architect and validate sophisticated, multimodal AI applications.
This module provides the foundation for multimodal search. You will learn how to configure a Weaviate instance and design a schema that can store both image and text embeddings for a single object, preparing your database for cross-modal queries.
Now that your database is configured, this module focuses on execution and validation. You will learn how to perform cross-modal queries to find text from an image and, critically, how to analyze the accuracy of your results using precision metrics.
What's included
2 videos1 reading2 assignments
Show info about module content
2 videos•Total 11 minutes
The Power of Visual Search•4 minutes
How-To: Run an Image-to-Text Search•7 minutes
1 reading•Total 4 minutes
Cross-Modal Queries and Measuring Precision•4 minutes
2 assignments•Total 38 minutes
Image Search Demo and Precision Report•30 minutes
Hands-On Learning: Execute a Cross-Modal Query•8 minutes
Earn a career certificate
Add this credential to your LinkedIn profile, resume, or CV. Share it on social media and in your performance review.
Coursera brings together a diverse network of subject matter experts who have demonstrated their expertise through professional industry experience or strong academic backgrounds. These instructors design and teach courses that make practical, career-relevant skills accessible to learners worldwide.
When will I have access to the lectures and assignments?
To access the course materials, assignments and to earn a Certificate, you will need to purchase the Certificate experience when you enroll in a course. You can try a Free Trial instead, or apply for Financial Aid. The course may offer 'Full Course, No Certificate' instead. This option lets you see all course materials, submit required assessments, and get a final grade. This also means that you will not be able to purchase a Certificate experience.
What will I get if I subscribe to this Specialization?
When you enroll in the course, you get access to all of the courses in the Specialization, and you earn a certificate when you complete the work. Your electronic Certificate will be added to your Accomplishments page - from there, you can print your Certificate or add it to your LinkedIn profile.
Is financial aid available?
Yes. In select learning programs, you can apply for financial aid or a scholarship if you can’t afford the enrollment fee. If fin aid or scholarship is available for your learning program selection, you’ll find a link to apply on the description page.