AI Instructor Live Labs Included

Claude Vision and PDF processing

Advanced
11h 15m
10 Lessons
Claude Vision and PDF processing Badge

View badge details

About This Course

Process images and PDFs with Claude chart extraction, receipt/invoice OCR, form understanding, multi-page document reasoning, and structured JSON extraction from visual sources. Learn the vision content-block shape, PDF beta, cost/latency tradeoffs, and when to route to Haiku vs Sonnet vs Opus.By the end of this course you will be able to build a document-intelligence pipeline for Orion Analytics that ingests scanned invoices, extracts line-items, validates totals, and hands off structured data to downstream systems.

Course Curriculum

10 Lessons
01
AI Lesson
AI Lesson

Vision content blocks and image inputs

1h 0m

How Claude accepts images: content-block shape, base64 vs URL, supported formats, size limits, and the "one image = ~1.15K tokens" cost model.

02
Lab Exercise
Lab Exercise

Describe a chart with Claude Vision - Lab Exercises

1h 15m 1 Exercises

Encode a chart image, send it via a base64 image block, and read Claude's structured description.

03
AI Lesson
AI Lesson

Structured extraction from images

1h 0m

Get JSON out of pictures — receipts, forms, screenshots. Prompt patterns for reliable schema-conforming extraction.

04
Lab Exercise
Lab Exercise

Extract a receipt to JSON - Lab Exercises

1h 15m 1 Exercises

Use tool_use to extract a scanned receipt into structured JSON with merchant, date, total, and line items.

05
AI Lesson
AI Lesson

PDF processing — pages, tokens, and the beta header

1h 0m

Send PDFs directly to Claude. Cost model, page limits, multi-page reasoning, and when to slice vs whole-doc.

06
Lab Exercise
Lab Exercise

Extract a 3-page invoice PDF - Lab Exercises

1h 15m 1 Exercises

Send a multi-page PDF invoice to Claude via the document content-block, extract totals and line items across pages.

07
AI Lesson
AI Lesson

Vision routing — Haiku vs Sonnet vs Opus

1h 0m

Choosing the right vision model. Cost/accuracy tradeoffs by task type: simple OCR (Haiku), forms/charts (Sonnet), complex layouts (Opus).

08
Lab Exercise
Lab Exercise

Compare Haiku, Sonnet, and Opus on the same image - Lab Exercises

1h 15m 1 Exercises

Send the same receipt image through all three model tiers, compare accuracy, latency, and cost side-by-side.

09
AI Lesson
AI Lesson

Vision + tool_use — decide then act

1h 0m

Combine vision with tools: Claude sees an image, decides what to do, then invokes a tool to act. Two-turn pattern for document-routing pipelines.

10
Lab Exercise
Lab Exercise

Orion invoice processing pipeline capstone - Lab Exercises

1h 15m 1 Exercises

Build an end-to-end document pipeline: vision router → structured extraction → validation. Works on both PDFs and images.

This course includes:

  • 24/7 AI Instructor Support
  • Live Lab Environments
  • 5 Hands-on Lessons
  • Completion Badge
Claude Vision and PDF processing Badge

Earn Your Badge

Complete all lessons to unlock the Claude Vision and PDF processing achievement badge.

Skill Level Advanced
Total Duration 11h 15m
Claude Vision and PDF processing Badge
Achievement Badge

Claude Vision and PDF processing

Awarded on completion of CLD-AI-110. The holder can process images and PDFs with Claude, extract structured JSON from documents, choose between Haiku/Sonnet/Opus by task complexity, and build document-intelligence pipelines.

Course Claude Vision and PDF processing
Criteria Complete all lessons and hands-on labs in CLD-AI-110 and pass the embedded assessments.

Skills You'll Earn

Vision content-block shape Base64 vs URL image sources Structured JSON extraction from images Multi-page PDF processing Model routing by task complexity Vision + tool_use combined pipelines

Complete all lessons in this course to earn this badge