WWDC25: Read documents using the Vision framework | Apple
Apple Developer
0:00 / 0:00
WWDC25: Read documents using the Vision framework | Apple
7 106 просмотров · 1 год назад
Apple Developer
329 тыс. подписчиков
7 106 просмотров · 1 год назад
Learn about the latest advancements in the Vision framework. We’ll introduce RecognizeDocumentsRequest, and how you can use it to read lines of text and group them into paragraphs, read tables, etc. And we’ll also dive into camera lens smudge detection, and how to identify potentially smudged images in photo libraries or your own camera capture pipeline.
Explore related documentation, sample code, and more:
Classifying Images with Vision and Core ML: https://developer.apple.com/documenta...
Vision: https://developer.apple.com/documenta...
Recognizing tables within a document: https://developer.apple.com/documenta...
Image Classification with Vision and CoreML: https://developer.apple.com/sample-co...
Discover Swift enhancements in the Vision framework: https://developer.apple.com/videos/pl...
Detect animal poses in Vision: https://developer.apple.com/videos/pl...
Explore 3D body pose and person segmentation in Vision: https://developer.apple.com/videos/pl...
Discover machine learning & AI frameworks on Apple platforms: https://developer.apple.com/videos/pl...
00:00 - Introduction
01:22 - Reading documents
13:35 - Camera lens smudge detection
17:59 - Hand pose update
More Apple Developer resources:
Video sessions: https://apple.co/VideoSessions
Documentation: https://apple.co/DeveloperDocs
Forums: https://apple.co/DeveloperForums
App: https://apple.co/DeveloperApp