The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) 2024年05月12日
How Deep Learning has Revolutionized OCR with Cha Zhang - #416
index_new5.html
../../../zaker_core/zaker_tpl_static/wap/tpl_guoji1.html

 

Today we’re joined by Cha Zhang, a Partner Engineering Manager at Microsoft Cloud & AI. 

Cha’s work at MSFT is focused on exploring ways that new technologies can be applied to optical character recognition, or OCR, pushing the boundaries of what has been seen as an otherwise ‘solved’ problem. In our conversation with Cha, we explore some of the traditional challenges of doing OCR in the wild, and what are the ways in which deep learning algorithms are being applied to transform these solutions. 

We also discuss the difficulties of using an end to end pipeline for OCR work, if there is a semi-supervised framing that could be used for OCR, the role of techniques like neural architecture search, how advances in NLP could influence the advancement of OCR problems, and much more. 

The complete show notes for this episode can be found at twimlai.com/go/416.

Fish AI Reader

Fish AI Reader

AI辅助创作,多种专业模板,深度分析,高质量内容生成。从观点提取到深度思考,FishAI为您提供全方位的创作支持。新版本引入自定义参数,让您的创作更加个性化和精准。

FishAI

FishAI

鱼阅,AI 时代的下一个智能信息助手,助你摆脱信息焦虑

联系邮箱 441953276@qq.com

相关标签

相关文章