{"id":8041,"date":"2026-09-23T10:00:00","date_gmt":"2026-09-23T10:00:00","guid":{"rendered":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/"},"modified":"2026-09-23T10:00:00","modified_gmt":"2026-09-23T10:00:00","slug":"the-ai-enabled-boom-in-document-transcription","status":"publish","type":"post","link":"https:\/\/dev95.site\/ar\/the-ai-enabled-boom-in-document-transcription\/","title":{"rendered":"The AI-Enabled Boom in Document Transcription"},"content":{"rendered":"<div id=\"dev95-629821892\" class=\"dev95-- dev95-entity-placement\"><script async=\"async\" data-cfasync=\"false\" src=\"https:\/\/pl27862732.profitableratecpmnetwork.com\/2ad7a50e0bbc23ac6801d7b77c501463\/invoke.js\"><\/script>\r\n<div id=\"container-2ad7a50e0bbc23ac6801d7b77c501463\"><\/div><\/div><div id=\"dev95-653877857\" class=\"dev95-before-content dev95-entity-placement\"><p style=\"text-align: center;\"><strong>Stop wasting time on dead links! \ud83d\uded1 This smart platform automatically detects your device and country to give you the exact best offer instantly. Check it out now 100% free! \ud83d\udc47<\/strong><br data-sfc-root=\"ep\" data-sfc-pl=\"|||[]\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: Arial, sans-serif, &quot;Noto Color Emoji&quot;; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\" \/><a href=\"https:\/\/www.profitableratecpmnetwork.com\/pvhx5mbcc?key=ff7d362db3c20bc86fad685cc94f1fc4\">\ud83d\udd17 <strong class=\"rQesXe MPyX\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-complete=\"true\" aria-owns=\"action-menu-parent-container\" data-copy-service-computed-style=\"font-family: Arial, sans-serif, &quot;Noto Color Emoji&quot;; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">[Click here]<\/strong><\/a><\/p>\n<\/div><div>\n<p class=\"wp-block-paragraph\"><em>Chlo\u00eb Farr with Jessica Jack<\/em> <em>and Jacob Polay<\/em><\/p><div id=\"dev95-447045987\" class=\"dev95- dev95-entity-placement\"><center>\r\n<script>\r\n  atOptions = {\r\n    'key' : '4ba6b6513c00e0ba76511f798ae56401',\r\n    'format' : 'iframe',\r\n    'height' : 50,\r\n    'width' : 320,\r\n    'params' : {}\r\n  };\r\n<\/script>\r\n<script src=\"https:\/\/www.highrevenueformat.com\/4ba6b6513c00e0ba76511f798ae56401\/invoke.js\"><\/script>\r\n\t<\/center><\/div>\n<p class=\"wp-block-paragraph\"><em>This post is part of a <a href=\"https:\/\/activehistory.ca\/ai-and-collaboration\/\">series on AI and Collaboration.<\/a><\/em><\/p>\n<div class=\"wp-block-image\">\n<figure class=\"aligncenter\"><img data-recalc-dims=\"1\" loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"683\" data-attachment-id=\"144089\" data-permalink=\"https:\/\/activehistory.ca\/blog\/2026\/09\/23\/ai-enabled-document-transcription\/jarmoluk-old-books-436498\/\" data-orig-file=\"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-scaled.jpg\" data-orig-size=\"2560,1707\" data-comments-opened=\"1\" data-image-meta='{\"aperture\":\"0\",\"credit\":\"\",\"camera\":\"\",\"caption\":\"\",\"created_timestamp\":\"0\",\"copyright\":\"\",\"focal_length\":\"0\",\"iso\":\"0\",\"shutter_speed\":\"0\",\"title\":\"\",\"orientation\":\"0\",\"alt\":\"\"}' data-image-title=\"jarmoluk-old-books-436498\" data-image-description=\"\" data-image-caption=\"&lt;p&gt;&amp;#8220;Old Books&amp;#8221; by jarmoluk via Pixabay. Free for use under the Pixabay Content License.&lt;\/p&gt;\n\" data-large-file=\"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg\" src=\"https:\/\/i0.wp.com\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg?resize=1024%2C683&#038;ssl=1\" alt=\"Colour photograph of old hardback books on a shelf.\" class=\"wp-image-144089\" srcset=\"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg 1024w, https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-300x200.jpg 300w, https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-768x512.jpg 768w, https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1536x1024.jpg 1536w, https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-2048x1365.jpg 2048w, https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-624x416.jpg 624w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"><figcaption class=\"wp-element-caption\"><em>\u201cOld Books\u201d by jarmoluk via Pixabay. Free for use under the <a href=\"https:\/\/pixabay.com\/service\/license-summary\/\">Pixabay Content License<\/a>.<\/em><\/figcaption><\/figure>\n<\/div>\n<p class=\"wp-block-paragraph\">For many people who use archives, the accessibility of archival material is a notable difficulty. Archivists and librarians have long been turning to software to help with that accessibility. One of their main tools is Optical Character Recognition (OCR). This type of software turns images of text into digital characters, which can then be accessed by anyone with an internet connection and can be easily processed and analyzed outside of an archive. But OCR started as a corporate solution to corporate problems, which meant it was not very good at dealing with the differences present in archival materials. Archivists have been trying to address this with bespoke technology, but the last few years have presented a new way to deal with the problem.<\/p>\n<p class=\"wp-block-paragraph\">OCR enabled by AI vision language models (VLM-OCR) has recently exploded as a research topic in multiple fields including computer science, digital humanities, library information science, linguistics, and beyond. This new movement in OCR began because AI developers hit a wall. They had trained their models on all available data on the web so they tried using AI-generated data instead, but the results were problematic, inconsistent, and risky. To solve this, they turned to physical documents to broaden their training base. However, they needed better OCR to make this possible. <a href=\"https:\/\/doi.org\/10.48550\/ARXIV.2502.18443\">The paper<\/a> by AllenAI that opened the field, \u201colmOCR: Unlocking Trillions of Tokens in PDFs with Vision Language Models,\u201d says as much in its opening lines: PDFs hold enormous volumes of \u201cnovel, high-quality\u201d training data, but their diversity of formats and layouts makes that content hard to extract faithfully.<\/p>\n<p><span id=\"more-144088\"><\/span><\/p>\n<p class=\"wp-block-paragraph\">AllenAI\u2019s statement highlights some of the motivations behind this technology. Many AI companies are now developing OCR models and releasing some of them openly in order to encourage adoption, build user communities, and establish their tools as part of the wider document-processing ecosystem. Open source models\u2014whose source code is publicly available, freely licensed to use, modify, and redistribute\u2014can reduce the cost of access to high-quality OCR. They enable galleries, libraries, archives, and museums (GLAM) institutions and researchers to run, evaluate, and adapt tools on their own infrastructure. However, this does not eliminate commercial competition: companies can still charge for hosted services, enterprise support, specialised fine-tuning, and licences for larger-scale commercial use. For example, Datalab\u2019s Surya makes its tool available under a semi-open license. The ability to fine-tune the tool\u2019s training is free for research, personal use, and smaller startups but they require broader commercial licensing for other users. While useful in some respects, Surya is not made for GLAM and humanities research and thus represents a limited use case of these semi-open corporate solutions.<\/p>\n<p class=\"wp-block-paragraph\">This kind of for-profit model system can also be seen in Transkribus, one of the most widely used OCR services for humanities research. This software is not fully open source, instead running on a credits-based system where the first 50 credits are free and then the cost increases. Customers can fine-tune or \u201ctrain\u201d their own models for improved transcription for their collections. Typically fine-tuning is most helpful on homogenous collections with a high volume of documents. However, the company retains the underlying model for themselves, including the clients\u2019 documents used for that development. This restricts the portability of the model outside of the Transkribus environment. Alternatively, Adobe PDF reader is a widely accessible OCR engine requiring little technical knowledge, but is available with a subscription fee, and performs poorly on historical documents. While these are but two examples, the expansive use of these models demonstrate the demand for OCR, but also the simultaneous need for this OCR to become more accessible and sustainable. It is in GLAM\u2019s interest to find ways to use OCR models without relying on for-profit services as open source models ensure ownership of data while minimizing overhead expenses for consistently underfunded institutions. This independence can be achieved through capacity- and knowledge-sharing, and collaborating on OCR processing to remove redundant work.<\/p>\n<p class=\"wp-block-paragraph\">The outputs of many OCR technologies were developed for training LLMs, but that are frequently of little use in the humanities. They can be used by people with specific data mining and data science purposes, like researchers doing Named Entity Recognition. But these are still rather niche, and the majority of archival researchers and historians who want to access these OCR outputs would instead benefit from the document processing of OCR resulting in searchable archives. This is another area where AI-driven OCR is helpful, as it has the capacity to easily transform OCR outputs into specific and bespoke formats for the needs of the researchers who are using these outputs.<\/p>\n<p class=\"wp-block-paragraph\">In my work as a GLAM researcher situated in libraries, I focus on making VLM-OCR accessible and useful for research and archival use. I\u2019m frequently asked \u201cWhat\u2019s the best VLM-OCR model right now?\u201d Before March 2026, it was usually pretty clear. The models were quite uniform, handling the same type of documents, providing the same output formats but with different levels of transcription accuracy based on the source document\u2019s language, scan quality, and text layout. Now, each model has their own distinct strengths. It is a welcome development that models are no longer competing on accuracy alone. For example, <a href=\"https:\/\/hunyuanocr.org\/\">Hunyuan OCR<\/a> provides coordinates for each word on the page, which in turn enables people to search for the word and see it highlighted right on the page. <a href=\"https:\/\/www.datalab.to\/blog\/chandra-2\">ChandraOCR-2<\/a> also does a great job of transcription, and it can detect images inside a document (photos, art, graphics, etc.) and keep them separate from the surrounding text, along with writing short descriptions of what\u2019s in each one. <a href=\"https:\/\/github.com\/datalab-to\/surya\">Surya OCR<\/a> is a very small model, meaning it can run on lower-quality hardware, and transcribes at a higher speed while occasionally sacrificing accuracy. Differentiating by strengths\u00a0 eases the burden on users, who no longer have to chase a 0.1% accuracy edge and can instead pick whichever model suits their purposes. Regular users seem to develop an intuition for this. It comes with experience, gained by testing different models across a range of document types and matching them to what the user needs from a transcription. In this sense, collaboration between these experienced researchers in the space is key to helping everyone access the models that best suit their needs.<\/p>\n<p class=\"wp-block-paragraph\">What emerges from this shift is an ecosystem of increasingly complementary AI-enabled OCR models, whose real value depends on the researchers who know how to use them. With these models, institutions do not need to bet everything on one company\u2019s roadmap or pricing model. They can now run the software on local hardware that ensures data ownership stays with the institution. And because the tools are open, GLAM professionals can pool their expertise built through hands-on testing, matching models to materials, and sharing their work across institutions. The models themselves are now good enough that further gains will be small and specialized. What needs improving is how we use them together. The people making these documents accessible need a shared and growing toolkit, built <em>with<\/em> and <em>for<\/em> each other to use, made easier by the support of LLMs. For chronically underfunded archives and libraries, that collaboration is worth as much as any accuracy gain. Better OCR is worth having. Building the capacity of archives and libraries to serve the people who rely on them is worth more.<\/p>\n<p class=\"wp-block-paragraph\"><strong><em>Chlo\u00eb Farr<\/em><\/strong><em>\u00a0is a researcher working at the intersection of artificial intelligence, archives, and digital humanities.<\/em> <em>Working out of the Open Science Lab at TIB \u2013 Leibniz Information Centre for Science and Technology, her<\/em> <em>research focuses on large-scale text recognition and analysis of historical documents, including newspapers, maps, and archival records. Learn more about Farr\u2019s work on<\/em> <a href=\"https:\/\/chloe-farr.github.io\/\"><em>GitHub<\/em><\/a><em>.<\/em><\/p>\n<p class=\"wp-block-paragraph\"><em><strong>Jessica Jack<\/strong>\u00a0is a PhD student in History at the University of Saskatchewan, developing applications for Large Language Models in historical research. They are doing so through studying settler land use in late 19th century and early 20th century Saskatchewan.<\/em><\/p>\n<p class=\"wp-block-paragraph\"><em><strong>Jacob Polay<\/strong>\u00a0is a PhD student in History at the University of Saskatchewan, studying the roles Large Language Models have in the historical method. His current research involves creating an information retrieval pipeline using artificial intelligence tools to unlock the early modern archive at scale.<\/em><\/p>\n<\/div>\n<div class=\"pvc_clear\"><\/div>\n<p id=\"pvc_stats_8041\" class=\"pvc_stats total_only  \" data-element-id=\"8041\" style=\"\"><i class=\"pvc-stats-icon large\" aria-hidden=\"true\"><svg aria-hidden=\"true\" focusable=\"false\" data-prefix=\"far\" data-icon=\"chart-bar\" role=\"img\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\" viewBox=\"0 0 512 512\" class=\"svg-inline--fa fa-chart-bar fa-w-16 fa-2x\"><path fill=\"currentColor\" d=\"M396.8 352h22.4c6.4 0 12.8-6.4 12.8-12.8V108.8c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v230.4c0 6.4 6.4 12.8 12.8 12.8zm-192 0h22.4c6.4 0 12.8-6.4 12.8-12.8V140.8c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v198.4c0 6.4 6.4 12.8 12.8 12.8zm96 0h22.4c6.4 0 12.8-6.4 12.8-12.8V204.8c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v134.4c0 6.4 6.4 12.8 12.8 12.8zM496 400H48V80c0-8.84-7.16-16-16-16H16C7.16 64 0 71.16 0 80v336c0 17.67 14.33 32 32 32h464c8.84 0 16-7.16 16-16v-16c0-8.84-7.16-16-16-16zm-387.2-48h22.4c6.4 0 12.8-6.4 12.8-12.8v-70.4c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v70.4c0 6.4 6.4 12.8 12.8 12.8z\" class=\"\"><\/path><\/svg><\/i> <img data-recalc-dims=\"1\" loading=\"lazy\" decoding=\"async\" width=\"16\" height=\"16\" alt=\"Loading\" src=\"https:\/\/i0.wp.com\/dev95.site\/wp-content\/plugins\/page-views-count\/ajax-loader-2x.gif?resize=16%2C16&#038;ssl=1\" border=0 \/><\/p>\n<div class=\"pvc_clear\"><\/div>\n<div id=\"dev95-1879269835\" class=\"dev95-after-content dev95-entity-placement\"><p style=\"text-align: center;\"><strong>Finally, a link that actually works for your region and device! \ud83c\udf0d Get instant access to the top exclusive offers tailored just for you right now. Don&#8217;t miss out, click here! \ud83d\udc47<\/strong><br data-sfc-root=\"ep\" data-sfc-pl=\"|||[]\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: Arial, sans-serif, &quot;Noto Color Emoji&quot;; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\" \/><a href=\"https:\/\/www.profitableratecpmnetwork.com\/uzeja8ahze?key=bc876be53d6ad0ff6370ab8ea030e479\"><strong>\ud83d\udd17 [Click here]<\/strong><\/a><\/p>\n<\/div>","protected":false},"excerpt":{"rendered":"<p>Chlo\u00eb Farr with Jessica Jack and Jacob Polay This post is part of a series on AI and Collaboration. \u201cOld Books\u201d by jarmoluk via Pixabay. Free for use under the Pixabay Content License. For many people who use archives, the<\/p>\n<div class=\"hosteria-entry-more\"><a href=\"https:\/\/dev95.site\/ar\/the-ai-enabled-boom-in-document-transcription\/\" class=\"no-underline font-light  group-hover:text-primary-800 dark:group-hover:text-primary-300 py-1\">Read more &gt;&gt;&gt;<\/a><\/div>\n<div class=\"pvc_clear\"><\/div>\n<p id=\"pvc_stats_8041\" class=\"pvc_stats total_only\" data-element-id=\"8041\" style=\"\"><i class=\"pvc-stats-icon large\" aria-hidden=\"true\"><svg aria-hidden=\"true\" focusable=\"false\" data-prefix=\"far\" data-icon=\"chart-bar\" role=\"img\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\" viewbox=\"0 0 512 512\" class=\"svg-inline--fa fa-chart-bar fa-w-16 fa-2x\"><path fill=\"currentColor\" d=\"M396.8 352h22.4c6.4 0 12.8-6.4 12.8-12.8V108.8c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v230.4c0 6.4 6.4 12.8 12.8 12.8zm-192 0h22.4c6.4 0 12.8-6.4 12.8-12.8V140.8c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v198.4c0 6.4 6.4 12.8 12.8 12.8zm96 0h22.4c6.4 0 12.8-6.4 12.8-12.8V204.8c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v134.4c0 6.4 6.4 12.8 12.8 12.8zM496 400H48V80c0-8.84-7.16-16-16-16H16C7.16 64 0 71.16 0 80v336c0 17.67 14.33 32 32 32h464c8.84 0 16-7.16 16-16v-16c0-8.84-7.16-16-16-16zm-387.2-48h22.4c6.4 0 12.8-6.4 12.8-12.8v-70.4c0-6.4-6.4-12.8-12.8-12.8h-22.4c-6.4 0-12.8 6.4-12.8 12.8v70.4c0 6.4 6.4 12.8 12.8 12.8z\" class=\"\"><\/path><\/svg><\/i> <img loading=\"lazy\" decoding=\"async\" width=\"16\" height=\"16\" alt=\"Loading\" src=\"https:\/\/dev95.site\/wp-content\/plugins\/page-views-count\/ajax-loader-2x.gif\" border=\"0\" \/><\/p>\n<div class=\"pvc_clear\"><\/div>","protected":false},"author":1,"featured_media":0,"comment_status":"open","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"fp_fajr_begins":"","fp_fajr_iqamah":"","fp_dhuhr_begins":"","fp_dhuhr_iqamah":"","fp_asr_begins":"","fp_asr_iqamah":"","fp_maghrib_begins":"","fp_maghrib_iqamah":"","fp_isha_begins":"","fp_isha_iqamah":"","fp_midnight":"","fp_midnight_name":"","fp_sunrise":"","fp_single_prayer_begins_title":"","fp_single_prayer_iqamah_title":"","fp_prayer_times_for_today":"","fp_hijra_date":"","fp_fajr_name":"","fp_dhuhr_name":"","fp_asr_name":"","fp_maghrib_name":"","fp_isha_name":"","fp_sunrise_name":"","fp_currentDate":"","fp_current_time":"","fp_current_title":"","fp_current_location":"","fp_masjid_name":"","fp_prayer_title":"","fp_next_prayer_iqamah_time":"","fp_next_prayer_iqamah_title":"","fp_next_prayer_begins_time":"","fp_next_prayer_begins_title":"","fp_next_prayer_title":"","_jetpack_newsletter_access":"","_jetpack_dont_email_post_to_subs":false,"_jetpack_newsletter_tier_id":0,"_jetpack_memberships_contains_paywalled_content":false,"_jetpack_feature_clip_id":0,"_jetpack_memberships_contains_paid_content":false,"footnotes":"","jetpack_post_was_ever_published":false},"categories":[37],"tags":[],"class_list":["post-8041","post","type-post","status-publish","format-standard","hentry","category-posts"],"a3_pvc":{"activated":true,"total_views":0,"today_views":0},"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v28.6 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>The AI-Enabled Boom in Document Transcription - Dev95<\/title>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/dev95.site\/ar\/the-ai-enabled-boom-in-document-transcription\/\" \/>\n<meta property=\"og:locale\" content=\"ar_AR\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"The AI-Enabled Boom in Document Transcription - Dev95\" \/>\n<meta property=\"og:description\" content=\"Chlo\u00eb Farr with Jessica Jack and Jacob Polay This post is part of a series on AI and Collaboration. \u201cOld Books\u201d by jarmoluk via Pixabay. Free for use under the Pixabay Content License. For many people who use archives, theRead more &gt;&gt;&gt;\" \/>\n<meta property=\"og:url\" content=\"https:\/\/dev95.site\/ar\/the-ai-enabled-boom-in-document-transcription\/\" \/>\n<meta property=\"og:site_name\" content=\"Dev95\" \/>\n<meta property=\"article:published_time\" content=\"2026-09-23T10:00:00+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg\" \/>\n<meta name=\"author\" content=\"dev95\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:label1\" content=\"\u0643\u064f\u062a\u0628 \u0628\u0648\u0627\u0633\u0637\u0629\" \/>\n\t<meta name=\"twitter:data1\" content=\"dev95\" \/>\n\t<meta name=\"twitter:label2\" content=\"\u0648\u0642\u062a \u0627\u0644\u0642\u0631\u0627\u0621\u0629 \u0627\u0644\u0645\u064f\u0642\u062f\u0651\u0631\" \/>\n\t<meta name=\"twitter:data2\" content=\"7 \u062f\u0642\u0627\u0626\u0642\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/\"},\"author\":{\"name\":\"dev95\",\"@id\":\"https:\\\/\\\/dev95.site\\\/#\\\/schema\\\/person\\\/b807805ffe2916206b04d0938bce0298\"},\"headline\":\"The AI-Enabled Boom in Document Transcription\",\"datePublished\":\"2026-09-23T10:00:00+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/\"},\"wordCount\":1402,\"commentCount\":0,\"publisher\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/#organization\"},\"image\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/activehistory.ca\\\/wp-content\\\/uploads\\\/2026\\\/09\\\/jarmoluk-old-books-436498-1024x683.jpg\",\"articleSection\":[\"Posts\"],\"inLanguage\":\"ar\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/\",\"url\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/\",\"name\":\"The AI-Enabled Boom in Document Transcription - Dev95\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/activehistory.ca\\\/wp-content\\\/uploads\\\/2026\\\/09\\\/jarmoluk-old-books-436498-1024x683.jpg\",\"datePublished\":\"2026-09-23T10:00:00+00:00\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#breadcrumb\"},\"inLanguage\":\"ar\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"ar\",\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#primaryimage\",\"url\":\"https:\\\/\\\/activehistory.ca\\\/wp-content\\\/uploads\\\/2026\\\/09\\\/jarmoluk-old-books-436498-1024x683.jpg\",\"contentUrl\":\"https:\\\/\\\/activehistory.ca\\\/wp-content\\\/uploads\\\/2026\\\/09\\\/jarmoluk-old-books-436498-1024x683.jpg\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/dev95.site\\\/the-ai-enabled-boom-in-document-transcription\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/dev95.site\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"The AI-Enabled Boom in Document Transcription\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/dev95.site\\\/#website\",\"url\":\"https:\\\/\\\/dev95.site\\\/\",\"name\":\"Dev95\",\"description\":\"Your comprehensive digital platform for knowledge, services, tools, and entertainment.\",\"publisher\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/dev95.site\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"ar\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/dev95.site\\\/#organization\",\"name\":\"Dev95\",\"url\":\"https:\\\/\\\/dev95.site\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"ar\",\"@id\":\"https:\\\/\\\/dev95.site\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/i0.wp.com\\\/dev95.site\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/rbrrbr-6.png?fit=512%2C512&ssl=1\",\"contentUrl\":\"https:\\\/\\\/i0.wp.com\\\/dev95.site\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/rbrrbr-6.png?fit=512%2C512&ssl=1\",\"width\":512,\"height\":512,\"caption\":\"Dev95\"},\"image\":{\"@id\":\"https:\\\/\\\/dev95.site\\\/#\\\/schema\\\/logo\\\/image\\\/\"}},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/dev95.site\\\/#\\\/schema\\\/person\\\/b807805ffe2916206b04d0938bce0298\",\"name\":\"dev95\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"ar\",\"@id\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/a70a73d950838b20cd80d7ebdc955737e802e8cd896044c5473b32b946c0662a?s=96&d=mm&r=g\",\"url\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/a70a73d950838b20cd80d7ebdc955737e802e8cd896044c5473b32b946c0662a?s=96&d=mm&r=g\",\"contentUrl\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/a70a73d950838b20cd80d7ebdc955737e802e8cd896044c5473b32b946c0662a?s=96&d=mm&r=g\",\"caption\":\"dev95\"},\"url\":\"https:\\\/\\\/dev95.site\\\/ar\\\/author\\\/mohammad\\\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"The AI-Enabled Boom in Document Transcription - Dev95","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/dev95.site\/ar\/the-ai-enabled-boom-in-document-transcription\/","og_locale":"ar_AR","og_type":"article","og_title":"The AI-Enabled Boom in Document Transcription - Dev95","og_description":"Chlo\u00eb Farr with Jessica Jack and Jacob Polay This post is part of a series on AI and Collaboration. \u201cOld Books\u201d by jarmoluk via Pixabay. Free for use under the Pixabay Content License. For many people who use archives, theRead more &gt;&gt;&gt;","og_url":"https:\/\/dev95.site\/ar\/the-ai-enabled-boom-in-document-transcription\/","og_site_name":"Dev95","article_published_time":"2026-09-23T10:00:00+00:00","og_image":[{"url":"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg","type":"","width":"","height":""}],"author":"dev95","twitter_card":"summary_large_image","twitter_misc":{"\u0643\u064f\u062a\u0628 \u0628\u0648\u0627\u0633\u0637\u0629":"dev95","\u0648\u0642\u062a \u0627\u0644\u0642\u0631\u0627\u0621\u0629 \u0627\u0644\u0645\u064f\u0642\u062f\u0651\u0631":"7 \u062f\u0642\u0627\u0626\u0642"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#article","isPartOf":{"@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/"},"author":{"name":"dev95","@id":"https:\/\/dev95.site\/#\/schema\/person\/b807805ffe2916206b04d0938bce0298"},"headline":"The AI-Enabled Boom in Document Transcription","datePublished":"2026-09-23T10:00:00+00:00","mainEntityOfPage":{"@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/"},"wordCount":1402,"commentCount":0,"publisher":{"@id":"https:\/\/dev95.site\/#organization"},"image":{"@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#primaryimage"},"thumbnailUrl":"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg","articleSection":["Posts"],"inLanguage":"ar","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/","url":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/","name":"The AI-Enabled Boom in Document Transcription - Dev95","isPartOf":{"@id":"https:\/\/dev95.site\/#website"},"primaryImageOfPage":{"@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#primaryimage"},"image":{"@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#primaryimage"},"thumbnailUrl":"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg","datePublished":"2026-09-23T10:00:00+00:00","breadcrumb":{"@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#breadcrumb"},"inLanguage":"ar","potentialAction":[{"@type":"ReadAction","target":["https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/"]}]},{"@type":"ImageObject","inLanguage":"ar","@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#primaryimage","url":"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg","contentUrl":"https:\/\/activehistory.ca\/wp-content\/uploads\/2026\/09\/jarmoluk-old-books-436498-1024x683.jpg"},{"@type":"BreadcrumbList","@id":"https:\/\/dev95.site\/the-ai-enabled-boom-in-document-transcription\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/dev95.site\/"},{"@type":"ListItem","position":2,"name":"The AI-Enabled Boom in Document Transcription"}]},{"@type":"WebSite","@id":"https:\/\/dev95.site\/#website","url":"https:\/\/dev95.site\/","name":"Dev95","description":"\u0645\u0646\u0635\u062a\u0643 \u0627\u0644\u0631\u0642\u0645\u064a\u0629 \u0627\u0644\u0634\u0627\u0645\u0644\u0629 \u0644\u0644\u0645\u0639\u0631\u0641\u0629 \u0648\u0627\u0644\u062e\u062f\u0645\u0627\u062a \u0648\u0627\u0644\u0623\u062f\u0648\u0627\u062a \u0648\u0627\u0644\u062a\u0631\u0641\u064a\u0647.","publisher":{"@id":"https:\/\/dev95.site\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/dev95.site\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"ar"},{"@type":"Organization","@id":"https:\/\/dev95.site\/#organization","name":"Dev95","url":"https:\/\/dev95.site\/","logo":{"@type":"ImageObject","inLanguage":"ar","@id":"https:\/\/dev95.site\/#\/schema\/logo\/image\/","url":"https:\/\/i0.wp.com\/dev95.site\/wp-content\/uploads\/2026\/07\/rbrrbr-6.png?fit=512%2C512&ssl=1","contentUrl":"https:\/\/i0.wp.com\/dev95.site\/wp-content\/uploads\/2026\/07\/rbrrbr-6.png?fit=512%2C512&ssl=1","width":512,"height":512,"caption":"Dev95"},"image":{"@id":"https:\/\/dev95.site\/#\/schema\/logo\/image\/"}},{"@type":"Person","@id":"https:\/\/dev95.site\/#\/schema\/person\/b807805ffe2916206b04d0938bce0298","name":"dev95","image":{"@type":"ImageObject","inLanguage":"ar","@id":"https:\/\/secure.gravatar.com\/avatar\/a70a73d950838b20cd80d7ebdc955737e802e8cd896044c5473b32b946c0662a?s=96&d=mm&r=g","url":"https:\/\/secure.gravatar.com\/avatar\/a70a73d950838b20cd80d7ebdc955737e802e8cd896044c5473b32b946c0662a?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/a70a73d950838b20cd80d7ebdc955737e802e8cd896044c5473b32b946c0662a?s=96&d=mm&r=g","caption":"dev95"},"url":"https:\/\/dev95.site\/ar\/author\/mohammad\/"}]}},"jetpack_sharing_enabled":true,"jetpack_featured_media_url":"","_links":{"self":[{"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/posts\/8041","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/comments?post=8041"}],"version-history":[{"count":0,"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/posts\/8041\/revisions"}],"wp:attachment":[{"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/media?parent=8041"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/categories?post=8041"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/dev95.site\/ar\/wp-json\/wp\/v2\/tags?post=8041"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}