SharePoint

Syntex: SharePoint’s New AI Feature Unfolded

By Andrew Chomik · 2021-01-14
Syntex: SharePoint’s New AI Feature Unfolded

SharePoint Syntex (formally part of Project Cortex) was released publicly at Microsoft Ignite 2020. Syntex is Microsoft’s first serious foray at infusing artificial intelligence functionality into SharePoint, the document management system many of us has come to use and (sort of) love in the corporate world over the last 15 years. It’s a critical tool for both managing documents and collaborating on content among teams. So with all the changes Microsoft has introduced into SharePoint in the last few years, the AI component was discernibly inevitable. And it’s a cool little tool – if you have the right use cases for it.


This blog post is dedicated to giving you the down-and-dirty intro to Syntex, including what it is, applicable scenarios, and why even bother with it. There’s already a lot written about how to configure and set it up (not the least of which is provided by Microsoft), so let’s focus on what matters from a business (and analyst) perspective.


How will it improve my current SharePoint experience?

In testing Syntex, I found that there were many good things that created genuine improvements on the SharePoint experience. After the initial configurations and learning the document processing model creation cycle, I can genuinely see how this might benefit organizations in the future – especially if said organizations are open to business change.


Automation of Repetitive Tasks

Rather than input document library metadata manually, Syntex basically allows users to upload documents and walk away. The model will fill in the library columns on its own. This saves time and effort.


A Little Bit of Auto-Pilot Mode for Information Architecture

The model design works with SharePoint to basically create new content types and site columns. All are available within the Content Type Gallery in the Admin Center. You can even let Syntex work with your Term Store to auto-tag documents based on your taxonomy. You can even auto-apply retention labels to the different content types as they come in.


Disseminate Information Faster

By tying Syntex models to document libraries, content and metadata is automatically indexed across your tenant to be searchable and usable. This means teams and business units can more easily find what they need.


Syntex Shortcomings

In my configuration and trial adventures of Syntex, it didn’t go all peaches ‘n cream. As with my usual experience with Microsoft 365, the vendor tends to over-market their products as simple, intuitive and far-reaching in terms of improving the digital workplace experience. There were a few things that left me wanting in terms of a better user experience, as well as a better consumer of this new product.


The accuracy of the documents you model is highly nuanced

As discussed, you could provide documents of many types and build highly trained, complex processing models with very accurate classifiers and extractors. But accuracy results will differ based on a number of factors. This is even more true on a customer-to-customer basis. Consumers need to be aware that what works for a colleague or a partner organization may not work as effectively for another. It’s all about sandboxing, sandboxing, and more sandboxing.


Syntex has a user learning curve and requires ongoing commitment to maintenance

Microsoft brands Syntex as low-code, easy to learn, and quick to jump in. I beg to differ, even as an analyst with a technical background. It will take practice and a deep understanding of both the model learning process, the user interface, and information architecture in general. For modeling to be successful, digital workplace teams will need to dedicate resources internally to learn, build, and curate processing models so that users can generate value out of Syntex and its capabilities.


Read the fine print for the licensing model of Syntex

The Syntex advertisements and marketing material has warts – or at least the way it’s contextualized. Signing up for a Syntex license only provides the Document Understanding processing model, not the Forms Processing Model – although you may be lead to believe it’s all one and the same. These are two different technologies with different investment costs. Read the fine print, do your research, and make sure you know what you’re getting with a Syntex license (and what you’re not).


Good Use Cases and Not-So-Good Use Cases

As with any piece of software or functionality, there are limitations with Syntex – and when I say “Syntex”, I mean the Document Understanding Processing model, not the Form Processing Model (for reasons explained above). Most people would agree that automation of documents and their classification is a good thing. But there are certain documents that make sense with Syntex, and some that don’t.


Document types that Syntex could understand quite well

  1. Contracts and letters, such as operating agreements, NDAs, statements of work
  2. Legal documents
  3. Incorporation documents
  4. Newspaper articles
  5. Digital images
  6. Accounting documents *
  7. Shipping documents *

Some things that might not place nicely with Syntex

  1. Handwritten documents / scribble notes
  2. Resumes and CVs
  3. Meeting minutes (without a template)
  4. Accounting documents *
  5. Shipping documents *

* Accounting and shipping documents are marked on both lists because these “semi-structured” documents could work quite well – or not, depending on layout consistency. It’s worth sandboxing with these documents as you build your AI models to see what results work for you.


Why am I not talking about the Form Processing Model?

As mentioned above, the Form Processing Model is great for AI usage to look at structured documents with layouts, but the difference between wanting it and enabling it is about $500 USD/month. This is no small cost to businesses operating on tight budgets. Additionally, the Form Processing Model is less of a structural integration into SharePoint than the Document Understanding Model – it’s handled by the PowerApps AI Builder instead, and doesn’t manage content types, content type columns, or the Power Automate part. You’d probably want to know exactly what you want to do with AI technology before committing to that kind of spend, especially if the Document Understanding Processing Model in Syntex can already handle most of your basic document automation needs.


Making the Most of Syntex

So what really works and what doesn’t? I have sandboxed with SharePoint Syntex considerably since its release. Here’s what I’ve found as the rules of thumb:

  1. Location-predictability is absolutely key. Ensure the right word markers, labels, and expected data are in predictable places. Regional consistency of your data in Syntex is key to its success.
  2. Become an Explanations superstar. Helping the model understand what you want to scan is really on you. By familiarizing yourself with the Explanations features (Phrases, Patterns, and Proximities), you can make the Classifier and Extractor steps much less daunting and dramatically improve your accuracy scores.
  3. Quantity matters. For best results, mix one part document type with at least five parts document samples – twenty for extra accuracy. Document Understanding thrives on learning patterns and consistencies.
  4. Pay attention to font type and font size. The type and size of font on your documents can impact model accuracy. Use clear, readable fonts that aren’t obscured or overlapped with other page elements, and if modeling handwriting, require print rather than cursive.
  5. Have a change management strategy when adopting Syntex. This requires a change in behavior and a change in “trust” toward automation technology. Communicate, demonstrate, and build in a way to capture feedback so processing models can be improved over time.
  6. Have a model champion on hand. Having someone dedicated as the Syntex Model Admin is good governance and good business – your documents and needs will change over time, and models will need ongoing maintenance to stay effective.

Much like a car, you’ll need to maintain your models to get the best mileage – and keep driving in the right direction. Templates that have been adjusted over time (e.g. new layouts) may cause problems, and you may need to build a new model to handle new document layouts.

← Back to all posts