Category Archives: Market Intelligence

GEN-AI IS NOT EMERGENT … AND CLAIMS THAT IT WILL “EVOLVE” TO SOLVE YOUR PROBLEMS ARE ALL FALSE!

A recent article in the CACM (Communications of the ACM) referenced a paper by Dan Carter last year that demonstrated that the claims of Wei et.al in their 2022 “Emergent Abilities of Large Language Models” were unsubstantiated and merely wrong interpretations of visual artifacts produced by computing graphs using an inappropriate semi-log scale.

Now, I realize the vast majority of you without advanced degrees in mathematics and theoretical computer science won’t understand the majority of technical details, but that’s okay because the doctor, who has advanced degrees in both, does, can verify the mathematical accuracy of Dan’s paper, and the conclusion:

LLMs — Large Language Models — the “backbone” of Gen-AI DO NOT have any emergent properties. As a result, they are no better than traditional deep learning neural networks, and are, at the present time, ACTUALLY WORSE since our lack of deep research and understanding means that we don’t have the same level of understanding of these models, and, thus, the ability to properly “train” them for repeatable behaviour or the ability to accurately “measure” the outputs with confidence.

And while our understanding of this new technology, like any new technology, will likely improve over time, the realities are thus:

  • no amount of computing power has ever hastened the development of AI technology since research began in the late 60s / early 70s (depending on what you accept as the first paper / first program), it’s always taken improvements in algorithms and the underlying science to make slow, steady progress (with most technologies taking one to two DECADES to mature to the point they are ready for wide-spread industrial use)
  • the technology currently takes 10 times the computing power (or more) to compute “results” that can be readily computed by existing, more narrow, techniques (often with more confidence in the results)
  • the technology is NOT well suited to the majority of problems that the majority of enterprise software companies (blindly jumping on the bandwagon with no steering wheel and no brakes for fear of missing out on the hype cycle that could cause a tech market crash unequally by any except the dot-com bust of the early 2000s) are trying to use it for (and yes, the doctor did use the word “majority” and not “all” because, while he despises it, it does have valid uses … in creative (writing, audio, and video) applications [not business or science applications] where it has almost unequalled potential compared to traditional ML designed for math and science based applications)

And the market realities that no one wants to tell you about are thus:

  • former AI evangelists and some of the original INVENTORS of AI are turning against the technology (out of a realization that it will never do what they hoped it would, that its energy requirements could destroy the planet if we keep trying, and/or that maybe there are some things we should just not be meddling with at our current stage of societal and technological evolution), including Weizenbaum and Hinton
  • Brands are now turning against AI … and even the Rolling Stone is writing about it
  • big tech and companies that depend on big tech (like Pharma) are starting to turn against AI … and CIOs are starting to drop Open AI and Microsoft CoPilot because, even when the cost is as low as $30 a user, the value isn’t there (see this recent article in Business Insider)

Now, the doctor knows there are still hundreds of marketers and sales people in our space who will consistently claim that the doctor is just a naysayer and against progress and innovation and AI and modern tech and blah blah blah because they, like their companies, have gone all in on the hype cycle and don’t want their bubble burst, but the reality is that

the doctor is NOT against “AI” or modern tech. the doctor, whose complete archives are available on Sourcing Innovation back to June 2006 when he started writing about Procurement Tech, has been a major proponent of optimization, analytics, machine learning, and “AI” since the beginning — his PhD is in advanced theoretical computer science, which followed a math degree — and, after actually studying machine learning, expert systems, and AI, he used to build optimization, analytics, and “AI” systems (including the first commercial semantic social search application on the internet)

what the doctor IS against is Gen-AI and all the false claims being made by the providers about its applicability in the enterprise back office (where it has very limited uses)

because the vast majority of the population does not have the math and computer science background to understand

  1. what is real and what is not
  2. what technologies (algorithms) will work for a certain type of problem and will not
  3. whether the provider’s implementation will work for their problem (variation)
  4. whether they have enough data to make it work

and, furthermore, this includes the vast majority of the consultants at the Big X and mid-sized consultancies who graduate from Business Schools with very basic statistics and data analytics training and a crash course in “prompt engineering” who can barely use the tech, couldn’t build the tech, and definitely couldn’t evaluate the efficacy and accuracy of the underlying algorithms.

The reality is that it takes years and years of study to truly understand this tech, and years more of day-in and day-out research to make true advancement.

For those of you who keep saying “but look at how well it works” and produce 20 examples to prove it, the reality is that it’s only random chance that it works.

With just a bit of simplification, we can describe these LLMs as essentially just super sophisticated deep neural networks with layers and layers of nodes that are linked together in new and novel configurations, with more feedback learning, and structured in a manner that gives them an ability to “produce” responses as a collection of “sub-responses” from elements in its data archive vs just returning a fixed response. As a result they can GENerate a reply vs just selecting from a fixed one. (And that’s why their natural language abilities seem far superior to traditional neural network approaches, which need a huge archive of responses to have a natural sounding conversation, because they can use “context” to compute, with high probability, the right parts of speech to string together to create a response that will sound human.)

Moreover, since these models, which are more distributed in nature, can use an order of magnitude more (computational) cores, they can process an order of magnitude more data. Thus, if there is ten to one hundred times the amount of data (and it’s good data), of course they are going to work reasonably well for expected queries at least 95% of the time (whereas a last generation NN without significant training and tweaking might only be 90% out of the box). If you then incorporate dynamic feedback on user validation, that may even get to 99% for a class of problems, which means that it will appear to be working, and learning, 99 times out of 100 instead of 19 out of 20. But it’s NOT! It’s all probabilities. It’s all random. You’re essentially rolling the bones on every request, and doing it with less certainty on what a good, or bad, result should look like. And even if the dice come “loaded” so that they should always roll a come out roll, there are so many variables that there are never any guarantee you won’t get craps.

And for those of you saying “those odds sound good“, let me make it clear. They’re NOT.

  • those odds are only for typical, expected queries, for which the LLM has been repeatedly (and repeatedly) trained on
  • the odds for unexpected, atypical queries could be as low as 9 in 10 … which is very, very, bad when you consider how often these systems are supposed to be used

But the odds aren’t the problem. The problem is what happens when the LLM fails. Because you don’t know!

With traditional AI, you either got no response, an invalid response with low confidence, or a rare (compared to Gen-AI) invalid response with high confidence, where the responses were always from a fixed pool (if non-numeric) or fixed range (if numeric). You knew what the worst case scenario would be if something went wrong, how bad that would be, how likely that was to happen, and could even use this information to set bounds and tweak the confidence calculation on a result to minimize the chance of this ever happening in a real world scenario.

But with LLMs, you have no idea what it will return, how far off the mark the result will be, or how devastating it will be for your business when that (eventually) happens (which, as per Murphy’s law, will be after the vendor convinces you to have confidence in it and you stop watching it closely, and then, out of the blue, it decides you need 1,000 custom configurations of a high end MacBook Pro in inventory [because 10 new sales support professionals need to produce better graphics] in a potentially recoverable case or it decides to change your currency hedge on a new contract to that of a troubled economy (like Greece, Brazil, etc.) because of a one day run on the trading markets in a market heading for a hyperinflation and a crash [and then you will need a wheelbarrow full of money to buy a loaf of bread — and for those who think it can’t happen, STUDY YOUR HISTORY: Germany during WWII, Zimbabwe in 2007, and Venezuela in 2018, etc.]). You just don’t know! Because that’s what happens when you employ technology that randomly makes stuff up based on random inputs from you don’t know who or what (and the situation gets worse when developers [who likely don’t know the first thing about AI] decide the best way to train a new AI is to use the unreliable output of the old AI).

So, if you want to progress, like the monks, leave that Genizah Artificial Idiocy where it belongs — in the genizah (the repository for discarded, damaged, or defective books and papers), and go find real technology built on real optimization, analytics, machine learning, and AI that has been properly researched, developed, tested, and verified for industrial use.

Advice For Dealing With The PROCUREMENT STINK from Leading Consultants!

Last week, the doctor asked fellow niche/independent consultants as to how we can help to dispel the PROCUREMENT STINK which is permeating the space as a result of poor choices, bad information, and sometimes bad actors, which include the reasons we described in that article as well as many more.

Why? Because it’s going to take a collective effort among analysts, consultants, and vendors to dispel the stink permeating the Procurement space, and no one on his or her own will have all the solutions. As expected, some of the greats chimed in with their thoughts and ideas and these thoughts and ideas need to be given center stage, so this is what we’re going to do today!

James Meads

Clarity and transparency on your business model is key, especially if you have revenue streams from solution providers.

As Patrick Van Osta echoed in the comments, the uphill path to recovery, I feel, is for consultants to reclaim the position of sole trusted advisor, and there’s no way we’re ever going to be trusted advisors if we are not clear and transparent in our operations and goals. If we’re hiding our intentions, or upsides, how will the client know whether or not our goals actually align with theirs?

Joël Collin-Demers

I’m 100% on-board with the need for transparency and taking decisions based on what’s best for the client long term. Your job is to make yourself redundant as soon as possible!

In Procurement, there’s always another project. ALWAYS. You don’t have to milk one for life, with your help and guidance, you can open the client’s eyes as to not only how much there is to do, but how much they can do better, for a great ROI, with your help. Just like there’s well over 25,000 (or 35,000) species of fish in the sea, there are tens of thousands of unique aspects to Procurement in a modern enterprise. And just like you have to know where to fish, what hook to use, and what bait to use to catch a type of fish, you need to know the equivalents for each category, methodology, and process.

Jon W. Hansen

Practitioners stop looking at technology as the “silver bullet” solution but instead focus on doing the real and hard work while Solution providers stop selling shiny paper and “falling in love” with your own technology. … and us consultants have to help the practitioners do the work, understand what they need, and steer clear of the vendor with the shiny new tech (that doesn’t actually do anything [more than cheaper, proven tech]).

Paul Martyn

Consultancies (and their clients) need to Provide performance based compensation with uniqueness. For example, provide specialist consultants with compensation that includes equity. In short, align compensation to customer value (revenue growth and retention). Because, right now, most of the good consultants that can generate the ROI a client should expect are not incentivized to do well on point-based projects (like an Affordable RFP), but instead are incentivized to work on, and sell, long-term “solution” oriented consulting that lines the firm’s (and not the clients’) pocketbook (i.e. keep doing the fishing vs. teaching the client). As a result, most of the good consultants move out of the roles they are needed in to the roles they are incentivized to take.

Vinnie Mirchandani

The web lulled a number of procurement (and IT) folks into expecting vendor, negotiation etc intelligence for cheap, if not free. Vendors are not afraid to spend on sales and marketing. Procurement needs to adopt a similar mindset to even the game.

The best things in life may be free, but the best things in business are not. (As the Arrogant Worms pointed out over three decades ago, you get NOTHING FOR NOTHING!) And if you don’t have the right tools that enable the right processes powered by the right intelligence, you’re not going to win the game. Remember that all of the best sports teams use high-tech sports tech backed by science and data analytics to help their athletes reach peak condition. Raw talent only gets you in the game. You need the right training to win, or, at least, the right guidance and tech to enable you as you learn.

There’s a lot of STINK out there now, but if you follow this advice, you’ll go a long way to removing it. After all, you can’t solve everything with a pressure washer.

PLEASE TELL ME: Why buys research cobbled together by “researchers” who don’t have a clue as to what they’re researching?

This press release just went live yesterday:

Sourcing and Procurement Operation Software Industry Future Trends Analysis

which announced a new “Sourcing and Procurement Operation Software market” research report from Orbis that opened with obvious (that we are a pivotal sector), stated a few more obvious facts around software delivery methods (could-based, traditional ASP based), broad market sectors (business, manufacturing, education, government, etc.), and top players that include:

  • GEP SMART – Source to Pay
  • Jaggaer – Source to Pay
  • Corcentric – Source to PayMENTS
  • Coupa – business spend management, sorry, margin multiplier maker based on Source to Pay

which are in every map, quadrant, wave, logo map, etc. … so no surprise there but …

  • Precoro – Procure to Pay

which only solves half the problem

  • Servicenow – workflow management
  • Kissflow – low code app development

which can build solutions, but doesn’t offer them out of the bark

  • Vendr – SaaS marketplace

where you can buy some of them

  • ClickUp – Project Management

which is not even remotely related to S2P at all!!! And if these are the top 9 vendors, I shudder at what other totally irrelevant, non-comparable, vendors were included!

A report such as this should ONLY include vendors that offer real, and core, Source-to-Pay functionality, and only if they break down the space into segments where included vendors are actually comparable!

And it shouldn’t be hard as there are over 600 such vendors in some core area of S2P … you don’t have to include generic workflow engines, project management, buy-an-app platforms, or generic project management just to hit 25!

Reports like these give analyst firms, and analysts, a bad name!

Why Won’t They Stop?!?

Procurement Organizations Need Automation, But that DOES NOT Necessarily Mean AI!

A number of leaders in our space, including Sarah Scudder in the comments to this post, have been noting to me that they are seeing AI resonate with companies of all sizes.

Sarah notes that:

1. She’s seeing AI agent automations resonate with smaller companies.

Smaller companies need automation desperately, but it’s important we educate smaller companies that doesn’t mean they need AI. We’ve had adaptive rules-based automation and tailored machine learning in this space for almost 20 years and they can get fantastic results without having to risk being pre-alpha testers for unproven AI while getting the solution they really need for a fraction of the cost of this new, relatively unproven, AI tech! (Remember, firms that dumped millions into this bandwagon need to recoup those millions fast before their investors abandon them, which means high prices for unproven tech!)

2. She’s seeing copilot intelligence resonate with bigger companies who understand risk.

Which makes sense for a small segment of the market who are ready for it because augmented intelligence and automated suggestions with yes/no approvals are great for organizations who

  1. understand risk and
  2. understand the categories/markets/domains they are applying the technology in, because a true expert will identify the 95% of the time it’s working just fine; the 3% of the time it’s probably okay (and not worth the effort to double check manually due to the risk threshold); and the 2% of the time they need to slam the breaks and take over.

However, that’s not a very large segment of the market. What most companies still need is better analytics, category intelligence, and guidance from category experts on how to use it and then where and when to integrate automation and co-pilot capabilities.

Furthermore, I’m also being told that:

3. Mid-Markets are looking for technology they can roll out to the organization at large to get tail-spend under control, manage intake, and/or relieve pressure on Procurement to focus on more strategic efforts.

Which resonates, but, again, this is an area where AI is typically not needed. Catalogs, be they hosted, punch-out, hybrid, etc. with the ability to also request/book standard, pre-negotiated, services, easy search, and easy RFQ where there is no standard item but the buyer has budget authority, the vendors are preferred, and the amount doesn’t hit a threshold is often enough. Maybe a natural language search to find the right policy documents or bring up the right products or forms, but that doesn’t require modern AI either — we’ve had that for quite some time as well.

And, as Sarah implies, while organizations of all sizes need help to overcome their excessive workload and limited market insight so that they can prioritize risk management and mitigation in their procurement activities, this doesn’t mean they need AI. Automation yes, advanced technology a definite yes, but AI, rarely! Remember that when building and recommending ACTUAL solutions and not just buzzwords.