← Back to Kantar Worldpanel fine-tunes GenAI models on Databricks to generate market-insight training data faster

Source proof for Kantar Worldpanel fine-tunes GenAI models on Databricks to generate market-insight training data faster

Source-bound proof

Verified source excerpts for every supported field

Each colour maps a published value to the exact source passage used to support it. Only bounded excerpts are public; administrators can inspect the complete captured source.

13 fields supported

Title

Kantar Worldpanel fine-tunes GenAI models on Databricks to generate market-insight training data faster

derived · high
Giving clients more accurate market insights, faster | Databricks Skip to main content Login Why Databricks Discover For App Develop…

Description

Kantar Worldpanel used the Databricks Data + AI Platform and MLflow to experiment with Llama, Mistral, GPT-4 and GPT-3.5 for a proof of concept linking receipt descriptions to product barcode names. GPT-4 produced the most accurate outputs (94%), which the team used to automatically generate a training dataset of about 120,000 receipt-to-barcode pairs in a couple of hours to fine-tune a smaller production model.

derived · high
…ixeira highlights the effectiveness of their GenAI models, stating, “We've experimented with Llama, Mistral, GPT-4 and GPT-3.5, all within the Databricks Data + AI Platform. Ultimately, GPT-4 provided better answers, with an accuracy of 94%.” This data accuracy translates directly into better insights for Kantar’s clien…

Company

Kantar Worldpanel

quote · high
…arameter model fine-tuned on Databricks 94% Accuracy of training data generated Kantar Worldpanel is a leading market research company specializing in consumer data analysis that helps clients make informed decisions. Kantar Worldpanel faced challenges with some of their legacy systems, which wer…

Industry

Technology & Software

classification · medium
…trusted market research, visit: https://www.kantar.com/ Share this post Details Industry : Technology and Software Use Case : Artificial Intelligence Cloud : Azure Product : Agent Bricks Ready t…

Problem

Kantar Worldpanel's legacy systems were inflexible, resource-intensive to maintain, and required specialized, outdated programming skillsets, limiting data democratization and experimentation with new AI-driven use cases.

derived · high
…le and scalable as what we’re currently getting from Databricks,” said Portêlo. Maintaining and managing them can be resource-intensive for the engineers. Requiring an outdated programming skillset not only limits the accessibility of…

Solution

Kantar Worldpanel used the Databricks Data + AI Platform and MLflow to manage the ML lifecycle, experimenting with Llama, Mistral, GPT-4 and GPT-3.5 to fine-tune a model linking receipt descriptions to product barcode names, downloading models via Databricks Marketplace, exploring Databricks AI Search for description comparisons, and using Unity Catalog to govern data sharing across teams.

derived · high
…that supports Kantar Worldpanel’s advanced AI and machine learning initiatives. The data science team leverages MLflow, an open source platform developed by Databricks, to manage the full machine learning lifecycle. This component allows them to track experiments, reproduce runs and deploy mode…

Business value

Kantar Worldpanel automatically generated a training dataset of about 120,000 receipt-to-barcode description pairs at 94% accuracy in just a couple of hours, letting manual coding teams focus on discrepant results and freeing engineering resources for core development, while streamlining data scientists' workflows.

derived · high
…ata faster, without using a lot of human resources. And we’ve done that — we’ve automatically generated a training dataset of about 120,000 pairs of receipt descriptions and barcode names with an accuracy of 94% in just a couple of hours,” said Portêlo. “We can allow our manual coding teams to focus on more discrepa…

AI capabilities

Generative AI

classification · high
…ixeira highlights the effectiveness of their GenAI models, stating, “We've experimented with Llama, Mistral, GPT-4 and GPT-3.5, all within the Databricks Data + AI Platform. Ultimately, GPT-4 provided better answers, with an accuracy of 94%.” This data…

Technology

Databricks, MLflow, Unity Catalog, GPT-4, Llama, Mistral

classification · high
…ixeira highlights the effectiveness of their GenAI models, stating, “We've experimented with Llama, Mistral, GPT-4 and GPT-3.5, all within the Databricks Data + AI Platform. Ultimately, GPT-4 provided better answers, with an accuracy of 94%.” This data…

Deployment model

Cloud

classification · high
…t Details Industry : Technology and Software Use Case : Artificial Intelligence Cloud : Azure Product : Agent Bricks Ready to get started? Try Databricks for free Learn more…

Deployment options

cloud

classification · high
…t Details Industry : Technology and Software Use Case : Artificial Intelligence Cloud : Azure Product : Agent Bricks Ready to get started? Try Databricks for free Learn more…

Headline outcome

derived · high
…re accurate market insights, faster 8B Parameter model fine-tuned on Databricks 94% Accuracy of training data generated Kantar Worldpanel is a leading market research company specializing in consumer…

Use case type

Analytics augmentation

classification · medium
…hanced data accuracy, streamlined workflows and optimized resource utilization. Now Kantar Worldpanel can experiment with advanced AI/ML models within Databricks, generating training data faster and more efficiently. This has helped them deliver new use cases such as providing their clients with…
Capture details
Captured
11 Sept 2026, 06:03 UTC
Extractor
fetch-strip@1
Snapshot hash
3648498d357f1539d64952bbdc8063d4a22b44ef00221ab25174cd3dc0e5147d