ManufacturingLarge Language ModelsOn-PremiseNVIDIA DGX SuperPODNVIDIA BlackwellNVIDIA NIMNVIDIA TensorRT-LLMNVIDIA Mission ControlNVIDIA NeMoNVIDIA RivaNVIDIA DGX Spark

MediaTek Accelerates AI Development With an AI Factory

MediaTek

MediaTek established an on-premises AI factory powered by NVIDIA DGX SuperPOD with NVIDIA Blackwell-based systems to accelerate enterprise AI efforts, including development of its Breeze series LLMs and a 480-billion-parameter traditional-Chinese model. The AI factory processes approximately 60 billion tokens per month for inference and completes over 24,000 model-training iterations monthly, training models exceeding 480 billion parameters within one week (versus 7-billion-parameter models in a week previously). Using NVIDIA NIM and TensorRT-LLM, MediaTek achieved a 40% improvement in inference speed and 60% increase in token throughput. NVIDIA Mission Control consolidated GPU provisioning and system monitoring, while AI-assisted code completion and an AI agent for chip design documentation reduced documentation time from weeks to days. NVIDIA Riva was integrated into NVIDIA DGX Spark for agentic voice control features like internet search, calendar and messaging.

Overview

MediaTek established an on-premises AI factory powered by NVIDIA DGX SuperPOD with NVIDIA Blackwell-based systems to accelerate enterprise AI efforts, including development of its Breeze series LLMs and a 480-billion-parameter traditional-Chinese model. The AI factory processes approximately 60 billion tokens per month for inference and completes over 24,000 model-training iterations monthly, training models exceeding 480 billion parameters within one week (versus 7-billion-parameter models in a week previously). Using NVIDIA NIM and TensorRT-LLM, MediaTek achieved a 40% improvement in inference speed and 60% increase in token throughput. NVIDIA Mission Control consolidated GPU provisioning and system monitoring, while AI-assisted code completion and an AI agent for chip design documentation reduced documentation time from weeks to days. NVIDIA Riva was integrated into NVIDIA DGX Spark for agentic voice control features like internet search, calendar and messaging.

The challenge

MediaTek needed a robust, scalable and cost-effective computing environment to handle thousands of model-training iterations monthly and billions of tokens processed for inference, while exploring newer models on local machines without exposing proprietary data, ensuring data security and compliance.

The solution

MediaTek established an on-premises AI factory powered by NVIDIA DGX SuperPOD with NVIDIA Blackwell-based systems, incorporating the DGX SuperPOD reference architecture into its data center designs. It leverages NVIDIA NIM and TensorRT-LLM for inference, NVIDIA Mission Control for GPU provisioning and system monitoring, NVIDIA NeMo to fine-tune LLMs, and NVIDIA Riva for agentic voice control (ASR/TTS) integrated into NVIDIA DGX Spark. MediaTek also integrated AI-assisted code completion and an AI agent for chip design documentation that extracts information from design flowcharts and state diagrams.

Large Language ModelsGenerative AIAgentic AICode Generation

Reported business value

MediaTek's AI factory processes approximately 60 billion tokens per month for inference and completes over 24,000 model-training iterations monthly, training models exceeding 480 billion parameters within one week versus 7-billion-parameter models in a week previously. NVIDIA NIM and TensorRT-LLM delivered a 40% improvement in inference speed and 60% increase in token throughput. Documentation time was reduced from weeks to days through AI-assisted generation.

Sources

Open any source and check the claim yourself — that is the point of the register.

Related entries

Other manufacturing entries in the register.

All entries
ManufacturingMachine LearningPublic Cloud

ArcelorMittal Enhances Steel Production Through Digital Innovation

ArcelorMittal partnered with IBM Consulting and Infosys to modernize operations through AI, cloud technology and SAP S/4HANA migration. At ArcelorMittal Eisenhüttenstadt, machine learning was introduced to predict and prevent surface defects on automotive steel sheets. At the Hamburg wire rod plant, AI optimized the trimming process by analyzing historical production data to determine optimal cutting points, reducing trim scrap by 20% and contributing to energy savings and lower CO2 emissions. ArcelorMittal also deployed a bio-inspired Ant Colony Optimization algorithm to calculate optimal production schedules, reducing downtime and material waste. At AM/NS India, IBM Consulting used IBM Rapid Move for SAP S/4HANA to migrate data and applications from outdated platforms to a single SAP instance across locations in Dubai, Indonesia and India.

96/100HighArcelorMittalPrimary source
ManufacturingComputer VisionUnknown

Michelin runs 200+ AI use cases across manufacturing, supply chain and innovation

French tire manufacturer Michelin has more than 200 AI use cases in production, led by group chief data and AI officer Ambica Rajagopal. Its in-house IRIS system, protected by over 20 patents, partially automates end-of-line visual tire defect inspection to improve inspector efficiency and workplace ergonomics while operators retain final accountability. Machine learning forecasting tools improve demand forecast accuracy and proactively detect stock shortages in the supply chain. Michelin scans the startup ecosystem and uses tools including Databricks and Dataiku, and has partnerships with Microsoft and Rockwell Automation to codevelop AI solutions. The company reports AI-project ROI exceeding €50 million per year, growing 30-40% annually for three consecutive years, governed by an internal data office and responsible-AI principles (people-centric, explainable, accountable).

92/100HighMichelinPrimary source
ManufacturingGenerative AIPublic Cloud

Schneider Electric fast-tracks innovation with Azure OpenAI Service

Schneider Electric bases customer-facing AI solutions on Azure OpenAI Service within Microsoft Cloud for Manufacturing. Its EcoStruxure Microgrid Advisor uses Azure OpenAI Service and Azure IoT for dynamic control of facility energy performance, EcoStruxure Resource Advisor Copilot helps customers manage energy usage, and the company is developing a PLC code generation copilot to automate programmable logic controller programming for manufacturing robots and IoT devices.

96/100HighSchneider ElectricPrimary source
ManufacturingLarge Language ModelsUnknown

Foxconn Develops Physical AI-Enabled Smart Factories With Digital Twins

Foxconn (Hon Hai Technology Group) uses physically accurate digital twins integrating NVIDIA Omniverse libraries and OpenUSD to design, deploy, and manage high-volume production facilities, including those producing NVIDIA GB200 Grace Blackwell Superchip systems. Its Fii Omniverse Digital Twin (FODT) platform creates virtual replicas of factories, enabling simulation-driven design, real-time monitoring, and optimized operations. Using NVIDIA PhysicsNeMo AI models, Foxconn achieves 150x faster computational fluid dynamics simulations for thermal analysis (minutes vs. hours). Standardized OpenUSD-based digital twin assets enable rapid migration of entire production lines between global factories (e.g., Taiwan to Mexico). Robot workcells and AGV logistics are simulated in FODT before physical deployment: complex robotic tasks such as screw tightening and cable insertion are simulated and refined with NVIDIA Isaac Sim, Isaac Lab, FoundationPose models, and NVIDIA cuMotion, while AGV path is optimized by connecting simulations with Material Control Systems and NVIDIA cuOpt. Foxconn also built video analytics AI agents with NVIDIA Metropolis and the NVIDIA AI Blueprint for video search and summarization to monitor factory floors, and developed FoxBrain, an AI platform powered by NVIDIA NeMo trained in four weeks, described as Taiwan's first large language model with advanced reasoning capabilities. Leo Guo, General Manager of Fii Robotic Group, said the company believes it can cut factory setup and planning time by about 50%.

92/100HighFoxconn (Hon Hai Technology Group)Primary source

Was this helpful?

Your feedback helps us improve our use case database