Learning On The Edge

October 17, 2022

Microcontrollers, miniature computers that can run simple commands, are the basis for billions of connected devices, from internet-of-things (IoT) devices to sensors in automobiles. But cheap, low-power microcontrollers have extremely limited memory and no operating system, making it challenging to train artificial intelligence models on ‘edge devices’ that work independently from central computing resources.

Training a machine-learning model on an intelligent edge device allows it to adapt to new data and make better predictions. For instance, training a model on a smart keyboard could enable the keyboard to continually learn from the user’s writing. However, the training process requires so much memory that it is typically done using powerful computers at a data center, before the model is deployed on a device. This is more costly and raises privacy issues since user data must be sent to a central server.

To address this problem, researchers at MIT and the MIT-IBM Watson AI Lab developed a new technique that enables on-device training using less than a quarter of a megabyte of memory. Other training solutions designed for connected devices can use more than 500 megabytes of memory, greatly exceeding the 256-kilobyte capacity of most microcontrollers (there are 1,024 kilobytes in one megabyte).

The intelligent algorithms and framework the researchers developed reduce the amount of computation required to train a model, which makes the process faster and more memory efficient. Their technique can be used to train a machine-learning model on a microcontroller in a matter of minutes.

This technique also preserves privacy by keeping data on the device, which could be especially beneficial when data are sensitive, such as in medical applications. It also could enable customization of a model based on the needs of users. Moreover, the framework preserves or improves the accuracy of the model when compared to other training approaches.

“Our study enables IoT devices to not only perform inference but also continuously update the AI models to newly collected data, paving the way for lifelong on-device learning. The low resource utilization makes deep learning more accessible and can have a broader reach, especially for low-power edge devices,” says Song Han, an associate professor in the Department of Electrical Engineering and Computer Science (EECS), a member of the MIT-IBM Watson AI Lab, and senior author of a paper describing the innovation.

Lightweight Training

A common type of machine-learning model is known as a neural network. Loosely based on the human brain, these models contain layers of interconnected nodes, or neurons, that process data to complete a task, such as recognizing people in photos. The model must be trained first, which involves showing it millions of examples so it can learn the task. As it learns, the model increases or decreases the strength of the connections between neurons, which are known as weights.

The model may undergo hundreds of updates as it learns, and the intermediate activations must be stored during each round. In a neural network, activation is the middle layer’s intermediate results. Because there may be millions of weights and activations, training a model requires much more memory than running a pre-trained model, Han explains.

Han and his collaborators employed two algorithmic solutions to make the training process more efficient and less memory-intensive. The first, known as sparse update, uses an algorithm that identifies the most important weights to update at each round of training. The algorithm starts freezing the weights one at a time until it sees the accuracy dip to a set threshold, then it stops. The remaining weights are updated, while the activations corresponding to the frozen weights don’t need to be stored in memory.

“Updating the whole model is very expensive because there are a lot of activations, so people tend to update only the last layer, but as you can imagine, this hurts the accuracy. For our method, we selectively update those important weights and make sure the accuracy is fully preserved,” Han says.

Their second solution involves quantized training and simplifying the weights, which are typically 32 bits. An algorithm rounds the weights so they are only eight bits, through a process known as quantization, which cuts the amount of memory for both training and inference. Inference is the process of applying a model to a dataset and generating a prediction. Then the algorithm applies a technique called quantization-aware scaling (QAS), which acts like a multiplier to adjust the ratio between weight and gradient, to avoid any drop in accuracy that may come from quantized training.

The researchers developed a system, called a tiny training engine, that can run these algorithmic innovations on a simple microcontroller that lacks an operating system. This system changes the order of steps in the training process so more work is completed in the compilation stage, before the model is deployed on the edge device.

“We push a lot of the computation, such as auto-differentiation and graph optimization, to compile time. We also aggressively prune the redundant operators to support sparse updates. Once at runtime, we have much less workload to do on the device,” Han explains.

A Successful Speedup

Their optimization only required 157 kilobytes of memory to train a machine-learning model on a microcontroller, whereas other techniques designed for lightweight training would still need between 300 and 600 megabytes.

They tested their framework by training a computer vision model to detect people in images. After only 10 minutes of training, it learned to complete the task successfully. Their method was able to train a model more than 20 times faster than other approaches.

Now that they have demonstrated the success of these techniques for computer vision models, the researchers want to apply them to language models and different types of data, such as time-series data. At the same time, they want to use what they’ve learned to shrink the size of larger models without sacrificing accuracy, which could help reduce the carbon footprint of training large-scale machine-learning models.

“AI model adaptation/training on a device, especially on embedded controllers, is an open challenge. This research from MIT has not only successfully demonstrated the capabilities, but also opened up new possibilities for privacy-preserving device personalization in real-time,” says Nilesh Jain, a principal engineer at Intel who was not involved with this work. “Innovations in the publication have broader applicability and will ignite new systems-algorithm co-design research.”

“On-device learning is the next major advance we are working toward for the connected intelligent edge. Professor Song Han’s group has shown great progress in demonstrating the effectiveness of edge devices for training,” adds Jilei Hou, vice president and head of AI research at Qualcomm. “Qualcomm has awarded his team an Innovation Fellowship for further innovation and advancement in this area.”

The work is funded by the National Science Foundation, the MIT-IBM Watson AI Lab, the MIT AI Hardware Program, Amazon, Intel, Qualcomm, Ford Motor Company, and Google.

For more information: web.mit.edu

HOME PAGE LINK

World First Combines Digital Holography with Coordinate Measuring Machines
July 31, 2026
Coordinate measurement technology is evolving from purely tactile devices to multisensor systems. For the first time worldwide, researchers at Fraunhofer IPM have fully integrated a high-precision digital holographic sensor into a coordinate measuring machine.
MAVOBASE LUX Delivers Standards-Compliant Illuminance Measurement
July 31, 2026
MAVOBASE LUX is an illuminance meter compliant with DIN 5032-7 Class C, designed for standard-compliant measurements in workplaces, production environments, and during maintenance and inspection tasks.
Verisurf Marks First Year with Sandvik by Accelerating Digital Manufacturing Innovation
July 30, 2026
Verisurf Software, Inc., a leading provider of Model-Based Enterprise (MBE) quality control, celebrates its first anniversary as a member of Sandvik. Since joining Sandvik in June 2025, Verisurf has expanded its innovation roadmap, strengthened its global reach, and accelerated the delivery of advanced digital manufacturing and quality assurance technologies.
Machine Safety Enters a New Era Under EU Machinery Regulation 2023/1230
July 30, 2026
The implementation of the new Machinery Regulation (EU) 2023/1230, effective from 20 January 2027, marks one of the most significant changes to European machinery safety legislation in recent years. While manufacturers remain responsible for compliance when machinery is first placed on the market, the regulation introduces expanded responsibilities for operators throughout the machine's operational lifecycle.
BMW’s Smart Battery Factory Combines AI, Digital Twins and Inline Inspection
July 30, 2026
With AI-supported production processes, digital twins and virtual reality applications, the new BMW plant is setting industry standards for the assembly of high-voltage batteries.
Smarter, Faster, Leaner: How Multi-Sensor Inspection Platforms Transform Production Lines
July 29, 2026
Part inspection is essential to manufacturing quality, but it also consumes production time, floor space, engineering attention, and capital. As manufacturers push for faster throughput and tighter tolerances, inspection systems must do more than find defects; they must reduce touch time, preserve valuable shop-floor space, and deliver reliable data without slowing production.
NIAR and Hexagon Expand Partnership to Advance Metrology Training and Regional Inspection Services
July 29, 2026
The National Institute for Aviation Research (NIAR) at Wichita State University and Hexagon Manufacturing Intelligence have expanded their long‑standing technology collaboration to increase regional access to advanced metrology training and contract inspection services.
July 2026 Metrology News Magazine
July 29, 2026
As digital transformation accelerates across manufacturing, the conversation is shifting beyond simply capturing measurement data to ensuring that information flows seamlessly across the entire product lifecycle. Metrology is becoming an increasingly connected discipline
Embraer and Hexagon Advance Predictive Quality for Next-Generation Aircraft Manufacturing
July 28, 2026
At the Farnborough International Airshow 2026 (FIA2026), Embraer and Hexagon Manufacturing Intelligence announced an extension of their long-term collaboration to advance predictive quality solutions for aircraft production and maintenance. The partnership brings together aerospace manufacturing expertise, precision metrology, automation, and artificial intelligence to transform how aircraft are inspected, assembled, and maintained.
RoboCT Brings Large-Scale Computed Tomography Out of the Lab
July 28, 2026
Comprehensive computed tomography (CT) inspection of large industrial components has traditionally been limited to specialized laboratory environments equipped with high-precision mechanical systems and significant infrastructure. That limitation is changing with RoboCT.

Smarter, Faster, Leaner: How Multi-Sensor Inspection Platforms Transform Production Lines
July 29, 2026
Part inspection is essential to manufacturing quality, but it also consumes production time, floor space, engineering attention, and capital. As manufacturers push for faster throughput and tighter tolerances, inspection systems must do more than find defects; they must reduce touch time, preserve valuable shop-floor space, and deliver reliable data without slowing production.
Closing the Digital Quality Loop – Integrating Metrology into Smart Manufacturing
July 20, 2026
As manufacturers accelerate their Industry 4.0 journeys, investment has largely focused on automation, machine connectivity, artificial intelligence and data analytics. Yet one critical element often remains disconnected from the digital manufacturing ecosystem, that of quality measurement.
Building Metrology 4.0 with Open Data Standards
July 14, 2026
As manufacturing continues its transition toward smart factories, digital twins, and autonomous production, metrology has evolved from a standalone quality function into an integral component of manufacturing intelligence. Yet despite remarkable advances in measurement technologies, one persistent challenge remains: interoperability.
From Model to Measurement: How Metrology Software Powers the MBD-Driven Digital Thread
July 1, 2026
As manufacturing accelerates toward Model-Based Enterprise (MBE), metrology software is evolving from a standalone inspection tool into a critical enabler of the digital thread. By connecting engineering intent, manufacturing execution, and quality validation, software platforms are helping organizations move beyond disconnected workflows toward fully traceable, model-driven manufacturing.
Why Tolerances Can Make or Break Production Cost
July 1, 2026
Tolerances are one of the most powerful and misunderstood parts of product design. A tolerance tells the manufacturer how much variation is acceptable. That simple instruction affects machining time, inspection effort, scrap risk, supplier choice, assembly quality, and production cost.