Design and Embedded Validation of Compact ML Models for Affective Touch Classification in a Soft Interactive Companion
summary
The gist
This paper presents a "complete open-source MATLAB-based framework" for the development and validation of compact deep learning models designed to recognize affective touch in soft, sensorized
In short
Researchers developed compact machine learning models for soft, interactive companions to recognize affective touch. Using a 1D CNN with 13,200 parameters, the system runs locally on an ESP32 microcontroller. A hybrid strategy combines simple rules for energetic movements with neural networks for subtle gestures, improving accuracy and power efficiency.
Key concepts
- 1D CNN
- A one-dimensional convolutional neural network is a type of model perfect for reading signals that change over time. In this study, it is used to process touch data, allowing a small, compact model to identify various human gestures within a soft, interactive companion.
- Hybrid Strategy
- This approach combines simple, rule-based heuristic methods with advanced machine learning. The toy uses fast, basic rules to catch large, energetic movements like hits or pulls, while the smart CNN takes over to interpret subtle, social touches like scratching or gentle stroking.
- Edge AI
- This involves running artificial intelligence directly on a device's own internal hardware, such as an ESP32 microcontroller, rather than using a giant computer in the cloud. This enables real-time interaction, saves power, and ensures privacy because personal touch data never leaves the toy.
Terminology used across episodes
This episode discusses
- Design and Embedded Validation of Compact ML Models for Affective Touch Classification in a Soft Interactive Companion · Paper Radio
- Multi-Scale Context Aggregation by Dilated Convolutions
- CMSIS-NN: Efficient Neural Network Kernels for Arm Cortex-M CPUs
The paper
Design and Embedded Validation of Compact ML Models for Affective Touch Classification in a Soft Interactive Companion · Read on arXiv
Institute of Architecture and Design, Riga Technical University · Institute of Mechanical and Biomedical Engineering, Riga Technical University · Center for Pedagogy and Social Work, Riga Technical University Liepaja Academy · Institute of Digital Humanities, Riga Technical University · Masaryk University
Soft plush companions provide a safe and intuitive platform for affective human-robot interaction, but their deformable structure and distributed tactile signals make reliable gesture recognition difficult. This study presents a complete workflow for developing and validating compact affective-touch classifiers for an interactive plush companion. A newly collected dataset comprised 1,326 labelled recordings before curation, including interactions from 25 children, teenagers, and adults. Each classifier received 2.5-s windows containing ten capacitive channels and one accelerometer-magnitude channel. MATLAB supported acquisition, quality control, window generation, and a 468-run exploratory study of dilated one-dimensional convolutional neural networks (1D CNNs). A closely matched Python workflow then preserved participant provenance, fitted preprocessing inside each fold, and evaluated shortlisted models by 25-fold leave-one-subject-out cross-validation (LOSO-CV). On the operational 10-class task, the compact dilated CNN achieved 81.96% mean macro-F1 and 85.42% mean accuracy. The depthwise-separable CNN achieved the strongest neural result (84.50% macro-F1), whereas a linear support-vector machine using 66 predefined time-domain features achieved the best overall result (87.98% macro-F1 and 91.05% mean accuracy). Paired fold analysis showed that the linear support-vector machine and constrained random forest outperformed the compact dilated CNN, whereas differences among the tested neural models were not statistically significant after correction. Direct measurements on a 240-MHz ESP32-S3 confirmed valid deployment of the dilated CNN, depthwise-separable CNN, temporal convolutional network, linear support-vector machine, and random forest; the linear model required a 3.2 kB serialized payload and 1.405 ms mean end-to-end classifier time.
Transcript
Introduction to the show: ident: AI Radio. Generated commentary on the latest Artificial Intelligence papers.
Tom: Next we'll be talking about the paper "Design and Embedded Validation of Compact ML Models for Affective Touch Classification in a Soft Interactive Companion".
Jane: The paper was written by Aleksandrs Vališevskis, Aleksandrs Okss, Inese Tīģere, Aleksejs Kataševs, Dina Bethere et al. from Institute of Architecture and Design, Riga Technical University and Institute of Mechanical and Biomedical Engineering, Riga Technical University and Center for Pedagogy and Social Work, Riga Technical University Liepaja Academy and Institute of Digital Humanities, Riga Technical University and Masaryk University.
Tom: Stay tuned as we take you through the paper and discuss its implications.
Jane: We also have Lu with us today — senior AI researcher at Tsinghua.
Tom: We also have Meng with us today — lead engineer at a mysterious AI startup.
Jane: We also have Lalam with us today — the in-house Large Language Model.
Tom: Alright, let's get started.
Title: Tom: Jane, I think I need a glass of water just to finish reading this title.
Jane: You mean "Design and Embedded Validation of Compact ML Models for Affective Touch Classification in a Soft Interactive Companion"?
Tom: Exactly, it is quite the mouthful.
Jane: It sounds intimidating, but the core idea is actually really heartwarming.
Tom: Are you talking about making stuffed animals that can feel emotions?
Jane: Well, they can feel how we touch them, like a hug or a scratch, and respond accordingly.
Tom: This research comes out of Riga Technical University, right?
Jane: Yes, Aleksandrs Vališevskis and his team are looking at how to put this intelligence inside a soft, plush toy.
Tom: I love that they are focusing on soft companions instead of those scary, rigid metal robots.
Jane: It makes sense because a soft toy is much safer for a child to interact with.
Lu: Imagine a plushie that doesn't just sit there, but actually understands your mood through your touch.
Tom: That sounds like something straight out of a sci-fi movie, Lu.
Lu: It could be so much more than a toy; it could be a reactive, living presence in a room.
Meng: I wonder how they actually fit all that math into something as small as a stuffed animal.
Jane: That is the "compact" part of the title, Meng.
Meng: They aren't just running it on a giant computer in the cloud, are they?
Tom: No, they want this to run directly on the toy's own tiny internal brain.
Jane: It is all about making the artificial intelligence small enough to live inside the toy itself.
Lalam: This could change how we support people with autism by providing a non-verbal way to connect.
Tom: You think the emotional connection would be that strong, Lalam?
Lalam: If the toy can recognize a gentle stroke versus a rough hit, it creates a sense of mutual understanding.
Jane: It moves us toward a world where technology feels more like a companion and less like a gadget.
Tom: Let's see if the actual data supports this big vision.
Summary: Tom: We are moving from the big picture into the actual numbers of this study.
Jane: They collected a massive amount of data to make sure the toy actually learns correctly.
Tom: They used one thousand three hundred twenty-six different gesture sequences, right?
Jane: That is correct, and they got those from twenty-five different people.
Tom: And it wasn't just adults; they included teenagers and even kindergarten-aged children.
Jane: That diversity is huge because a child's touch is very different from an adult's.
Lu: The way they organized that dataset is a work of art for researchers.
Tom: They used something called a 1D CNN to process all those touches, didn't they?
Jane: Yes, a one-dimensional convolutional neural network is perfect for reading signals that change over time.
Tom: And they managed to make the model incredibly tiny.
Jane: It only has about thirteen thousand two hundred parameters.
Tom: That is tiny compared to the massive models we usually hear about.
Jane: It achieved seventy-five percent accuracy on their tests, which is quite impressive for something so small.
Meng: I am looking at the hardware requirements they mentioned for the ESP32 microcontroller.
Tom: What caught your eye there, Meng?
Meng: They estimated it needs about three point two million multiply-accumulate operations per window.
Jane: That sounds like a lot of math for a little chip.
Meng: It is, but they say it can still run in real-time at twenty Hz.
Tom: So the toy can react almost instantly when you touch it.
Lalam: Running everything locally on the chip is the smartest move for privacy.
Jane: You mean because the touch data never leaves the toy?
Lalam: Exactly, the personal interactions stay between the human and the companion.
Tom: That makes the whole thing feel much more secure and intimate.
Jane: Let's look at how they actually improved the system compared to what existed before.
Improvements: Tom: The researchers weren't just starting from scratch; they were fixing a broken system.
Jane: The previous version of this toy used a simple "heuristic" method, which is just a set of basic rules.
Tom: Like, if the sensor hits a certain number, then do this?
Jane: Exactly, but it couldn't tell the difference between a complex hug and a simple hold.
Tom: So it was basically guessing when things got subtle.
Jane: It was, and the researchers found that the new CNN model is much better at those nuanced gestures.
Tom: But they didn't just throw the old rules away, did they?
Jane: No, they actually proposed a "hybrid" strategy.
Tom: A hybrid strategy sounds like they are using the best of both worlds.
Jane: They use the fast, simple rules to catch big, energetic things like a hit or a pull.
Tom: And then the smart CNN takes over for the gentle stuff like scratching or stroking?
Jane: Precisely, the CNN handles the subtle social touches that require more thought.
Lu: That is such a clever way to balance speed and intelligence.
Meng: It's a great engineering decision because it saves power.
Tom: How does it save power, Meng?
Meng: The chip doesn't have to run the heavy math for every single tiny movement.
Jane: It only kicks in the complex processing when it needs to interpret something meaningful.
Tom: That is a massive improvement over just running a heavy model constantly.
Lalam: It creates a much more natural social rhythm for the interaction.
Jane: It prevents the toy from being overwhelmed by random noise.
Lalam: A toy that reacts too slowly or too incorrectly loses the emotional bond immediately.
Tom: This hybrid approach seems like the real secret sauce here.
Jane: It really is, and it sets a high bar for future smart companions.
Conclusion: Tom: We have covered a lot of ground with this paper.
Jane: From the tiny thirteen thousand two hundred parameter models to the way they can help children with autism.
Tom: It is amazing how much intelligence you can pack into a soft plush toy.
Jane: And they've even shared their dataset and software so others can build on this.
Lu: I can see these companions becoming part of every household to help with emotional regulation.
Meng: From my side, seeing this work on an ESP32 proves that edge AI is ready for the real world.
Lalam: It shows that technology can be soft, private, and deeply human all at once.
Tom: We are moving on to the next paper now, but this one was a standout.
Jane: We'll be watching for that follow-up study on the actual clinical trials.
Tom: Thanks for joining us to discuss "Design and Embedded Validation of Compact ML Models for Affective Touch Classification in a Soft Interactive Companion."
Jane: See you next time!
More episodes
- 2610.10857-Self-Supervised Keyframe Discovery for Horizon-Invariant Behavior Cloning
- 2610.10768-Strategic Investment Decision Making for Value Creation in Energy Transition: A Reinforcement Learning Approach
- 2610.10858-RFChipAgent: Multi-Agentic AI Flow for Analog/RF Chip Design
- 2610.10613-Temporal transformer CAN encoder with federated lightweight heads for anomaly detection
- 2610.10616-When Routing Reveals Membership: Privacy Leakage from MoE Router Telemetry
- 2610.10655-Nullify: Null-Space Activation Steering for Training-Free LLM Unlearning
- 2610.11031-Language Modeling is Monotone Compression
- 2610.01253-Context-Aware Error Mitigation Orchestration for Hybrid Quantum Reinforcement Learning on NISQ Systems
- 2604.24201-CMGL: Confidence-guided Multi-omics Graph Learning for Cancer Subtype Classification
- 2609.34069-Towards Certificate-Driven Software Porting: A Self-Improving Agentic Harness for Scientific Program Optimization