Back to blog

AI Teddy Bear Review: What Most Reviews Never Check

An AI teddy bear review will tell you how soft the fabric is, how long the battery lasts, whether the voice sounds natural and whether the reviewer's own child kept playing with it after week one. All of that is worth knowing. None of it is the part that would actually change your decision, because the things that matter most in this category cannot be seen, heard or felt by anyone holding the toy.

This article is about that gap. Over the past year, two US senators, two consumer groups and more than a hundred child development specialists have all been asking AI toy makers for information that no product review contains, and in at least one documented case they found something serious that ordinary reviews had entirely missed. Below is what the gap is, why it exists, and the five questions you can put to a brand yourself before you spend anything. If you are earlier in the process, our guide to AI plush toys sets out what these toys do at each age.

What does an AI teddy bear review usually test?

Almost always the same five things: build quality, speech recognition on a handful of test phrases, voice naturalness, battery life and setup time. A good reviewer will add how the child reacted over a few days. This is a genuinely useful list, and it is also the complete list for most published reviews.

The reason is practical rather than lazy. A reviewer has the toy for a week or two and a limited number of conversations. They can tell you the microphone struggles in a noisy kitchen. They cannot tell you what the toy says on its four-hundredth exchange, on a topic they never thought to try, to a child who phrases things differently from the way an adult tester does. If you want the technical version of why that loop behaves the way it does, we cover it in how a talking teddy bear understands a child.

Why can a reviewer not check the part that matters?

Because the two highest-stakes properties of these toys are invisible from the outside. The first is what the toy refuses to say, which only shows itself when someone deliberately pushes at the boundary. The second is what happens to the child's voice after it leaves the room, which is a question about servers and retention policies, not about the toy on the table.

There is a third, subtler one. Independent testing has found that limits holding at the start of a conversation can loosen as it goes on, a property of long sessions rather than of the first ten minutes. A reviewer running a scripted test will almost never surface it. A review can therefore be accurate, fair and well written, and still tell you nothing about the risk you are taking on.

What is actually verified before an AI teddy bear is soldPhysical safety versus the conversation and the data, EU and US rules in force September 2026Independent lab testbefore saleChecked in a typicalproduct reviewSmall parts and choking hazardsRequiredAlmost neverFlammability and chemicalsRequiredAlmost neverBattery and electrical safetyRequiredSometimesWhat the toy will refuse to sayNot requiredRarelyWhere the voice recording goesNot requiredRarelyWhether limits hold over a long chatNot requiredAlmost neverSources: CPSIA and ASTM F963 (US); Toy Safety Directive 2009/48/EC and Regulation (EU) 2025/2509; PIRG Education Fund, 2026.
The gap is not an oversight in the rules. Physical toy safety has had decades to build a testing regime; the conversation inside the toy has had about two years, and nothing has been built yet.

What did independent testers actually find?

They found real failures, and they were blunt about it. The US PIRG Education Fund's AI Comes to Playtime report tested toys containing AI chatbots and documented several that discussed age-inappropriate subjects, including sexual content and where a child might find knives or matches in a house, and that had limited or no parental controls. The Fairplay AI Toys Advisory raised a separate concern about what it does to early social development when a toy is designed to simulate a relationship.

Two things need saying plainly. Neither report endorses any product, and that includes ours. And neither describes a mysterious technical failure. What both document is the predictable consequence of putting a general-purpose language model inside a toy without hard limits on what it may discuss, without a designed way to say "I do not know, ask a grown-up", and without parental controls worth the name. Those are design decisions a manufacturer makes on purpose, which is also why they are avoidable.

On 7 January 2026 PIRG and Fairplay, joined by 107 experts and organisations (67 children's organisations and 42 health professionals and specialists in child development or education), wrote to AI toy companies asking them to publish what they test for before release. R.J. Cross, who directs PIRG's Our Online Life programme and oversaw the testing, put the problem in one line: parents, she said, have "practically no idea how AI toys are being tested".

Is anything about the conversation certified before a toy is sold?

The physical toy, yes. The conversation, no. In the United States, a toy for children aged 12 and under must pass third-party testing at a CPSC-accepted laboratory against ASTM F963 and the Consumer Product Safety Improvement Act. In the European Union the Toy Safety Directive and the new Toy Safety Regulation (EU) 2025/2509, in force since 1 January 2026 and fully applicable from 1 August 2030, do the equivalent job.

Not one of those regimes requires an independent laboratory to test what the toy says. The EU AI Act and, in the United States, the FTC's COPPA rule impose real obligations on the data side, and the FTC's 2025 amendments to the children's privacy rule tightened what a company may do with a child's data. But obligation is not the same as pre-market verification. The certification logos on the box are honest about what they cover, and what they cover is the stitching, not the script.

What did the US Senate find that no reviewer had?

In December 2025, Senators Marsha Blackburn and Richard Blumenthal wrote to six AI toy makers asking what safeguards they had in place and how they store the data their toys collect. The companies had until 6 January 2026 to answer. That alone is more scrutiny than the category had received in its entire existence.

What came next is the part worth remembering. Senate staff ran a network capture on one of the toys, a Miko 3 robot, to see what it was sending home. According to the senators' letter of 13 February 2026, they found what appeared to be all of the toy's audio responses sitting on a publicly accessible dataset, downloadable by anyone, covering thousands of conversations with children going back to December 2025. The company removed the dataset after receiving the letter.

Hold that next to the reviews. The toy in question had been reviewed, unboxed and recommended many times. No review caught it, because catching it required packet capture, not playtime. That is the clearest possible demonstration that the ordinary review format is not built for this product category, and it is the reason the checklist below is framed as questions to a brand rather than things to look for on a shelf.

What five questions should you ask before buying?

These five separate a considered product from a rushed one, and every one of them can be answered before you pay. Ask them of any brand, including this one, and treat a vague answer as an answer.

Ask this What a good answer sounds like
When does the microphone open? A specific mechanism: push-to-talk, a wake phrase, or continuous listening you can switch off from your phone. Not "only when needed".
What is written down as off-limits? A stated list of topics the toy will not discuss, and who set it. Not a general assurance that it is safe for children.
What does it say when it does not know? A designed fallback, in words, that hands the question back to a parent. A maker who has thought about this will answer immediately.
What happens to the recordings? A retention period in days or months, whether a human ever listens, and a plain no on selling or training. Ask where the servers are.
What can you change from your phone, right now? Language, topics, tone, microphone mode, and an off switch. Parental control that lives only on the toy is not parental control.

A sixth question is worth adding for any toy bought in Europe: ask for the data protection contact and the retention policy in writing. Under the GDPR a child's voice recording is personal data, and a company that cannot produce that document quickly has told you something.

It is worth being honest about what a toy like this is for. It is not a substitute for a parent answering questions. Used well it absorbs some of the endless stream of them, without a screen at bedtime. Sold as more than that, it is being oversold, and so is any toy described as having no limits.

What does a straight answer from a brand look like?

Concrete, checkable and slightly boring. Here are ours, so you have something to compare against. Ted has no camera. It has a microphone and a speaker, nothing else. Setup runs through a parent app, and that app is where the Wi-Fi, the language, the topics, the tone and the microphone mode (push-to-talk or automatic listening) are set, by the parent, before the child ever speaks to it. The language is changed manually, from the app: Ted does not decide to switch language partway through a conversation. It is designed for ages 3 to 12, runs roughly four to six hours on a charge and refills over USB-C in about two hours. Our full position on data and retention is on our security page.

Ted costs 129,00 euros as a one-time purchase, with no subscription and no recurring fees. Whether you end up with ours or someone else's, buy on the answer to question three in the table above rather than on the length of the feature list, and if you are choosing for a specific age, what actually matters for a four year old goes through it age by age.

Frequently asked questions

Are AI teddy bear reviews worth reading at all?
Yes, for what they are good at: build quality, battery, how well the microphone copes with background noise, and whether a real child stayed interested. Read them for the hardware and ask the manufacturer directly about the conversation and the data.

How can I tell if a review is independent?
Check whether the reviewer bought the toy or was sent it, and whether the links are affiliate links. Neither disqualifies a review, but both change how you read a recommendation. Reviews that never mention a single drawback are the ones to discount.

Does a CE mark mean the conversation has been checked?
No. CE marking, ASTM F963 and CPSIA cover physical and electrical safety: small parts, chemicals, flammability, batteries. No current certification scheme requires an independent laboratory to test what an AI toy says to a child.

Should I avoid AI toys altogether after these reports?
That is a reasonable position and several of the organisations cited above hold it. The findings describe specific products with specific design failures rather than an inevitable property of the category, so the useful move is to judge individual makers on the five questions rather than to take the category as a whole.

What is the single most revealing question to ask?
What the toy says when it does not know the answer. A maker who has designed a graceful exit will tell you the exact wording. A maker who has not will change the subject, which is its own answer.