

Why Behavioral Science Is the Missing Piece of AI Safety
AI Safety
This article expands on ideas first introduced in my Harvard Business Review article,"Balancing Digital Safety and Innovation."
Read the Harvard Business Review article →
I’ve spent years studying why people get deceived online. The answer is never just about technology.
People get scammed for many reasons — security systems fail, platforms don’t catch bad actors in time, and enforcement lags behind the speed of fraud. But even when the technology works, people still get deceived. They trust someone they shouldn’t have. They miss a warning. They take a risk they knew was dangerous because, in that situation, the emotional pull felt stronger than the logic against it.
Technology is necessary. But it’s not sufficient. And that gap — between what the system can detect and what human psychology makes us vulnerable to — is where most of the harm happens.”
The Industry’s Expensive Blind Spot
For decades, the technology industry has embraced a familiar philosophy: ship fast, learn from failures, and improve over time. This approach has driven remarkable innovation, but it has also revealed a fundamental weakness. Some failures cannot simply be fixed after deployment because the harm has already been done.
Recent legal actions and growing public scrutiny of major technology platforms have reinforced an important lesson: reactive safety is no longer enough. Organizations are increasingly expected to anticipate foreseeable risks, protect users by design, and take responsibility for the unintended consequences of their products.
This shift reflects a broader change in how we think about innovation. Safety should not be treated as a feature added after launch or a response to public criticism. It should be considered a core design principle from the very beginning.
That is the foundation of Safety by Design—an approach that integrates safety, trust, and human behavior into every stage of product development, rather than attempting to repair problems after they emerge.
What Uber’s Delivery Cyclists Taught Me About Human Behavior
In 2020, I joined Uber Japan as Head of Safety. Uber Eats was growing rapidly, and delivery cyclists were making headlines for traffic violations — running red lights, riding without helmets, ignoring pedestrians.
The obvious response would have been stricter rules or more aggressive enforcement. Instead, I started by asking a different question: why do people behave this way even when they know it’s dangerous?
The research revealed two distinct problems. The first was psychological: cyclists under time pressure experience a narrowing of attention. The delivery notification creates urgency. In that moment, the abstract risk of an accident loses out to the immediate pressure of the job. The second was cultural: helmet-wearing simply wasn’t an established norm in Japan at the time. Many cyclists weren’t making a conscious decision to skip the helmet — the habit didn’t exist in the first place. You can’t enforce a behavior that people have never internalized.
This meant that awareness campaigns and stricter rules alone wouldn’t be enough. We needed to build the norm itself. We partnered with police departments across Japan to run in-person traffic safety workshops for delivery cyclists — practical sessions designed to build new habits, not just communicate rules. At the same time, we applied behavioral economics principles to the app itself, building what became the world’s first in-app traffic safety checklist — appearing at the exact moment a courier was about to start a delivery, not before, not after. The friction was minimal. The timing was everything. The feature was later adopted across multiple Uber markets internationally.
The lesson I took from this wasn’t about any single intervention. It was about the difference between designing for the user you imagine and designing for the user who actually exists — shaped by culture, constrained by time, and reliably human.
Rethinking Safety at Match Group
After Uber, I led the launch of the Safety by Design program at Match Group, the parent company of Tinder, Plenty of Fish, and more than 30 dating apps worldwide. At the time, safety across online platforms followed a common pattern: react, respond, remediate. A talented team could work tirelessly and still be perpetually behind — because the entire system was built around responding to harm rather than preventing it.
We reframed the question. Instead of “how do we respond when something goes wrong,” we asked “how do we make it structurally harder for things to go wrong in the first place?”
One outcome was that Pairs became the first major dating app to require facial verification for every user. It wasn’t a popular decision internally — friction in onboarding is never welcomed by growth teams. But the principle held. Tinder and others eventually followed.
The harder problem was education. Facial verification catches impostors. It doesn’t help users who don’t know romance scams exist — and therefore have no reason to be suspicious when someone takes a sudden romantic interest in them online. Scammers don’t hack systems. They exploit cognitive biases: the need for connection, the tendency to trust people who seem to understand us, the difficulty of believing someone would invest months in a relationship purely to steal from us.
My subsequent research on romance scams, shared with organizations including the FBI and Google, led to the same conclusion: people cannot protect themselves from risks they do not recognize. Awareness is not a soft add-on to safety—it is a core design requirement.


Safety Doesn’t Constrain Innovation. It Focuses It.
The pushback I hear most often is that embedding safety into product development slows things down. Having built safety programs at Uber and Match Group, I've come to believe the opposite is true.
Companies that take safety seriously are forced to understand their users at a level most product teams never reach. They have to ask not just “what do users want to do” but “what do users actually do, and why do they sometimes do things that harm themselves or others?” That quality of understanding doesn’t just produce safer products. It produces better ones.
Safety is not a cost center. It’s a source of product insight — and, increasingly, a source of competitive advantage. Trust is hard to build and easy to lose. The companies that design for it from day one are building something that can’t be easily replicated.
Why AI Makes This More Urgent, Not Less
As AI systems become more capable, the temptation is to assume that better technology will solve safety problems that worse technology couldn’t. I think this gets the problem backwards.
Most AI failures aren’t purely technical. They emerge at the intersection of AI capability and human behavior. People overtrust AI-generated outputs. They fail to notice when an AI is confidently wrong. And increasingly, bad actors use AI not to break systems but to manipulate the people using them — more convincing scam messages, more sophisticated social engineering, more personalized deception at scale.
The more powerful AI becomes, the more the risk surface shifts toward human psychology. Which means the more important it becomes to design AI systems around how people actually think, decide, and err — not around how we wish they would.
The future of AI safety will depend not only on better models, but also on a deeper understanding of how people think, decide, and behave.
_______________________________________________
The future of AI safety will depend not only on better models, but also on a deeper understanding of how people think, decide, and behave.
At Behavioral AI Lab, we believe AI safety is ultimately as much a human challenge as it is a technical one. By combining behavioral science with AI, we can build systems that are not only more capable, but also safer, more trustworthy, and better aligned with the people they serve.
In a future article, I'll share insights from my keynote at the United Nations and explore what it takes to build AI safety that works—not only technically, but behaviorally.
If you're interested in collaborating with Behavioral AI Lab, or would like to discuss research, advisory services, or speaking opportunities, we'd love to hear from you.
BLOG & INSIGHTS
Exploring the Human Side
of Artificial Intelligence.


Why Behavioral Science Is the Missing Piece of AI Safety
AI Safety
This article expands on ideas first introduced in my Harvard Business Review article,"Balancing Digital Safety and Innovation."
Read the Harvard Business Review article →
I’ve spent years studying why people get deceived online. The answer is never just about technology.
People get scammed for many reasons — security systems fail, platforms don’t catch bad actors in time, and enforcement lags behind the speed of fraud. But even when the technology works, people still get deceived. They trust someone they shouldn’t have. They miss a warning. They take a risk they knew was dangerous because, in that situation, the emotional pull felt stronger than the logic against it.
Technology is necessary. But it’s not sufficient. And that gap — between what the system can detect and what human psychology makes us vulnerable to — is where most of the harm happens.”
The Industry’s Expensive Blind Spot
For decades, the technology industry has embraced a familiar philosophy: ship fast, learn from failures, and improve over time. This approach has driven remarkable innovation, but it has also revealed a fundamental weakness. Some failures cannot simply be fixed after deployment because the harm has already been done.
Recent legal actions and growing public scrutiny of major technology platforms have reinforced an important lesson: reactive safety is no longer enough. Organizations are increasingly expected to anticipate foreseeable risks, protect users by design, and take responsibility for the unintended consequences of their products.
This shift reflects a broader change in how we think about innovation. Safety should not be treated as a feature added after launch or a response to public criticism. It should be considered a core design principle from the very beginning.
That is the foundation of Safety by Design—an approach that integrates safety, trust, and human behavior into every stage of product development, rather than attempting to repair problems after they emerge.
What Uber’s Delivery Cyclists Taught Me About Human Behavior
In 2020, I joined Uber Japan as Head of Safety. Uber Eats was growing rapidly, and delivery cyclists were making headlines for traffic violations — running red lights, riding without helmets, ignoring pedestrians.
The obvious response would have been stricter rules or more aggressive enforcement. Instead, I started by asking a different question: why do people behave this way even when they know it’s dangerous?
The research revealed two distinct problems. The first was psychological: cyclists under time pressure experience a narrowing of attention. The delivery notification creates urgency. In that moment, the abstract risk of an accident loses out to the immediate pressure of the job. The second was cultural: helmet-wearing simply wasn’t an established norm in Japan at the time. Many cyclists weren’t making a conscious decision to skip the helmet — the habit didn’t exist in the first place. You can’t enforce a behavior that people have never internalized.
This meant that awareness campaigns and stricter rules alone wouldn’t be enough. We needed to build the norm itself. We partnered with police departments across Japan to run in-person traffic safety workshops for delivery cyclists — practical sessions designed to build new habits, not just communicate rules. At the same time, we applied behavioral economics principles to the app itself, building what became the world’s first in-app traffic safety checklist — appearing at the exact moment a courier was about to start a delivery, not before, not after. The friction was minimal. The timing was everything. The feature was later adopted across multiple Uber markets internationally.
The lesson I took from this wasn’t about any single intervention. It was about the difference between designing for the user you imagine and designing for the user who actually exists — shaped by culture, constrained by time, and reliably human.
Rethinking Safety at Match Group
After Uber, I led the launch of the Safety by Design program at Match Group, the parent company of Tinder, Plenty of Fish, and more than 30 dating apps worldwide. At the time, safety across online platforms followed a common pattern: react, respond, remediate. A talented team could work tirelessly and still be perpetually behind — because the entire system was built around responding to harm rather than preventing it.
We reframed the question. Instead of “how do we respond when something goes wrong,” we asked “how do we make it structurally harder for things to go wrong in the first place?”
One outcome was that Pairs became the first major dating app to require facial verification for every user. It wasn’t a popular decision internally — friction in onboarding is never welcomed by growth teams. But the principle held. Tinder and others eventually followed.
The harder problem was education. Facial verification catches impostors. It doesn’t help users who don’t know romance scams exist — and therefore have no reason to be suspicious when someone takes a sudden romantic interest in them online. Scammers don’t hack systems. They exploit cognitive biases: the need for connection, the tendency to trust people who seem to understand us, the difficulty of believing someone would invest months in a relationship purely to steal from us.
My subsequent research on romance scams, shared with organizations including the FBI and Google, led to the same conclusion: people cannot protect themselves from risks they do not recognize. Awareness is not a soft add-on to safety—it is a core design requirement.


Safety Doesn’t Constrain Innovation. It Focuses It.
The pushback I hear most often is that embedding safety into product development slows things down. Having built safety programs at Uber and Match Group, I've come to believe the opposite is true.
Companies that take safety seriously are forced to understand their users at a level most product teams never reach. They have to ask not just “what do users want to do” but “what do users actually do, and why do they sometimes do things that harm themselves or others?” That quality of understanding doesn’t just produce safer products. It produces better ones.
Safety is not a cost center. It’s a source of product insight — and, increasingly, a source of competitive advantage. Trust is hard to build and easy to lose. The companies that design for it from day one are building something that can’t be easily replicated.
Why AI Makes This More Urgent, Not Less
As AI systems become more capable, the temptation is to assume that better technology will solve safety problems that worse technology couldn’t. I think this gets the problem backwards.
Most AI failures aren’t purely technical. They emerge at the intersection of AI capability and human behavior. People overtrust AI-generated outputs. They fail to notice when an AI is confidently wrong. And increasingly, bad actors use AI not to break systems but to manipulate the people using them — more convincing scam messages, more sophisticated social engineering, more personalized deception at scale.
The more powerful AI becomes, the more the risk surface shifts toward human psychology. Which means the more important it becomes to design AI systems around how people actually think, decide, and err — not around how we wish they would.
The future of AI safety will depend not only on better models, but also on a deeper understanding of how people think, decide, and behave.
_______________________________________________
The future of AI safety will depend not only on better models, but also on a deeper understanding of how people think, decide, and behave.
At Behavioral AI Lab, we believe AI safety is ultimately as much a human challenge as it is a technical one. By combining behavioral science with AI, we can build systems that are not only more capable, but also safer, more trustworthy, and better aligned with the people they serve.
In a future article, I'll share insights from my keynote at the United Nations and explore what it takes to build AI safety that works—not only technically, but behaviorally.
If you're interested in collaborating with Behavioral AI Lab, or would like to discuss research, advisory services, or speaking opportunities, we'd love to hear from you.
BLOG & INSIGHTS
