How AI decision models could change content moderation
On Tuesday, Musubi announced a lightweight decision model made for real-time moderation called PolicyLM-1.7B, released with open weights.
As decision models spread across the industry, a company called Musubi has a new idea for how to put them to work: moderating content. On Tuesday, Musubi announced a lightweight decision model made for real-time moderation called PolicyLM-1.7B , released with open weights.
The idea is to take a content policy written in plain English and apply it to messages in under 50 milliseconds. Musubi’s model is designed to be similar in cost and speed to the AI classifier systems that power moderation on most social platforms, but because it has the flexibility of a modern LLM, it can apply complex policies without special training. Even more important, the model won’t need new training when the policy changes, allowing for human policy-setters to iterate as much as they need.
As Musubi co-founder and chief AI officer Filip Jankovic sees it, it gives platform managers a way to label content proactively.
“Product teams just want a better understanding of what’s happening on their platform, especially as the amount of content is exponentially increasing,” Jankovic says. “Being able to label all of that in a very scalable, customizable way is extremely useful.”
Decision models have become a hot topic in the AI world since the release of TypeSafe AI’s Jev in September , which was shortly followed by competing decision models from OpenAI and Amazon . Instead of outputting text, a decision model outputs outcome probabilities, though in this case the model outputs a binary judgement: Either the content is in the category or it isn’t. By limiting the model’s output to a set of predetermined choices, decision models are able to run faster and cheaper than large language models, while still maintaining the flexibility of the transformer architecture.
One early use case is reining in misbehavior by AI agents, so it’s only natural to apply the same technology to human misbehavior.
Notably, Jankovic says his interest in decision models predates Jev, tracing it back to a 2024 project called GLiNER (Generalist Model for Named Entity Recognition) that deployed many of the same techniques.
Still, Musubi isn’t wary of the comparison. If anything, the company is eager to use the new interest in decision models to shine a light on content moderation. “If Jev caught your eye, PolicyLM-1.7B is the same kind of model, trained specifically for content moderation, that you can run yourself,” the product announcement reads.
When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence.
Get 50% off a second pass The Disrupt experience is meant to be shared. Get your pass and bring a colleague, partner, or peer at 50% off. Cover more ground by making connections, building momentum, and discovering what’s next in the startup ecosystem.
At 19, founder raises $11M for Ghost, maker of a $3,499 computer for personal AI
Federal judge calls Flock ‘indiscriminate mass surveillance’
Amazon responds to data center backlash, says it no longer uses NDAs
OpenAI safety employee resigns, claiming the company’s ‘culture is broken’
Google thinks SpaceX’s Starship has to launch 1,800 times before space data centers get off the ground
World’s first enhanced geothermal power plant completed in just 23 months