How AI decision models could change content moderation


As decision models spread across the industry, a company called Musubi has a new idea for how to put them to work: moderating content. On Tuesday, Musubi announced a lightweight decision model made for real-time moderation called PolicyLM-1.7B, released with open weights.

The idea is to take a content policy written in plain English and apply it to messages in under 50 milliseconds. Musubi’s model is designed to be similar in cost and speed to the AI classifier systems that power moderation on most social platforms — but because it has the flexibility of a modern LLM, it can apply complex policies without special training. Even more important, the model won’t need new training when the policy changes, allowing for human policy-setters to iterate as much as they need.

As Musubi co-founder and chief AI officer Filip Jankovic sees it, it gives platform managers a way to label content proactively.

“Product teams just want a better understanding of what’s happening on their platform, especially as the amount of content is exponentially increasing,” Jankovic says. “Being able to label all of that in a very scalable, customizable way is extremely useful.”

Decision models have become a hot topic in the AI world since the release of TypeSafe AI’s Jev in September, which was shortly followed by competing decision models from OpenAI and Amazon. Instead of outputting text, a decision model outputs outcome probabilities, though in this case the model outputs a binary judgement: Either the content is in the category or it isn’t. By limiting the model’s output to a set of predetermined choices, decision models are able to run faster and cheaper than large language models, while still maintaining the flexibility of the transformer architecture.

One early use case is reining in misbehavior by AI agents — so it’s only natural to apply the same technology to human misbehavior.

Notably, Jankovic says his interest in decision models predates Jev, tracing it back to a 2024 project called GLiNER (Generalist Model for Named Entity Recognition) that deployed many of the same techniques.

Still, Musubi isn’t wary of the comparison. If anything, the company is eager to use the new interest in decision models to shine a light on content moderation. “If Jev caught your eye, PolicyLM-1.7B is the same kind of model, trained specifically for content moderation, that you can run yourself,” the product announcement reads.

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.



Source link

  • Related Posts

    Today’s NYT Strands Hints, Answers and Help for Oct. 7, #948

    Looking for the most recent NYT Strands puzzle answers? CNET publishes daily answers and hints for The New York Times Mini Crossword, Connections, Connections: Sports Edition and Strands puzzles. Strands…

    Continue reading
    Sebastian Maniscalco’s SiriusXM channel is hurting up-and-coming talent, comics say

    Comedian Sebastian Maniscalco is facing backlash from fellow comics who claim his takeover of SiriusXM’s Raw Comedy channel is harming up-and-coming talent, as reported earlier by Deadline. Many established comics,…

    Continue reading

    Leave a Reply

    Your email address will not be published. Required fields are marked *

    You Missed

    Today’s NYT Strands Hints, Answers and Help for Oct. 7, #948

    Today’s NYT Strands Hints, Answers and Help for Oct. 7, #948

    Delta Airbus A350 Winglet Wedged In Air Canada Boeing 777’s Tail After LAX Collision

    Delta Airbus A350 Winglet Wedged In Air Canada Boeing 777’s Tail After LAX Collision

    Yellow labs make first leaf pile jump of this fall season

    Yellow labs make first leaf pile jump of this fall season

    Cancelled MMO spin-off Borderlands Online is playable once again, and the team behind the unofficial project thankfully didn’t use any genAI to get it running

    Cancelled MMO spin-off Borderlands Online is playable once again, and the team behind the unofficial project thankfully didn’t use any genAI to get it running

    Sereact, Zalando and CEVA Logistics start AI-powered, automated returns handling in Germany and Poland

    Sebastian Maniscalco’s SiriusXM channel is hurting up-and-coming talent, comics say

    Sebastian Maniscalco’s SiriusXM channel is hurting up-and-coming talent, comics say