All posts

Translated from Spanish · See the original

How do Community Notes work on X?

1/ The algorithm of truth — How do Community Notes work on X?

They’re those notes that show up under some tweets to clarify or correct information. The interesting thing is they work surprisingly well.

How does it all work? Let me explain ⤵️

2/ It all revolves around an algorithm that measures whether a note is “helpful” or not.

First, users rate notes as “helpful,” “somewhat helpful,” or “not helpful.” Then the algorithm uses these votes to train a model that evaluates both how helpful the notes are and the “quality” (or trustworthiness) of the voters. 🧠

3/ But votes alone aren’t enough: you need cross-partisan consensus. That is, support from both left- and right-leaning users. If there’s no agreement between both sides, the note isn’t published. 🤝

This consensus acts as a quality indicator: if it convinces everyone, it’s probably reliable information.

4/ Also, your influence depends on your track record.

If your votes tend to align with the consensus, your weight in the system goes up. If you’re always disagreeing with everyone, your impact goes down.

This prevents biased users from distorting the results. 🎯

(Obviously, this isn’t risk-free)

5/ Abuse is penalized: if a note is marked as offensive by cross-partisan consensus, those who rated it as “helpful” lose credibility. 🛑

That way, the system discourages toxic behavior and rewards constructive participation.

6/ Algorithm summary:
• Only notes with consensus between left and right get published.

• Reliable users have more weight.

• Abuse is punished.

It’s simple in concept, but it raises some questions:
• Is this balance sustainable?

• Isn’t reducing everything to “left” and “right” a bit simplistic? 🤔

7/ In any case, the system is a fascinating experiment. It’s not perfect, but it seems like a good foundation for managing misinformation on large platforms. And realistically, it works pretty well in practice.

What do you think? 💬

8/ If you want to dig deeper, this post from Less Wrong explains it better than I can:
lesswrong.com/posts/sx9wTyCp5kgy8xGac/co…

Views: 446Likes: 7Replies: 2Reposts: 0View on X

Enjoyed this post?

Leave me your email and I'll let you know when I publish something new.

Related posts

It’s sooo easy to tell which apps have been vibecoded. So easy. There are details that scream it right in your face from minute one. Little things… but they immediately give you away: 🎨 The design. It’s not that it’s ugly. It’s that it’s the default design. You can almost guess which model has

Views: 93.6KLikes: 802Replies: 64Reposts: 32View on X

If Fable is already almost unusable even when you’re paying €100/€200 a month, I’m scared to see what happens when the extended limits period ends. Anthropic needs compute now. Well, it’s needed it for months, but now people are realizing just how much you can get done with Codex and even Grok.

Views: 10.5KLikes: 155Replies: 7Reposts: 3View on X