-
Tournesol: Permissionless Collaborative Algorithmic Governance with Security Guarantees
Authors:
Lê Nguyên Hoang,
Romain Beylerian,
Bérangère Colbois,
Julien Fageot,
Louis Faucon,
Aidan Jungo,
Alain Le Noac'h,
Adrien Matissart,
Oscar Villemaud
Abstract:
Recommendation algorithms play an increasingly central role in our information ecosystem. Yet, so far, they are mostly designed, parameterized and updated unilaterally by private groups or governmental authorities, based on insecure data from increasingly many fake accounts. In this paper, we present an end-to-end permissionless collaborative algorithmic governance pipeline with security guarantee…
▽ More
Recommendation algorithms play an increasingly central role in our information ecosystem. Yet, so far, they are mostly designed, parameterized and updated unilaterally by private groups or governmental authorities, based on insecure data from increasingly many fake accounts. In this paper, we present an end-to-end permissionless collaborative algorithmic governance pipeline with security guarantees, which is deployed on the open-source platform https://tournesol.app. Our pipeline has essentially four steps. First, voting rights are assigned to the contributors, based on Sybil-resilient email domains and on a novel secure trust propagation algorithm. Second, a generalized Bradley-Terry model turns contributors' pairwise alternative comparisons into scores. Third, contributors' scores are collaboratively scaled, by an adaptation of the robust sparse voting solution Mehestan. Finally, scaled scores are post-processed and securely aggregated into human-readable global scores, which are used for recommendation and display. We believe that our pipeline lays an appealing foundation for any collaborative, effective, scalable, fair, interpretable and secure algorithmic governance.
△ Less
Submitted 15 August, 2023; v1 submitted 30 October, 2022;
originally announced November 2022.
-
Tournesol: A quest for a large, secure and trustworthy database of reliable human judgments
Authors:
Lê-Nguyên Hoang,
Louis Faucon,
Aidan Jungo,
Sergei Volodin,
Dalia Papuc,
Orfeas Liossatos,
Ben Crulis,
Mariame Tighanimine,
Isabela Constantin,
Anastasiia Kucherenko,
Alexandre Maurer,
Felix Grimberg,
Vlad Nitu,
Chris Vossen,
Sébastien Rouault,
El-Mahdi El-Mhamdi
Abstract:
Today's large-scale algorithms have become immensely influential, as they recommend and moderate the content that billions of humans are exposed to on a daily basis. They are the de-facto regulators of our societies' information diet, from shaping opinions on public health to organizing groups for social movements. This creates serious concerns, but also great opportunities to promote quality info…
▽ More
Today's large-scale algorithms have become immensely influential, as they recommend and moderate the content that billions of humans are exposed to on a daily basis. They are the de-facto regulators of our societies' information diet, from shaping opinions on public health to organizing groups for social movements. This creates serious concerns, but also great opportunities to promote quality information. Addressing the concerns and seizing the opportunities is a challenging, enormous and fabulous endeavor, as intuitively appealing ideas often come with unwanted {\it side effects}, and as it requires us to think about what we deeply prefer.
Understanding how today's large-scale algorithms are built is critical to determine what interventions will be most effective. Given that these algorithms rely heavily on {\it machine learning}, we make the following key observation: \emph{any algorithm trained on uncontrolled data must not be trusted}. Indeed, a malicious entity could take control over the data, poison it with dangerously manipulative fabricated inputs, and thereby make the trained algorithm extremely unsafe. We thus argue that the first step towards safe and ethical large-scale algorithms must be the collection of a large, secure and trustworthy dataset of reliable human judgments.
To achieve this, we introduce \emph{Tournesol}, an open source platform available at \url{https://tournesol.app}. Tournesol aims to collect a large database of human judgments on what algorithms ought to widely recommend (and what they ought to stop widely recommending). We outline the structure of the Tournesol database, the key features of the Tournesol platform and the main hurdles that must be overcome to make it a successful project. Most importantly, we argue that, if successful, Tournesol may then serve as the essential foundation for any safe and ethical large-scale algorithm.
△ Less
Submitted 29 May, 2021;
originally announced July 2021.
-
Iterative Classroom Teaching
Authors:
Teresa Yeo,
Parameswaran Kamalaruban,
Adish Singla,
Arpit Merchant,
Thibault Asselborn,
Louis Faucon,
Pierre Dillenbourg,
Volkan Cevher
Abstract:
We consider the machine teaching problem in a classroom-like setting wherein the teacher has to deliver the same examples to a diverse group of students. Their diversity stems from differences in their initial internal states as well as their learning rates. We prove that a teacher with full knowledge about the learning dynamics of the students can teach a target concept to the entire classroom us…
▽ More
We consider the machine teaching problem in a classroom-like setting wherein the teacher has to deliver the same examples to a diverse group of students. Their diversity stems from differences in their initial internal states as well as their learning rates. We prove that a teacher with full knowledge about the learning dynamics of the students can teach a target concept to the entire classroom using O(min{d,N} log(1/eps)) examples, where d is the ambient dimension of the problem, N is the number of learners, and eps is the accuracy parameter. We show the robustness of our teaching strategy when the teacher has limited knowledge of the learners' internal dynamics as provided by a noisy oracle. Further, we study the trade-off between the learners' workload and the teacher's cost in teaching the target concept. Our experiments validate our theoretical results and suggest that appropriately partitioning the classroom into homogenous groups provides a balance between these two objectives.
△ Less
Submitted 12 November, 2018; v1 submitted 8 November, 2018;
originally announced November 2018.