Confirmed authors

Writers with articles

71 people whose names appear in publisher-supplied bylines for retained AI ethics articles, papers and essays. Their profiles and writing update after the daily ingest.

71 people with confirmed bylines169 authored records retainedUpdated after daily ingest

Latest writing by these authors

publisher-supplied bylines

Articles, papers and essays attributed to people in this directory by the source metadata. Open the original publisher for the authoritative byline and text.

By Vincent ConitzerarXiv cs.CYyesterdaySafety & alignment

Position: We Need Practical AI Alignment Methods to Mirror Human Reasoning

arXiv:2608.12372v1 Announce Type: cross Abstract: AI systems are increasingly employed as decision aids, decision delegates, or autonomous decision-makers. This position paper argues that in many settings, particularly high-stakes decision-making, we need accurate cognitively-aligned AI systems that reason similarly to their users, and faithfully communicate their reasoning. We review evidence that cognitive alignment improves understandability and trustworthiness, and provide new survey data sh
By Yoshua BengioarXiv4d agoSafety & alignment

How to Verify Consistency of Probabilistic Claims

When a probabilistic predictor answers many conditional-probability queries, are its answers self-consistent, and can this be verified in polynomial time? This problem is of interest for AI safety, where safety is derived from honesty about probabilistic predictions of unwanted outcomes potentially caused by an AI action. We construct an interactive PCP as follows. Let a predictive model be specified by a probability circuit P and a circuit Q which outputs confidence in predictions. Together, P
By Abeba BirhanearXiv4d agoSafety & alignmentAgents & autonomy

Toward a Theory of Value in AI Alignment

Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms spanning from toxic speech and hallucinations to AI agents executing unauthorized actions. Within the field of AI safety, these harmful instances are often framed as the alignment problem, or of models being misaligned with human values. Researchers have responded by pursuing applied and theoretical AI value alignment efforts, often without spec
By Sigal SamuelVox Future Perfect6d agoChildren & education

Is it wrong to send your kid to private school?

Editor’s note, August 9, 8 am ET: We’re bringing you some of our best-loved Your Mileage May Vary columns while Sigal Samuel is on parental leave. The one below was originally published in April. This unconventional advice column offers you a unique framework for thinking through moral dilemmas. It’s based on value pluralism: the idea that each of […]
By Roman YampolskiyHuggingFace Daily Papers8d agoAgents & autonomy

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits that become the runtime for later work. Core evolution proceeds in two modes. In recursive free evolution, improvement is itself a task, and completing one evolution cycle can schedule the next. In experience-driven core evolution, ordinary work and social interaction expose bugs, rough edges, and inefficient context construction that lead to revi
By Dawn SongarXiv cs.AI2d agoAgents & autonomy

Vero: Can AI Agents Build Formally Verified Software Repositories?

AI agents are increasingly used for programming, but do not provide any guarantee on the correctness of generated code. Verified code generation, in which an agent produces both an implementation and a machine-checked proof of its specification, offers a stronger path toward trustworthy AI-generated software. Existing benchmarks in this direction either focus on individual functions or only evaluate proof generation with provided implementations. It is still an open question whether agents can m
By Vincent ConitzerarXiv cs.AI3d agoAgents & autonomy

Do LLMs Take Care of Their Own? Similarity Signals Can Induce Cooperation

As LLM-based agents with user-instructed goals are becoming widely deployed, they increasingly encounter each other in strategic interactions, and face challenges of finding mutually beneficial outcomes. Prior literature has argued that cooperation problems such as the Prisoner's Dilemma are resolvable in settings where agents know they follow very similar decision making patterns, as for example in monocultural AI ecosystems. Following that line of work, this paper introduces the first framewor
By Andrew SelbstarXiv4d agoRegulation

What We Know about Responsible AI Practices in Industry: A Half Decade of Empirical Research

Responsible AI (RAI) has become a central concern for technology companies, regulators, and the public. How industry practitioners interpret, implement, and sustain RAI work directly shapes the design and deployment of AI systems. As empirical scholarship examining RAI practices in industry has rapidly expanded, findings are dispersed across studies that focus on different roles, organizational contexts, and interventions. This work synthesizes current knowledge through a literature review of 16
By Solon BarocasarXiv cs.HC7d agoChildren & education

Beyond "I Can't Help With That": How Child Safety Experts Evaluate AI Chatbot Safety

Youth increasingly turn to AI chatbots for social and emotional support, raising concerns about how these systems respond, especially in high-stakes situations. However, existing child safety evaluations of AI lack grounding in real-world harms that youth experience, rely on unvalidated assumptions about what counts as an appropriate output (e.g., refusal), and typically focus on detecting adversarial prompts or surface-level harms in outputs only. Thus, these evaluations can fail to detect resp

A

B

SB Solon Barocas Principal Researcher Microsoft Research (adjunct, Cornell University) Foundational work on 'disparate impact' in machine learning and co-author of the 'Fairness and Machine Learning' textbook. 3 bylined records EB Emily M. Bender Professor of Linguistics University of Washington Co-author of 'Stochastic Parrots' and a leading critic of LLM hype and the 'AI' framing itself. 1 bylined record YB Yoshua Bengio Professor & Founder, LawZero University of Montreal / Mila / LawZero Turing Award laureate who now leads international AI-risk assessment and founded a nonprofit for 'safe-by-design' AI. 3 bylined records SB Stella Biderman Open language-model researcher Many Builders reference roster Open language-model work and research arguing that AI science must examine training dynamics, not only post-training fixes. Source-linked from Many Builders 3 bylined records AB Abeba Birhane Assistant Professor & Founding Director, AI Accountability Lab Trinity College Dublin Auditing large-scale training datasets and exposing harmful content and bias in vision-language data. 6 bylined records RB Rishi Bommasani Society Lead, Center for Research on Foundation Models Stanford University Leading transparency and evaluation work on foundation models and their societal impact. 4 bylined records NB Nick Bostrom Founder & Principal Researcher, Macrostrategy Research Initiative Macrostrategy Research Initiative Author of 'Superintelligence,' framing existential risk from advanced AI for a broad audience. 1 bylined record DB danah boyd Partner Researcher, Microsoft Research; Founder, Data & Society Microsoft Research / Georgetown University Foundational research on the social implications of data-driven and algorithmic systems. 1 bylined record MB Miles Brundage Founder, AI Verification and Evaluation Research Institute (AVERI) AVERI (formerly OpenAI) Former OpenAI AGI-readiness lead now advocating for independent third-party auditing of frontier models. 2 bylined records RB Ryan Burnell AI capability-evaluation researcher Many Builders reference roster Research applying cognitive-science methods to measurement of progress toward general-purpose AI. Source-linked from Many Builders 1 bylined record

C

D

E

F

G

H

I

K

L

M

N

O

P

R

S

T

W

Y

Z

What counts as a writer here

This view includes only profiles with at least one exact full-name match in source-supplied author metadata. It is a practical route into confirmed writing, not a ranking or a complete bibliography. A publisher byline can still be incomplete or wrong, and identical names can refer to different people, so verify authorship and identity at the original publisher.

The list and its article counts update after daily ingestion. Name mentions in headlines or summaries are excluded from authorship and remain separated on individual profiles.