23:26 UTC
AK

Atoosa Kasirzadeh

Staff Research Scientist, Google DeepMind; Assistant Professor (on leave)

Google DeepMind / Carnegie Mellon University

Philosophy of AI ethics and safety, including work on the ethics of generative models and sociotechnical risk.

Homepage / profile → Safety & alignment coverage →

Latest in the feed

A Roadmap to Impactful Pluralistic Alignment Research

Pluralistic value alignment---the goal of building AI systems that represent and serve diverse human values and perspectives---has emerged as an active research agenda. Yet, there's no public evidence that it has shaped the training or evaluation of the AI systems people actually use. We audit the public behavior documents and evaluations of frontier labs, finding none name pluralism as a goal, and as of this writing, no clear indication that production models are explicitly trained or tested fo
arXiv 6d ago Research Safety & alignmentTransparency

Distributed Denial of Science: How Indirect Data Poisoning of AI Systems Can Industrialize Scientific Fraud

Scientific fraud is the instrument of doubt that malicious entities can use to establish controversy in science. Historically, it required the resources of a company: deep pockets, ghostwritten articles, and corrupt academics. Today, Artificial Intelligence (AI) is increasingly automating scientific research, so we ask: Can a remote adversary weaponize the honest use of AI in science to compromise scientific integrity? We envision and empirically evaluate a new attack, indirect data poisoning, i
arXiv 18d ago Research Military & security

Bridging the Gap in the Responsible AI Divides

Tensions between AI Safety (AIS) and AI Ethics (AIE) have increasingly surfaced in AI governance and public debates about AI, leading to what we term the "responsible AI divides". We introduce a model that categorizes four modes of engagement with the tensions: radical confrontation, disengagement, compartmentalized coexistence, and critical bridging. We then investigate how critical bridging, with a particular focus on bridging problems, offers one of the most viable constructive paths for adva
arXiv 137d ago Research RegulationSafety & alignment

Science in the age of large language models

OpenAlex 1191d ago Research