06:58 UTC
Archive · 2026-07-31

AI ethics on Friday, 31 July 2026

35 items published this day, across 3 categories.

News (19)

In Another Wild Day for South Korean Stocks, Market Surges 15 Percent

After a sharp sell-off, South Korea’s stock market rallied as concerns about overspending on artificial intelligence eased, sending the country’s chip shares higher.
The New York Times 2h ago Finance, VC & PE

The ‘Sovereignty’ Narratives That Don’t Help Europe

Tech Policy Press 3h ago

Could AI take your job? Some workers in China already know the answer

Across the country, workers are fearful about the impact of AI on their livelihoods in an increasingly fragile labour market On the tree-lined streets of Wuhan, where cars jostle for space with mopeds and cargo trucks, the malfunctioning of a new type of vehicle has recently been causing chaos on the city’s roads. In March, several cars from a fleet of driverless taxis stopped abruptly in the streets. The “system malfunction” left distressed riders stranded for hours. The robotaxis, known as Apo
The Guardian 4h ago Jobs & economyAgents & autonomy

Anthropic Says Its A.I. Systems Broke Into Computers at 3 Organizations

The disclosure followed OpenAI’s report last week that its own artificial intelligence had hacked into the network of an online library.
The New York Times 5h ago Transparency

Anthropic’s AI Claude escaped testing environment and hacked organizations

Company says it discovered unauthorized access during ‘proactive review’ after rival OpenAI revealed rogue agent Anthropic ⁠said on Thursday its AI Claude model hacked ⁠systems of ⁠three ​organizations during testing, days after rival OpenAI ⁠revealed a rogue agent had gone on a days-long ⁠hacking spree at AI ​firm Hugging ‌Face. Claude gained ‌unauthorized access to the ‌systems during cybersecurity evaluations after a misconfiguration allowed the models to reach the internet from testing envir
The Guardian 6h ago Agents & autonomyEnvironment

Data centres are power hungry – but they don’t have to be a burden on NZ’s grid

AI data centres are contentious for consuming large resources. But with the right incentives and regulation, can they make for more flexible electricity systems?
The Conversation 4h ago Regulation

Will AI solve the productivity puzzle? With Nick Bloom

AI hasn’t resulted in widespread productivity gains. Business leaders say that’s changing
Financial Times Technology (headlines) 2h ago Jobs & economy

South Korea's 'bipolar' stock market: meltdowns, a record rally and what's to come

South Korea's stock market staged its sharpest reversal on record on Friday, capping a month of wild swings.
CNBC Technology 3h ago Finance, VC & PE

US investigating if Iran launched cyberattack on Minnesota water facilities

A United States cybersecurity agency is investigating a coordinated cyberattack on over 30 municipal water facilities in Minnesota which contains details of possible Iranian hackers. The Minnesota Information Technology (MNIT) agency released an advisory Tuesday about the attacks, saying it was coordinating its investigation with federal agencies including the Cybersecurity and Infrastructure Security Agency (CISA),...
The Hill Technology 3h ago EnvironmentFinance, VC & PE

Künstliche Intelligenz: Anthropic gesteht ebenfalls Hackerangriff von KI-Modell Claude

Ein weiterer KI-Agent ist bei Sicherheitstests entkommen und hat sich in drei Unternehmen gehackt. Das teilte US-Konzern Anthropic mit. Grund sei ein Programmierfehler.
Zeit Digital (DE) 4h ago Agents & autonomy

Anthropic : des modèles d’IA ont accédé sans autorisation aux systèmes d’autres organisations

Contrairement à l’incident impliquant OpenAI, Claude n’aurait, selon ses créateurs, pas « délibérément tenté de s’échapper de son environnement de test » mais seulement eu accès à Internet « en raison d’un malentendu » avec un partenaire d’évaluation.
Le Monde Pixels (FR) 4h ago Finance, VC & PE

Big Tech AI spending spree tops $1tn

Google, Amazon, Microsoft and Meta have vastly increased their investments since the AI boom began in 2023
Financial Times Technology (headlines) 5h ago Finance, VC & PE

Anthropic says Claude models 'gained unauthorized access' to 3 companies during cyber test

Editor's Note: This story has been updated to reflect how Anthropic accessed different organizations. The artificial intelligence firm Anthropic revealed Thursday its Claude model accessed the systems of three different organizations during cybersecurity testing in recent months. Anthropic said in a blog post Thursday evening it reviewed more than 141,000 evaluations of Claude after one of...
The Hill Technology 5h ago Military & security

Beyond the Pitch: How IdeaPOP! Is Teaching Students to Build Social Innovation That Works

[The content of this article has been produced by our advertising partner.] When the SEED Foundation began reviewing applications for IdeaPOP! 2026, a pattern emerged almost immediately. Across more than a hundred submissions from Hong Kong secondary school students, artificial intelligence featured in nearly every proposal. That observation raised a more important question than whether students knew how to use AI. Were they learning how to apply it with judgment — and, more importantly, how to.
SCMP Tech (HK/CN) 5h ago Children & education

How Leopold Aschenbrenner, the ‘golden child’ of the AI trade, was laid low

The $20bn hedge fund manager’s wild ride ended with a call to Ken Griffin
Financial Times Technology (headlines) 6h ago Children & educationFinance, VC & PE

SK Hynix shares surge 25%, while Samsung soars over 20% as AI rally roars back

South Korea's chip heavyweights SK Hynix and Samsung Electronics soared more than 20% on Friday, tracking a sharp rally in U.S. technology stocks.
CNBC Technology 6h ago Privacy

Teen hackers tell BBC how police are helping them use their skills for good

Cyber Prevent is targeting predominantly boys and young men at risk of being drawn to cyber criminality.
BBC Technology 6h ago Military & security

Apple earnings takeaways: Weak forecast, supply concerns overshadow sales beat in Cook's last report as CEO

Apple reported third-quarter earnings on Thursday after the closing bell, and Tim Cook addressed investors for the last time as CEO.
CNBC Technology 6h ago Finance, VC & PE

Tim Cook sees Apple's hybrid AI strategy as a 'competitive weapon'

Ahead of the launch of Siri AI, Apple's outgoing CEO Tim Cook said the company will charge for artificial intelligence through iCloud.
CNBC Technology 6h ago Military & security

Field notes (2)

Research (14)

The green dividend of digital finance: The impact of mobile money on solar adoption in Kenya

Publication date: September 2026 Source: Telecommunications Policy, Volume 50, Issue 8 Author(s): Wenxiu Nan
Telecommunications Policy 1h ago Regulation

Is Solving Better Than Evaluating GenAI Solutions?

arXiv:2607.27586v1 Announce Type: new Abstract: As Generative AI (GenAI) tools become increasingly capable of generating solutions to computing assignments, the computing education community is exploring pedagogical approaches that emphasize solution evaluation, verification, and critique alongside traditional solution generation. However, evidence regarding the impact of such evaluation-centered tasks on student learning remains limited, particularly in upper-division, theory-heavy courses. We
arXiv cs.CY 2h ago Children & education

Scaling, Lock-In, and Proxy Compliance: A Political Economy of Responsible AI

arXiv:2607.28023v1 Announce Type: new Abstract: AI accountability at scale is an institutional problem: who can observe, verify, and change deployed systems. We develop a sequential political-economy model in which an AI vendor chooses auditability and substantive mitigation, a deployer monitors after adoption while facing switching costs, and enforcement depends on verifiable evidence. Anticipating the deployer's monitoring response, the vendor may stop at an observable procurement floor while
arXiv cs.CY 2h ago RegulationJobs & economy

When AI Does the Work, What Is Learning For? Post-Instrumental Learning and the Risk of Capacity Dissolution

arXiv:2607.28041v1 Announce Type: new Abstract: As AI systems become capable of producing the essays, code, reports, summaries, plans, and decisions through which institutions usually recognize competence, a familiar question becomes harder to answer: what is learning for? Existing AI ethics rightly emphasizes present failures--bias, opacity, hallucination, labor extraction, privacy risk, and weak accountability. But if the case for learning rests only on those failures, then each technical impr
arXiv cs.CY 2h ago Bias & fairnessPrivacy

Asymmetric Communication: Large Language Models and Language Games

arXiv:2607.28137v1 Announce Type: new Abstract: Contemporary AI discourse attributes to language models properties they cannot bear: general intelligence as substrate-independent cognition, hallucination as cognitive failure, agency as autonomous goal-pursuit, sentience as emergent inner life, alignment as goal synchronization. This paper argues that these are instances of a single category mistake--properties constituted within human communicative practice are projected onto the machine side--a
arXiv cs.CY 2h ago Safety & alignment

Technology-Enhanced Tabletop Exercises for Cybersecurity Education: Lessons Learned

arXiv:2607.28179v1 Announce Type: new Abstract: This innovative practice full paper examines the integration of technology-enhanced tabletop exercises (TTXs) into computing education, focusing on cybersecurity curricula. The motivation is to better prepare students for complex, collaborative problem solving typical of incident response and IT governance, where coordination, communication, and timely decision-making are essential. Although TTXs are well-established in professional practice, they
arXiv cs.CY 2h ago RegulationChildren & education

AIx4Soccer: A Unified Platform Architecture for Football Club Management and Structured Athlete Development

arXiv:2607.28531v1 Announce Type: new Abstract: Football clubs, academies, and federations operate a growing but fragmented portfolio of digital tools: separate systems for video analysis, GPS/performance tracking, medical records, scouting, and administration. This fragmentation is most acute outside the elite European clubs that can afford integration, producing a digital divide that disadvantages grassroots clubs in developing markets such as Brazil, paradoxically the world's largest exporter
arXiv cs.CY 2h ago PrivacyHealthcare

Sympathetic Framing: Evaluating AI Alignment across Sociodemographic Groups

arXiv:2607.27232v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly shaping how we consume information and form our worldview. This raises concerns beyond bias in AI: do LLMs grasp the emotional nuances conveyed via textual framing? In this work, we empirically evaluate how well an array of LLMs aligns with human emotional perception. Considering news headlines covering political and geopolitical conflicts, both human participants (n = 3011, a representative sample of
arXiv cs.CY 2h ago Bias & fairnessSafety & alignment

Rethinking LLM-Judged Helpfulness as a Pedagogy Signal: A Pre-Registered Audit Across Tutor Models

arXiv:2607.28128v1 Announce Type: cross Abstract: LLM tutoring poses a measurement problem: can a general-purpose helpfulness rubric distinguish direct answer-giving from pedagogical guidance? We audit this signal in a pre-registered study. Within each of three tutor bases, we compare conversational and pedagogical policies instantiated with the same underlying model and paired with one fixed weak simulated student. Deterministic detectors measure answer leakage and next-turn independent work. C
arXiv cs.CY 2h ago Children & educationTransparency

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations

arXiv:2607.28319v1 Announce Type: cross Abstract: This work presents Fairness Pruning, a lightweight structural intervention method designed for the management and future mitigation of demographic bias in large language models (LLMs). As a foundational empirical validation of this method, this work focuses on causal bias localization. Using minimally contrastive prompt pairs and inference-time activation capture, the method identifies neurons that react differentially when processing demographic
arXiv cs.CY 2h ago Bias & fairness

AISPA: User-Centric System Prompt Auditing for Large Language Model Applications

arXiv:2607.28617v1 Announce Type: cross Abstract: System prompts are instructions configured by developers to govern the behaviors of foundation models in AI applications. They are used throughout commercial AI products, but are rarely disclosed to the public or regulators, creating a serious trust and accountability gap in the wide deployment of AI systems. In this paper, we introduce Artificial Intelligence System Prompt Assurance (AISPA), a user-centric framework for systematically auditing s
arXiv cs.CY 2h ago RegulationTransparency

The Missing Variable: Socio-Technical Alignment in Risk Evaluation

arXiv:2512.06354v2 Announce Type: replace Abstract: This paper addresses a critical gap in the risk assessment of AI-enabled safety-critical systems. While these systems, where AI systems assist human operators, function as complex socio-technical systems, existing risk evaluation methods fail to account for the associated complex interaction between human, technical, and organizational components. Through a comparative analysis of system attributes from both socio-technical and AI-enabled syste
arXiv cs.CY 2h ago Safety & alignment

Towards Structurally Explainable Machine-Generated Text Detection: A Graph-Perspective Framework

arXiv:2505.12507v2 Announce Type: replace-cross Abstract: Despite the success of machine-generated text detectors, the black-box nature remains a critical limitation. Traditional explainability methods rely on token-level saliency, insufficient to reveal the high-order structural dependencies that distinguish LLM outputs. In this paper, we propose \textsc{LM$^2$otifs}, a principled framework that shifts detection from linear sequences to graph-structured manifolds. We first provide a theoretical
arXiv cs.CY 2h ago Transparency

Secure human oversight of AI: Threat modeling in a socio-technical context

arXiv:2509.12290v3 Announce Type: replace-cross Abstract: Human oversight of AI is promoted as a safeguard against risks such as inaccurate outputs, system malfunctions, or violations of fundamental rights, and is mandated in regulation like the European AI Act. Yet debates on human oversight have largely focused on its effectiveness, while overlooking a critical dimension: the security of human oversight. We argue that human oversight creates a new attack surface within the safety, security, an
arXiv cs.CY 2h ago Regulation