23:20 UTC
Company · updated daily

Google & DeepMind

Google's AI ethics record spans DeepMind's Frontier Safety Framework, Gemini's rollout controversies, ongoing antitrust scrutiny, and the AI Overviews accuracy disputes — tracked daily.

Hearsay: Vision-Language Medical Diagnoses Without an Image

arXiv:2607.26886v1 Announce Type: cross Abstract: When asked to describe a medical image that was never attached, frontier vision-language models do not abstain: they confabulate a diagnosis. We show that this confabulation is not random. It is structured by who the patient is said to be. Across chest X-ray, brain MRI, and dermatology, Claude Opus-4.7, GPT-5.4, and Gemini-3.1-Pro are each queried with only a demographic descriptor and no image, and changing the descriptor systematically shifts t
arXiv cs.CY 19h ago Research Healthcare

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

arXiv:2503.10647v2 Announce Type: replace-cross Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dimensions: consistency under rephrased inputs, susceptibility to irrelevant prompt content, and responsiveness to added clinical context. We designed 52 clinical scenarios and modified each under controlled conditions. For consistency, scenarios were rephrased with demographic, wording, and exam
arXiv cs.CY 19h ago Research Healthcare

Discover what’s next for AI, from the SaaS reckoning to the agent security gap, at TechCrunch Disrupt 2026

At TechCrunch Disrupt 2026, the AI Stage is back to dig into the single hottest topic in the community for the past few years, presented by Google for Startups.
TechCrunch yesterday News Agents & autonomy

Performance of 5 Large Language Models in Perioperative Consultation for Pediatric Hypospadias: Cross-Sectional Comparative Study

Background: Hypospadias is a common congenital malformation requiring surgery. Caregivers face substantial perioperative information needs, and large language models (LLMs) offer a potential health education channel, but their performance in pediatric urology and the relation between citation accuracy and clinical content safety lack systematic evaluation. Objective: This study aimed to evaluate 5 LLMs (ChatGPT-4o, Gemini-2.5-Pro, OpenEvidence, Zhipu Qingyan, and DeepSeek) for pediatric hypospad
JMIR (Journal of Medical Internet Research) yesterday Research HealthcareChildren & education

Hearsay: Vision-Language Medical Diagnoses Without an Image

When asked to describe a medical image that was never attached, frontier vision-language models do not abstain: they confabulate a diagnosis. We show that this confabulation is not random. It is structured by who the patient is said to be. Across chest X-ray, brain MRI, and dermatology, Claude Opus-4.7, GPT-5.4, and Gemini-3.1-Pro are each queried with only a demographic descriptor and no image, and changing the descriptor systematically shifts the diagnosis returned. Claude concentrates sharply
arXiv cs.AI yesterday Research Healthcare

Google makes Gemini Spark AI agent available to Hongkongers as it lowers geofences

Google on Wednesday launched its artificial intelligence agent Gemini Spark in the Hong Kong market, giving local users direct access to a smart assistant to manage complex digital workflows. The launch came months after the American tech giant’s decision in March to lift regional geofences for generative AI services, starting with the Gemini chatbot. Hong Kong users can now access Gemini without using a virtual private network or third-party platform. The roll-out of the Spark agent echoes an..
SCMP Tech (HK/CN) yesterday News Agents & autonomy

Google's SynthID watermark is hard to break, but it doesn't solve AI disinformation

Deciding what's real on the Internet won't be easy in the future.
Ars Technica yesterday News Misinformation

Google DeepMind dismantles Nobel-winning AlphaFold team in strategy shift

Landmark project that solved protein folding gives way to a wider race to build AI systems for scientific discovery
Financial Times Technology (headlines) yesterday News Biotech

AI company employees petition US government for regulation

A thousand workers from OpenAI, Anthropic, Google, Meta and more have signed the letter.
Engadget AI 2d ago News Regulation

Despite AI hype, Google's data shows workers aren't automating themselves away

Analysis of 15 million real AI interactions finds most tasks at most jobs are unaffected.
Ars Technica 2d ago News Jobs & economy

AI leaders sign a statement asking the government to do something about automated AI

Employees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowdown of sorts for frontier AI development - or at least a speed-up of global coordinated governance efforts. "Al could help create a dramatically better […]
The Verge 2d ago News Regulation

Amazon, Meta and Microsoft face skeptical investors this week after Google report sparked sell-off

Alphabet's free cash flow has turned negative, and the company lifted its capital spending forecast. Slower-growing cloud rivals report this week.
CNBC Technology 2d ago News Finance, VC & PE

AI’s finally expensive enough to make Wall Street nervous

It's earnings season, and investors got an unpleasant surprise from Google: an increase on its spending estimate, to as much as $205 billion - from the last quarter's projection of up to $190 billion. Even the lower end of Google's new projected range - $195 billion - is much more than the company had previously […]
The Verge 2d ago News Finance, VC & PE

Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA

In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing on geometry clipping. In this evaluation, a custom exploration agent navigates a game level to collect visual observations, while the automatic annotation pipeline provides frame-level clipping labels. This setup allows us to evaluate recent VLMs on a controlled anomaly detection task without manual annotation. We benchmark six recent VLMs (Gemini
arXiv cs.AI 2d ago Research Agents & autonomy

Gemini API Managed Agents: 3.6 Flash, hooks, and more

We’re announcing even more new capabilities in Managed Agents in Gemini API so developers can build reliable, production-ready agents.
Google AI Blog 2d ago Field notes Agents & autonomy

I noticed a customer review I think might be fake. In Australia, are businesses allowed to do this?

False testimonials are prohibited under Australian consumer law, writes policy professional Kat George, but they are also very hard to police Read more Australian customer service questions While looking at the website for NannyLane Australia, which bills itself as “the Uber for Nannies”, I noticed some glowing testimonials apparently written by happy parents who’d used their services. But a simple reverse Google image search of “James & Emily T” from Sydney led me to a stock image of a couple,
The Guardian 2d ago News Regulation

(g+) Artificial Intelligence: AI companies spend record sums on Washington lobbying

Rising expenditure from OpenAI, Anthropic, Google and Microsoft reflects growing battle over federal policy Von Michael Taffe ( Wirtschaft , KI )
Golem (DE) 2d ago News Regulation

Your old Google Pixel smartphone could be repurposed in a data center

Keeping phones out of landfills and doing useful things is a good thing.
Engadget AI 2d ago News Environment

Nvidia invests in Ilya Sutskever's AI lab, shifting SSI away from Google chips

Nvidia is pouring what it calls a "substantial" sum into Safe Superintelligence (SSI), the AI lab run by Ilya Sutskever, OpenAI's former chief scientist. The article Nvidia invests in Ilya Sutskever's AI lab, shifting SSI away from Google chips appeared first on The Decoder .
The Decoder 2d ago News Finance, VC & PE

Google Android Antitrust Ruling Shows the Next Battle is Over Who Controls Trust

Tech Policy Press 2d ago News

Why has Nitin Gadkari sued Meta, X and Google over AI deepfakes?

The Bombay High Court will hear Gadkari's plea seeking Rs 11 crore in damages and sweeping takedown orders against AI-generated content. The post Why has Nitin Gadkari sued Meta, X and Google over AI deepfakes? appeared first on MEDIANAMA .
MediaNama (IN) 2d ago News Misinformation

Socioeconomic Inference in LLM Medical Triage: Same Symptoms, Different ZIP Code

arXiv:2607.22605v1 Announce Type: new Abstract: We investigate whether large language models alter medical triage recommendations for identical symptoms when only the patient's socioeconomic status (SES) varies. Using three deployment-tier models (Gemini 3.5 Flash, Claude Sonnet 4.6, GPT-5.4-mini), we hold a single neurological symptom profile fixed and vary the SES signal along two channels: explicit (insurance status, occupation, housing) and implicit (a US ZIP code, with no other socioeconomi
arXiv cs.CY 2d ago Research HealthcareFinance, VC & PE

An opinionated guide to which AI to use to do stuff

An opinionated guide to which AI to use to do stuff It's interesting watching the evolution of Ethan Mollick's guide over time. A year ago it was still all about chat - ChatGPT, Claude, Gemini - with o3, Claude 4 Opus, and Gemini 2.5 Pro as the models and Deep Research as a useful alternative mode. Today it's much more about agentic systems - "where the AI is capable of doing the equivalent of many hours of real human work in one go". Gemini has fallen off Ethan's list, since Google still doesn’
Simon Willisons Weblog 3d ago Field notes Agents & autonomy

Google promised not to show ads based on your emails. AI could undo that

Since 2017, Google has promised not to show ads based on users’ Gmail messages. While that policy isn’t changing today, Google may be laying the groundwork for personalized ads based on AI insights from users’ inboxes. Google’s support documents already acknowledge that it may show ads based on users’ interactions with AI, and with new Personal Intelligence AI features in Google Search, those interactions can now include data from Gmail. Let’s say, for instance, that you’re researching new cars.
Fast Company Tech 3d ago News Regulation

Pluralistic: How the EU can punish Google (despite Trump) (27 Jul 2026)

Today's links How the EU can punish Google (despite Trump): Grant me the courage to change the things I can. Hey look at this: Delights to delectate. Object permanence: Chilling Effects; Billy Bragg v Myspace; Glenn Beck calls murdered Norwegian children "Hitler Youth"; Photog sues Getty for $1b copyfraud; Best paid CEOs perform worst; Olympics v "Olympics"; Alberta tar sands v "hot lesbians"; IoT security apocalypse; 3 Little Pigs in pidgin; The nasty party; Copyright extortionist infringed fel
Pluralistic (Cory Doctorow) 3d ago Field notes Copyright & IPChildren & education

AI companies spend record sums on Washington lobbying

Rising expenditure from OpenAI, Anthropic, Google and Microsoft reflects growing battle over federal policy
Financial Times Technology (headlines) 3d ago News Regulation

Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science

arXiv:2607.22513v1 Announce Type: new Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor transparent. We tested how four major LLM families (Claude, Grok, GPT, Gemini) evaluate ethnonationalist pseudo-science derived from Frank Salter's biosocial framework across four temporal snapshots (October 2025-February 2026), via both API and web interfaces. Grok's Fast versions (which power the de
arXiv cs.CY 3d ago Research Transparency

US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns

The Trump administration is planning targeted bans on Chinese AI models rather than a blanket ban. After public pressure, OpenAI and Google DeepMind signed an open letter opposing regulation of open-weight models, yet OpenAI and Anthropic continue to lobby privately for those same restrictions amid security concerns and powerful business interests. The article US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns appeared first on Th
The Decoder 4d ago News Regulation

How to make Google Maps and Apple Maps show you the actual fastest route, every time

When it comes to driving, most of us want to get to our destination as quickly (and safely) as possible. After all, given how bad sitting behind the wheel is for your health , who wants to be in a car any longer than they have to be? Most of us might expect that, when it comes to getting directions, the world’s two biggest mapping apps, Google Maps and Apple Maps , would show us the route that gets us there as fast as possible. But that’s not always the case. Here’s how to make Google Maps and A
Fast Company Tech 5d ago News Healthcare

Trump vows to investigate EU over fining of US tech companies

The US president says fines against Google, as well as Apple, Meta and Amazon, should be "entirely reversed."
BBC Technology 6d ago News Finance, VC & PE

Bond market anxiety is growing over AI capex budgets

Google, Amazon and Meta are seeing credit spreads widen as fixed-income investors demand more reward on companies they lend to.
CNBC Technology 6d ago News Finance, VC & PE

EU strikes Google with $1 billion fine in major crackdown on antitrust regulations even as Trump threatens retaliation

Despite warnings from US President Trump about penalties on American tech companies, the European Union issued a $1 billion fine on the tech giant over its Google Play search.
Fortune AI 6d ago News Regulation

Trump fires back at EU over Google's $1B fine, launches probe

President Trump on Friday slammed the European Union for fining Google over allegedly violating its digital competition law, stating the "illegal and highly discriminatory practice" will be probed in a trade investigation. "The European Union is at it again and, as usual, taking direct aim at GREAT American Companies!" Trump wrote on Truth Social. "This...
The Hill Technology 6d ago News Bias & fairnessRegulation

Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science

Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor transparent. We tested how four major LLM families (Claude, Grok, GPT, Gemini) evaluate ethnonationalist pseudo-science derived from Frank Salter's biosocial framework across four temporal snapshots (October 2025-February 2026), via both API and web interfaces. Grok's Fast versions (which power the default user experience on X) consistently assigne
arXiv cs.AI 6d ago Research Transparency

Team uses AlphaFold AI to redesign gene-editing proteins to make them safer

Google's AlphaFold can help ID what parts of a gene editing protein enable mistakes.
Ars Technica 6d ago News Biotech

Meta is making its AI chatbot more like an assistant

Meta is upgrading its AI chatbot with new productivity features in a bid to compete with rivals like Gemini, ChatGPT, and Claude. The update will allow Meta AI to tap into your calendar to help you plan events and generate daily briefings, as well as perform in-depth research that you can steer as it progresses. […]
The Verge 6d ago News Jobs & economy

AWS, Google Cloud, Microsoft Azure, and Cloudflare now all offer agent sandboxes. None built them the same way.

Google Cloud announced it had put Cloud Run sandboxes into public preview earlier this month at WeAreDevelopers World Congress in The post AWS, Google Cloud, Microsoft Azure, and Cloudflare now all offer agent sandboxes. None built them the same way. appeared first on The New Stack .
The New Stack AI 6d ago News Agents & autonomy

If open weight models are the future, U.S. AI companies are going to have a hard time

Top executives at leading Western AI companies are increasingly warning about the safety and national security risks posed by Chinese open-weight frontier models. What they tend not to mention is that these models are improving rapidly and, because they are freely available, pose a serious threat to Western labs’ business models. The most powerful models from U.S. labs such as OpenAI , Anthropic , and Google DeepMind are closed, meaning the parameters that shape their outputs are kept secret. Ge
Fast Company Tech 6d ago News Military & security

[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model

A HUGE win for BFL!
Latent Space 6d ago Field notes Agents & autonomy

Do language families matter? Evaluating LLMs for sentiment analysis through a hierarchical cross-lingual lens

Social media sentiment analysis has become one of the most significant instruments for understanding the opinion of the population in the spheres of healthcare, politics, and education. Yet, large language models (LLMs) remain unevenly distributed in their linguistic coverage, failing to adequately serve a large portion of the world's languages. This study evaluates five state-of-the-art LLMs: GPT-4o, Gemini 2.0 Flash, DeepSeek-V3, Mistral Large, and Claude 3.7 Sonnet on three-class sentiment cl
Frontiers in Artificial Intelligence 6d ago Research HealthcareChildren & education
Older →