Company · updated daily
Google & DeepMind
Google's AI ethics record spans DeepMind's Frontier Safety Framework, Gemini's rollout controversies, ongoing antitrust scrutiny, and the AI Overviews accuracy disputes — tracked daily.
Hearsay: Vision-Language Medical Diagnoses Without an Image
arXiv:2607.26886v1 Announce Type: cross Abstract: When asked to describe a medical image that was never attached, frontier vision-language models do not abstain: they confabulate a diagnosis. We show that this confabulation is not random. It is structured by who the patient is said to be. Across chest X-ray, brain MRI, and dermatology, Claude Opus-4.7, GPT-5.4, and Gemini-3.1-Pro are each queried with only a demographic descriptor and no image, and changing the descriptor systematically shifts t
The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness
arXiv:2503.10647v2 Announce Type: replace-cross Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dimensions: consistency under rephrased inputs, susceptibility to irrelevant prompt content, and responsiveness to added clinical context. We designed 52 clinical scenarios and modified each under controlled conditions. For consistency, scenarios were rephrased with demographic, wording, and exam
Discover what’s next for AI, from the SaaS reckoning to the agent security gap, at TechCrunch Disrupt 2026
At TechCrunch Disrupt 2026, the AI Stage is back to dig into the single hottest topic in the community for the past few years, presented by Google for Startups.
Performance of 5 Large Language Models in Perioperative Consultation for Pediatric Hypospadias: Cross-Sectional Comparative Study
Background: Hypospadias is a common congenital malformation requiring surgery. Caregivers face substantial perioperative information needs, and large language models (LLMs) offer a potential health education channel, but their performance in pediatric urology and the relation between citation accuracy and clinical content safety lack systematic evaluation. Objective: This study aimed to evaluate 5 LLMs (ChatGPT-4o, Gemini-2.5-Pro, OpenEvidence, Zhipu Qingyan, and DeepSeek) for pediatric hypospad
Hearsay: Vision-Language Medical Diagnoses Without an Image
When asked to describe a medical image that was never attached, frontier vision-language models do not abstain: they confabulate a diagnosis. We show that this confabulation is not random. It is structured by who the patient is said to be. Across chest X-ray, brain MRI, and dermatology, Claude Opus-4.7, GPT-5.4, and Gemini-3.1-Pro are each queried with only a demographic descriptor and no image, and changing the descriptor systematically shifts the diagnosis returned. Claude concentrates sharply
Google makes Gemini Spark AI agent available to Hongkongers as it lowers geofences
Google on Wednesday launched its artificial intelligence agent Gemini Spark in the Hong Kong market, giving local users direct access to a smart assistant to manage complex digital workflows. The launch came months after the American tech giant’s decision in March to lift regional geofences for generative AI services, starting with the Gemini chatbot. Hong Kong users can now access Gemini without using a virtual private network or third-party platform. The roll-out of the Spark agent echoes an..
Google's SynthID watermark is hard to break, but it doesn't solve AI disinformation
Deciding what's real on the Internet won't be easy in the future.
Google DeepMind dismantles Nobel-winning AlphaFold team in strategy shift
Landmark project that solved protein folding gives way to a wider race to build AI systems for scientific discovery
AI company employees petition US government for regulation
A thousand workers from OpenAI, Anthropic, Google, Meta and more have signed the letter.
Despite AI hype, Google's data shows workers aren't automating themselves away
Analysis of 15 million real AI interactions finds most tasks at most jobs are unaffected.
AI leaders sign a statement asking the government to do something about automated AI
Employees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowdown of sorts for frontier AI development - or at least a speed-up of global coordinated governance efforts. "Al could help create a dramatically better […]
Amazon, Meta and Microsoft face skeptical investors this week after Google report sparked sell-off
Alphabet's free cash flow has turned negative, and the company lifted its capital spending forecast. Slower-growing cloud rivals report this week.
AI’s finally expensive enough to make Wall Street nervous
It's earnings season, and investors got an unpleasant surprise from Google: an increase on its spending estimate, to as much as $205 billion - from the last quarter's projection of up to $190 billion. Even the lower end of Google's new projected range - $195 billion - is much more than the company had previously […]
Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA
In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing on geometry clipping. In this evaluation, a custom exploration agent navigates a game level to collect visual observations, while the automatic annotation pipeline provides frame-level clipping labels. This setup allows us to evaluate recent VLMs on a controlled anomaly detection task without manual annotation. We benchmark six recent VLMs (Gemini
Gemini API Managed Agents: 3.6 Flash, hooks, and more
We’re announcing even more new capabilities in Managed Agents in Gemini API so developers can build reliable, production-ready agents.
I noticed a customer review I think might be fake. In Australia, are businesses allowed to do this?
False testimonials are prohibited under Australian consumer law, writes policy professional Kat George, but they are also very hard to police Read more Australian customer service questions While looking at the website for NannyLane Australia, which bills itself as “the Uber for Nannies”, I noticed some glowing testimonials apparently written by happy parents who’d used their services. But a simple reverse Google image search of “James & Emily T” from Sydney led me to a stock image of a couple,
(g+) Artificial Intelligence: AI companies spend record sums on Washington lobbying
Rising expenditure from OpenAI, Anthropic, Google and Microsoft reflects growing battle over federal policy Von Michael Taffe ( Wirtschaft , KI )
Your old Google Pixel smartphone could be repurposed in a data center
Keeping phones out of landfills and doing useful things is a good thing.
Nvidia invests in Ilya Sutskever's AI lab, shifting SSI away from Google chips
Nvidia is pouring what it calls a "substantial" sum into Safe Superintelligence (SSI), the AI lab run by Ilya Sutskever, OpenAI's former chief scientist. The article Nvidia invests in Ilya Sutskever's AI lab, shifting SSI away from Google chips appeared first on The Decoder .
Google Android Antitrust Ruling Shows the Next Battle is Over Who Controls Trust
Why has Nitin Gadkari sued Meta, X and Google over AI deepfakes?
The Bombay High Court will hear Gadkari's plea seeking Rs 11 crore in damages and sweeping takedown orders against AI-generated content. The post Why has Nitin Gadkari sued Meta, X and Google over AI deepfakes? appeared first on MEDIANAMA .
Socioeconomic Inference in LLM Medical Triage: Same Symptoms, Different ZIP Code
arXiv:2607.22605v1 Announce Type: new Abstract: We investigate whether large language models alter medical triage recommendations for identical symptoms when only the patient's socioeconomic status (SES) varies. Using three deployment-tier models (Gemini 3.5 Flash, Claude Sonnet 4.6, GPT-5.4-mini), we hold a single neurological symptom profile fixed and vary the SES signal along two channels: explicit (insurance status, occupation, housing) and implicit (a US ZIP code, with no other socioeconomi
An opinionated guide to which AI to use to do stuff
An opinionated guide to which AI to use to do stuff It's interesting watching the evolution of Ethan Mollick's guide over time. A year ago it was still all about chat - ChatGPT, Claude, Gemini - with o3, Claude 4 Opus, and Gemini 2.5 Pro as the models and Deep Research as a useful alternative mode. Today it's much more about agentic systems - "where the AI is capable of doing the equivalent of many hours of real human work in one go". Gemini has fallen off Ethan's list, since Google still doesn’
Google promised not to show ads based on your emails. AI could undo that
Since 2017, Google has promised not to show ads based on users’ Gmail messages. While that policy isn’t changing today, Google may be laying the groundwork for personalized ads based on AI insights from users’ inboxes. Google’s support documents already acknowledge that it may show ads based on users’ interactions with AI, and with new Personal Intelligence AI features in Google Search, those interactions can now include data from Gmail. Let’s say, for instance, that you’re researching new cars.
Pluralistic: How the EU can punish Google (despite Trump) (27 Jul 2026)
Today's links How the EU can punish Google (despite Trump): Grant me the courage to change the things I can. Hey look at this: Delights to delectate. Object permanence: Chilling Effects; Billy Bragg v Myspace; Glenn Beck calls murdered Norwegian children "Hitler Youth"; Photog sues Getty for $1b copyfraud; Best paid CEOs perform worst; Olympics v "Olympics"; Alberta tar sands v "hot lesbians"; IoT security apocalypse; 3 Little Pigs in pidgin; The nasty party; Copyright extortionist infringed fel
AI companies spend record sums on Washington lobbying
Rising expenditure from OpenAI, Anthropic, Google and Microsoft reflects growing battle over federal policy
Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science
arXiv:2607.22513v1 Announce Type: new Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor transparent. We tested how four major LLM families (Claude, Grok, GPT, Gemini) evaluate ethnonationalist pseudo-science derived from Frank Salter's biosocial framework across four temporal snapshots (October 2025-February 2026), via both API and web interfaces. Grok's Fast versions (which power the de
US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns
The Trump administration is planning targeted bans on Chinese AI models rather than a blanket ban. After public pressure, OpenAI and Google DeepMind signed an open letter opposing regulation of open-weight models, yet OpenAI and Anthropic continue to lobby privately for those same restrictions amid security concerns and powerful business interests. The article US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns appeared first on Th
How to make Google Maps and Apple Maps show you the actual fastest route, every time
When it comes to driving, most of us want to get to our destination as quickly (and safely) as possible. After all, given how bad sitting behind the wheel is for your health , who wants to be in a car any longer than they have to be? Most of us might expect that, when it comes to getting directions, the world’s two biggest mapping apps, Google Maps and Apple Maps , would show us the route that gets us there as fast as possible. But that’s not always the case. Here’s how to make Google Maps and A
Trump vows to investigate EU over fining of US tech companies
The US president says fines against Google, as well as Apple, Meta and Amazon, should be "entirely reversed."
Bond market anxiety is growing over AI capex budgets
Google, Amazon and Meta are seeing credit spreads widen as fixed-income investors demand more reward on companies they lend to.
EU strikes Google with $1 billion fine in major crackdown on antitrust regulations even as Trump threatens retaliation
Despite warnings from US President Trump about penalties on American tech companies, the European Union issued a $1 billion fine on the tech giant over its Google Play search.
Trump fires back at EU over Google's $1B fine, launches probe
President Trump on Friday slammed the European Union for fining Google over allegedly violating its digital competition law, stating the "illegal and highly discriminatory practice" will be probed in a trade investigation. "The European Union is at it again and, as usual, taking direct aim at GREAT American Companies!" Trump wrote on Truth Social. "This...
Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science
Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor transparent. We tested how four major LLM families (Claude, Grok, GPT, Gemini) evaluate ethnonationalist pseudo-science derived from Frank Salter's biosocial framework across four temporal snapshots (October 2025-February 2026), via both API and web interfaces. Grok's Fast versions (which power the default user experience on X) consistently assigne
Team uses AlphaFold AI to redesign gene-editing proteins to make them safer
Google's AlphaFold can help ID what parts of a gene editing protein enable mistakes.
Meta is making its AI chatbot more like an assistant
Meta is upgrading its AI chatbot with new productivity features in a bid to compete with rivals like Gemini, ChatGPT, and Claude. The update will allow Meta AI to tap into your calendar to help you plan events and generate daily briefings, as well as perform in-depth research that you can steer as it progresses. […]
AWS, Google Cloud, Microsoft Azure, and Cloudflare now all offer agent sandboxes. None built them the same way.
Google Cloud announced it had put Cloud Run sandboxes into public preview earlier this month at WeAreDevelopers World Congress in The post AWS, Google Cloud, Microsoft Azure, and Cloudflare now all offer agent sandboxes. None built them the same way. appeared first on The New Stack .
If open weight models are the future, U.S. AI companies are going to have a hard time
Top executives at leading Western AI companies are increasingly warning about the safety and national security risks posed by Chinese open-weight frontier models. What they tend not to mention is that these models are improving rapidly and, because they are freely available, pose a serious threat to Western labs’ business models. The most powerful models from U.S. labs such as OpenAI , Anthropic , and Google DeepMind are closed, meaning the parameters that shape their outputs are kept secret. Ge
[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model
A HUGE win for BFL!
Do language families matter? Evaluating LLMs for sentiment analysis through a hierarchical cross-lingual lens
Social media sentiment analysis has become one of the most significant instruments for understanding the opinion of the population in the spheres of healthcare, politics, and education. Yet, large language models (LLMs) remain unevenly distributed in their linguistic coverage, failing to adequately serve a large portion of the world's languages. This study evaluates five state-of-the-art LLMs: GPT-4o, Gemini 2.0 Flash, DeepSeek-V3, Mistral Large, and Claude 3.7 Sonnet on three-class sentiment cl