Alphabet, Google's parent company, is reportedly developing a new AI chip aimed at making its Gemini models run far more efficiently.

According to a report surfaced by Bing News, the chip is being designed specifically to help Gemini — Google's flagship family of AI models — operate "much more efficiently." The effort is described as still in development rather than a shipping product.

A separate item carried through Google News, citing bloomingbit, frames the project as a "server AI chip" intended to "ease computing bottlenecks." In plain terms, that points to a chip built for the data centers where AI models actually run, rather than for phones or laptops.

The two descriptions line up: a piece of custom silicon meant to speed up and cut the cost of running Gemini at scale, while relieving the strain that heavy AI workloads put on server infrastructure.

Why does this matter? Running large AI models is enormously expensive, and demand for the specialized chips that power them has strained supply across the industry. If Google can design its own hardware tuned to its own models, it stands to lower its running costs, reduce its dependence on outside chip suppliers, and squeeze more performance out of every server. For everyday users, more efficient chips can translate into faster, cheaper, and more widely available AI features — making this quiet hardware work a meaningful piece of the broader race to make artificial intelligence sustainable to operate.