Skip to contentSkip to stories

Updated

#Google

Showing low-relevance items too. Hide low-relevance items

May 19

May 19Tue
  1. koray kavukcuogluAI score72

    Google's Gemini 3.5 Flash beats Gemini 3.1 Pro on coding and agentic benchmarks

    AIGoogle's Gemini 3.5 Flash outperforms Gemini 3.1 Pro on Terminal-Bench 2.1 (76.2%), GDPval-AA (1656 Elo), and MCP Atlas (83.6%). The post also claims it is 4x faster than other frontier models, or 12x in Antigravity, and reports 83.6% on MMMU-Pro for multimodal performance.

    Why it matters: The post gives specific benchmark scores against Gemini 3.1 Pro, letting readers compare coding, agentic, and multimodal results directly.

    Image from @koraykv's post
  2. koray kavukcuogluAI score72

    Google rolls out Gemini 3.5 Flash globally across consumer, developer, and enterprise platforms

    AIGoogle is rolling out Gemini 3.5 Flash globally for consumers in the Gemini app and Search AI Mode. It is also available to developers through the Gemini API, Google Antigravity, and Google AI Studio, and to businesses on the Gemini Enterprise Agent Platform.

    Why it matters: The post shows where each Gemini 3.5 Flash access path goes, from consumer apps to developer and enterprise platforms, which helps readers pick the right entry point.

  3. koray kavukcuogluAI score62

    Google introduces Gemini 3.5 Flash, used with agents to rebuild AlphaZero

    AIAt Google I/O, Google introduced Gemini 3.5 Flash, which the author says has become part of the daily research cycle. The author says a team of agents in Antigravity 2.0 recreated the original AlphaZero paper and built a playable web version from two prompts, coding the reinforcement learning pipeline in JAX/Flax and training a ResNet model via self-play on multi-TPU pods.

    Video from @koraykv's post

May 1

May 1Fri
  1. ReflectionAI score38

    Reflection joins AI coalition on responsible U.S. government deployment

    AIReflection has joined a coalition including AWS, Microsoft, OpenAI, Google, and Nvidia on a framework governing how the U.S. government licenses and deploys AI. The agreement, which includes a non-binding memorandum of understanding with the DoW, commits to safety, red-teaming, and ongoing evaluation and explicitly prohibits unlawful mass surveillance and autonomous weapon use. Reflection says it will keep its commitment to open source while customizing its models for scientists in national labs.

Apr 30

Apr 30Thu
  1. koray kavukcuogluAI score12

    Alphabet named to TIME's 2026 Most Influential Companies list

    AIAlphabet has been recognized in TIME's 2026 Most Influential Companies list, with Google's Koray Kavukcuoglu crediting teams for turning AI breakthroughs into everyday user features. TIME's cover story says Sundar Pichai's 2016 "AI-first" push, including custom chips, Cloud, YouTube, and deep AI research, has paid off.

Apr 23

Apr 23Thu

Apr 22

Apr 22Wed
  1. koray kavukcuogluAI score49

    Google unveils 8th-generation TPUs, with 8t for training and 8i for inference

    AIAt Google Cloud Next this week, Google introduced its 8th-generation TPUs, split into two variants: 8t for massive-scale training and 8i for low-latency inference. Google presents the launch as a milestone in its accelerator roadmap, aimed at optimizing the full AI stack. The post links to a blog with further details on the systems architecture.

Mar 26

Mar 26Thu

Mar 11

Mar 11Wed

Mar 3

Mar 3Tue

Mar 2

Mar 2Mon

Feb 28

Feb 28Sat

Feb 26

Feb 26Thu
  1. Nano Banana 2.1AI score67

    Google introduces Nano Banana 2, its best image generation and editing model

    AINano Banana announces Nano Banana 2, which it describes as its best image generation and editing model yet. The model can be tried in the Gemini app, Google AI Studio, and other places the post does not specify.

    Why it matters: The post names the access points for Nano Banana 2, which helps readers see where the image generation and editing model can be tried.

Feb 25

Feb 25Wed
  1. Quoc LeAI score53

    Google's Aletheia math agent solves 6 of 10 FirstProof problems

    AIQuoc Le announced that Aletheia, a math research agent, autonomously solved 6 of 10 FirstProof problems, the best result in the inaugural challenge. The post says this exceeds last year's IMO-gold achievement and points to a paper and thread for full details. The accompanying figure shows 10 unmodified problems, 6 candidate solutions per agent, and expert evaluation yielding 6 solved problems on a best-of-2 basis.

  2. Quoc LeAI score65

    Aletheia Agent Solves 6 of 10 FirstProof Math Problems Autonomously

    AIGoogle researchers used the Aletheia agent, powered by Gemini 3 Deep Think, to attempt 10 FirstProof challenge problems without modification. The agent operated fully autonomously and solved 6 of the 10 problems, according to the post, with methodology and expert evaluations described in the linked arXiv paper.

    Why it matters: The post gives the autonomous setup and expert-evaluated results for an AI agent on FirstProof math problems, useful for judging how far such systems go on research-level math.

    Image from @quocleix's post

Feb 24

Feb 24Tue

Feb 19

Feb 19Thu
  1. Yi TayAI score78

    Google releases Gemini 3.1 Pro, reporting 77.1% on ARC-AGI-2

    AIGoogle has released Gemini 3.1 Pro, reporting 77.1% on ARC-AGI-2 and more than twice the score of Gemini 3 Pro on that benchmark. The model is rolling out to developers in preview through the Gemini API and Google AI Studio, to enterprises via Vertex AI and Gemini Enterprise, and to consumers in the Gemini app and NotebookLM.

    Why it matters: The post pairs the release with a benchmark table comparing Gemini 3.1 Pro against Gemini 3 Pro, Claude Sonnet 4.6, Claude Opus 4.6, and GPT-5.2 on reasoning and coding tasks.

Feb 18

Feb 18Wed

Feb 16

Feb 16Mon
  1. Nano Banana 2.1AI score14

    Nano Banana shares a highly detailed prompt example for image generation

    AINano Banana (@NanoBanana) posts a sample prompt showing how specific image-generation instructions can be, describing a San Francisco cafe scene with a couple by the window, a sleeping black-and-white French bulldog, and a mirrored "Nano's Cafe" gold lettering on the glass. The post offers the detailed prompt as an illustration of how far users can go with specificity, covering clothing, props, lighting, and poses.

    Image from @NanoBanana's post

Feb 14

Feb 14Sat

Feb 12

Feb 12Thu

Feb 11

Feb 11Wed
  1. Yi TayAI score67

    Aletheia math research agent produces two papers and solves open Erdős problems

    AIYi Tay introduces Aletheia, a math research agent powered by an advanced version of Gemini Deep Think. The post says it produced two publishable papers, one fully automatic and one human-AI collaboration, and solved multiple open Erdős problems. The attached image shows a Google DeepMind paper titled "Towards Autonomous Mathematics Research" with a generator, verifier, and reviser loop.

    Image from @YiTayML's post

Feb 9

Feb 9Mon

Feb 2

Feb 2Mon

Jan 9

Jan 9Fri

Dec 18, 2025

Dec 18, 2025Thu

Dec 17, 2025

Dec 17, 2025Wed