Skip to content

#Trend

Oct 8

TodayOct 8Thu7 items
  1. Tessl BlogAI score42

    Tessl Says Merge Rate Shows Whether AI Adoption Is Real

    Tessl argues that an AI-native organization collapses the handoff between people who own outcomes and the work itself, so product managers and designers can execute changes through agents. It says PR count and token spend are insufficient measures, and that merge rate better shows whether the new workflow is working. The article also says the boundary should follow decision authority, with engineers still owning architecture and data models.

  2. Tessl BlogAI score38

    AI DevCon NYC Focuses on Software Factories for Scaling Agentic Development

    AI DevCon New York, running November 2–4 at Industry City in Brooklyn, centers its program on software factories, the systems needed to make agentic development repeatable, trustworthy and scalable. The article argues that moving from one developer using an agent to an engineering organization requires layers covering context and skills, harnesses and tools, orchestration, verification and evaluation, and feedback.

  3. MIT News · AIAI score34

    MIT's Sasha Rakhlin outlines how universities should respond to AI in research and training

    MIT Statistics and Data Science Center director Sasha Rakhlin argues that AI progress is fastest where results can be verified quickly, citing a model reaching gold-medal level at the International Mathematical Olympiad a year before models produced new research results. He says departments should reconsider how they reward work, emphasizing question-asking, replication, and disclosure of AI's role in a researcher's contributions. He also urges universities to build shared lab infrastructure that captures failed experiments and tacit expertise.

  4. Epoch AI · The Epoch BriefAI score49

    Epoch AI's October 2026 Brief Covers AI Agents, Falling Costs, and China's Chip Exposure

    Epoch AI estimates the AI chips shipped through 2027 could support about 30 to 170 million concurrent frontier-model agents, or nearly 2 billion with cheaper models. Its researchers find the cost of a fixed level of AI performance has fallen about 47% per quarter over the past three years. The newsletter also reports China's semiconductor supply-chain exposure is 2.7 times that of the US.

  5. Anthropic NewsroomAI score46

    Anthropic Updates Claude Usage Policy, Effective November 12, 2026

    Anthropic has published a 2026 update to its Usage Policy, taking effect November 12, mostly to clarify existing rules for longer, more autonomous Claude work. The changes consolidate deceptive-campaign prohibitions into a new section, narrow the elections rules to voter deception and disruption, and explicitly ban weapons-related software and surveillance tools. Requirements for high-risk uses and for models connected to autonomous physical hardware were also tightened.

  6. MIT News · AIAI score24

    MIT's Christina Delimitrou uses machine learning to make data centers more efficient

    MIT associate professor Christina Delimitrou is applying machine learning to make large-scale data centers more efficient, secure, and reliable, rethinking how servers and networking equipment operate. Her group redesigns outdated cloud systems, manages shared hardware resources, and creates streamlined server architectures so operators can extract more computing power from existing hardware. She also uses AI to help programmers find and fix problems in cloud-based applications, reducing downtime that hampers performance and drains resources.

  7. Artificial Analysis ArticlesAI score50

    Harvey LAB-AA v1.1 adds hallucination checks to legal AI benchmark

    Harvey LAB-AA v1.1 adds hallucination checks that audit every model deliverable against task source documents, with material hallucinations zeroing a task's score. GPT-6 Astra averaged 0.03 material hallucinations per task across 120 tasks, while Gemini 3.8 Flash averaged 13.96. Harvey uses GPT-6 Sol (high) as the hallucination checker, separate from its three-judge rubric panel.

Oct 7

Oct 7Wed
  1. Waymo BlogAI score42

    Sober Drivers Still Face Nearly 4x Nighttime Fatal Crash Risk, Waymo Study Finds

    Waymo research found that even fully sober human drivers face nighttime fatal crash risk 3.1 to 3.9 times higher than daytime risk, pointing to systemic hazards beyond impairment. The study used an exposure reconstruction model across the 50 most populous U.S. urban areas, showing removing alcohol-involved drivers lowers the average urban fatal crash rate by 23%, from 1.42 to 1.10 per 100 million miles.

  2. Waymo BlogAI score46

    Waymo Closes $5 Billion Debt Financing to Fund Expansion

    Waymo closed a $5 billion term loan, its first debt financing, with PIMCO, Blackstone, and Sixth Street as lead syndicated lenders and Goldman Sachs as sole lead bookrunner. The debt complements a $16 billion equity investment earlier this year and gives the company added financial flexibility to expand its fully autonomous ride-hailing service across the United States and internationally.

  3. Google ResearchAI score62

    Google Research finds AI boosts patent drafting but junior lawyers' gains vanish without it

    A Google Research field experiment with 133 patent lawyers found AI tool access raised drafting scores by 0.34 to 0.38 standard deviations over three months. When the tool was removed for a redlining task, only senior lawyers kept an advantage of 0.45 SD, while junior lawyers showed no discernible improvement. The authors argue that tools which boost current output must not stop junior professionals from building the judgment that senior experts rely on.

    AIWhy it matters: The field experiment separates AI's short-term productivity gains from skill retained after the tool is removed, which matters for training junior professionals.

  4. AWS Machine Learning BlogAI score38

    Agentic Automation Business Cases Need to Count More Than Saved Hours

    AWS Machine Learning Blog argues that the traditional hours-saved ROI model, built for rule-based RPA, misses most of the value of agentic automation. It proposes an Agentic Value Model covering time savings, exception handling, decision quality, and change resilience, with value counted only when tied to a defined P&L mechanism and owner.

  5. Azure BlogAI score34

    Microsoft Uses AI Agents to Speed Azure Cloud Infrastructure Supply Chain Planning

    Microsoft's Azure Hardware Systems and Infrastructure team is applying AI agents across its infrastructure lifecycle, starting with cloud supply chain demand planning. Following a "Lean before AI" approach, the company reports that multi-agent workflows cut planning work that took five to seven business days to hours, with approximately 50% less manual effort and cycle time down up to 75% in selected workflows. Microsoft says it is extending the approach to fulfillment, logistics, and fleet operations while keeping human judgment central.

  6. ElevenLabs BlogAI score14

    Contact center automation guide explains AI tools for faster customer support

    Contact center automation uses AI to handle customer support workflows with little or no human intervention, including voice, chat, and email. Unlike traditional IVR systems, AI contact center software understands intent, retrieves customer data, and routes complex cases to human agents. The guide cites Klarna, Rohlik, and Getmobil deployments of ElevenAgents, with Klarna offering voice support to 35 million US customers.

Oct 6

Oct 6Tue
  1. Google · Innovation & AIAI score42

    Google Study Tests AI-Guided Blind Sweep Ultrasounds for Pregnant Women in Kenya and Chicago

    Google researchers, working with Northwestern Medicine and Jacaranda Health, trained healthcare workers to perform "blind sweep" ultrasounds analyzed by machine learning models. The models estimated gestational age and fetal presentation as accurately as a trained sonographer in a study of 1,000 mothers each in Nairobi and Chicago. The AI processes results on the device, so it needs no electricity supply or Wi-Fi.

  2. Google ResearchAI score51

    Google's PDFM location embeddings improve five global public health tasks

    Google Research reports that Population Dynamics Foundation Model (PDFM) embeddings, built from search trends, mobility, built environment, and weather signals, were tested by partners across five public health tasks. The embeddings improved results in cross-border MMR vaccination coverage, dengue forecasting, postpartum depression screening, and cholera outbreak prediction, and matched census inputs for cardiovascular mortality nowcasting.

  3. MIT News · AIAI score22

    MIT launches MIT for America initiative to strengthen STEM education nationwide

    MIT has launched MIT for America, a strategic initiative to strengthen STEM education across the United States, from kindergarten through community college. Its programs target mathematical thinking and problem solving, living and working with AI, and hands-on design and fabrication. The initiative builds on the MIT for America Calculus Project, which pairs MIT students and alumni with school districts for calculus tutoring.

  4. Luma AI NewsAI score22

    Claymation AI Prompts for Stop-Motion Looks Without a Physical Rig

    The article explains how to write AI video prompts that produce authentic claymation and stop-motion looks without physical sculpting or frame-by-frame photography. It stresses specifying material properties such as polymer clay with visible thumbprints, movement rhythm such as a 12fps animation feel, and negative prompts such as "no photorealism" to suppress glossy 3D defaults. It also includes 15 example prompts organized by material, texture, and category.

Oct 4

Oct 4Sun
  1. Epoch AIAI score62

    OpenAI researchers' coding-agent usage is doubling about monthly, Epoch AI reports

    OpenAI researchers' daily coding-agent usage, valued at API prices, rose from under $1 in January 2026 to $601 for the median researcher by mid-August. The 90th-percentile researcher reached over $7,000 per day, and both groups show doubling times of roughly one month. Epoch notes these are API-list values, not OpenAI's internal costs.

    AIWhy it matters: The figures show internal coding-agent usage growing fast enough to matter for research cost, though they measure API-list value rather than OpenAI's actual spending.

Oct 2

Oct 2Fri
  1. Epoch AI · The Epoch BriefAI score62

    Epoch AI estimates 2026 compute could run hundreds of millions of AI agents

    Epoch AI estimates that compute built from projected 2025 to 2027 high-bandwidth memory shipments could support tens to hundreds of millions of frontier AI agents, or billions of cheaper ones. Running nonstop, the top-tier agents would match the working hours of 140 million to 700 million full-time employees, and the central DeepSeek V4 Pro estimate of about 1.9 billion agents would match 8 billion workers.

    AIWhy it matters: The estimate converts memory shipments into agent capacity and revenue ranges, showing how hardware supply could translate into labor and sales if demand keeps up.

  2. Cloudflare Blog · AIAI score36

    Civil society groups automate their work on Cloudflare with $7.5 million in credits

    Dozens of civil society organizations have built AI-powered tools on Cloudflare's developer services using more than $7.5 million in Cloudflare credits. Cloudflare says its serverless architecture, Workers AI and AI Gateway let non-technical teams build and scale applications without dedicated GPU infrastructure, while providing built-in security protections.

Oct 1

Oct 1Thu

Sep 30

Sep 30Wed
  1. Anthropic ResearchAI score62

    Anthropic study finds robots can do most physical tasks but rarely cost-effectively

    Anthropic's research rates how well present-day robots can perform US job tasks, finding they can do 74% of physical tasks, or 34% of working hours, mostly in limited settings. Robots are cost-competitive for only 0.3% of job tasks, and at a 3% annual price decline it would take about 40 years to reach 10%. The report also finds robot-exposed jobs tend to pay less and be more physically demanding than LLM-exposed jobs.

    AIWhy it matters: The report separates current robot capability from cost, showing that physical automation is technically broad but economically narrow for now.

Sep 29

Sep 29Tue

Sep 28

Sep 28Mon
  1. Microsoft ResearchAI score30

    Microsoft Research Asia – Singapore marks one year advancing AI research, partnerships and talent

    Microsoft Research Asia – Singapore, opened July 24, 2025 as Microsoft's first Southeast Asian research lab, reports progress after its first year. The lab's work spans next-generation AI models and agentic systems, domain-specific AI for real-world impact, AI-native research practices, and ecosystem and talent development. Its healthcare collaborations on multimodal and agentic AI for clinical decision-making are being deployed through partnerships across Singapore's healthcare ecosystem.

  2. Google · Gemini appAI score38

    See what 4 builders are making with Gemini 3.8 Flash

    Google says Gemini 3.8 Flash, its most intelligent workhorse model, improves on 3.7 Flash in software engineering, agentic tasks, and multistep reasoning by running extra reasoning steps and calling tools iteratively. The post highlights four community builds, including a model rocket simulation, an animated ink-painting effect, a 3D dinosaur skeleton, and an interactive automatic transmission simulation. Developers can try the model through Google Antigravity and Google AI Studio.

Sep 25

Sep 25Fri

Sep 24

Sep 24Thu
  1. GitHub Blog · AI & MLAI score46

    GitHub Copilot app's canvases argue chat is the wrong AI interface

    GitHub argues that chat is often the wrong interface for AI work and proposes customizable "canvases" inside the GitHub Copilot app. Canvases are full-stack applications running without browser chrome that can communicate bi-directionally with the Copilot agent and execute code locally. The post cites examples including a Connect 4 game, a Winget package manager UI, and a SQLite database interface.

  2. Epoch AI · The Epoch BriefAI score45

    Huawei Trails Nvidia by About Four Years in AI Chip Performance and Output

    Huawei will likely remain about four years behind Nvidia in AI chip performance and production through 2030, Epoch AI estimates. Its flagship Ascend 950 delivers roughly half the performance of Nvidia's 2022 H100, and Huawei is projected to produce about 1.5 million chips in 2026 versus Nvidia's roughly 6 million, leaving it about 25 times behind in total compute.

  3. Google Cloud · AI & Machine LearningAI score25

    Latin American midsize businesses adopt Google Cloud Gemini Enterprise to build AI agents

    AI adoption among Latin American small and medium-sized businesses has surged, with Google Cloud AI tool users growing 8x year-over-year across the region and 9x in Brazil. Companies such as AdGoat, Angelus, and BunkerDB are using Gemini Enterprise and Cloud infrastructure to automate content analysis, project management, and marketing workflows. BunkerDB reports cutting creative turnaround times from weeks to hours and reducing cost per lead by up to 25%.

  4. Google · Innovation & AIAI score62

    Google's Project Suncatcher will test TPUs in orbit on a prototype satellite

    Google's Project Suncatcher will launch a prototype satellite on the Transporter-18 rideshare mission with SpaceX to test how its TPUs handle spaceflight. Initial ground tests showed the Trillium TPUs survived vibration and a radiation dose greater than a five-year space mission would deliver. Google says cooling with heat pipes and radiators and laser links between satellites in 2027 remain open engineering challenges.

    AIWhy it matters: The source reports concrete radiation, vibration, and cooling test results for TPUs, showing what space-based AI compute still has to solve.

Sep 23

Sep 23Wed
  1. Waymo BlogAI score39

    Waymo Outlines Vision for Fully Autonomous Passenger Service in London

    Waymo has set out its vision for a fully autonomous passenger service in London, seeking approval under the Automated Passenger Services permitting scheme, which requires Department for Transport approval and Transport for London consent. The company cites a 94% reduction in serious or fatal injury crashes where it operates and says the service would complement public transport and support the Mayor's Transport Strategy.

  2. Microsoft ResearchAI score60

    Microsoft Research shows offloading robot AI inference improves performance and battery life

    Microsoft Research reports that running physical AI inference on onboard GPUs can limit robot performance and battery life, while offloading inference to edge or cloud GPUs improved results in mobile manipulation tests. In its evaluation, smaller onboard GPUs slowed mapping and planning by up to 383% compared with an A100, and large onboard GPUs such as Jetson Thor drained robot batteries by up to 160%.

    AIWhy it matters: The study measures how offloading robot inference to edge or cloud GPUs changes task success, battery life, and model size, offering evidence for infrastructure design.

Sep 18

Sep 18Fri
  1. Google · AI blogAI score29

    Google adds Nobel laureate Philippe Aghion and new directors to AI & Economy research team

    Google's AI & Economy Research Program has added Nobel laureate Philippe Aghion as an Academic Advisor and Ajay Agrawal as a Visiting Fellow, with Anu Madgavkar and Daniel Rock named Directors. The program covers the future of work, productivity and growth, global technology diffusion, and AI's impact on scientific discovery, and builds on the recently launched AI & Economy ATLAS v1.0.

Sep 16

Sep 16Wed
  1. Waymo BlogAI score55

    Waymo plans fully autonomous ride-hailing in Singapore by 2028

    Waymo announced plans to launch its fully autonomous, all-electric ride-hailing service in Singapore through the Waymo app in 2028, in partnership with the Ministry of Transport and Land Transport Authority. The rollout begins with initial Jaguar I-PACE vehicles arriving in the coming months, followed by a 2027 readiness phase of manual driving to adapt the Waymo Driver to local conditions.

Sep 15

Sep 15Tue
  1. Google · AI blogAI score14

    Google spotlights AI projects for disease, disaster prediction, education, and economic opportunity

    Google is showcasing how partners are applying AI to societal challenges, including making disease detectable, treatable, and preventable, predicting natural disasters, expanding learning, and unlocking economic opportunities. The source describes these efforts as measurable real-world impact but provides no specific models, figures, or benchmarks.

Sep 9

Sep 9Wed
  1. Google DeepMind · YouTubeAI score38

    How AI is transforming weather prediction, featuring WeatherNext 3

    Google DeepMind's Peter Battaglia discusses how machine learning is changing global weather forecasting, including early warnings for storms such as Hurricane Melissa. The episode covers traditional physics-based models versus AI models and probabilistic forecasting, and highlights WeatherNext 3 as Google DeepMind's most advanced global weather AI model yet.

Sep 6

Sep 6Sun
  1. Google DeepMind · The KeywordAI score24

    Google DeepMind backs 16 Asia-Pacific green AI projects in inaugural accelerator

    Google DeepMind selected 16 organizations for its inaugural Accelerator: AI for the Planet (APAC) cohort to scale AI-powered environmental solutions. Participants span biodiversity monitoring, sustainable agriculture, and climate and carbon projects across countries including New Zealand, Singapore, Indonesia, India, and Japan. Over three months, they will receive access to Google's AI stack, including frontier models, plus mentorship from company experts.

Aug 27

Aug 27Thu
  1. Epoch AI · The Epoch BriefAI score62

    Anthropic and OpenAI's 2026 revenue growth raises the question of how long it lasts

    Combined annualized revenue for OpenAI and Anthropic reached $105 billion by August 2026, up 3.5 times from $30 billion at the start of the year. The author argues the key question is whether this growth comes from continued capability progress or from diffusion that will saturate. At the 3 times annual pace, frontier AI revenue would take about six years to reach today's world economy size.

    AIWhy it matters: The piece tests whether OpenAI and Anthropic's hypergrowth reflects a temporary coding-agent spike or durable progress, using revenue scale to frame the question.

Aug 24

Aug 24Mon
  1. Epoch AI · The Epoch BriefAI score58

    Epoch AI says US GDP underestimates AI growth by missing Nvidia's value

    Epoch AI argues US GDP growth over the last year was underestimated by about 0.3 percentage points because value from fabless chipmakers like Nvidia goes unrecorded. The report says no goods export, IP export, service export, or merchanting category captures Nvidia's value-add, and the Bureau of Economic Analysis confirmed the analysis. If Nvidia's growth continues, the gap could reach almost two percentage points per year by 2028.

Aug 14

Aug 14Fri
  1. Epoch AI · The Epoch BriefAI score42

    Epoch AI lists nine big AI questions its benchmarks aim to answer

    Epoch AI outlines nine open questions about AI capabilities, including whether AI can take over full jobs and whether benchmark scores are correlated. The author says Epoch's benchmarking work is built to help answer them, citing examples such as MirrorCode, Remote Labor Index, and the Epoch Capabilities Index (ECI). The post notes that benchmark scores are highly correlated across domains, and that ECI growth trends can help detect whether AI capability progress has accelerated.