FLUX 3 Image is now available on Vercel AI Gateway
AIVercel says FLUX 3 Image, from Black Forest Labs, is live on AI Gateway for generating and editing images. It supports up to 10 reference images and 4K output.

Updated
Updated
Items with an AI score under 20 are hidden. Show low-relevance items
AIVercel says FLUX 3 Image, from Black Forest Labs, is live on AI Gateway for generating and editing images. It supports up to 10 reference images and 4K output.

AIGPT-6 Sol (Daybreak Blue, max) has been added to the Artificial Analysis Cyber Index as a trusted-access model and ranks #1 on the Index. Compared with the publicly available GPT-6 Sol, it shows its largest gains on CyberGym-E2E, the benchmark where the most safety refusals are observed.

AIGPT-6 Sol (Daybreak Blue, max) ranks first on the Artificial Analysis Cyber Index at a Cost per Task of $1.77. That is significantly cheaper than other leading models, including Grok 4.7 (xhigh) at $11.67 per task.

AIArtificial Analysis added trusted-access models to its Cyber Index, and GPT-6 Sol (Daybreak Blue, max) now leads the leaderboard. The model is available only through OpenAI's Daybreak program and records no safety blocks, improving 32 points over the publicly available GPT-6 Sol (max). It costs $1.77 per task, below Grok 4.7 (xhigh) at $11.67 per task.

This story has a top pick“GPT-6 Sol Daybreak Blue leads the Artificial Analysis Cyber Index”
AIOpenAI has improved steering so the model reacts much faster to user adjustments, letting people course-correct in real time. The company is also releasing GPT-6.1 Sol ultrafast, which it says works very well alongside the steering improvement.
AIBryce Vincent Gowryluk, 16, had to be rescued by North Shore Rescue after following route directions from the AI chatbot Claude that took him to the base of the Widowmaker Arete, a 1,700-foot wall in British Columbia. He had asked Claude to plan a route from Grouse Mountain to Crown Mountain and back, and he was ill-equipped for the rock climbing the wall requires. Rescue manager Paul Markey warned hikers not to rely blindly on AI for route planning.
AIOpenRouter now offers Sol Ultrafast, a version of GPT-6.1 Sol, available through its platform. According to OpenAI Devs, the Ultrafast mode delivers near-Astra intelligence at up to 8x the speed of Sol Standard and is rolling out in the API, Codex, and ChatGPT Work.
AIAnthropic will classify abusive behavior toward Claude as a violation of its Usage Policy starting November 12, 2026. The change is announced in an official Anthropic post shared by Andrew Curran.

AIGoogle Research is presenting ContinuousBench, a standardized benchmark for measuring knowledge transfer and data contamination in differentially private (DP) synthetic data generation. Alex Bie is giving an encore presentation at the Google booth #107 at COLM 2026 today, at 2:00 PM PT.
AIGoogle proposed FlowAgent, a ReAct-style agent that generates and validates fixes for pre-submit test failures and shows them in its code review tools. Two abstention filters, before and after execution, suppress weak suggestions; in a manual review of 195 real failures, 67.18% of fixes were correct. After the Google-wide launch, it suggested fixes on 295,508 changes, with developers previewing 65,069 and applying 28,554.

AIMustafa Suleyman, Microsoft AI chief, responded to Anthropic's policy change with only an emoji post, offering no detailed comment. The quoted background post says abusive behavior toward Claude will violate Anthropic's Usage Policy effective November 12, 2026.
AISherwin Wu, an OpenAI employee, says the updated Harvey LAB-AA v1.1 benchmark, announced by Artificial Analysis with Harvey, is more useful than the original LAB results. The new Hallucination-Gated All-Pass Rate credits a task only when every rubric criterion passes and no material hallucination appears. Grok 4.7 (xhigh) leads at 9.4%, while GPT-6 Astra (max) at 8.6% has very few material hallucinations.
Why it matters: The update adds a hallucination gate to a legal benchmark, showing that models with high all-pass rates can rank much lower once material errors count.
AIAnthropic launched two beta features for Claude: Dashboards, which turns connected data sources like BigQuery, Databricks, Snowflake, or Salesforce into auto-updating live dashboards from text prompts, and Motion, which creates animated explainer videos from text, diagrams, and images. Dashboards is available to paid users and Motion to Team and Enterprise plans, while Docs, Slides, and Design leave beta and work across all plans, including free accounts.
AILuma says users can bring Claude Motion animations into Luma to resize them for different formats and refine them for shipping. Claude Motion is in beta on Claude Team and Enterprise plans.

AITibo, an OpenAI team member, says GPT-6.1 Sol ultrafast is rolling out today in the API, Codex, and ChatGPT Work. He says it offers near-Astra intelligence at up to 8x the speed of Sol Standard. The post also says improved steering now lets the model react faster to user adjustments in real time.
Why it matters: The post specifies the new Ultrafast option, its availability across API, Codex, and ChatGPT Work, and its speed claim relative to Sol Standard.
AIThe Alzheimer's Translation Challenge, led by Primamente and AlzData with NVIDIA, Hugging Face, Nebius, Talisman Therapeutics, Ultima Genomics, Cellanome, Prime Intellect, and Boltz, is now open for registration. The competition starts in spring 2027, according to the post.
AIParticipants will use a dataset to build AI models, evaluated through a series of evals ranging from general benchmarks to more complex tasks. Top teams will be shortlisted and have their experimental hypotheses tested in Prima Mente's wet lab.
AIThe Alzheimer's Translation Challenge is built on a new atlas of 150M cells, covering neurons, astrocytes, and microglia across different genetic backgrounds under combinatorial perturbations with multi-modal readouts. The data will be made available through the AD workbench and Prima Mente's modeling platform.
AIGoodfire and PrimaMente report discovering a new class of Alzheimer's biomarkers. They are now opening a global AI competition to find new treatments, called the Alzheimer's Translation Challenge, led by PrimaMente and AlzData.
AIThe Arena Alignment Index, built from over 90K real-world agent sessions across 27 models, ranks OpenAI's GPT-6.1-Sol first with a score of 87.9, ahead of Claude-Opus-5.5 at 83.2 and Grok-4.7 at 82.7. GPT-6.1-Sol also posted the lowest observed rates across the index's three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion. The index's authors report that newer models consistently outperform their predecessors across all four labs, suggesting broad progress in agent safety.
AIFired OpenAI safety researchers published the full letter they sent to the company's board. The author, quoted in a linked post, believes they were dismissed for prioritizing safety over OpenAI's near-term corporate interests.
AISam from LangChain discusses where decision models fit inside an agent harness and how to use Jev with LangChain. The post links to a LangChain blog on building a harness with Jev, which the background post says Jev from typesafeai popularized alongside OpenAI's Decisions API and Databricks' ai_decide function.
AIAnthropic has updated Claude's usage policy for the first time in over a year, banning sustained and needless abusive or cruel behavior toward Claude. The company says ordinary frustration, pushback, dark creative themes, and model testing are not covered, and that the rule applies only in extreme cases. Violations can lead to warnings, throttling, restriction, suspension, or termination of access.
AIClaude Docs, Slides, and Design are out of beta and available on every plan, including Free, starting today. The source also says a team and Claude can edit the same doc, deck, or design together.
AIClaude Dashboards lets users connect a data platform or CRM tool and ask questions in plain language. Claude writes the query and builds a dashboard that updates as the data changes, and every chart shows its underlying query. The feature is in beta on paid plans.

AIAnthropic's Claude Motion converts reports, charts, or product walkthroughs into short animations. Claude writes each animation as code rather than using a video model, so users can edit any word, number, or timing and export an MP4. The feature is in beta on Team and Enterprise plans.

AIAnthropic's Claude has launched Claude Dashboards and Claude Motion in beta. Users can ask Claude to turn their data into live dashboards and their ideas into animated explainers.

AIThree OpenAI safety and alignment employees, Tomek Korbak, Jasmine Wang, and Mikita Balesni, were fired last week and have published an open letter to OpenAI's safety and governance committees. The letter argues that OpenAI cannot make AI safe on its own, calls for open debate, third-party collaboration, and clear internal procedures, and says the firing and its handling bear directly on safety oversight.

AIOpenAI's Codex 0.162.0 release adds tools for creating and listing managed Git worktrees from trusted local projects when the worktrees feature is enabled. The update also lets users pin tasks in the agent Command Center, copy transcript blocks with /copy, and make URLs clickable in approval headers, questions, and warnings, along with several Linux and Windows sandbox fixes.
AIMicrosoft will replace per-user OneDrive allowances on Microsoft 365 Family and Premium plans with a shared 2 TB storage pool, down from 6 TB across six accounts. Heavy users could face bills that double or triple, since each extra 1 TB adds $10 a month, while family members will be able to share a single AI usage allowance. New and upgraded subscriptions get shared storage starting Oct. 8, 2026, and existing subscribers move at their first renewal on or after May 2, 2027.
AIGoogle's recently announced Gemini Agent for Gemini Business is reportedly set to offer Gemini Argon 4, Gemini Flash 3.8, Claude Opus 5, and Claude Sonnet 5.5. If accurate, it would mark the first time Claude models appear on Google's platform alongside Google's own models, which the post frames as a way for Google to compete for enterprise customers.
AIThe U.S. Department of Labor said it suspended Microsoft and Adobe from its Permanent Labor Certification program, citing multiple active federal investigations. Labor Secretary Keith Sonderling also said no new applications will be accepted for Cognizant, Infosys, Capgemini, Tata, Wipro and HCL. Microsoft said the vast majority of its U.S. employees are Americans and that 80% of its roughly 6,000 H-1B petitions last fiscal year were to extend or change the status of existing employees.
AIArena has secured a $200M Series B, with Lightspeed doubling down on its investment. The company is also launching the Alignment Index, which measures how closely AI behavior aligns with human values in real-world settings. Arena reports annualized revenue above $100M since its Series A, with millions of people helping evaluate frontier models through real-world use.
AIGoogle Research is hosting a live demonstration of EmbeddingGemma 2 at its COLM booth #107 today at 12:00pm. The open multimodal model unifies text, image, audio, and video representations, with Sahil Dua available to connect with attendees.
AIJohn Groetzinger, writing in a personal capacity rather than for Cisco, argues that enterprise skills need packaging, evaluation, syncing, and distribution rather than scattered markdown files. He describes using skills to make cheaper models viable, converting curated TAC knowledge-base articles into maintained skills, and rolling out an eval framework across teams. He also describes syncing a repository README to Confluence with a deterministic script.
AIOpenAI is rolling out GPT-6.1 Sol Ultrafast, which generates tokens up to 8x faster than Sol Standard. On the API it costs $12 per 1M input tokens and $60 per 1M output tokens, and the main post says it is just 1.2x the cost of Astra.
AILiveKit is offering free access to its Simulations product through October, letting teams check what their agent can do and find gaps before deployment. The product also lets teams test any model against their own scenarios before switching models.
AIOpenAI is rolling out GPT-6.1 Sol Ultrafast on ChatGPT Work, Codex, and the API. The Ultrafast mode is priced at $12 per million input tokens and $60 per million output tokens, and it runs 8x faster than Sol Standard.
AIOpenAI is rolling out GPT-6.1 Sol Ultrafast, which generates tokens up to 8x faster than Sol Standard. On the API it is priced at $12 per 1M input tokens and $60 per 1M output tokens. The mode is also available today in Codex and ChatGPT Work.
AIArtificial Analysis has released Harvey LAB-AA, an evaluation built on Harvey's LAB dataset and developed in collaboration with Harvey. Full results are published on the Artificial Analysis evaluations page, alongside Harvey's commentary on the benchmark and human expert preferences.