Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 1

Oct 1Thu
  1. Perplexity DevelopersOfficialAI score41

    Perplexity launches Decisions API powered by pplx-decider-v1-27b

    AIPerplexity introduced its Decisions API, powered by pplx-decider-v1-27b, a multimodal model that outputs a probability distribution over a fixed set of answers rather than text. The company says the API costs $0.04 per million input tokens and scores 85.71% across benchmarks.

    Image from @perplexitydevs's post
  2. Microsoft CopilotOfficialAI score34

    Microsoft Copilot adds GPT-6.1 Sol and Claude Sonnet 5.5 models

    AIMicrosoft Copilot begins rolling out OpenAI's GPT-6.1 Sol and Anthropic's Claude Sonnet 5.5 today, joining Claude Opus 5.5 and GPT-6 Sol added earlier this month. Users can pick the model suited to each task, with Work IQ grounding responses in their files, meetings, and chats within existing permissions. The rollout starts today in Copilot Cowork and Copilot Studio, with Word, Excel, PowerPoint, and Chat following in phases over the coming week.

  3. Grok BotOfficialAI score36

    Grok Bot can now suggest ways to help proactively

    AIGrok Bot can now offer suggestions for ways to help without the user needing to ask first. The post does not provide further details about how these proactive suggestions work.

    Video from @bot's post
  4. Google GemmaOfficialAI score54

    Google Gemma credits StudentBench study comparing AI and expert human GRE tutors

    AIGoogle Gemma relays a StudentBench study reporting that AI tutors matched expert human tutors on immediate GRE learning gains. The author reports 2,383 students and a cost of 7 cents per AI tutor hour versus $75 for an expert human hour. The post also says the top AI tutor beat expert human tutors on average in 5 of 7 academic topics, and that the data and paper are publicly available.

  5. DatabricksOfficialAI score20

    Databricks Smart Routing assigns each coding task to a suitable model

    AIDatabricks' Smart Routing evaluates each coding task separately and selects the lowest-cost model capable of handling it, balancing quality, latency, and cost. In a demo, Omnigent splits an app build into planning, backend, and frontend work, routes each part to a different model, and runs some tasks in parallel.

    Video from @databricks's post
  6. Harrison ChaseXAI score22

    Harrison Chase outlines a four-step approach to model routing

    AIHarrison Chase says model routing is a provocative term that lacks a clear definition, but offers a practical approach. His four steps are to understand tasks, understand the models, build the router inside the harness, and track outcomes, aiming to lower costs without a performance hit.

    Image from @hwchase17's post
  7. Vaibhav (VB) SrivastavXAI score12

    OpenAI shares a dots demo with improved WiFi

    AIOpenAI's Vaibhav Srivastav posted a short "we're so back" message, linking a demo of "dots" that the OpenAIDevs account described as now running with better WiFi. The post gives no technical details about what dots is, its capabilities, or its performance figures.

  8. Microsoft AIOfficialAI score36

    Microsoft's MAI models now available through Vercel AI Gateway

    AIMicrosoft AI's MAI models are now accessible to developers via Vercel, including the newest releases MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash. The partnership brings these models into Vercel's AI Gateway as another route for building Microsoft AI models into applications.

  9. Google WorkspaceOfficialAI score32

    Google Sheets canvas turns spreadsheets into interactive mini-apps via prompts

    AIGoogle Workspace says Sheets canvas can turn static spreadsheet data into interactive tools such as Kanban boards, dashboards, and visual workflows from a simple prompt. Derek Snyder, Director of Product Marketing for Google Workspace, demonstrates the feature in the latest AI Boost Bite video.

    Video from @GoogleWorkspace's post
  10. GoodfireOfficialAI score58

    Goodfire Says AI Biosecurity Risks Are Next After Cybersecurity Risks

    AIGoodfire says AI cybersecurity risks are already here and that biosecurity risks are next, as models improve at biology. The post presents this as both an opportunity for science and medicine and a reason for stronger security. It quotes Demis Hassabis announcing SynthID for biology, a watermarking approach for AI-generated proteins, published in Nature with SynthID Bio tools open sourced.

  11. Goodfire ResearchOfficialAI score60

    Goodfire proposes protein embedding monitors for biosecurity risks in AI agents

    AIGoodfire Research developed sequence-aware monitors using protein language model embeddings to flag concerning biological sequences in dual-use AI agent tasks. On a custom benchmark, the monitors outperformed frontier model safeguards with fewer refusals on benign requests, and they held up better against paraphrasing and fragmentation attacks. The paraphrase results rely on in-silico estimates and do not establish whether the redesigned proteins keep biological activity, and the monitors run in milliseconds per sequence.

    Why it matters: The post gives a concrete benchmark setup and fragmentation results, showing how sequence embeddings can separate dual-use biology requests that task-based safeguards handle poorly.

  12. Guillermo RauchXAI score22

    Vercel brings Microsoft AI's speech models to AI Gateway on day zero

    AIVercel has made Microsoft AI's MAI-Voice-2.1 for long-form and fast-reply speech and MAI-Transcribe-2-Streaming for transcription available on AI Gateway starting today. Guillermo Rauch of Vercel said the team is excited to bring the models to Vercel at launch.

  13. Vercel DevelopersOfficialAI score32

    Vercel adds Microsoft AI speech and transcription models to AI Gateway

    AIVercel says it partnered with Microsoft AI to make MAI-Voice-2.1 and MAI-Transcribe-2-Streaming available through AI Gateway today. MAI-Voice-2.1 handles long-form and fast-reply speech, while MAI-Transcribe-2-Streaming provides transcription.

  14. Prime IntellectOfficialAI score34

    Qwen3.6 reward rises 2.8x via GRPO on Hosted Training

    AIPrime Intellect reports that after about 100 GRPO steps on Hosted Training, Qwen3.6's reward on held-out problems rose from 0.127 to 0.361, a 2.8x gain. Qwen3.5, trained the same way, reached 0.356, suggesting the method works across model families. Both post-trained models finished well ahead of other open models and narrowed the gap to Claude Opus 4.8, with Qwen3.6 activating only 3B parameters per token.

    Image from @PrimeIntellect's post
  15. Mustafa SuleymanXAI score40

    Microsoft AI launches MAI-Transcribe-2-Streaming, claiming top real-time transcription accuracy

    AIMicrosoft AI launched MAI-Transcribe-2-Streaming, which Artificial Analysis ranks #1 of 38 models for final transcript accuracy at 2.5% WER, returned 0.13s after end of speech. Artificial Analysis lists its streaming price at $0.54 per hour of audio, at the higher end among leading streaming models. Microsoft's post claims the model is 55% faster and 60% cheaper than ElevenLabs and invites developers to build agents on its platform.

  16. ZyphraOfficialAI score20

    Zyphra's Results Explain How NoPE Models Encode Position

    AIZyphra says its results clarify how state-of-the-art NoPE models encode position and which inductive biases support generalization. It adds that global NoPE could enable models to extrapolate to contexts longer than those seen in training, potentially indefinitely.