Skip to contentSkip to stories

Updated

#Product update

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 29

Sep 29Tue
  1. Azure BlogOfficialAI score40

    SQL Server on Azure Local Becomes Generally Available for Connected and Disconnected Use

    AIMicrosoft has made SQL Server on Azure Local generally available for connected and disconnected deployments, letting organizations run SQL Server in their own datacenters and edge locations. Disconnected operations continue locally where external connectivity is restricted or unavailable. Eligible existing SQL Server licenses can be used, and Foundry Local on Azure Local, currently in preview, brings AI inference alongside SQL Server data.

  2. Meta NewsroomOfficialAI score34

    Meta Launches Forum, a Standalone App for Browsing Facebook Groups

    AIMeta is testing Forum, a standalone iOS and Android app in the US that syncs with users' Facebook Groups to consolidate their conversations in one place. The update adds a new top-contributor role replacing previous badges, an AI-powered Ask feature that surfaces group posts and comments, and topic labels for exploring interests.

  3. Max ZeffXAI score45

    OpenAI re-opens $200 Pro subscriptions with halved effective API value

    AIOpenAI says it will reopen its $200 Pro subscription to new subscribers tomorrow while changing usage calculation, netting out at half the dollar-value in API spend versus the old plan. The company says it will not reintroduce the 5-hour limit and that subscribers should get more work done than a month ago, as it passes model efficiency gains on through API price cuts. This week it introduced GPT-6 Sol and GPT-6 Luna at 50% of their previous prices.

  4. DatabricksOfficialAI score22

    Databricks rolls out frontier models to employees on Day 1 via Unity Gateway

    AIDatabricks says it aims to give its employees the best models on launch day, quickly adopting new releases such as Opus 5.5 and GPT-6 Sol while tracking real-world usage and cost. Its AI engineering team uses Unity Gateway to manage access, spend, and model selection across thousands of employees, and to decide which models join its AI stack.

    Image from @databricks's post
  5. Microsoft ResearchOfficialAI score75

    Microsoft Research introduces Quine, a multimodal biology world model and research harness

    AIMicrosoft Research introduced Quine, an experimental research system combining a multimodal world model of biology with an interactive harness that connects models, scientific tools, literature, and researchers. In a pancreatic cancer study with the Broad Institute, Quine prioritized compounds that shifted tumor cell states, and several top-ranked candidates were validated in wet-lab assays. Access is initially limited to the Quine Fellows program and select collaborations, and the system is intended for research use only, not clinical use.

    Why it matters: The post shows how a multimodal biology world model is wired into a harness, grounded in one wet-lab cancer example and a limited fellows-program access path.

  6. Ahead of AI (Sebastian Raschka)BlogAI score43

    Language Models for Text Classification: From Bag-of-Words to Jev

    AISebastian Raschka traces text classification from bag-of-words models such as naive Bayes and logistic regression through pre-transformer neural networks, then sets up an analysis of the recently released Jev AI model. The article frames Jev as a general-purpose classifier that trades specialized accuracy for speed, cost, and breadth of tasks.

  7. Tibor BlahoXAI score53

    OpenAI reopens Pro $200 plan with usage calculation halved

    AIOpenAI is reopening its Pro $200 subscription to new subscribers while changing how usage is calculated, so the plan nets out at about half the API spend of the old Pro $200 plan. The quoted post says the 5-hour limit will not return and that GPT-6 Sol and GPT-6 Luna API prices were cut 50% this week. The author adds that Pro's earlier generosity was unsustainable and cites a January 2025 Sam Altman post saying OpenAI was losing money on Pro subscriptions.

    Image from @btibor91's post
  8. X.PINXAI score38

    ByteDance's Doubao fast-tracks codenamed "Spell" personal AI agent

    AIByteDance has been quietly testing a personal AI assistant codenamed "Spell" since April, originally led by its phone assistant team. Spurred by the rapid growth of overseas personal agents such as Muse and Instinct, Doubao is now accelerating the rollout and plans to integrate "Spell" into the Doubao app.

    Image from @thexpin's post
  9. X.PINXAI score46

    Tencent launches LightVela, a cloud-hosted Hermes Agent inside WeChat and QQ

    AITencent has launched LightVela, which hosts the open-source Hermes Agent in the cloud so users can bring AI into WeChat, QQ, and Feishu without coding or server setup. The post contrasts this with Meta's AI agent Muse, which reportedly topped US app charts in September with over 2.5 million downloads in 13 days. The author argues personal AI assistants will deeply integrate into daily life, noting the pace of change is very fast.

    Image from @thexpin's post
  10. OpenBMBOfficialAI score34

    MiniCPM-o 4.5 now runs in SGLang Omni v0.1.7 for developers

    AIOpenBMB announced that MiniCPM-o 4.5 is now supported in SGLang Omni v0.1.7, giving developers more flexibility to run and build with the model. The background release notes add that MiniCPM-o 4.5 brings multimodal input and speech output to the runtime. MiniCPM-o and MiniMax-Music3 also gained Intel XPU support in the same release.

  11. vLLMOfficialAI score58

    IQuest-Q1 320B MoE coding model gets day-0 support in vLLM

    AIvLLM announced day-0 support for IQuest-Q1, a 320B-parameter MoE model with 15B active per token, 256 experts with 8 active, and a 524,288-token context. The post credits existing vLLM features such as the hybrid KV cache coordinator, sinks attention path, and EAGLE speculative decoding with probabilistic draft sampling. The linked material includes a Docker image and vllm serve commands, with and without recursive MTP.

    Image from @vllm_project's post
  12. Azure BlogOfficialAI score75

    Microsoft announces Fabric IQ in Copilot, Power BI agentic app creation, and new Fabric and SQL updates

    AIMicrosoft announces new Microsoft Fabric and SQL Server updates at FabCon and SQLCon in Barcelona, including Fabric IQ integration with Microsoft Copilot Chat and Cowork, now generally available. Power BI agentic app creation enters preview in the coming weeks for Pro and Premium Per User customers, with Fabric Apps database capabilities up to 1 GB per app at no additional cost.

    Why it matters: The post lists dozens of Fabric and SQL updates tied to Copilot and agents, with specific availability and pricing terms for Power BI customers that help readers judge what applies to them.

  13. SGLangOfficialAI score53

    SGLang adds Day-0 support for IQuest-Q1 with a single-node serve command

    AISGLang says it has Day-0 support for IQuest-Q1, an open-source sparse MoE model with 320B total and 15B active parameters for coding and agentic tasks. The post includes a single-node serving command for H200 GPUs in BF16, using tensor parallelism of 8, EAGLE speculative decoding, and the iquest_q1 reasoning and tool-call parsers. The image marks the command as not verified.

    Image from @sgl_project's post
  14. Together AIOfficialAI score22

    Qwen3.8-Flash gets 40% off through month's end on Together AI

    AITogether AI is offering 40% off Qwen3.8-Flash through the rest of the month, a window it suggests for running evaluations. Alibaba's Qwen3.8-Flash is designed for high-volume applications such as coding and coworking assistants, with an emphasis on quality at low cost.

    Image from @togethercompute's post
  15. Mastra BlogOfficialAI score42

    Mastra Adds Memory Hooks to Observe and Modify Agent Memory Cycles

    AIMastra has added memory hooks that let developers monitor or alter an agent's observational memory cycles. Lifecycle hooks such as onObservationStart and onReflectionEnd report on each cycle, including token usage for spotting cost spikes, while transform hooks like beforeObservation and afterReflection can prune, remove, or redact memory data.

  16. Manus BlogOfficialAI score50

    Manus Flex lets users connect their own API keys to the Manus workspace

    AIManus is launching Manus Flex, a module that lets users power Manus agents with their own API key from a supported inference provider. Model inference is billed directly by that provider, while other services used in Manus tasks still consume Manus credits. OpenRouter, Fireworks, and Modal are announced as initial inference partners for the Flex Inference Partner Program.

  17. Luma AI NewsOfficialAI score22

    AI Photo Editing Prompt Formula Preserves Color, Light, and Skin in Campaign Edits

    AIThe article presents a four-part prompt structure (action verb, target element, desired result, protection instructions) for AI photo editing, saying it preserves approved work across platforms. It identifies three common failure causes: unmatched light direction, stacked edits in one prompt, and vague visual language. It states that simple skin retouching takes 2-3 minutes versus 15-30 minutes manually.

Sep 28

Sep 28Mon
  1. KrASIA · Big TechNewsAI score47

    Alibaba unveils Zhenwu V900 AI chip, targets 20 GW data center capacity by 2032

    AIAlibaba unveiled the Zhenwu V900 AI chip at its 2026 Apsara Conference, claiming three times its predecessor's performance and support for clusters of up to 500,000 cards. The company is pursuing data center capacity beyond 20 gigawatts by 2032 and has committed RMB 380 billion in capital spending over three years.

  2. KhazixXAI score38

    Khazix open-sources AIHOT, the AI news site, with its full pipeline and prompts

    AIKhazix (Shuzi Shengming Kazike) says the monthly-active-million AI news site AIHOT is now open source on GitHub, including its collection workflow, curation scoring, clustering mechanism, and all production prompts. He says the release is meant to hand the project to others, since readers have asked for versions for industries such as gaming, law, HR, and finance.

  3. KreaOfficialAI score22

    Seedance 2.5 Draft Mode now available on Krea

    AIKrea has launched Draft Mode for Seedance 2.5, letting users experiment with 480p generations before switching to 1080p once a scene is right. The post directs readers to try the feature on Krea's platform.

    Video from @krea_ai's post
  4. vLLM BlogOfficialAI score54

    vLLM guide explains disaggregated serving for prefill and decode

    AIThe vLLM blog guide explains how separating prefill and decode, and moving tokenization to a CPU-only render tier, can keep token streams from stalling under load. In a two-L40S test on Qwen2.5-7B, collocated p99 inter-token latency reached 169 ms at 0.4 req/s while disaggregated serving stayed between 25 and 52 ms. The guide notes that the gain depends on fast KV cache transfer, and it includes setup code for NIXL-based serving and the render/derender API.

  5. Amp NewsOfficialAI score34

    Amp Adds Plaid Speed for GPT-6 Astra Modes at 6x Speed and Cost

    AIAmp now supports Plaid speed for modes that use GPT-6 Astra, using OpenAI's ultrafast tier to run inference up to 6× faster at 6× cost per token. Plaid works only with Amp-provided inference, not linked ChatGPT subscriptions, and subagents and non-Plaid inference fall back to fast or standard speed.

  6. DatabricksOfficialAI score38

    Claude Sonnet 5.5 now available on Databricks across AWS, Azure, GCP

    AIDatabricks now offers Anthropic's Claude Sonnet 5.5 on AWS, Azure, and GCP, governed through Unity Gateway. The post says Sonnet 5.5 is more efficient than Sonnet 5 for coding and agentic use and reaches Opus 5-level accuracy on document understanding, parsing, and search. It joins Claude Opus 5.5, Claude Fable 5.1, and 60+ other open-source and frontier models on the platform.

    Video from @databricks's post
  7. SpaceXAIOfficialAI score46

    Grok 4.7 is now available on Amazon Bedrock

    AIaccording to the SpaceXAI account, which is operated by xAI and Grok. The post gives no further details on pricing, context length, or benchmarks.

    Video from @SpaceXAI's post
  8. Lydia Hallie ✨XAI score22

    Claude Code Projects default effort level and override setting

    AIAnthropic's Lydia Hallie asks users who raised the main chat's effort in Claude Code Projects to explain why, since the default is low because it mainly coordinates threads. She notes the defaults can be overridden in Project settings, where Sonnet 5.5 is also available.

    Image from @lydiahallie's post
  9. Simon WillisonXAI score55

    Sonnet 5.5 becomes the free-tier model on claude.ai

    AISimon Willison says Claude Sonnet 5.5 now powers the free tier on claude.ai, so free users can run the kinds of experiments he describes. He contrasts this with ChatGPT's free tier, which he says still runs the less capable GPT-5.6 Luna.

  10. Grok BotOfficialAI score45

    SpaceXAI launches Team Bots public beta for Teams and Enterprise

    AISpaceXAI says its Team Bots, which prep account teams, coordinate engineering work, answer data questions, triage customer feedback, and run hiring loops, are now in public beta for Teams and Enterprise customers. The post links to a full announcement at

  11. François CholletXAI score32

    K3-Node: a Keras 3 GNN library running on JAX, PyTorch, and TF

    AIK3-Node is a graph neural network library built natively on Keras 3, with models that run on JAX, PyTorch, and TensorFlow with hardware acceleration including Apple Silicon and TPU. According to the post, it achieves 100% public API parity with PyG and incorporates foundation models and architectures from Spektral and StellarGraph.

  12. LlamaIndex 🦙OfficialAI score30

    LlamaIndex says frontier VLMs still struggle parsing tax and W-series forms

    AILlamaIndex argues that frontier vision-language models still fail on real forms such as W-2s, 1040s, W-9s, and scanned W-4s, because forms require detecting every field, preserving section hierarchy, linking values to their exact boxes, and reading handwriting and checkmarks. The company's blog post details these failure modes and presents a custom cookbook for LlamaParse as a cheaper way to handle such forms.

    Image from @llama_index's post
  13. Perplexity DevelopersOfficialAI score44

    Perplexity adds reusable custom agents to its Agent API

    AIPerplexity says developers can now build custom reusable agents in its Agent API using Profiles, Skills, and managed connectors. Agents are configured once in the API Portal and can then be reused across applications and workflows.

    Video from @perplexitydevs's post
  14. Google AIOfficialAI score44

    Google Labs expands experimental CC agent into a family group assistant

    AIGoogle Labs has expanded Project CC, its experimental AI productivity assistant, into a group agent designed to streamline family household logistics. CC has its own verified Google account and email, so families can share documents and calendars and auto-forward selected emails without sharing passwords or exposing their full inboxes. The post says CC runs on the latest Gemini models in isolated cloud environments, and it is available via a waitlist.