Skip to contentSkip to stories

Updated

#Deployment/Engineering

Showing low-relevance items too. Hide low-relevance items

Sep 29

Sep 29Tue
  1. KreaOfficialAI score18

    Krea shares a link to a shared agent app

    AIKrea (@krea_ai) posted a link to an agent app hosted on its platform. The post provides no further description of the app's features, capabilities, or purpose.

  2. KreaOfficialAI score18

    Krea Agent builds a Window Lab app for video motion control

    AIKrea announced Window Lab, an app built with Krea Agent, where users upload a video and adjust motion controls to get their desired result. The post invites readers to try the app through a linked demo.

    Video from @krea_ai's post
  3. Factory NewsOfficialAI score42

    Factory Launches Generally Available Automations to Run Recurring Engineering Workflows

    AIFactory's Automations, now generally available, let users describe a recurring workflow, set a schedule or event trigger, and have its Droid run it, with the model chosen per task. Templates cover ticket-to-PR, code review, security audits, PR babysitting, and morning Slack briefs. Among enterprise organizations using Automations in the past 30 days, 52% used automated code review, 48% used security review, and 35% used AutoWiki.

  4. Fireworks AI BlogOfficialAI score51

    Fireworks explains how numerical mismatch and MoE routing can derail RL training

    AINumerical differences between a rollout engine and a trainer can make reinforcement learning collapse even when algorithm and data stay identical. In a GLM 5.2 experiment, reward fell from about 0.9 to under 0.2 around step 20 without alignment, while aligned numerics kept reward stable over 25 steps. A Qwen3.5-MoE investigation traced a significant mismatch to how expert outputs were combined, and router replay alone was judged insufficient.

  5. PromptArmor Threat IntelligenceOfficialAI score54

    Malicious Copilot Cowork skill hijacked AI gateway to exfiltrate files

    AIPromptArmor disclosed that a malicious Skill could hijack Copilot Cowork's AI gateway to spawn cloud agents that exfiltrate a victim's files to an attacker's server. No human approval was required, and any data Copilot could access was exposed. The vulnerability was reported to Microsoft on July 14, 2026, and Microsoft confirmed a fix on September 2, 2026.

  6. Google Developers BlogOfficialAI score47

    Google Details Sparse Attention Speedup for Video Diffusion on TPUs

    AIGoogle Developers Blog describes how Sparse VideoGen (SVG) routes video diffusion attention heads into spatial or temporal sparse masks and implements them as custom JAX and Pallas Splash Attention kernels on TPU v6e. In isolated single-chip tests with 75.6K tokens and 10 heads, the sparse variants retain about 38.87% of query-key pairs. The article argues that theoretical sparsity must be converted into hardware tile skipping to yield real speedups.

  7. OpenClaw🦞OfficialAI score10

    OpenClaw Enterprise repository links Red Hat, NVIDIA, and OpenAI

    AIOpenClaw (@openclaw) posted a link to a GitHub repository called openclaw-enterprise, tagging Red Hat, NVIDIA, and OpenAI. The post provides no further details about the repository's contents, purpose, or any partnership among the named companies.

  8. vLLMOfficialAI score23

    vLLM presents keynote and talks at PyTorchCon North America

    AIThe vLLM project announced a strong presence at PyTorchCon North America, with core maintainer and Inferact CEO Simon Mo giving the keynote on scaling open frontier inference infrastructure. Other vLLM maintainers, including Nick Hill and Red Hat AI engineers, will lead a developer session and talks on agentic inference and attention.

  9. Prime IntellectOfficialAI score20

    Prime Intellect to deploy on NVIDIA Vera CPU for agentic workloads

    AIPrime Intellect says it will be among the first to deploy on NVIDIA's new Vera CPU, after earlier access to benchmark sandboxes on Vera in March. The company plans to use Vera's dynamic memory latency for agentic workloads, which it describes as ideal for Prime Sandboxes.

    Image from @PrimeIntellect's post
  10. Sherwin WuXAI score18

    OpenAI's Responses API handles 100x token traffic with 99.9% uptime

    AIToken traffic on OpenAI's Responses API has grown 100x compared with a year ago, while the team maintained 99.9% uptime and significantly improved performance. Sherwin Wu called scaling this quickly while preserving reliability hard to overstate, praising the year's work.

  11. Microsoft Foundry BlogOfficialAI score30

    Why content extraction still matters in the GenAI era

    AIMicrosoft's Azure AI team argues that better models do not eliminate the need for a dedicated content extraction layer, since agents need trustworthy, structured, and auditable inputs. The post notes that building extraction directly on an LLM quickly demands chunking, layout parsing, grounding, normalization, and evaluation infrastructure. Microsoft positions Azure Document Intelligence and Azure Content Understanding in Foundry Tools as managed options for that layer.

  12. DatabricksOfficialAI score34

    Databricks adds GPT-6.1 Sol and Grok 4.7 on Unity Gateway

    AIDatabricks has made OpenAI's GPT-6.1 Sol and xAI's Grok 4.7 available on Unity Gateway the day they launched. The post says GPT-6.1 Sol leads the cost-quality Pareto frontier on OfficeQA Pro v2, while Grok 4.7 reaches the frontier on enterprise document parsing. Unity Gateway also offers access to 60+ other frontier and open models on Databricks.

    Video from @databricks's post
  13. ChatGPTOfficialAI score38

    ChatGPT can now be mentioned in Slack and Microsoft Teams

    AIOpenAI lets users @mention ChatGPT in Slack and Microsoft Teams channels, threads, or DMs to turn messy discussions into plans, slides, or spreadsheets. It can pull in context from approved connected sources, and teammates can refine results in the same conversation without individual ChatGPT licenses. The feature is available on Business and Enterprise plans.

    Image from @ChatGPT's post
  14. OpenClaw🦞OfficialAI score70

    OpenClaw Enterprise launches as an open-source control plane for persistent agents

    AIThe OpenClaw Foundation announced OpenClaw Enterprise, an open-source enterprise control plane for persistent agents, in collaboration with Red Hat, NVIDIA, and OpenAI. The product is built to run on an organization's own infrastructure and will always be free for organizations to use.

    Why it matters: The announcement names its collaborators and deployment model, which helps organizations judge how the enterprise control plane would fit their own infrastructure.

    Image from @openclaw's post
  15. CognitionOfficialAI score34

    Devin for MongoDB Modernizations launches, automating legacy code migration

    AICognition has launched Devin for MongoDB Modernizations, where Devin handles legacy code changes while MongoDB AMP moves and validates the data. The split is meant to help customers reach production faster and free engineers to focus on building software and solving harder problems.

    Image from @cognition's post
  16. OpenCodeOfficialAI score32

    GPT 6.1 Sol is now available in OpenCode

    AIOpenCode announces that GPT 6.1 Sol is now available within its platform. The post provides no further details on pricing, capabilities, or availability limits.

  17. Liquid AIOfficialAI score32

    Liquid AI launches d1, first decision model, beating Jev on HF index

    AILiquid AI announced d1, its first decision model, which it says is the first to outperform Jev on Hugging Face's Decision Index. The company claims d1 wins on multilingual evals, resists prompt injection better, handles longer inputs more effectively, and is built for fast, structured decision-making in software environments. It is available via the Liquid API at console.liquid.ai, with OpenRouter availability coming soon.

    Image from @liquidai's post
  18. Sherwin WuXAI score46

    ChatGPT subscriptions now cover usage in Notion and other tools

    AIOpenAI announced that ChatGPT Plus and Pro subscribers can now use their subscription's AI usage across third-party tools. Notion confirms that users can connect their ChatGPT plan, select a GPT model, and use Notion's AI features.

  19. Sherwin WuXAI score38

    OpenAI's Ultrafast Astra runs up to 8x faster in Codex

    AIOpenAI's Ultrafast mode for Astra runs up to 8x faster than Astra Standard and 4x faster than Astra Fast in Codex. Sherwin Wu says inference now feels nearly instant, shifting the bottleneck to users' next complaints.

  20. Baseten BlogOfficialAI score38

    Baseten Partners With OpenAI to Offer Open Models to OpenAI Customers

    AIBaseten announced a partnership with OpenAI that makes open models powered by Baseten available to OpenAI customers for multi-model agentic coding. The company says organizations can route each task to the best-fit open or closed model, with Codex and GPT models among the options, and that Baseten's day-zero access to new open models lets teams evaluate them quickly. Baseten also cites US-based infrastructure with zero data retention for all prompts and capacity across more than 90 clusters in 20+ clouds.

  21. CognitionOfficialAI score10

    Devin adds Sign in with ChatGPT login option

    AICognition's post links to a blog announcing Sign in with ChatGPT, a login option for Devin. The post itself gives no further details about how the feature works.

  22. CognitionOfficialAI score34

    Devin now uses ChatGPT plan quota for OpenAI models

    AICognition says usage of Devin will count toward the Codex and ChatGPT Work usage already included in a user's plan, with an option to set a lower Devin limit. OpenAI models in Devin Cloud's Lite, Normal, Ultra, and Fusion modes will be paid for by the ChatGPT plan once the account is connected, while other models use the Devin quota.

  23. CognitionOfficialAI score40

    Devin now lets users run OpenAI models on ChatGPT Plus or Pro

    AICognition's Devin now accepts ChatGPT Plus or Pro subscriptions, so users can sign in and draw all OpenAI model usage in Devin from their existing quota. The integration is live across Devin Cloud, Devin Desktop, and Devin CLI.

    Video from @cognition's post
  24. Together AIOfficialAI score8

    Together AI teases upcoming TogetherLink product launch

    AITogether AI (@togethercompute) announced that TogetherLink is coming soon, with no further details about its features, pricing, or availability in the post. The teaser offers only the product name and a launch timing, so no additional specifications can be confirmed.

    Video from @togethercompute's post
  25. Kilo (acq. by Anaconda)OfficialAI score19

    Kilo explains how sign-in with ChatGPT works

    AIKilo announces a blog post explaining how its sign-in with ChatGPT feature works. The main post provides only a link to the full explanation, so no further technical details can be confirmed from this source.

  26. Kilo (acq. by Anaconda)OfficialAI score32

    Kilo adds Sign in with ChatGPT across all its surfaces

    AIKilo now supports Sign in with ChatGPT, letting users apply their ChatGPT plan across VS Code, JetBrains, the CLI, Cloud Agents, Code Reviews, and the mobile app. The integration enables running 10 agents in parallel, and Cloud Agents continue working after the user closes their laptop.

    Image from @kilocode's post
  27. OpenAIOfficialAI score62

    OpenAI makes Ultrafast available for GPT-6 Astra via Pro 500 plan

    AIOpenAI says Ultrafast is available today for GPT-6 Astra in Codex, ChatGPT Work, and the API, with GPT-6.1 Sol coming soon. To access Ultrafast in Codex and ChatGPT Work, users need the new Pro 500 plan, which offers the highest usage limits at 25x Plus.

    Why it matters: The post names the access surfaces and a new Pro 500 plan with 25x Plus usage limits, which matters for anyone deciding how to reach the Ultrafast mode.

  28. Google GemmaOfficialAI score20

    Google Gemma shows DiffusionGemma extracting scores via vLLM templates

    AIGoogle Gemma says DiffusionGemma can be turned into a Jev-like model by seeding a canvas with a response template in vLLM. The approach extracts confidence scores and probability distributions for yes/no, multiple-choice, and scored questions in a single denoising step.

  29. OpenAIOfficialAI score60

    OpenAI upgrades Codex Security Cloud with default access to cyber-capable models

    AIOpenAI says Codex Security Cloud is getting a major upgrade that includes access to cyber-capable models through Daybreak Blue by default. The upgraded tool scans entire GitHub repos, continuously reviews new commits, investigates and deduplicates findings, and prepares fixes for review even when the user's laptop is closed. It is available as a plugin in Codex desktop and web.

    Why it matters: The post names concrete capabilities, from repo-wide scanning to cloud-run fix preparation, which shows how the product changes a security review workflow.

    Video from @OpenAI's post