Skip to contentSkip to stories

Updated

#Coding

Showing low-relevance items too. Hide low-relevance items

Sep 30

Sep 30Wed
  1. DeepSeek HarnessAI score62

    DeepSeek Harness v0.2 preview launches as a desktop app for macOS and Windows

    AIDeepSeek releases the DeepSeek Harness v0.2 preview with a desktop app for macOS and Windows. The release adds a plugin manager for installing, disabling, and uninstalling plugins without terminal commands, plus an experimental creator mode that generates plugins from user descriptions. The company says DeepSeek Harness is now the most widely used coding agent among users of the official DeepSeek API by DAU and daily sessions.

  2. Google GeminiAI score45

    Gemini skills now proactive, stackable, and support reference files

    AIGemini can build custom skills from chats and apply a saved skill automatically when a prompt matches it. Multiple skills can be stacked for larger tasks, such as combining a personal writing style skill with a brand guidelines skill. Starting today, skills can include reference files such as plain text documents, PDFs, or images, with sharing and Google Drive file support coming soon.

  3. KhazixAI score9

    Blogger shares a checklist for keeping a new Claude account stable

    AIThe author, whose earlier device was flagged so the account got banned within about half an hour, reports a new Claude account has run stably for six days. The shared tips include logging in with a Google account, using a home static IP, a clean new device, timezone set to Taiwan, paying via Google Play, starting at the $20 Max tier, and running Claude on a single always-on Mac Mini accessed remotely.

  4. Cloudflare Blog · AIAI score72

    Cloudflare launches Auto Router in AI Gateway to cut AI token spend

    AICloudflare has released Auto Router in public beta through AI Gateway, where setting the model to cloudflare/auto routes each request to a model judged capable enough for the task. Internal tests showed up to 30% cost savings against frontier models, and on a 97-task internal benchmark cloudflare/auto scored 86.6% at $0.0084 per success versus 96.6% at $0.0210 for Claude Opus 5.5. The router is free during beta.

    Why it matters: The source gives a benchmark table of success rates and costs per trial, showing how routing trades quality against price for a gateway deployment.

  5. Karl's AI WattsAI score38

    Can you keep your session after switching models in magpie?

    AIKarl's AI Watts asks whether a menu-bar tool can switch models while preserving the existing conversation, so users avoid re-explaining their project each time. The post frames this as the reason they want to keep the menu bar tool, which the quoted post describes as magpie, a menu-bar switcher for 20+ agents including Claude Code and Codex that also offers a local gateway.

Sep 29

Sep 29Tue
  1. Factory NewsAI score42

    Factory Launches Generally Available Automations to Run Recurring Engineering Workflows

    AIFactory's Automations, now generally available, let users describe a recurring workflow, set a schedule or event trigger, and have its Droid run it, with the model chosen per task. Templates cover ticket-to-PR, code review, security audits, PR babysitting, and morning Slack briefs. Among enterprise organizations using Automations in the past 30 days, 52% used automated code review, 48% used security review, and 35% used AutoWiki.

  2. Tibor BlahoAI score78

    OpenAI's DevDay 2026 brings dots agents, GPT-6.1 Sol, and Ultrafast speed tier

    AIOpenAI announced more than 20 updates at DevDay 2026, including dots always-on agents, GPT-6.1 Sol, Ultrafast token generation, ChatGPT Space, and a $500/month Pro 500 plan. GPT-6.1 Sol is priced at $2 input and $10 output per 1M tokens and is available in the API as gpt-6.1-sol. Ultrafast generates tokens up to 8x faster in Codex and up to 6x faster in the API.

    Why it matters: The post lists dozens of OpenAI DevDay 2026 changes across models, agents, plans, and APIs, useful for scanning what shipped and who gets access.

    Image from @btibor91's post
  3. Baseten BlogAI score38

    Baseten Partners With OpenAI to Offer Open Models to OpenAI Customers

    AIBaseten announced a partnership with OpenAI that makes open models powered by Baseten available to OpenAI customers for multi-model agentic coding. The company says organizations can route each task to the best-fit open or closed model, with Codex and GPT models among the options, and that Baseten's day-zero access to new open models lets teams evaluate them quickly. Baseten also cites US-based infrastructure with zero data retention for all prompts and capacity across more than 90 clusters in 20+ clouds.

  4. BAAI · new models on Hugging FaceAI score62

    BAAI releases AREX-2, a 27B agent model for self-improving long-horizon tasks

    AIBAAI released AREX-2, a 27B-parameter long-horizon agent model that improves solutions over multiple test-time rounds by proposing, measuring, reflecting, and revising. It was trained on machine-learning and algorithmic-programming tasks with verifiable feedback, and the source reports that this self-improvement transfers to deep research. The model is Apache License 2.0 licensed and has a 262,144-token context length.

    Why it matters: The source compares AREX-2 against closed and open models on coding and deep-research benchmarks, showing how test-time self-improvement is measured across task types.

  5. OpenAIAI score60

    OpenAI upgrades Codex Security Cloud with default access to cyber-capable models

    AIOpenAI says Codex Security Cloud is getting a major upgrade that includes access to cyber-capable models through Daybreak Blue by default. The upgraded tool scans entire GitHub repos, continuously reviews new commits, investigates and deduplicates findings, and prepares fixes for review even when the user's laptop is closed. It is available as a plugin in Codex desktop and web.

    Video from @OpenAI's post
  6. Replit BlogAI score62

    Replit Agent lets the core model choose subagents and effort instead of a router

    AIReplit explains how its Agent lets the core model pick subagent tier and effort mid-task rather than relying on an external router. On DeepSWE and Terminal-Bench, Replit Agent scored 72% at $2.11 per task and 49% at $2.53 per task, beating a single long-lived worker sidekick setup by 11 and 16 points. The company says Astra on its own scores higher only at more than twice the cost.

    Why it matters: The post gives a concrete harness design with benchmark cost-score comparisons, helping builders weigh delegation strategies against routers and single-worker setups.

  7. Alex HeathAI score34

    Factory CEO Matan Grinberg says AGI is already here

    AIFactory CEO Matan Grinberg, whose AI coding startup builds Droid agents, argues AGI is already here and explains why the company bets on many competing models. The discussion covers balancing model performance against token costs and why companies should avoid depending on a single AI provider. It also touches on hiring, the open-versus-closed AI debate, and competition with Cognition.

    Video from @alexeheath's post