Hugging Face urges open sharing of science via Paper Pages
AIHugging Face is asking users to try Paper Pages at to help keep scientific research open. The post provides no further technical details, figures, or feature specifics about the platform.
Updated
Updated
Items with an AI score under 20 are hidden. Show low-relevance items
AIHugging Face is asking users to try Paper Pages at to help keep scientific research open. The post provides no further technical details, figures, or feature specifics about the platform.
AILightOn has released LightOnOCR-3, a family of OCR models in 0.8B and 4B versions that it says lead benchmarks including OlmOCR-Bench and ParseBench, with the 0.8B model positioned as the sub-1B option. The models recognize text, handwriting, images, charts and document structure in one pass, process documents up to twice as fast as LightOnOCR-2, and are released under the Apache 2.0 license.

AISupermemory introduces MemoryRepo.dev, an open-source implementation of Cognition's dreaming memory system built on Cloudflare Artifacts, Durable Objects, Alchemy, and Effect. The project follows Cognition's Devin memory design, which builds a memory graph across sessions and prunes stale records overnight. Supermemory says it will incorporate learnings from this research into its own product.
AILithos AI says it is open-sourcing lithos-metal, which uses megakernels and DSpark speculative decoding. The post claims Qwen3.8-27B reaches a peak of over 200 tokens per second per user on a single Apple M5 Max. It says users can try the tool with any coding agent in one command, and links to the code on GitHub and a technical blog.
AICline says Step 5 Preview is now free in its coding tool and scores ahead of Kimi K3 and GLM-5.3 on DeepSWE. The company describes it as one of the strongest open-weights coding models available. StepFun's background announcement describes Step 5 Preview as a 600B total / 27B active MoE model with 1M context and vision, and says open weights arrive on Oct 15.

AICursor says it is excited to support open-source software alongside SpaceXAI, and it plans to bring more projects on board soon. Background from DHH indicates SpaceXAI joined the Omacom Foundation as a Founding Corporate Patron, contributing $1,500,000 in @grok tokens for maintaining and developing Omarchy.
AITessl's blog post argues that repository automation needs Continuous AI, a third pillar alongside CI and CD for scheduled, auditable AI workflows that improve repositories over time. The article describes GitHub Agentic Workflows, which harden agentic workflow specifications into GitHub Actions that can run coding agents such as Claude Code, Copilot CLI, Gemini CLI, or Codex-style agents. It emphasizes read-only agent steps, restricted outputs, and human review of pull requests.
AIRSIGym provides a research agent with training, inference, evals, and sandboxes as callable services, so it spends its budget on experiments rather than rebuilding infrastructure. With Opus 5 as the researcher, the improved system rose from 17.67% to 50.33% on SWE-bench Verified. The post also highlights a way to measure co-evolution between harnesses and models.
AIUnsloth now supports OS-level sandboxing on Linux via bwrap, on Mac via seatbelt, and on Windows via Microsoft's MXC. Per-tool-call latency is under 100ms across all three, and its software-style sandboxing with regex AST checks adds about 3ms. The Windows integration was built in collaboration with Microsoft.
AISuno announced that Albums are now live, letting users combine songs into a full release, set artwork, arrange the tracklist, and publish when ready. Existing playlists can be converted into Albums without rebuilding them from scratch.
AIOpenClaw released its v2026.9.9 patch, adding support for GPT-6.1 Sol in Codex and Claude Haiku 5.5. The update also improves recovery from failed updates, fixes missing iMessage replies, and resolves scheduled-job issues. The release credits 90 contributors.
AIJetBrains has released Mellum2.1, a 12B mixture-of-experts thinking model with 2.5B active parameters, under Apache 2.0 on Hugging Face. Post-training reinforcement learning in real software repositories raised SWE-bench Verified from 2.0 to 47.0, according to JetBrains' self-reported results. Qwen3.5-9B still leads on SWE-bench Pro, GPQA Diamond and AIME, and GGUF builds start at 7.0 GB for local use.
AIUnsloth now supports OS-level sandboxing on Windows by integrating Microsoft's open-source mxc repository for sandboxed code execution. The integration adds under 100 ms of overhead, according to the post. A setup guide is available in Unsloth's documentation.

AIGoodfire Research describes probe-based cyber monitors for Kimi K3 and GLM 5.3 deployed on a production inference stack. The probe filters suspicious exchanges before an LLM judge reviews them, reaching about 93% recall at a 5.5% benign-session interruption rate at roughly 50x lower judge cost. In FAR.AI's red-teaming, the monitor reduced universal jailbreaks to zero across 140 tested strategies.
AIJulien Chaumond, Hugging Face's account, mocked the idea of needing US or Chinese rivals when European partners exist. The remark responds to Berlin-based search engine Ecosia dropping French AI partner Mistral in favor of open-source models, including those from China.
AIStepFun's flagship model, Step 5 Preview, is available free for one week through Nous Portal, the platform from NousResearch. The post invites users to try it with Hermes Agent and share what they build.
AICarbon-A is an open model that predicts gene locations directly from DNA, and it has been used to annotate genomes from over 22,000 species. The release includes a database of 566 million gene candidates, about 16 times the gene annotations in the RefSeq dataset. Wet-lab RNA experiments supported 239 candidates missing from RefSeq across cats, Syrian hamsters, chickens, and Arabidopsis.
Why it matters: The source ties an open gene-annotation model to specific wet-lab checks and gene counts, helping readers judge how far its predictions extend beyond well-studied genomes.
AIOpenAI reportedly posted solutions to 90 of the top 500 open math problems, using an average of three hours of Pro-level compute per question. Anthropic released Claude Haiku 5.5 at $0.10 input and $0.50 output per million tokens, and the author says Jay Clayton was named AI Czar to head a new taskforce.
AIThomas Wolf says Carbon-A, an open model that finds genes directly in DNA, has been released with a database of 566.34 million candidate genes across 22,617 species. The team reports wet-lab validation of several new genes in cats, chickens and arabidopsis, and RNA evidence for 239 genes missing from reference annotations of common species.
This story has a top pick“Carbon-A open model and database predict 566 million gene candidates across 22,617 species”
AIMicrosoft has released MxC, an open-source sandboxing library that runs on Windows, macOS, and Linux. It builds on processcontainer, bubblewrap, and seatbelt as its underlying mechanisms. The project is aimed at developers struggling with fragmented sandbox tools for AI agents.
AICamilla Moraes from GitHub's Tiny Wins team joins the latest GitHub Podcast to discuss the AI slop problem, which has become one of the biggest pain points for open source maintainers. The episode also covers listening to maintainers and deciding what to fix next.

AIPerplexity's pplx-decider v1.1 scored 643 of 669 on clinical decisions versus 628 for Jev, according to a quoted post. The quoted post says it costs 42% less per decision at about the same speed and that its weights can be downloaded and run inside hospitals.
AIGoogle's SynthID Detector is now publicly available, letting users check whether an image, video, or audio file was generated by supported tools. Per the post, it scans for watermarks from Google and partners, including Nano Banana 2.1, OpenAI, NVIDIA, and Kakao, with Apple support coming soon. Uploaded files are deleted right after scanning.
AIStepFun is offering its flagship Step 5 Preview model free in OpenCode for one week, targeting agentic and professional work. The preview supports a 1M context window, multimodal input, and zero data retention.
AIOpenCode is offering free access to its Step 5 Preview for one week. The announcement highlights a 1M context window, multimodal support, and zero data retention.
AIMistral CEO Arthur Mensch replied that his company serves such customers too, pointing to nuclear energy as a possible obstacle. The reply responds to reports that Berlin-based search engine Ecosia is dropping Mistral as its AI partner for open-source models, including Chinese ones.
AISemiAnalysis argues that many businesses, especially low-margin ones, are offloading simpler software and white-collar tasks to increasingly capable open-source models. It frames the durability of frontier labs as depending on whether new tasks enabled by smarter frontier intelligence will outgrow the work moved to cheaper models. The post asks whether an economy could absorb 100 million superintelligent PhD-level experts quickly while still earning high ROI.
AIReJev, an independent community project, applied LoRA post-training to OpenBMB's MiniCPM5-2B for bounded agent decisions: state, question, and candidate options yield one choice. On its sealed 1,892-sample holdout, accuracy rose from 51.11% to 80.50% (+29.39 percentage points) with 0% invalid outputs, at about $5.31 in cumulative Modal billing including earlier experimental overhead. The authors describe this as an early, task-specific result, not parity with Jev.

AIAmazon Web Services launched an open-source Physical AI Toolchain that combines AWS services with NVIDIA's Physical AI software to cover data generation, model training, simulation, edge deployment, and continuous improvement for robots. AWS uses Amazon SageMaker for training and AWS IoT Greengrass for distributing models to edge devices, while NVIDIA contributes Isaac Sim, Isaac Lab, Isaac GR00T, and Cosmos. The toolchain is hardware-neutral and does not directly replace RoboMaker, which was shut down in 2025.
AIHugging Face CEO Clément Delangue says the field urgently needs more public traces of AI agents attacking and defending systems, since defenders cannot learn from activity they cannot see. He invites anyone holding such traces who faces pressure to keep them private to contact him directly.
AIllama.cpp v0.6.0 adds support for multimodal Jev-like models and a performance upgrade for Metal, according to Merve Noyan. The post says models can be served with a single command, `llama serve -hf ggml-org/Clef-Flash-GGUF`, and points to a trending list of decision models on Hugging Face.
AIIBM's torch-spyre integration makes Spyre, its dataflow inference accelerator, a native PyTorch device by mapping PyTorch's device, allocator, stream, and event abstractions onto the Spyre runtime and firmware. Tensors stay resident on device="spyre" between operations, and FX graphs remain in the Inductor compiler path. The approach gives eager and compiled execution one path with lower launch overhead.
AIJetBrains released Mellum2.1, a 12B mixture-of-experts model with 2.5B active parameters under the Apache 2.0 license, built for coding agents. Post-training shifted to reinforcement learning across thousands of environments and millions of sandboxed runs, and the model is available on Hugging Face. The source reports gains over Mellum2 on LiveCodeBench, AIME, GPQA Diamond, BFCL v4, IFEval, and SWE-bench Verified, and says it serves almost twice the tokens of Qwen3.5-9B under heavy load.
Why it matters: The post shows how reinforcement learning in real sandboxed environments changed a compact open model's repository work, with benchmark gains against Mellum2 and two peers.
AITencent Cloud has open-sourced Octop, a self-hosted multi-agent AI assistant platform aimed at families and small teams, with multi-user accounts and data kept on the user's own machine. The full text describes it as a single Python process that bundles the backend, web dashboard, CLI, IM gateway, cron jobs, and multi-agent runtime, with state rebuilt from SQLite on restart.

AIY Combinator CEO Garry Tan has open-sourced his complete Claude Code configuration as the gstack repository on GitHub. The post calls it one of the best GitHub repos ever, though it provides no details on the setup's contents or features.

AIvLLM v0.31.0 is released with 717 commits from 307 contributors, including 96 first-time contributors. Highlights include DeepSeek-V4.1-Flash support, a vllm preload command that keeps weights in GPU memory across restarts, and Model Runner V2 with draft-model speculative decoding. The release also adds large-scale serving, scheduling, and HiSparse fixes, with full notes linked on GitHub.

AIUniPat AI's PaperBenchX benchmark found the strongest model, GPT-6 Astra, fully reproduced only 13.98% of 93 real research-paper tasks across 12 scientific fields. Reproduction was judged by regenerating outputs in an isolated environment, with 3,168 expert-verified scoring items. UniPat has open-sourced 12 test tasks and kept 81 tasks closed to preserve long-term evaluation validity.
AILast week, several of South Korea's largest banks were hit by a cyberattack. A CrowdStrike report reportedly indicates the entire attack may have been carried out by one person. The attacker reportedly combined the open-source AI penetration tool ARTEX, DeepSeek v4.1-Flash, GLM-5.3, Grok 4.6, and Claude Code.

AIHuawei engineers presented XMFS, an experimental Linux kernel prototype filesystem, at the Linux Plumbers Conference in Prague on October 5. It aims to let applications reach cross-node shared memory on CXL 3.0 or Huawei unified bus servers through standard POSIX file calls. The code exists only on openEuler, not in the mainline Linux kernel.
AIPerplexity has released pplx-embed-v2-late, a pair of ColBERT-style multimodal embedding models in 0.6B and 9B sizes that retrieve text, images and rendered PDF pages in a shared embedding space. Both are available on Hugging Face under the MIT license, while a hosted API endpoint is planned but not yet live.