Skip to contentSkip to stories

Updated

Open source

Showing low-relevance items too. Hide low-relevance items

Jul 8

Jul 8Wed
  1. Cognition Blog (Devin, Windsurf)OfficialAI score47

    Cognition Tests Trustworthiness of SWE-1.7, Built on Kimi K2.7 Code

    AICognition says its SWE-1.7 model, developed from the open-source Kimi K2.7 Code base, performs as well as or better than leading U.S. frontier models on its new trustworthiness evaluation suite. The suite combines 145 politically sensitive questions, sampled in English and Chinese, with realistic coding scenarios to measure propaganda, censorship, and security behavior. Cognition says SWE-1.7 improves substantially over the base Kimi K2.7 Code model, though the company says the benchmarks are still in development.

Jul 5

Jul 5Sun

Jul 3

Jul 3Fri
  1. Arthur MenschXAI score34

    Mistral argues enterprises need open models and their own data for AI growth

    AIMistral CEO Arthur Mensch says enterprises should use open-source models because closed providers that force data retention gain leverage over their business. He argues companies should store data in open systems, control AI access rules, and build continuous training loops to shrink costs and create hard-to-copy systems. Mistral offers its Studio control plane and Forge training platform, deployed on customer infrastructure or through zero-data-retention hosting.

  2. Xiaomi MiMo · new models on Hugging FaceOfficialAI score22

    XiaomiMiMo releases MiMo-V2.5-DFlash model weights on Hugging Face

    AIA model repository named XiaomiMiMo/MiMo-V2.5-DFlash is listed on Hugging Face with 311B parameters and tensor types F32, BF16, and F8_E4M3. The README is empty, and the page reports 434 downloads last month and no Inference Provider deployment.

Jul 1

Jul 1Wed
  1. Mistral AI · new models on Hugging FaceOfficialAI score54

    Mistral AI releases Leanstral 1.5, an open-source Lean 4 code agent model

    AIMistral AI released Leanstral 1.5 on Hugging Face as an open-source code agent model for Lean 4 proof assistant tasks. The model uses 119B total parameters with 6.5B activated per token, a 256k context length, and accepts text and image input. The source gives setup paths through Mistral Vibe and a local vLLM server, with recommended settings of temperature 1.0 and reasoning effort set to high for complex prompts. The model is licensed under Apache 2.0.

  2. Jim FanXAI score51

    Jim Fan introduces ASPIRE, a self-evolving robot skills library for continual learning

    AIJim Fan announces ASPIRE, a system where coding agents use multimodal sensory traces from simulation and real robots to run evolutionary search over control programs and add the results to a growing skills library. The post claims up to a roughly 10x reduction in transfer learning tokens for sim-to-real and single-arm to bimanual transfer, and says the full stack will be open-sourced.

Jun 30

Jun 30Tue
  1. Jim FanXAI score60

    ASPIRE lets robots build an evolving skills library that transfers across tasks

    AIJim Fan introduces ASPIRE, a system in which coding agents observe multimodal sensory traces and run evolutionary search over control programs to distill skills into a growing library. The post says ASPIRE shares know-how rather than pixels or weights across the sim-to-real gap, reducing transfer learning tokens by up to about 10x. The author also says the full stack will be open-sourced and provides a gallery of 150+ tasks and 90+ skills.

    Video from @DrJimFan's post
  2. Xiaomi MiMoOfficialAI score22

    Xiaomi MiMo praised as developers build on open-weights models

    AIXiaomi MiMo's account celebrated growing developer adoption of its open-weights models, crediting Cline for building on MiMo. Cline's linked post announced a $9.99/month subscription offering 2-5x discounted access to GLM-5.2 and other open-weight models including DeepSeek, Kimi, MiniMax, MiMo, and Qwen, with a $1.99 promo for sign-ups via npm i -g cline.

Jun 28

Jun 28Sun
  1. PaddlePaddleOfficialAI score46

    PaddlePaddle announces Unlimited-OCR now runs in vLLM

    AIUnlimited-OCR, Baidu's long-context OCR model, now runs in vLLM, with a recipe provided for developers to try it. The background post says it parses entire books in one pass using Reference Sliding Window Attention (R-SWA), which keeps the KV cache fixed during decoding, and claims 35% faster throughput than DeepSeek-OCR at 6K output tokens.

  2. DeepSeek · new models on Hugging FaceOfficialAI score14

    DeepSeek publishes eagle3_gemma4_12b_ttt7 model on Hugging Face

    AIDeepSeek has posted the eagle3_gemma4_12b_ttt7 model on Hugging Face, listed at 2B parameters in BF16 tensor format with 621 downloads last month. The model card is empty and no Inference Provider currently deploys it, while it appears in the DeepSpec collection of 12 items.

  3. DeepSeek · new models on Hugging FaceOfficialAI score14

    DeepSeek posts eagle3_qwen3_14b_ttt7 draft model on Hugging Face

    AIDeepSeek published a model named eagle3_qwen3_14b_ttt7 on Hugging Face, listed at 2B parameters in BF16 with the Safetensors format. The model has no model card, is not deployed by any Inference Provider, and recorded 400 downloads last month.

Jun 27

Jun 27Sat
  1. PaddlePaddleOfficialAI score36

    PaddleFormers 1.2 adds DeepSeek-V4 training with 128K+ context support

    AIPaddleFormers 1.2 is released with support for training DeepSeek-V4 and 128K+ long-context training. The update adds Context Parallel, Packing, Document Mask Attention, and the Muon optimizer, plus ultra-fused mHC, CSA, and HCA operators, DeepEP/HybridEP communication, and lossless FP8 training with AutoSubbatch memory balancing. The project is presented as fully open-source and is available on GitHub.

  2. Ahead of AI (Sebastian Raschka)BlogAI score37

    Local Coding Agents: Setting Up Qwen3.6 with Open-Source Harnesses

    AISebastian Raschka's tutorial shows how to build a fully local coding agent by pairing an open-weight LLM served through an inference runtime with an open-source harness that can read files, edit code, and run commands. He recommends Qwen-Code for Qwen3.6, citing Nvidia's Polar paper, which found Qwen models performed best in Qwen-Code. The Qwen3.6 35B-A3B model is about 22 GB to download and needs roughly 30–40 GB of RAM.

Jun 26

Jun 26Fri
  1. PaddlePaddleOfficialAI score32

    PP-OCRv6 Ep.4 benchmarks show 3.9x CPU speedup and 0.13s A100 OCR

    AIPaddlePaddle's PP-OCRv6 Tech Deep Dive Ep.4 benchmarks the OCR models across A100, V100, Intel Xeon CPU, and Apple M4 setups. PP-OCRv6_tiny processes an image in 0.13s on A100, while PP-OCRv6_tiny with OpenVINO runs 3.9x faster than PP-OCRv5_mobile on Intel CPU. The post recommends Medium for high-concurrency APIs, Small for CPU document systems, Tiny for mobile or embedded devices, and Medium or Small for multilingual business use.

    Image from @PaddlePaddle's post
  2. Qwen · new models on Hugging FaceOfficialAI score44

    Qwen3-ForcedAligner-0.6B-hf Adds Timestamp Alignment for Speech Transcripts

    AIQwen released Qwen3-ForcedAligner-0.6B-hf, a Transformers-format forced aligner that predicts timestamps for arbitrary units within up to 5 minutes of speech in 11 languages. The model accepts transcripts from any ASR system, and the documentation shows it paired with Qwen3-ASR-0.6B and NVIDIA Parakeet CTC. Until it ships in an official Transformers release, users must install Transformers from source.

Jun 25

Jun 25Thu

Jun 23

Jun 23Tue
  1. PaddlePaddleOfficialAI score38

    PP-OCRv6 lightweight OCR model challenges large VLMs with 34.5M params

    AIPaddlePaddle introduced PP-OCRv6, a lightweight OCR architecture built on the LCNetV4 backbone, in the first episode of its tech deep dive series. The post says PP-OCRv6_medium reaches 86.2% detection Hmean and 83.2% recognition accuracy, surpassing PP-OCRv5_server while running faster. Three model specs—Tiny, Small, and Medium—target edge CPU devices, balanced deployment, and industrial high-accuracy pipelines.

    Image from @PaddlePaddle's post

Jun 20

Jun 20Sat

Jun 19

Jun 19Fri
  1. Andrew NgXAI score72

    Andrew Ng says Anthropic and U.S. export controls on Fable expose AI access risks

    AIAndrew Ng argues that Anthropic's restrictions on building competing LLMs and a U.S. Commerce Department license requirement for foreign nationals led Anthropic to disable Fable access worldwide. He says this shows governments and providers can quickly cut off access to frontier AI, which may push nations and businesses toward sovereignty efforts and open-source alternatives, though training frontier models remains difficult.

    Why it matters: The post links Anthropic's usage restrictions and a U.S. export license requirement to renewed interest in AI sovereignty and open-source alternatives, which bears on how builders assess provider dependence.

    Image from @AndrewYNg's post

Jun 18

Jun 18Thu
  1. Cohere · new models on Hugging FaceOfficialAI score43

    Cohere Releases Open-Source 2B Arabic Speech Recognition Model Transcribe Arabic

    AICohere and Cohere Labs released Cohere Transcribe Arabic, an open-source 2B-parameter Arabic automatic speech recognition model under Apache 2.0. It is optimized for Arabic, Arabic dialects, English, and Arabic-English code-switched speech, using a Conformer encoder-decoder architecture supported natively in Transformers. The model's average WER of 25.87 and CER of 11.80 on the Open Universal Arabic ASR Leaderboard, as of 07.07.2026, is reported in the source.

Jun 17

Jun 17Wed
  1. Jim FanXAI score64

    ENPIRE lets Codex agents run autonomous research on a robot fleet

    AINVIDIA GEAR's ENPIRE gives eight Codex agents a fleet of robots, GPUs, and a token budget to solve physical tasks with minimal human oversight. The author reports tasks such as tying zip-ties, organizing fine pins, and installing GPUs, and a faster time-to-solution with eight parallel robots than with fewer. Safety uses a kinematic limit that resets a robot leaving its envelope, a torque-limited gripper, and a frozen reward function classifier. The team says everything will be open-sourced.

    Video from @DrJimFan's post
  2. PaddlePaddleOfficialAI score40

    PaddleOCR 3.7 adds ONNX Runtime backend and PP-OCRv6 models

    AIPaddleOCR 3.7 adds an ONNX Runtime inference backend, switchable with a single parameter and no code changes. The release supports CPU via OpenVINO and GPU via CUDA and TensorRT, and introduces PP-OCRv6 Tiny, Small, and Medium models that run up to 3.9× faster on CPU.

    Image from @PaddlePaddle's post

Jun 16

Jun 16Tue
  1. Jim FanXAI score62

    Jim Fan's ENPIRE lets Codex agents run autonomous research on robot fleets

    AIJim Fan introduces ENPIRE, which gives eight Codex agents a fleet of robots, GPUs, and a token budget to solve physical tasks autonomously. The post reports that the system can tie zip-ties, organize fine pins, and install GPUs, and that eight robots exploring in parallel improve faster than fewer. The team plans to open-source everything.

    Video from @DrJimFan's post
  2. Arthur MenschXAI score44

    Mistral says its upcoming models will all be open-weight

    AIMistral states that this model and upcoming ones will be open-weight. The company argues that open weights are critical for customer confidence and for research and developer communities. It contends that systems reachable only through someone else's interface cannot be owned, inspected, audited, or improved, especially if data recording can no longer be turned off.

  3. Z.ai (GLM) · new models on Hugging FaceOfficialAI score72

    Z.ai releases GLM-5.2 with 1M-token context and MIT open-source license

    AIZ.ai has released GLM-5.2, its flagship model for long-horizon tasks, which it says substantially improves on GLM-5.1 and supports a 1M-token context. The model adds IndexShare, which cuts per-token FLOPs by 2.9× at 1M context, and is released under the MIT open-source license.

    Why it matters: The source gives benchmark tables against named rival models and deployment settings, useful for judging where GLM-5.2 sits among current flagship models.

Jun 15

Jun 15Mon
  1. Zed BlogOfficialAI score38

    Zed Guild Cohort 1 Ends with 148 Merged Pull Requests from 33 Contributors

    AIZed's 12-week Guild program, its first cohort run this spring, had 33 active contributors merge 148 pull requests into the open-source editor. The top contributor, feitreim, merged 23 PRs, including fixes for Vim mode screen flickering and terminal ANSI rendering, and won a trip to Rust Week in Utrecht. Zed plans to organize Cohort 2 work into tighter groups around specific parts of the codebase.

  2. ByteDance · new models on Hugging FaceOfficialAI score24

    Sa2VA-LLaVA-1.5-7B: ByteDance's SAM2-Grounded Segmentation and Chat Model

    AIByteDance has released Sa2VA-LLaVA-1.5-7B on Hugging Face, a model built on LLaVA-1.5-7B with a SAM2 grounding encoder that performs dense image and video referring segmentation alongside open-ended chat. The checkpoint is self-contained and loads with trust_remote_code=True without extra packages, and it is positioned as a LISA-comparable baseline within the Sa2VA family. Reported results include 80.3 cIoU on RefCOCO val and 54.8 J&F on MeViS (val_u).

Jun 13

Jun 13Sat
  1. Moonshot AI (Kimi) · new models on Hugging FaceOfficialAI score88

    Moonshot AI releases open-weight Kimi K3 with 2.8T parameters and 1M context

    AIMoonshot AI released Kimi K3 on Hugging Face as an open-weight, native multimodal agentic model with 2.8T total parameters and 104B activated parameters. It supports a 1-million-token context window and text and image input, with weights released under the Kimi K3 License. The model card reports benchmark results for coding, agentic, and vision tasks against several closed models, and recommends vLLM, SGLang, or TokenSpeed for inference.

    Why it matters: The release pairs open weights with a 2.8T-parameter MoE architecture and benchmark tables against several named closed models, useful for comparing frontier capability claims.

Jun 12

Jun 12Fri
  1. PaddlePaddleOfficialAI score41

    PaddleOCR releases PP-OCRv6 with models from 1.5M to 34.5M parameters

    AIPaddlePaddle has released PP-OCRv6, a new OCR model series in Tiny, Small, and Medium sizes at 1.5M, 7.7M, and 34.5M parameters. The models reportedly improve detection accuracy by 4.9% and recognition accuracy by 5.1% over PP-OCRv5, with up to 5.2× faster CPU inference via OpenVINO. The unified model supports 50 languages and new scenarios including PCB, CAD drawings, digital tubes, and dot-matrix text, under Apache 2.0.

    Image from @PaddlePaddle's post

Jun 11

Jun 11Thu
  1. Moonshot AI (Kimi) · new models on Hugging FaceOfficialAI score62

    Moonshot AI releases Kimi K2.7 Code, a coding-focused agentic model

    AIMoonshot AI published Kimi-K2.7-Code, a coding-focused agentic model built on Kimi K2.6, with a 1T-parameter MoE architecture and 32B activated parameters. The model card reports about 30% fewer thinking tokens than K2.6 and benchmark results against GPT-5.5 and Claude Opus 4.8, with weights and code released under a Modified MIT License.

    Why it matters: The model card gives benchmark comparisons against GPT-5.5 and Claude Opus 4.8 on coding and agentic tasks, useful for judging its position among current coding models.

Jun 10

Jun 10Wed
  1. Xiaomi MiMoOfficialAI score67

    Xiaomi releases open-source MiMo Code V0.1 terminal coding assistant

    AIXiaomi MiMo has released MiMo Code V0.1, an open-source AI coding assistant for the terminal under the MIT license. It ships with MiMo V2.5, a multimodal model offered free for a limited time with a million-token context window. The tool automatically loads existing Claude Code skills, MCP servers and commands, and reuses API configuration, and it supports providers including Anthropic, OpenAI, DeepSeek, Kimi and GLM.

    Why it matters: The post specifies MiMo Code's Claude Code compatibility and MIT license, which bear directly on whether existing coding-agent setups can migrate without rework.

    Image from @XiaomiMiMo's post