๐Ÿค– GitHub AI ็ƒญ้—จๆŠฅๅ‘Š

๐Ÿ“… ็”Ÿๆˆๆ—ถ้—ด: 2026/8/5 22:17:28  |  ๐Ÿ”‘ API Token: โœ… ๅทฒ้…็ฝฎ (5000ๆฌก/ๅฐๆ—ถ)

๐ŸŒŸ ไธ€ใ€AI ็ƒญ้—จไป“ๅบ“ Top 25

ๅŸบไบŽ GitHub Stars ๆŽ’ๅ็š„ๅฝ“ๅ‰ๆœ€็ƒญ้—จ AI ็›ธๅ…ณไป“ๅบ“

#ไป“ๅบ“StarsForks่ฏญ่จ€็ฎ€ไป‹
1rasbt/LLMs-from-scratchโญ 100.6k๐Ÿด 15.5kJupyter NotebookImplement a ChatGPT-like LLM in PyTorch from scratch, step by step
2microsoft/AI-For-Beginnersโญ 62.0k๐Ÿด 12.0kJupyter Notebook12 Weeks, 24 Lessons, AI for All!
3ashishpatel26/500-AI-Machine-learning-Deep-learning-Computer-vision-NLP-Projects-with-codeโญ 36.0k๐Ÿด 7.4kN/A500 AI Machine learning Deep learning Computer vision NLP Projects with code
4explosion/spaCyโญ 33.8k๐Ÿด 4.7kPython๐Ÿ’ซ Industrial-strength Natural Language Processing (NLP) in Python
5Lightning-AI/pytorch-lightningโญ 31.3k๐Ÿด 3.8kPythonPretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.
6AMAI-GmbH/AI-Expert-Roadmapโญ 31.2k๐Ÿด 2.6kJavaScriptRoadmap to becoming an Artificial Intelligence Expert in 2022
7harvard-edge/cs249r_bookโญ 27.7k๐Ÿด 3.5kPythonMachine Learning Systems
8recommenders-team/recommendersโญ 21.8k๐Ÿด 3.3kPythonBest Practices on Recommendation Systems
9huggingface/datasetsโญ 21.8k๐Ÿด 3.3kPython๐Ÿค— The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data ...
10onnx/onnxโญ 21.3k๐Ÿด 4.0kPythonOpen standard for machine learning interoperability
11stefan-jansen/machine-learning-for-tradingโญ 20.3k๐Ÿด 5.5kJupyter NotebookCode for Machine Learning for Trading, 3rd edition โ€” from data sourcing to live execution.
12owainlewis/awesome-artificial-intelligenceโญ 15.7k๐Ÿด 2.5kPythonA curated list of Artificial Intelligence (AI) courses, books, video lectures and papers.
13tensorzero/tensorzeroโญ 11.7k๐Ÿด 958RustTensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation,...
14kornia/korniaโญ 11.3k๐Ÿด 1.2kPython๐Ÿ Geometric Computer Vision Library for Spatial AI
15voxel51/fiftyoneโญ 11.0k๐Ÿด 806TypeScriptRefine high-quality datasets and visual AI models
16CVHub520/X-AnyLabelingโญ 10.0k๐Ÿด 1.1kPythonX-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotati...
17HenryNdubuaku/maths-cs-ai-compendiumโญ 7.2k๐Ÿด 889TypeScriptBecome a cracked AI/ML researcher/engineer with this unconventional textbook covering maths, computi...
18flwrlabs/flowerโญ 7.1k๐Ÿด 1.2kPythonFlower: A Friendly Federated AI Framework
19deeppavlov/DeepPavlovโญ 7.0k๐Ÿด 1.2kPythonAn open source library for deep learning end-to-end dialog systems and chatbots.
20amusi/AI-Job-Notesโญ 6.1k๐Ÿด 666N/AAI็ฎ—ๆณ•ๅฒ—ๆฑ‚่Œๆ”ป็•ฅ๏ผˆๆถต็›–ๅ‡†ๅค‡ๆ”ป็•ฅใ€ๅˆท้ข˜ๆŒ‡ๅ—ใ€ๅ†…ๆŽจๅ’ŒAIๅ…ฌๅธๆธ…ๅ•็ญ‰่ต„ๆ–™๏ผ‰
21louisfb01/start-machine-learningโญ 5.3k๐Ÿด 700N/AA complete guide to start and improve in machine learning (ML), artificial intelligence (AI) in 2026...
22sktime/pytorch-forecastingโญ 5.0k๐Ÿด 883PythonTime series forecasting with PyTorch
23BoltzmannEntropy/interviews.aiโญ 4.9k๐Ÿด 325N/AIt is my belief that you, the postgraduate students and job-seekers for whom the book is primarily m...
24rasbt/reasoning-from-scratchโญ 4.9k๐Ÿด 746Jupyter NotebookImplement a reasoning LLM in PyTorch from scratch, step by step
25alirezadir/Production-Level-Deep-Learningโญ 4.7k๐Ÿด 686N/AA guideline for building practical production-level deep learning systems to be deployed in real wor...

๐Ÿ†• ไบŒใ€่ฟ‘30ๅคฉๆ–ฐ่ฏž็”Ÿ็š„ๆ˜Žๆ˜Ÿ AI ้กน็›ฎ

ๆœ€่ฟ‘30ๅคฉๅ†…ๅˆ›ๅปบไธ” Stars ๅขž้•ฟๆœ€ๅฟซ็š„ AI ้กน็›ฎ

#ไป“ๅบ“Starsๅˆ›ๅปบๆ—ฅๆœŸ่ฏญ่จ€็ฎ€ไป‹
1xai-org/grok-buildโญ 24.2k๐Ÿ“… 2026/07/15RustSpaceXAI's coding agent harness and TUI. Fullscreen, mouse interactive, extensible.
2yc-software/qmโญ 11.5k๐Ÿ“… 2026/07/30TypeScriptMultiplayer agent harness for work
3unicity-aos/aos-ceโญ 8.6k๐Ÿ“… 2026/07/13RustAOS Community Edition: the open agent operating system.
4trycompai/crmโญ 5.7k๐Ÿ“… 2026/08/01TypeScriptAn open-source, agentic-first CRM.
5drumih/turbo-fieldfareโญ 5.0k๐Ÿ“… 2026/07/17SwiftGemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
6petergyang/no-ai-slopโญ 4.0k๐Ÿ“… 2026/07/07PythonRemoves 20+ patterns of AI slop from any piece of writing.
7Vincentwei1021/video-shotcraftโญ 3.6k๐Ÿ“… 2026/07/19TypeScriptAI video skill for Claude Code & Codex โ€” cinematic product videos with Remotion: 106 shot recipe car...
8slvDev/esp32-aiโญ 3.5k๐Ÿ“… 2026/07/23Python
9jakubkrehel/skillsโญ 3.1k๐Ÿ“… 2026/07/10N/AA collection of agent skills that help with various parts of building a great interface. From animat...
10kvcache-ai/AgentENVโญ 2.9k๐Ÿ“… 2026/07/23RustAgentENV (AENV) is a distributed platform for running agent environments at scale.
11lopopolo/harness-engineeringโญ 2.5k๐Ÿ“… 2026/07/19Python๐ŸŽ Ryan Lopopoloโ€™s anthology, field guide, and agent context bundle for harness engineering
12FareedKhan-dev/kimi-k3-in-cโญ 2.4k๐Ÿ“… 2026/08/01CA 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99:...
13MIgHTy-alIeN/MEV-Ethereum-Trading-Botโญ 2.3k๐Ÿ“… 2026/07/17SolidityAn arbitrage bot is a smart contract connected to an external automation script that controls its op...
14AlephAITech/WorkBuddyGuideโญ 2.0k๐Ÿ“… 2026/07/10PythonA practical, open-source guide to mastering WorkBuddy through real-world workflows.ๅผ€ๆบ็š„ WorkBuddy ๅฎžๆˆ˜่“...
15QwenAudio/qwen-audio-agentโญ 1.9k๐Ÿ“… 2026/07/27JavaScriptA realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime f...
16penecho/penechoโญ 1.9k๐Ÿ“… 2026/07/14JavaScriptThink with AI beyond the chat box. A shared canvas for handwriting, equations, diagrams, and spatial...
17makecindy/cindyโญ 1.8k๐Ÿ“… 2026/07/23TypeScriptConsider it done. The open-source AI agent that works out of the box ยท ๆƒณๅˆฐ๏ผŒๅฐฑ่ƒฝๅšๅˆฐใ€‚ๅผ€ๆบใ€ๅผ€็ฎฑๅณ็”จ็š„ AI Agentใ€‚
18Jakubantalik/thinking-orbsโญ 1.7k๐Ÿ“… 2026/07/21TypeScriptDotted thought-orb loading indicators for AI & agent UIs, 9 tuned types, two sizes, auto dark/light
19AminBlg/SimpleEnglishโญ 1.7k๐Ÿ“… 2026/07/21PythonAgent skill: make LLMs write docs in ASD-STE100 Simplified Technical English โ€” no AI slop
20QoderAI/better-harnessโญ 1.6k๐Ÿ“… 2026/07/21JavaScriptBetter Harness turns project and session evidence into loop-level insights, prioritized improvements...
21genspark-ai/genofficeโญ 1.6k๐Ÿ“… 2026/07/31TypeScriptAn AI-native office suite for macOS and Windows: word processor, spreadsheet, presentations, and PDF...
22v-modal/vmodal_sdk_flutterโญ 1.4k๐Ÿ“… 2026/07/16DartV- Modal AI: MultiModal Video Search - SDK Flutter
23Kritt-ai/open-krittโญ 1.4k๐Ÿ“… 2026/07/21JavaScriptOrchestrate AI agents to find real vulnerabilities in code.
24nethical6/conversation-steganographyโญ 1.2k๐Ÿ“… 2026/07/17GoUse LLMs to hide messages inside normal looking conversations
25cosmtrek/mindwalkโญ 1.1k๐Ÿ“… 2026/07/09GoA visualization tool that replays coding-agent sessions on a 3D map of your codebase.

๐Ÿš€ ไธ‰ใ€่ฟ‘7ๅคฉ้ฃž้€Ÿๅขž้•ฟ็š„ๆ–ฐ้กน็›ฎ

ๆœ€่ฟ‘7ๅคฉๅ†…ๅˆ›ๅปบ็š„ๆฝœๅŠ›้กน็›ฎ (Stars > 10)

โš ๏ธ ๆš‚ๆ— ๆ•ฐๆฎ

๐Ÿ“ฆ ๅ››ใ€ๆ˜Žๆ˜Ÿไป“ๅบ“ๆœ€ๆ–ฐ Release ๅŠจๆ€

่ฟฝ่ธช็š„ 52 ไธชๆ˜Žๆ˜Ÿไป“ๅบ“ไธญๆœ‰ๆœ€ๆ–ฐ Release ็š„้กน็›ฎ

๐Ÿ“Œ Significant-Gravitas/AutoGPT

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# ๐Ÿš€ Release `autogpt-platform-beta-v0.7.0`

**Date:** August 2026

---

## ๐Ÿ”ฅ What's New?

### New Features
- **#13687** - Expert-scoped sessions and identity context in Copilot backend (by @0ubbe)
- **#13689** - Experts marketplace section, team page, and per-expert threads (by @0ubbe)
- **#13699** - Compact wallet popover for the new layout (by @Abhi1992002)
- **#13330** - Replace Supabase Auth with Better Auth (by @ntindle)
- **#13627** - Single-source LLM model catalog (cutover + Kimi K3) (by @ntindle)
- **#13764** - Voice brain-dump onboarding step (by @Abhi1992002)

### UI/UX Improvements
- **#13751** - Replace Phosphor icons with Hugeicons stroke-rounded (by @Abhi1992002)

### Bug Fixes

๐Ÿ“Œ qdrant/qdrant

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# Change log

## Features :point_up:

* https://github.com/qdrant/qdrant/milestone/50 - TurboQuant 4-bit as a datatype of primary vector storage. Only store 4-bit quantized vectors and spare disk space on original vectors. [[docs](https://qdrant.tech/documentation/manage-data/vectors/#turbo4)]
* https://github.com/qdrant/qdrant/pull/9669, https://github.com/qdrant/qdrant/pull/9684, https://github.com/qdrant/qdrant/pull/9950 - Unify definition of memory usage strategy for collection components. Use `"memory": "cold" / "cached" / "pinned"` to define memory behavior for each individual collection component. Allows for more fine-grained control over memory usage and performance. [[docs](https://qdrant.tech/documentation/ops-configuration/memory-tiers/)]
* https://github.com/qdrant/qdrant/pull/9683 - Allow `"match": {"prefix": "..."}` in `filter` to match keywords by prefix, must be enabled in keyword index. [[docs](https://qdrant.tech/documentation/search/filtering/#prefix-match)]
* https://github.com/qdrant/qdrant/pull/9661 - Per-query IDF corpus for sparse vector search [[docs](https://qdrant.tech/documentation/search/text-search/full-text-search/#per-tenant-idf-statistics)]
* https://github.com/qdrant/qdrant/pull/9899 - Slice filtering condition: sliced scroll / deterministic sampling [[docs](https://qdrant.tech/documentation/search/filtering/#slice)]
* https://github.com/qdrant/qdrant/pull/10035 - Global quota API [[docs](https://qdrant.tech/documentation/ops-configuration/quotas/)]
* https://github.com/qdrant/qdrant/pull/9338 - Add routing token for deterministic read routes [[docs](https://qdrant.tech/documentation/scaling/consistency-guarantees/#read-affinity)]

## Improvements :point_left:

* https://github.com/qdrant/qdrant/pull/9113 - Use batched reads in vector and payload storage
* https://github.com/qdrant/qdrant/pull/9409 - Utilize `io_uring` for payload storage
* https://github.com/qdrant/qdrant/pull/9448 - Lower default update queue length, reduce from 1M to 200
* https://github.com/qdrant/qdrant/pull/9500 - Use Entry API to avoid redundant map double-lookups, improving performance
* https://github.com/qdrant/qdrant/pull/9332 - Enable single file mmap vector storage by default for immutable segments
* https://github.com/qdrant/qdrant/pull/9376 - Add option to explicitly disable BM25 stemmer; deprecate "none" hack

๐Ÿ“Œ nltk/nltk

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
Version 3.10.2 2026-08-05

* Remove inisec.py and document PYTHONSAFEPATH instead
* Skip draft step in release workflow
* Fix symlink escape in FramenetCorpusReader (CWE-59)
* Guard tempfile.gettempdir() when building pathsec allowed roots
* add tests for transitive_closure

Thanks to the following contributors to 3.10.2:
Litesh Ghute, Eric Kafe, Evan Kiefer, tarann26 and Rav Singh Chandan

๐Ÿ“Œ streamlit/streamlit

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
<!-- Release notes generated using configuration in .github/release.yml at 1.61.0 -->

## What's Changed
### Breaking Changes ๐Ÿ› 
* [chore] Remove deprecated `use_column_width` parameter from `st.image` by @lukasmasuch in https://github.com/streamlit/streamlit/pull/15786
* [fix] Deprecate string file paths in st.html/st.iframe in favour of pathlib.Path by @lukasmasuch in https://github.com/streamlit/streamlit/pull/16150
### New Features ๐ŸŽ‰
* [feat] Add lazy loading for st.dataframe by @lukasmasuch in https://github.com/streamlit/streamlit/pull/15756
* [feature] Add `icon` parameter to `st.metric` (#12298) by @sfc-gh-dbyttow in https://github.com/streamlit/streamlit/pull/15805
* [feature] Add background refresh for st.cache_data and st.cache_resource by @lukasmasuch in https://github.com/streamlit/streamlit/pull/16057
* [feature] Add WebSocket Host allow-list by @lukasmasuch in https://github.com/streamlit/streamlit/pull/16147
* Add seconds granularity and `format` for hour cycle to `st.time_input` by @mayagbarnes in https://github.com/streamlit/streamlit/pull/16126
* Improve form support and enter to submit to `st.time_input` by @mayagbarnes in https://github.com/streamlit/streamlit/pull/16128
* Infer download_button file_name and mime from file object name by @SoMika00 in https://github.com/streamlit/streamlit/pull/16061
* [feature] Enforce disabled widget parameter server-side by @lukasmasuch in https://github.com/streamlit/streamlit/pull/16209
* Improve tag accessibility in `st.multiselect` by @mayagbarnes in https://github.com/streamlit/streamlit/pull/16200
### Bug Fixes ๐Ÿ›
* [fix] Popover in sidebar rendered off-screen by @sfc-gh-dbyttow in https://github.com/streamlit/streamlit/pull/16087
* [fix] Escape now clears typed search query in st.selectbox by @sfc-gh-dbyttow in https://github.com/streamlit/streamlit/pull/16088
* Fix AppTest sys.path handling for script paths by @ishaanlabs-gg in https://github.com/streamlit/streamlit/pull/15775

๐Ÿ“Œ openai/openai-python

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## [2.53.0](https://github.com/openai/openai-python/compare/v2.52.1...v2.53.0) (2026-08-03)


### Features

* **api:** Add gpt-5.5 and tool name/namespace to Responses types ([#3569](https://github.com/openai/openai-python/issues/3569)) ([dd1202d](https://github.com/openai/openai-python/commit/dd1202d5dacff985861289c1d9c46996ded2d2a5))


### Bug Fixes

* **ci:** avoid NumPy source builds and duplicate HTTPX coverage ([#3573](https://github.com/openai/openai-python/issues/3573)) ([b58332f](https://github.com/openai/openai-python/commit/b58332f8a0717f7b1effb1788a594011cee6e02f))

๐Ÿ“Œ BerriAI/litellm

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## Verify Docker Image Signature

All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0).

**Verify using the pinned commit hash (recommended):**

A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \
  ghcr.io/berriai/litellm:v1.95.0
```

**Verify using the release tag (convenience):**

Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules:

```bash
cosign verify \

๐Ÿ“Œ mlflow/mlflow

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
MLflow 3.15.1 is a patch release that includes bug fixes and documentation updates.

### Bug fixes:

- [Model Registry] Skip `env_pack` on ARM client images (#24762, @qyc)
- [Scoring / Tracking] Harden version parsing against missing/non-PEP440 versions on Databricks Serverless (#24799, @PattaraS)

### Documentation updates:

- [Docs] Clarify scorer versioning documentation (#24769, @nihalmenon)

๐Ÿ“Œ ultralytics/ultralytics

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## ๐ŸŒŸ Summary

**v8.4.115 transitions Ultralytics from legacy HUB integrations to the streamlined [Ultralytics Platform](https://platform.ultralytics.com/) experience, with simpler authentication and a leaner training codebase.** ๐Ÿš€

## ๐Ÿ“Š Key Changes

- ๐Ÿ” **Introduced validated Platform CLI authentication**
  - Log in with `yolo login API_KEY`
  - Remove credentials with `yolo logout`
  - API keys are checked against the Platform before being saved.

- ๐Ÿ”„ **Added settings migration to schema `0.0.7`**
  - Existing compatible settings, such as custom dataset and run directories, are preserved.
  - Legacy HUB configuration and incompatible HUB API keys are removed automatically.
  - Users with old credentials are directed to create a Platform API key.

- ๐Ÿงน **Removed the legacy `ultralytics.hub` package**
  - HUB authentication, remote training sessions, model loading, exports, dataset utilities, callbacks, and HUB-specific exceptions have been retired.
  - HUB-related API references, documentation pages, navigation entries, and the HUB example notebook were also removed.

๐Ÿ“Œ gradio-app/gradio

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
### Features

-   [#13685](https://github.com/gradio-app/gradio/pull/13685) [`6302098`](https://github.com/gradio-app/gradio/commit/6302098bc300df0dd36f2cbea904bafca03b208d) - `gr.Workflow`: auto-create input/output nodes for model nodes (as already happens for Space nodes), replace the "Input"/"Output" buttons with a single "Component" button whose direction is derived from wiring, rename "Data" to "Dataset", let node error messages be copied, and document the `oauth_token` parameter in the View API panel.  Thanks @abidlabs!
-   [#13688](https://github.com/gradio-app/gradio/pull/13688) [`321361f`](https://github.com/gradio-app/gradio/commit/321361fd8de9942e7046dcb54987d18a2a091e7e) - Workflow: resizable nodes, full-screen image view, and webcam/mic capture.  Thanks @abidlabs!
-   [#13697](https://github.com/gradio-app/gradio/pull/13697) [`3ac9d5d`](https://github.com/gradio-app/gradio/commit/3ac9d5db5892e88f57260215793feeaa06bdde59) - Fix release CI regressions for assets and bundles.  Thanks @abidlabs!

### Fixes

-   [#13692](https://github.com/gradio-app/gradio/pull/13692) [`3676c45`](https://github.com/gradio-app/gradio/commit/3676c45acfc12456de097996fe5adab2132e2d30) - publish the `Prism` global before its grammar files load, so the docs pages stop failing to hydrate with `ReferenceError: Prism is not defined`.  Thanks @abidlabs!
-   [#13687](https://github.com/gradio-app/gradio/pull/13687) [`fd79d09`](https://github.com/gradio-app/gradio/commit/fd79d0999e66a3331a59d2ac80a7b61ec3e90f46) - Harden authentication and file redirect boundaries.  Thanks @abidlabs!

๐Ÿ“Œ langchain-ai/langgraph

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
Changes since checkpointsqlite==3.1.0

* release(checkpoint-sqlite): 3.1.1 (#8481)
* fix(checkpoint-postgres,checkpoint-sqlite): scope namespace matching to segment boundaries (#8478)
* chore(deps): bump the minor-and-patch group in /libs/checkpoint-sqlite with 4 updates (#8249)
* chore(deps): bump langsmith from 0.8.0 to 0.8.18 in /libs/checkpoint-sqlite (#8177)
* docs: standardize package `README.md` structure (#8064)
* chore: migrate Python type checking to ty (#8002)
* chore(deps): bump the minor-and-patch group in /libs/checkpoint-sqlite with 3 updates (#7961)
* release(checkpoint): 4.1.1 (#7890)
* chore(deps): bump langsmith from 0.7.31 to 0.8.0 in /libs/checkpoint-sqlite (#7786)
* chore(deps): bump idna from 3.11 to 3.15 in /libs/checkpoint-sqlite (#7862)

๐Ÿ“Œ langchain-ai/langchain

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
Changes since langchain-core==1.5.2

release(core): 1.5.3 (#39145)
fix(core): fall back to `LANGSMITH_API_KEY` for gateway (#39115)

๐Ÿ“Œ keras-team/keras

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# Keras 3.12.4

Keras 3.12.4 is a security patch release that hardens dataset loading and model file handling against insecure deserialization and decompression-bomb attacks.

## Security Fixes

* **Restrict unpickling when loading IMDB and Reuters datasets** โ€” Replaces `np.load(allow_pickle=True)` with a restricted unpickler that only permits numpy array reconstruction, preventing arbitrary code execution via crafted `.npz` files (CWE-502). ([#23047](https://github.com/keras-team/keras/pull/23047)) by @LinZiyuu
* **Verify all intermediary H5 groups when navigating H5 files** โ€” Manually resolves nested H5 group paths to verify group types at each step, preventing potential path traversal. ([#23168](https://github.com/keras-team/keras/pull/23168)) by @hertschuh
* **Reject decompression-bomb members on the `.keras` asset extraction path** โ€” Adds per-member decompression-ratio checks before extracting `.keras` archives to disk, preventing disk-exhaustion attacks via crafted archives. ([#23101](https://github.com/keras-team/keras/pull/23101)) by @LinZiyuu
* **Restrict unpickling when loading CIFAR datasets** โ€” Replaces bare `cPickle.load` in CIFAR-10/100 batch loading with the numpy-only `RestrictedUnpickler`, blocking arbitrary code execution via pickle gadgets. ([#23252](https://github.com/keras-team/keras/pull/23252)) by @SABITHSAHEB

---

## Contributors

Thank you to all the contributors who made this release possible! ๐ŸŽ‰

* @LinZiyuu โ€” Security hardening for IMDB, Reuters, and `.keras` asset extraction ([#23047](https://github.com/keras-team/keras/pull/23047), [#23101](https://github.com/keras-team/keras/pull/23101))
* @hertschuh โ€” H5 group verification ([#23168](https://github.com/keras-team/keras/pull/23168))
* @SABITHSAHEB โ€” CIFAR dataset pickle restriction ([#23252](https://github.com/keras-team/keras/pull/23252))

๐Ÿ“Œ milvus-io/milvus

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## v3.0.0

Release date: July 29, 2026

| Milvus Version | Python SDK Version | Node.js SDK Version | Java SDK Version | Go SDK Version |
| -------------- | ------------------ | ------------------- | ---------------- | -------------- |
| 3.0.0          | 3.0.1              | 3.0.3               | 3.0.5            | 3.0.0          |

Milvus 3.0.0 is officially released! Building on the lake-native architecture introduced in [3.0-beta](https://milvus.io/docs/release_notes.md#v30-beta), this release completes what the beta started: External Collection covers more lakehouse workflows; schema supports online add / backfill / drop; the sparse index is rebuilt around SINDI; StructArray and faceted search round out the retrieval engine; FAISS passthrough, and TEXT extend index and modality choices; and Woodpecker runs as a standalone service.

If you are new to the 3.0 line, the Core 3.0 features recall section below summarizes the capabilities introduced in 3.0-beta; the [3.0-beta release notes](https://milvus.io/docs/release_notes.md#v30-beta) have the full write-ups.

### What's new in 3.0.0 (since 3.0-beta)

#### External Collection: more complete lakehouse workflows

3.0-beta introduced External Collection: reference lake files in place, build indexes, and search them without copying data into Milvus. This release extends it toward complete lakehouse retrieval workflows. External fields can now feed function output fields such as BM25 sparse vectors, MinHash signatures, and text embeddings, so text and model-derived retrieval fields are built inside Milvus without copying the source table. Refresh also supports additive schema evolution: when the external table gains new columns, Milvus patches the affected segments instead of rebuilding the collection.

This release also adds a `milvus-table` external format that treats Milvus Snapshot metadata and Storage V3 manifests as an external source, so a collection snapshot can itself be served as an external table โ€” batch and serving systems get a shared, manifest-backed view of the same data.

๐Ÿ“Œ anthropics/anthropic-sdk-python

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## 0.120.2 (2026-07-28)

Full Changelog: [v0.120.1...v0.120.2](https://github.com/anthropics/anthropic-sdk-python/compare/v0.120.1...v0.120.2)

### Bug Fixes

* **mcp:** support mcp sdk v2 alongside v1 ([#300](https://github.com/anthropics/anthropic-sdk-python/issues/300)) ([177f88c](https://github.com/anthropics/anthropic-sdk-python/commit/177f88ccd7f966e47b654cb19ad0e9cfa4c58ac2))

๐Ÿ“Œ huggingface/peft

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# Highlights

<img width="1250" height="560" alt="peft-v0 20 0" src="https://github.com/user-attachments/assets/6194fae9-f3c3-48ca-b43b-40c774bfe6eb" />

This release adds no less than nine new PEFT methods and puts a lot of work into the surrounding infrastructure, for example adding a new image generation benchmark for the method comparison suite and greatly improving the documentation structure.

## New Methods

### HiRA

@hqsiswiliam added ["HiRA: Parameter-Efficient Hadamard High-Rank Adaptation for Large Language Models"](https://openreview.net/forum?id=TwJrTz9cRS) to PEFT (#2668). Instead of adding the low-rank product `BA` to the base weight, HiRA multiplies it elementwise (Hadamard product) with the frozen base weight. Because the base weight itself is full rank, the resulting update is no longer constrained to be low rank, while the trainable parameter count stays the same as LoRA's.

### GLoRA

@not-lain contributed GLoRA: ["One-for-All: Generalized LoRA for Parameter-Efficient Fine-Tuning"](https://arxiv.org/abs/2306.07967) in #3098. It is a flexible PEFT method that extends LoRA with configurable weight, activation, and bias adaptation, delivering richer fine-tuning with no extra inference cost. Use it when you need per-layer flexibility or stronger adaptation than vanilla LoRA. Skip it for non-Linear layers (e.g. Conv/Embedding) or when standard LoRA is already sufficient and simplicity matters.

### BEFT

@whubaichuan added ["BEFT: Bias-Efficient Fine-Tuning of Language Models"](https://arxiv.org/abs/2509.15974v2) in #3195. BEFT builds on the observation that fine-tuning bias terms alone can be competitive in low-data regimes, but goes further: rather than training *all* biases, it targets the value projection by default, as the authors found this to be most efficient. This brings the trainable parameter count down to roughly 0.01% of the total parameters.

๐Ÿ“Œ google-deepmind/mujoco

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# Version 3.11.0 (July 27, 2026)

## Engine

1. [4787c809](https://github.com/google-deepmind/mujoco/commit/4787c809) Added [geom/surfacevel](https://mujoco.readthedocs.io/en/stable/XMLreference.html#body-geom-surfacevel): the velocity of a geom's surface as seen by contacts, given as a velocity field with a constant component and a rotational component about the geom frame origin. This allows conveyor belts, treadmills and turntables to be modeled with static geoms and no degrees of freedom: friction drives touching bodies along the motion of the surface, with the field projected onto each contact's tangent plane. Surface velocities compose correctly with each other and with body motion. Note that the contact rows of `mjData.efc_vel`, and the constraint-state sensors that read them, report the velocity relative to the moving surface rather than to the geom, since that is the quantity the constraint acts on; for geoms without `surfacevel` the two are identical. Contact-point visualization draws an arrow along the surface velocity at contacts with moving surfaces.

   [![Watch video](https://img.youtube.com/vi/PdSdrqhSiZA/mqdefault.jpg)](https://youtu.be/PdSdrqhSiZA)

2. [a264d0bc](https://github.com/google-deepmind/mujoco/commit/a264d0bc) Added [geom/adhesion](https://mujoco.readthedocs.io/en/stable/XMLreference.html#body-geom-adhesion) and [pair/adhesion](https://mujoco.readthedocs.io/en/stable/XMLreference.html#contact-pair-adhesion): an adhesive force associated with a contact, useful for modeling sticky materials. Contacts can pull with up to the given force before breaking, and the friction budget becomes $\mu(f_N + \text{adhesion})$. Combined with [gap](https://mujoco.readthedocs.io/en/stable/XMLreference.html#body-geom-gap), adhesive contacts apply "adhesion at a distance", useful for modeling magnets. Resting penetration is unaffected by adhesion. [mj_contactForce](https://mujoco.readthedocs.io/en/stable/APIreference/APIfunctions.html#mj-contactforce) reports the net interface force, whose normal component can now be negative.

   [![Watch video](https://img.youtube.com/vi/GioWwB36XHI/mqdefault.jpg)](https://youtu.be/GioWwB36XHI)

3. [f0fa3d82](https://github.com/google-deepmind/mujoco/commit/f0fa3d82) Replaced midpoint integration of free bodies with [gyroscopic derivatives](https://mujoco.readthedocs.io/en/stable/computation/index.html#gefreebody) in the `implicitfast` [integrator](https://mujoco.readthedocs.io/en/stable/computation/index.html#geintegrators): the bias-force derivative of every standalone free body is applied via a local unsymmetric solve of its decoupled block, making `implicitfast` identical to `implicit` for such bodies. Unlike midpoint integration, which required vacuum and no constraints, this applies in all environments (contacts, fluid, constraints), and is compatible with discrete-time inverse dynamics. Spinning free bodies no longer gain energy, but tumbling motion is now mildly damped; models requiring long-horizon energy conservation of tumbling bodies in vacuum should use `RK4`. The [invdiscrete](https://mujoco.readthedocs.io/en/stable/XMLreference.html#option-flag-invdiscrete) flag no longer has any effect on forward dynamics.

4. [5618666a](https://github.com/google-deepmind/mujoco/commit/5618666a) Added [body/simple](https://mujoco.readthedocs.io/en/stable/XMLreference.html#body-simple) attribute ("false"/"auto") to disable the *simple body* mass matrix optimization. This is useful for domain randomization, where model parameters may change post-compilation.

5. [14c0b0c9](https://github.com/google-deepmind/mujoco/commit/14c0b0c9) [mj_setConst](https://mujoco.readthedocs.io/en/stable/APIreference/APIfunctions.html#mj-setconst) now recomputes the `mjModel.{body,geom,site}_sameframe` flags, to account for changes in body/geom/site frames after compilation.

6. [2444defc](https://github.com/google-deepmind/mujoco/commit/2444defc) Added support for [multiccd](https://mujoco.readthedocs.io/en/stable/computation/index.html#comulticcd) with arbitrarily large meshes.

๐Ÿ“Œ open-webui/open-webui

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
### Added

- ๐ŸŽจ **Redesigned interface.** Open WebUI has been visually rebuilt from the ground up. All aspects of the User Interface, from the chat view to the admin panel. Now with a narrower conversation column, lighter typography, tidier spacing, consistent menus and dropdowns, clearly outlined text boxes, and settings rearranged. [Commit](https://github.com/open-webui/open-webui/commit/aedb6bef4e2eb12234c02085a545ff395d96db18), [Commit](https://github.com/open-webui/open-webui/commit/b3255a36569f295766271b8a2b0bd969b4083b9f), [Commit](https://github.com/open-webui/open-webui/commit/ba067258dea2229a9956077b3b0d7b1c68b56f66), [Commit](https://github.com/open-webui/open-webui/commit/8dd862d3383978f21111e63fb2d6029711abed9a), [Commit](https://github.com/open-webui/open-webui/commit/263bbc77d803e83b9af4b04c0cae29705af5f072), [Commit](https://github.com/open-webui/open-webui/commit/f8ea15b84a274712dca33daa970f63ed7368043e), [Commit](https://github.com/open-webui/open-webui/commit/9f17c5960a0e47a09773da4bba12997a31222fc8), [Commit](https://github.com/open-webui/open-webui/commit/6772b1cb4f4e0d3dc166956014e6e7b9bddc721a), [Commit](https://github.com/open-webui/open-webui/commit/d3fd860c131846a9458888f9c256a9a29f3767f2), [Commit](https://github.com/open-webui/open-webui/commit/f1584b5a3764f72de2de6caad507e7c39ad19c23), [Commit](https://github.com/open-webui/open-webui/commit/2e8d92c7b1a9bb8d35f4a27ba3c73368d735c480), [Commit](https://github.com/open-webui/open-webui/commit/e58a4633b15ae53d33fc3b46cb97c76d86be325f), [Commit](https://github.com/open-webui/open-webui/commit/04b146f2cec7e6a01e9d3590eb83655c128fa3c7), [Commit](https://github.com/open-webui/open-webui/commit/e5e2cd78769639b2df83776f1b991966f922f8b4), [Commit](https://github.com/open-webui/open-webui/commit/3316ba76aabe5429596ffd130fd36be4d5c3aa6c), [Commit](https://github.com/open-webui/open-webui/commit/6fcb38fe2e0aded9b85f655cd8f3279e9f4e765e), [Commit](https://github.com/open-webui/open-webui/commit/421da674468f638f72cc5266c5a3874aa3bca3b7), [Commit](https://github.com/open-webui/open-webui/commit/d0bea60581eaa07d41f92ad8f86007e83247e061), [Commit](https://github.com/open-webui/open-webui/commit/21e180182a5096481d4cbb1a8f94212c0a515a40), [Commit](https://github.com/open-webui/open-webui/commit/d027a32ed134ae104f2f142ba45ff38e56215c5f), [Commit](https://github.com/open-webui/open-webui/commit/437c06c4795a72700295d7690d5fd65d1153372c), [Commit](https://github.com/open-webui/open-webui/commit/1bf05ebc7d135d74969438824778c9f73243ba8d), [Commit](https://github.com/open-webui/open-webui/commit/fd07e3a8e3e619f3712067765b416f0925fa80d3), [Commit](https://github.com/open-webui/open-webui/commit/d3ea51fd466a8741afc4dfd4f0d0f2f77fb6467f), [Commit](https://github.com/open-webui/open-webui/commit/4da2ff2655d9abb851805da127cf60b4d9ad1aa7), [Commit](https://github.com/open-webui/open-webui/commit/2fcb36267f034f2b83f936bfacedc20b680a2710), [Commit](https://github.com/open-webui/open-webui/commit/1428a4ddce4998cb3664a5ce37e176442dd426fa), [Commit](https://github.com/open-webui/open-webui/commit/bc8d24c951e9a2c973fc2dd1f2832a2b0855bc0e), [Commit](https://github.com/open-webui/open-webui/commit/704d07e9a20a830aad7bfc5131b0d92621cf0691), [Commit](https://github.com/open-webui/open-webui/commit/e88d2e053c2a63cce3823c9b6006f4a184c4fef2), [Commit](https://github.com/open-webui/open-webui/commit/6940297486d4a5de127efd1a5148b0adcebe87e3), [Commit](https://github.com/open-webui/open-webui/commit/9ca8cf528af1c49da3f0a2bc3c6ca95c1dedbcf5), [Commit](https://github.com/open-webui/open-webui/commit/c4efa81d08c425678c810c51b4d62716e1e57117), [Commit](https://github.com/open-webui/open-webui/commit/bb12b1a18b77d80829cedb2d5bf965808222415b), [Commit](https://github.com/open-webui/open-webui/commit/49abfbdd155dc22882fdcb09989e4f4964db16ee), [#27178](https://github.com/open-webui/open-webui/pull/27178), [Commit](https://github.com/open-webui/open-webui/commit/dcc7fb1e8ef144205531829f8a56e52171c4d63d), [Commit](https://github.com/open-webui/open-webui/commit/5c505c1119fec6170c1bc092ed162f86262a887e)
- ๐Ÿค– **Sub-agents.** Administrators can now enable sub-agents, which let a model hand parts of a task to background helper agents that run their own tool-driven conversations and report results back into the chat, tuned through new "ENABLE_SUBAGENTS", concurrency, iteration, and system-prompt settings. [Commit](https://github.com/open-webui/open-webui/commit/7088d245bb45fc69c0b22748563b9f3c6f0daa73), [Commit](https://github.com/open-webui/open-webui/commit/2f37e853d1259a901f736a823bad29dcc2c3b130), [Commit](https://github.com/open-webui/open-webui/commit/959558fd82eb2a3c980231acd500b73ba4b698b3), [Commit](https://github.com/open-webui/open-webui/commit/3005b7bc71fcbd5abc6e73c3e4caa4ea781cdb76)
- ๐Ÿ“‚ **Folder pages.** Opening a folder now takes you to its own page, where its chats load a page at a time, can be sorted by title or last updated, and you can start a new chat straight from the folder. [Commit](https://github.com/open-webui/open-webui/commit/409fb39717be9ab7becd9e8c01801a08c5bae318)
- โฒ๏ธ **Chat timers.** The assistant can now set a timer that brings a prompt back into the conversation later, after a delay or at a set time, and can drop it automatically if you read the chat or reply before it fires. [Commit](https://github.com/open-webui/open-webui/commit/b23ddeb2800098c6352203ec8fbe9fca40ba415c)
- ๐Ÿ”” **Notification targets.** Notifications now have their own settings tab where you can send them to several webhook destinations, each picking which events it wants, from chats finishing or failing to channel messages and calendar alerts, with a test button and a choice between always notifying or only when you are away, and any webhook you already had is carried over for you. [Commit](https://github.com/open-webui/open-webui/commit/c55e373b994d3a14c99a97f44261422012f63266), [Commit](https://github.com/open-webui/open-webui/commit/cf235738f5a44db415012b3b0ebc1f6e752f5439), [Commit](https://github.com/open-webui/open-webui/commit/200d447f6289faca42f2a666bbabae2c7f3ebadf), [#24750](https://github.com/open-webui/open-webui/issues/24750)
- ๏ฟฝ๏ฟฝ๏ฟฝ๏ฟฝ๏ธ **Full replies in channels.** A reply from the assistant in a channel is now saved and shown in full, with its reasoning, tool calls and other structured parts, where it previously came through blank. [Commit](https://github.com/open-webui/open-webui/commit/498cdab9a548d7d2fd19c389204ee26236fc7efe), [#26720](https://github.com/open-webui/open-webui/pull/26720), [#27409](https://github.com/open-webui/open-webui/pull/27409), [#26707](https://github.com/open-webui/open-webui/issues/26707), [#26656](https://github.com/open-webui/open-webui/issues/26656)
- ๐Ÿ“ฃ **Notifications from the assistant.** The assistant can now send you a notification itself when something is worth your attention, so a long task can reach you after you have moved on to something else. [Commit](https://github.com/open-webui/open-webui/commit/c55e373b994d3a14c99a97f44261422012f63266), [Commit](https://github.com/open-webui/open-webui/commit/200d447f6289faca42f2a666bbabae2c7f3ebadf)
- ๐ŸŒŽ **Share a chat with anyone holding the link.** A shared chat can now be set to Open so it opens without signing in, with visitors no longer bounced to the sign-in page on their way to it, which administrators must first allow through a new "Chats Open Sharing" permission that stays off by default, and such pages ask search engines not to index them. [Commit](https://github.com/open-webui/open-webui/commit/1f0dc90abe879a55f654f2333e29fb0f630831c7), [Commit](https://github.com/open-webui/open-webui/commit/0e0d08382ac0d05b1ad98c47e8a2e37df2a185bb)
- ๐Ÿ”– **Chat variables.** A model's system prompt can now declare fields such as text boxes and dropdown lists that you fill in for a conversation, with the values saved alongside the chat and carried over when it is forked or cloned. [Commit](https://github.com/open-webui/open-webui/commit/bef8ae4b2f05ca49ed88a02ab7a3cdc11b62c4f1), [Commit](https://github.com/open-webui/open-webui/commit/4e869011cd5040b5d6a197fc83d5f50d2425dbc2), [Commit](https://github.com/open-webui/open-webui/commit/1e88367cc837b39c0e9958fefbe053803336dce2), [Commit](https://github.com/open-webui/open-webui/commit/8cbb7f765cfc9c9b3237a6c5593cd93f849033f0), [Commit](https://github.com/open-webui/open-webui/commit/b35e2d265a4e4a2f2a075b31917d48be1dd9ef19), [Commit](https://github.com/open-webui/open-webui/commit/239cb740077a14e452ad002a1e671a09ab558e40), [#26915](https://github.com/open-webui/open-webui/discussions/26915)
- ๐Ÿ—„๏ธ **LDAP group synchronization.** Administrators can now map LDAP groups to Open WebUI groups from the authentication settings, with optional automatic creation of missing groups, so a user's group memberships are kept in step with the directory each time they sign in. [#27263](https://github.com/open-webui/open-webui/pull/27263), [#18015](https://github.com/open-webui/open-webui/issues/18015)
- ๐Ÿ‘ฅ **Restrict sharing with groups.** Admins can now stop resources from being shared with entire groups through a new "USER_PERMISSIONS_ACCESS_GRANTS_ALLOW_GROUPS" permission, which stays enabled by default so existing group sharing keeps working untouched. [Commit](https://github.com/open-webui/open-webui/commit/4ed19d504bd30c0fc801e9228d9816669ec1c09c), [Commit](https://github.com/open-webui/open-webui/commit/f84dabe3d97ff701c097055023f28f3f2f7ebd07), [Commit](https://github.com/open-webui/open-webui/commit/77da3d8c81b9a6fda4354619de94f5d433328d8e), [Commit](https://github.com/open-webui/open-webui/commit/84e4d6ef8277f4b4f3ac4d355b3219e9b5a37268), [#27124](https://github.com/open-webui/open-webui/pull/27124)
- ๐Ÿค **Shared folder collaboration.** People with access to a shared folder can now use its files and system prompt as knowledge in chat and, with write access, rename and manage the folder, all according to their read or write permission. [Commit](https://github.com/open-webui/open-webui/commit/797293c74957bd79e42262d1dc0fd637a45d0357), [Commit](https://github.com/open-webui/open-webui/commit/caa2457c17e592587b804f21054cc000944af75c), [Commit](https://github.com/open-webui/open-webui/commit/009715cd63d1c8b5320aba68e9a70afcde519016), [Commit](https://github.com/open-webui/open-webui/commit/53ccd718a53de25bb6d61476a6617bfb3f130a44)
- ๐Ÿ‘๏ธ **Chat previews in the sidebar.** Hovering a chat in the sidebar now shows a compact preview of its recent messages, so you can find the conversation you want without opening it. [Commit](https://github.com/open-webui/open-webui/commit/d0f7da4f45b8831b09b2ab3ec8f91aa354d90ba3), [Commit](https://github.com/open-webui/open-webui/commit/aaf2834db758bfec69408ab4cabcf324965c221c), [Commit](https://github.com/open-webui/open-webui/commit/1513ddaf58fe18029086880461d7cad0649a699c), [Commit](https://github.com/open-webui/open-webui/commit/93bd05271c07c249978d69abf3297fd2841900f9)
- ๐Ÿ•— **Local message timestamps.** Message timestamps now appear on hover in your device's local date and time format, with the full weekday and date shown in a tooltip. [Commit](https://github.com/open-webui/open-webui/commit/797293c74957bd79e42262d1dc0fd637a45d0357), [Commit](https://github.com/open-webui/open-webui/commit/f84dabe3d97ff701c097055023f28f3f2f7ebd07)
- ๐Ÿ“‡ **User variables.** You can now store your own values in account settings, such as your role or how you like answers written, and a model's system prompt can insert them wherever they are needed. [Commit](https://github.com/open-webui/open-webui/commit/bd5d7b2e879511882429075d804d9222956f4a1a), [Commit](https://github.com/open-webui/open-webui/commit/212eec408ca320edfa2271e604d10e14b6a9bc1a), [Commit](https://github.com/open-webui/open-webui/commit/793a43d9c48225925929eb312fe2b70d5914d1da)
- ๐Ÿงบ **Automations that file their chats away.** An automation can now be pointed at one of your folders, from the dialog, the editor or by asking the assistant, so each run lands there instead of loose in your chat list, and the folder is cleared automatically if it is later deleted. [Commit](https://github.com/open-webui/open-webui/commit/f798d05586a140f1a6b51f1e51b2b2a63d079d45), [Commit](https://github.com/open-webui/open-webui/commit/bab71ed08b5af6f4a8ff2daa02792baae9edab03), [Commit](https://github.com/open-webui/open-webui/commit/db5c092299471444c356216d5ef39b382ba1aa1e)
- ๐Ÿ”ต **See what you have not read yet.** Folders in the sidebar now carry a count of chats with something new in them, a folder's own page marks unread chats with a dot, shows a spinner on any still generating, clears the dot as you open one, and keeps itself up to date as replies finish elsewhere, unread chats sort to the top of a folder, and you can mark a single chat unread again mark everything in a folder read, or mark every chat read at once from the sidebar. [Commit](https://github.com/open-webui/open-webui/commit/f798d05586a140f1a6b51f1e51b2b2a63d079d45), [Commit](https://github.com/open-webui/open-webui/commit/f867825bf3b7699bc2bd967ef46b2bb63e48b098), [Commit](https://github.com/open-webui/open-webui/commit/1de36d600f7191c28a98bf4b347b44cf8f1bef43), [Commit](https://github.com/open-webui/open-webui/commit/85c47fb467177ed811ba77dfe62461dfbe8e2548), [Commit](https://github.com/open-webui/open-webui/commit/b7489bbc6c4e376c017edffd8da3c2eb4e6c1c8e), [Commit](https://github.com/open-webui/open-webui/commit/3cd72ee6a8e93dc39a4d4c173117056e25326c8a), [Commit](https://github.com/open-webui/open-webui/commit/6f93ecd4fd77b0d51a5fbbd2fc3fd6d151036a55), [Commit](https://github.com/open-webui/open-webui/commit/e5a08d52208e8b1ed07ff94906e27d174146b1ca), [Commit](https://github.com/open-webui/open-webui/commit/8ddf119570b3c0b04b673d41cb2363370c23939b)

๐Ÿ“Œ ollama/ollama

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## What's Changed

* Fixed an MLX Metal bug that could reduce output quality for NVFP4 models, particularly Laguna.

**Full Changelog**: https://github.com/ollama/ollama/compare/v0.32.4...v0.32.5

๐Ÿ“Œ janhq/jan

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## Migration

**Settings and credentials now live in a backend-managed store.**

Provider settings and credentials previously persisted in the webview's `localStorage`. In 0.8.4, Jan reads and writes them from a backend-managed store instead (secrets go to the OS keyring). Your existing settings and API keys are copied over automatically on first launch - no manual action is needed, and from this version on all new changes are saved to the backend store.

Your old `localStorage` data is **not deleted** - it's kept as a snapshot so you can downgrade to a pre-0.8.4 build and still have your previous settings. Any changes you make in 0.8.4+ live only in the backend store, so a downgrade sees the older snapshot, not your latest settings.

---

## What's Changed
* chore(flatpak): add 0.8.3 release notes to appdata by @qnixsynapse in https://github.com/janhq/jan/pull/8343
* docs: update changelogs for v0.8.3 and update documentation by @qnixsynapse in https://github.com/janhq/jan/pull/8336
* Merge release/v0.8.3 to main by @qnixsynapse in https://github.com/janhq/jan/pull/8344
* i18n(ja): update Japanese translations (#8264) by @mahirhir in https://github.com/janhq/jan/pull/8348
* i18n(ja): complete common.json translations and fix terminology (#8264) by @mahirhir in https://github.com/janhq/jan/pull/8349
* i18n(ja): complete settings.json translations (#8264) by @mahirhir in https://github.com/janhq/jan/pull/8352
* fix(markdown): force code blocks LTR in RTL language context by @qnixsynapse in https://github.com/janhq/jan/pull/8350
* feat(chat): add toggle for folding interim text into reasoning trace by @qnixsynapse in https://github.com/janhq/jan/pull/8363
* fix(rag): repair embeddings, stop chat-model reloads, and parse more PDFs by @qnixsynapse in https://github.com/janhq/jan/pull/8360

๐Ÿ“Œ ray-project/ray

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# Highlights

- **Ray Data**: We added fixes for several `to_pandas` regressions introduced in 2.56: an opt-out flag (`RAY_DATA_ENABLE_ARROW_BACKED_PANDAS_CONVERSION`) for Arrow-backed conversion, an int64/`double[pyarrow]` overflow crash on concatenation, and a `TensorDtype.__from_arrow__` crash on empty tensor columns (#64793, #64794).
- **Ray Core**: We added early detection for system-slice memory pressure: the memory monitor now snapshots the user and system cgroup slices together and logs an error when the system slice exceeds reserved system memory, warning users to raise `--system-reserved-memory` before it causes node deaths (#64492).
- **Ray Serve**: We added protobuf 7 compatibility and a routing fix for LLM direct streaming, so body-aware routers like `PrefixCacheAffinityRouter` no longer hang when `RAY_SERVE_LLM_ENABLE_DIRECT_STREAMING=1` (#64592, #64488).

# Ray Data

## ๐Ÿ”จ Fixes
- Fixed two Arrow-backed `to_pandas` regressions: added `DataContext.enable_arrow_backed_pandas_conversion` as an opt-out, and reconciled divergent numeric column types before concatenation to avoid int64/`double[pyarrow]` overflow crashes (#64793, #64768).
- Fixed a `TensorDtype.__from_arrow__` crash on zero-size tensor elements by using an explicit row count instead of numpy's `-1` dimension inference (#64794, #64767).
- Fixed a crash in hash partition caused by read-only hash arrays (#64584, #64552, #64559).
- Nullified `_input_dependencies` in `_get_args` so exporting operator args no longer triggers an exponential `sanitize_for_struct` call chain over fused operators (#64412, #64316).

# Ray Serve

## ๐Ÿ”จ Fixes
- Added protobuf `>=7` compatibility to `_proto_to_dict` by binding to `FieldDescriptor.is_repeated` when the deprecated `label` attribute is absent (#64592, #64362).

# Ray LLM

๐Ÿ“Œ jax-ml/jax

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
* New features
  * Added a doc on defining custom derivative rules with the experimental
    hijax API (`hijax-custom-derivatives`), along with
    `jax.experimental.hijax` helpers for deriving `VJPHiPrimitive` autodiff
    rules from a `jvp` or `lin` rule: `linearize_from_jvp` with
    `apply_derived_linearization`, `vjp_fwd_from_jvp` with `transpose_jvp`,
    `vjp_fwd_from_lin` with `transpose_linearized`, and `jvp_from_lin`.
  * Added `jax.custom_remat` to the top-level `jax` namespace, for
    per-function control of rematerialization under the new `jax_remat3`
    implementation.
  * `jax.checkpoint_policies` is now a submodule rather than a namespace
    object (so `from jax.checkpoint_policies import ...` now works; attribute
    access is unchanged), and it additionally exposes the name-based policy
    classes `SaveOnlyTheseNames`, `SaveAnyNamesButThese`, and
    `SaveAndOffloadOnlyTheseNames`.
  * Added `jax.Inline` enum for specify inlining policies to
    `jax.jit`.

* Breaking changes
  * The deprecated module j`ax.cloud_tpu_init` was removed. This did nothing and

๐Ÿ“Œ huggingface/transformers

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# Patch release v5.14.1

This patch solves a few issues which appeared when integrating Inkling model, most notably an issue affecting models using EncoderDecoderCache during assisted generation. It also fixes an issue that could appear during prefill with StaticCache and sdpa without padding for Inkling which uses a position_bias. 
It contains the following commits:

- Fix sdpa prefill with position_bias (#47359) by @Cyrilvallez
- Fix assisted decoding for models with EncoderDecoder cache & OlmoHybrid (#47361) by @Cyrilvallez
- [FP8] Bump kernels version (#47344) by @vasqu 
- Fix deepgemm on multiple devices (#47323) by @IlyasMoutawwakil 

๐Ÿ“Œ pytorch/pytorch

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# PyTorch 2.13.0 Release Notes

- [Highlights](#highlights)
- [Backwards Incompatible Changes](#backwards-incompatible-changes)
- [Deprecations](#deprecations)
- [New Features](#new-features)
- [Improvements](#improvements)
- [Bug fixes](#bug-fixes)
- [Performance](#performance)
- [Documentation](#documentation)
- [Developers](#developers)

# Highlights

<table>
  <tr><td><strong>FlexAttention</strong> lands on Apple Silicon (MPS), with up to ~12x speedup over SDPA on sparse patterns, and gains a deterministic backward path on CUDA for reproducible gradient computation.</td></tr>
  <tr><td><strong>CuTeDSL "Native DSL" backend</strong> gives Inductor a second high-performance code path (alongside Triton) for key GPU operations, with faster compilation. [Prototype]</td></tr>
  <tr><td><strong><code>nn.LinearCrossEntropyLoss</code></strong> combines the final prediction and loss computation to cut peak GPU memory by up to 4x for large-vocabulary language model training.</td></tr>
  <tr><td><strong>torchcomms</strong>, a new communications backend for PyTorch Distributed, improves fault tolerance, scalability, and debuggability for large-cluster training.</td></tr>
  <tr><td><strong>FSDP2</strong> now overlaps reduce-scatter and all-gather communications via a dedicated process group (opt-in), increasing distributed training throughput.</td></tr>

๐Ÿ“Œ huggingface/diffusers

  • ็‰ˆๆœฌ: v0.39.0
  • ๅ็งฐ: Diffusers 0.39.0: New image and video pipelines, core library improvements, and more
  • ๅ‘ๅธƒๆ—ฅๆœŸ: 2026/07/03
  • ้“พๆŽฅ: ๆŸฅ็œ‹ๅฎŒๆ•ด Release โ†—
๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## New Pipelines

### Cosmos 3

[**Cosmos 3**](https://huggingface.co/docs/diffusers/main/api/pipelines/cosmos3) is NVIDIA's unified world foundation model (WFM) for Physical AI โ€” a single omni-model built on a Mixture-of-Transformers (MoT) architecture that combines world generation, physical reasoning, and action generation, replacing the separate Predict, Reason, and Transfer models from earlier Cosmos releases. A single `Cosmos3OmniTransformer` runs a Qwen-style language model in parallel with a diffusion generation pathway, joined by a 3D multimodal RoPE. This release also lands video-to-video and action-conditioned generation, and a sound encoder.

- PR: [https://github.com/huggingface/diffusers/pull/13818](https://github.com/huggingface/diffusers/pull/13818)
- Docs: [https://huggingface.co/docs/diffusers/main/api/pipelines/cosmos3](https://huggingface.co/docs/diffusers/main/api/pipelines/cosmos3)

Thanks to @atharvajoshi10, @yzhautouskay, and @MaciejBalaNV for the contributions.

### Ideogram 4

[**Ideogram 4**](https://huggingface.co/docs/diffusers/main/api/pipelines/ideogram4) is a flow-matching text-to-image model that uses a multimodal text encoder and an asymmetric classifier-free guidance scheme: a dedicated `unconditional_transformer` produces the negative branch with zeroed text features, while the main `transformer` consumes the full packed text + image sequence. The pipeline ships with structured prompt upsampling and LoRA loading support.

- PR: [https://github.com/huggingface/diffusers/pull/13859](https://github.com/huggingface/diffusers/pull/13859)
- Docs: [https://huggingface.co/docs/diffusers/main/api/pipelines/ideogram4](https://huggingface.co/docs/diffusers/main/api/pipelines/ideogram4)

Thanks to @JinLiIdeogram for the contribution.

๐Ÿ“Œ DLR-RM/stable-baselines3

  • ็‰ˆๆœฌ: v2.9.0
  • ๅ็งฐ: v2.9.0: Updated dependencies (pandas is now optional, gymnasium 1.3.0 support, torch>=2.8)
  • ๅ‘ๅธƒๆ—ฅๆœŸ: 2026/06/16
  • ้“พๆŽฅ: ๆŸฅ็œ‹ๅฎŒๆ•ด Release โ†—
๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
### Breaking Changes:
- Relaxed Gymnasium version range (from `"gymnasium>=0.29.1,<1.3.0"` to `"gymnasium>=0.29.1,<2.0"`)
- `pandas` and `matplotlib` are no longer core dependencies; they are now optional and only required for loading results and plotting (moved to `stable-baselines3[extra]`).
- Moved `read_json` and `read_csv` helper functions to test files
- Raised `torch` minimum version from 2.3 to 2.8 to mitigate https://github.com/advisories/GHSA-887c-mr87-cxwp

### Bug Fixes:
- Fixed deprecated error Taxi-v3 from gymnasium v1.3.0 in tests

### [SB3-Contrib]
- Optimized tests (faster to run)
- Fixed dead link for `RecurrentPPO`.

### [RL Zoo]


### [SBX] (SB3 + Jax)

- Added support for `rollout_buffer_class` and `rollout_buffer_kwargs` arguments in `PPO` and `OnPolicyAlgorithmJax` constructors, as in Stable Baselines3. (@Trenza1ore)
- Updated Jax dependency

๐Ÿ“Œ huggingface/accelerate

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## FSDP2 Improvements

This release brings a large batch of FSDP2 fixes and quality-of-life improvements: correct dtype handling on load, sharding of embeddings/norms, QLoRA crash prevention, and a more robust auto-wrap policy.

  - Fsdp2 fully_shard embedding and norm by @SunMarc in #4015
  - Fix fsdp2 load full state dict dtype mismatch by @SunMarc in #4021
  - Fix region compilation fsdpv2 by @SunMarc in #4022
  - [FSDP2] Cast model to uniform dtype before fully_shard to fix mixed-dtype AssertionError by @roycho96 in #3985
  - [FSDP2] Auto-exclude non-floating frozen Params4bit from fully_shard to prevent QLoRA crash by @roycho96 in #3987
  - fix(FSDP2): auto-wrap policy ignoring _no_split_modules fallback by @JohnGiorgi in #3999
  - fix: use key-based matching in fsdp2_load_full_state_dict by @roycho96 in #3982
  - fix: add missing model_has_params4bit guard to fsdp2_load_full_state_dict call by @roycho96 in #3981
  - Fix to-fsdp2: drop REMOVED / NOT_YET_IMPLEMENTED FSDP1 keys instead of leaking them by @lollinng in #4065
  - Prevent double-wrapping models in prepare_model() by @joshuaswanson in #3977

## AMD ROCm support

Accelerate now works end-to-end on AMD ROCm devices. Thanks @Abdennacer-Badaoui!

- Make accelerate work end-to-end on AMD ROCm by @Abdennacer-Badaoui in #4025

๐Ÿ“Œ google-deepmind/gemma

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
- Fix `dialog` dependency requirement to be `>= 1.1.0`.

๐Ÿ“Œ openai/tiktoken

๐Ÿ“Œ chroma-core/chroma

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
Version: `1.5.9`
Git ref: `refs/tags/1.5.9`
Build Date: `2026-05-05T05:55`
PIP Package: `chroma-1.5.9.tar.gz`
Github Container Registry Image: `:1.5.9`
DockerHub Image: `:1.5.9`

## What's Changed
* [ENH](frontend): block functions on topology dbs by @rescrv in https://github.com/chroma-core/chroma/pull/6836
* [ENH](faults): Add Tilt fault injection CLI by @rescrv in https://github.com/chroma-core/chroma/pull/6881
* [CHORE]  Debug TimeoutError in test_add.py by @rescrv in https://github.com/chroma-core/chroma/pull/6905
* [ENH]: Enable rebuilds for sharded collections by @tanujnay112 in https://github.com/chroma-core/chroma/pull/6916
* [ENH]: Group by support with sharding by @sanketkedia in https://github.com/chroma-core/chroma/pull/6909
* [CHORE]: Denormalize tenant and database into collection_compaction_cursors table by @tanujnay112 in https://github.com/chroma-core/chroma/pull/6940
* [CHORE]  Use normalized record sets for test add by @rescrv in https://github.com/chroma-core/chroma/pull/6935
* [ENH]: Add workflow to build and publish service container images by @jasonvigil in https://github.com/chroma-core/chroma/pull/6944
* [ENH] - Updates language around Chroma Cloud to be more representative. by @tjkrusinskichroma in https://github.com/chroma-core/chroma/pull/6952
* [ENH]: Add change stream to collection compaction cursors by @tanujnay112 in https://github.com/chroma-core/chroma/pull/6955
* [BUG] Switch to storing DOCKERHUB_USERNAME as var by @jasonvigil in https://github.com/chroma-core/chroma/pull/6962
* [CHORE]: Standardize Tilt CI image build on root docker-bake.hcl by @jasonvigil in https://github.com/chroma-core/chroma/pull/6958

๐Ÿ“Œ explosion/spaCy

  • ็‰ˆๆœฌ: release-v3.8.14
  • ๅ็งฐ: v3.8.14: Bug fix for model downloading in environments without pip on PATH
  • ๅ‘ๅธƒๆ—ฅๆœŸ: 2026/03/29
  • ้“พๆŽฅ: ๆŸฅ็œ‹ๅฎŒๆ•ด Release โ†—
๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
- Fix `spacy download` failing in environments where `pip` is not on PATH but is available as a Python module (e.g., some virtual environments and containers)

๐Ÿ“Œ tensorflow/tensorflow

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
# Release 2.21.0

## TensorFlow

### Breaking Changes

* Support for Python 3.9 has been removed starting with TF 2.21.
* The TensorBoard (TB) dependency has been removed starting with TF 2.21.

### Major Features and Improvements

* `tf.lite`
    * Adds int8 and int16x8 support for SQRT operator.
    * Adds int16x8 support for EQUAL and NOT_EQUAL operators.
    * Addsย support for int2 type.
    * Adds support for int2/int4 in tfl.cast .
    * Adds support for SRQ int2 in tfl.fully_connected.
    * Adds support for int4 in tfl.slice.
    * Adds support for uint4 type.

๐Ÿ“Œ huggingface/text-generation-inference

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## What's Changed
* misc(gha): expose action cache url and runtime as secrets by @mfuntowicz in https://github.com/huggingface/text-generation-inference/pull/2964
* feat: support max_image_fetch_size to limit by @drbh in https://github.com/huggingface/text-generation-inference/pull/3339
* Maintenance mode by @LysandreJik in https://github.com/huggingface/text-generation-inference/pull/3344
* Maintenance mode by @LysandreJik in https://github.com/huggingface/text-generation-inference/pull/3345
* fix(num_devices): fix num_shard/num device auto compute when NVIDIA_VISIBLE_DEVICES == "all" or "void" by @oOraph in https://github.com/huggingface/text-generation-inference/pull/3346


**Full Changelog**: https://github.com/huggingface/text-generation-inference/compare/v3.3.6...v3.3.7

๐Ÿ“Œ microsoft/autogen

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## What's Changed
* Fix docs dotnet core typo by @lach-g in https://github.com/microsoft/autogen/pull/6950
* Fix loading streaming Bedrock response with tool usage with empty argument by @pawel-dabro in https://github.com/microsoft/autogen/pull/6979
* Support linear memory in RedisMemory by @justin-cechmanek in https://github.com/microsoft/autogen/pull/6972
* Fix message ID for correlation between streaming chunks and final mesโ€ฆ by @smalltalkman in https://github.com/microsoft/autogen/pull/6969
* fix: extra args not work to disable thinking by @liuyunrui123 in https://github.com/microsoft/autogen/pull/7006
* Add thinking mode support for anthropic client by @SrikarMannepalli in https://github.com/microsoft/autogen/pull/7002
* Fix spurious </think> tags caused by empty string reasoning_content in streaming by @Copilot in https://github.com/microsoft/autogen/pull/7025
* Fix GraphFlow cycle detection to properly clean up recursion state by @Copilot in https://github.com/microsoft/autogen/pull/7026
* Add comprehensive GitHub Copilot instructions for AutoGen development by @Copilot in https://github.com/microsoft/autogen/pull/7029
* Fix Redis caching always returning False due to unhandled string values by @Copilot in https://github.com/microsoft/autogen/pull/7022
* Fix OllamaChatCompletionClient load_component() error by adding to WELL_KNOWN_PROVIDERS by @Copilot in https://github.com/microsoft/autogen/pull/7030
* Fix finish_reason logic in Azure AI client streaming response by @litterzhang in https://github.com/microsoft/autogen/pull/6963
* Add security warnings and default to DockerCommandLineCodeExecutor by @ekzhu in https://github.com/microsoft/autogen/pull/7035
* Fix: Handle nested objects in array items for JSON schema conversion by @kkutrowski in https://github.com/microsoft/autogen/pull/6993
* Fix not supported field warnings in count_tokens_openai by @seunggil1 in https://github.com/microsoft/autogen/pull/6987
* Fix(mcp): drain pending command futures on McpSessionActor failure by @withsmilo in https://github.com/microsoft/autogen/pull/7045
* Add missing reasoning_effort parameter support for OpenAI GPT-5 models by @Copilot in https://github.com/microsoft/autogen/pull/7054
* Update version to 0.7.5 by @ekzhu in https://github.com/microsoft/autogen/pull/7058

๐Ÿ“Œ Aider-AI/aider

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
- Added support for all GPT-5 models.
- Added support for Grok-4 via `xai/grok-4` and `openrouter/x-ai/grok-4` model names.
- Added support for `gemini/gemini-2.5-flash-lite-preview-06-17` model, by Tamir Zahavi-Brunner.
- `/clear` now prints โ€œAll chat history cleared.โ€ so you know it worked, by Zexin Yuan.
- `/undo` output now shows only the first line of each commit message, making it easier to read.
- Added support for `openrouter/moonshotai/kimi-k2` model, by Jack Harrington.
- Display model announcements with no-arg `/model` command.
- Fixed an issue where new settings for an existing model didn't replace the old ones, by Andrew Grigorev.
- Fixed analytics to support the latest PostHog SDK event-capture API.
- Bumped dependencies to pick up latest litellm==1.75.0.

- Aider wrote 88% of the code in this release.

๐Ÿ“Œ openai/whisper

๐Ÿ“Œ AUTOMATIC1111/stable-diffusion-webui

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## 1.10.1

### Bug Fixes:
* fix image upscale on cpu ([#16275](https://github.com/AUTOMATIC1111/stable-diffusion-webui/pull/16275))

๐Ÿ“Œ lllyasviel/Fooocus

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## What's Changed
* fix: resolve colab unsupported image type issue by @mashb1t in https://github.com/lllyasviel/Fooocus/pull/3506

**Full Changelog**: https://github.com/lllyasviel/Fooocus/compare/v2.5.4...v2.5.5

๐Ÿ“Œ lm-sys/FastChat

๐Ÿ“ ็‚นๅ‡ปๅฑ•ๅผ€ๆ›ดๆ–ฐๆ—ฅๅฟ—
## Highlights
- Added SGLang worker for vision language models, lower latency and higher throughput https://github.com/lm-sys/FastChat/pull/2928
- Vision langauge WebUI https://github.com/lm-sys/FastChat/pull/2960
- OpenAI-compatible API server now supports image input https://github.com/lm-sys/FastChat/pull/2928
- Added LightLLM worker for higher throughput https://github.com/lm-sys/FastChat/blob/main/docs/lightllm_integration.md
- Added Apple MLX worker https://github.com/lm-sys/FastChat/pull/2940

## What's Changed
* fix specify local path issue use model from www.modelscope.cn by @liuyhwangyh in https://github.com/lm-sys/FastChat/pull/2934
* support openai embedding for topic clustering by @CodingWithTim in https://github.com/lm-sys/FastChat/pull/2729
* Remove duplicate API endpoint by @surak in https://github.com/lm-sys/FastChat/pull/2949
* Update Hermes Mixtral by @teknium1 in https://github.com/lm-sys/FastChat/pull/2938
* Enablement of REST API Usage within Google Colab Free Tier by @ggcr in https://github.com/lm-sys/FastChat/pull/2940
* Create a new worker implementation for Apple MLX by @aliasaria in https://github.com/lm-sys/FastChat/pull/2937
* feat: support Model Yuan2.0, a new generation Fundamental Large Language Model developed by IEIT System by @cauwulixuan in https://github.com/lm-sys/FastChat/pull/2936
* Fix the pooling method of BGE embedding model by @staoxiao in https://github.com/lm-sys/FastChat/pull/2926
* SGLang Worker by @BabyChouSr in https://github.com/lm-sys/FastChat/pull/2928
* Update mlx_worker to be async by @aliasaria in https://github.com/lm-sys/FastChat/pull/2958
* Integrate LightLLM into serve worker by @zeyugao in https://github.com/lm-sys/FastChat/pull/2888
* Copy button by @surak in https://github.com/lm-sys/FastChat/pull/2963

๐Ÿ”ฅ ไบ”ใ€็ƒญ้—จ AI ็›ธๅ…ณ Issues

GitHub ไธŠ่ฟ‘ๆœŸๆœ€ๅ—ๅ…ณๆณจ็š„ AI ็›ธๅ…ณ Issues

โš ๏ธ ๆš‚ๆ— ็ƒญ้—จ Issues

๐Ÿ“Š ๅ…ญใ€AI ้ข†ๅŸŸๅˆ†็ฑปๆฆ‚่งˆ

๐Ÿง  LLM & ๅคง่ฏญ่จ€ๆจกๅž‹

  • langchain-ai/langchain โ€” ๆœ€ๆ–ฐ็‰ˆ: langchain-core==1.5.3 (2026/07/30)
  • open-webui/open-webui โ€” ๆœ€ๆ–ฐ็‰ˆ: v0.11.0 (2026/07/27)
  • ollama/ollama โ€” ๆœ€ๆ–ฐ็‰ˆ: v0.32.5 (2026/07/27)
  • janhq/jan โ€” ๆœ€ๆ–ฐ็‰ˆ: v0.8.4 (2026/07/23)
  • huggingface/transformers โ€” ๆœ€ๆ–ฐ็‰ˆ: v5.14.1 (2026/07/16)
  • openai/whisper โ€” ๆœ€ๆ–ฐ็‰ˆ: v20250625 (2025/06/26)
  • lm-sys/FastChat โ€” ๆœ€ๆ–ฐ็‰ˆ: v0.2.36 (2024/02/11)

๐ŸŽจ ๅ›พๅƒ็”Ÿๆˆ

  • huggingface/diffusers โ€” ๆœ€ๆ–ฐ็‰ˆ: v0.39.0 (2026/07/03)
  • AUTOMATIC1111/stable-diffusion-webui โ€” ๆœ€ๆ–ฐ็‰ˆ: v1.10.1 (2025/02/09)
  • lllyasviel/Fooocus โ€” ๆœ€ๆ–ฐ็‰ˆ: v2.5.5 (2024/08/12)

๐Ÿ”ง ๆทฑๅบฆๅญฆไน ๆก†ๆžถ

  • keras-team/keras โ€” ๆœ€ๆ–ฐ็‰ˆ: v3.12.4 (2026/07/30)
  • jax-ml/jax โ€” ๆœ€ๆ–ฐ็‰ˆ: jax-v0.11.0 (2026/07/17)
  • pytorch/pytorch โ€” ๆœ€ๆ–ฐ็‰ˆ: v2.13.0 (2026/07/09)
  • tensorflow/tensorflow โ€” ๆœ€ๆ–ฐ็‰ˆ: v2.21.0 (2026/03/07)

๐Ÿค– AI Agent

  • Significant-Gravitas/AutoGPT โ€” ๆœ€ๆ–ฐ็‰ˆ: autogpt-platform-beta-v0.7.0 (2026/08/05)
  • microsoft/autogen โ€” ๆœ€ๆ–ฐ็‰ˆ: python-v0.7.5 (2025/09/30)
  • Aider-AI/aider โ€” ๆœ€ๆ–ฐ็‰ˆ: v0.86.0 (2025/08/10)

๐Ÿ“ ๅ‘้‡ๆ•ฐๆฎๅบ“

  • qdrant/qdrant โ€” ๆœ€ๆ–ฐ็‰ˆ: v1.19.0 (2026/08/05)
  • milvus-io/milvus โ€” ๆœ€ๆ–ฐ็‰ˆ: v3.0.0 (2026/07/29)
  • chroma-core/chroma โ€” ๆœ€ๆ–ฐ็‰ˆ: 1.5.9 (2026/05/05)

๐Ÿ› ๏ธ MLOps & ๅทฅๅ…ท

  • streamlit/streamlit โ€” ๆœ€ๆ–ฐ็‰ˆ: 1.61.0 (2026/08/05)
  • mlflow/mlflow โ€” ๆœ€ๆ–ฐ็‰ˆ: v3.15.1 (2026/08/03)
  • gradio-app/gradio โ€” ๆœ€ๆ–ฐ็‰ˆ: gradio@6.22.0 (2026/07/31)
  • ray-project/ray โ€” ๆœ€ๆ–ฐ็‰ˆ: ray-2.56.1 (2026/07/18)