1unsloth/Laguna-S-2.1-GGUF57,536 DOWNLOADS · 174 LIKESUnsloth published GGUF builds of poolside's Laguna S 2.1 using its Dynamic 2.0 imatrix quantization. The repo is aimed at running the model locally through llama.cpp or Unsloth Studio, with UD-Q4_K_XL among the offered quants.Unsloth 发布了 poolside 旗下 Laguna S 2.1 的 GGUF 版本,采用其 Dynamic 2.0 imatrix 量化方案。该仓库面向本地部署,可通过 llama.cpp 或 Unsloth Studio 运行,提供的量化档位包括 UD-Q4_K_XL。Unsloth が poolside の Laguna S 2.1 を GGUF 化し、独自の Dynamic 2.0 imatrix 量子化を適用して公開した。llama.cpp や Unsloth Studio でのローカル実行を想定しており、量子化バリアントには UD-Q4_K_XL などが含まれる。
2baseten/GLM-5.2-Vision-NVFP4494 DOWNLOADS · 92 LIKESBaseten gave GLM-5.2 sight by attaching the MoonViT vision encoder from Kimi-K2.6 through a trained PatchMerger projector. Both the text backbone and the vision tower stay frozen and byte-identical to their upstream releases, so only the projector is new. This checkpoint is quantized to NVFP4 for SGLang serving.Baseten 通过训练一个 PatchMerger 投影层,把 Kimi-K2.6 的 MoonViT 视觉编码器接到原本没有视觉输入的 GLM-5.2 上。文本主干与视觉塔均保持冻结,且与各自上游版本逐字节一致,新增的只有投影层。该检查点已量化为 NVFP4,面向 SGLang 部署。Baseten は、Kimi-K2.6 の視覚エンコーダ MoonViT を学習済みの PatchMerger プロジェクタ経由で接続し、GLM-5.2 に視覚入力を与えた。テキスト側のバックボーンと視覚タワーはいずれも凍結され、上流のリリースとバイト単位で同一のまま、新しいのはプロジェクタだけである。本チェックポイントは SGLang での配信向けに NVFP4 量子化されている。
3google/gemma-4-31B-it12,511,030 DOWNLOADS · 3,363 LIKESGoogle DeepMind's instruction-tuned Gemma 4 at 31B parameters takes text and image input and ships under Apache 2.0. The Gemma 4 line carries a context window of up to 256K tokens and keeps the family's multilingual coverage. At over 12 million downloads it is the most pulled model on today's list.谷歌 DeepMind 的 Gemma 4 指令微调版本,参数量 31B,支持文本与图像输入,以 Apache 2.0 许可发布。Gemma 4 系列的上下文窗口最长可达 256K token,并延续了该家族的多语言能力。其下载量超过 1200 万次,是今日榜单中被拉取最多的模型。Google DeepMind による Gemma 4 の指示チューニング版で、パラメータ数は 31B、テキストと画像の入力に対応し、Apache 2.0 で提供される。Gemma 4 系列は最大 256K トークンのコンテキストウィンドウを備え、ファミリー共通の多言語対応も維持している。ダウンロード数は 1200 万超で、本日の一覧では最多である。
4Qwen/Qwen3.6-35B-A3B6,413,105 DOWNLOADS · 2,506 LIKESQwen shipped the first open-weight variant of Qwen3.6, a 35B mixture-of-experts model with 3B active parameters that accepts image and text input. It follows the Qwen3.5 series released in February and is distributed in Hugging Face Transformers format, with vLLM, SGLang and KTransformers listed as compatible runtimes.Qwen 发布了 Qwen3.6 系列的首个开放权重版本:一个总参数 35B、激活参数 3B 的混合专家模型,支持图像与文本输入。它承接今年 2 月发布的 Qwen3.5 系列,以 Hugging Face Transformers 格式分发,并兼容 vLLM、SGLang 和 KTransformers 等运行时。Qwen は Qwen3.6 系列で初となるオープンウェイト版を公開した。総パラメータ 35B、活性化 3B の Mixture-of-Experts モデルで、画像とテキストの入力に対応する。2 月公開の Qwen3.5 系列を受け継ぐもので、Hugging Face Transformers 形式で配布され、vLLM や SGLang、KTransformers での実行に対応するとされる。
5nvidia/nemotron-3.5-asr-streaming-0.6b840,691 DOWNLOADS · 938 LIKESNvidia released a 0.6B streaming speech recognition model in the Nemotron 3.5 line, built around cache-aware decoding for low-latency transcription. It is the multilingual extension of the earlier English-only nemotron-speech-streaming-en-0.6b.英伟达在 Nemotron 3.5 系列下发布了一个 0.6B 的流式语音识别模型,采用 cache-aware 解码以降低转写延迟。它是此前仅支持英语的 nemotron-speech-streaming-en-0.6b 的多语言扩展版本。エヌビディアは Nemotron 3.5 系列として、0.6B 規模のストリーミング音声認識モデルを公開した。低遅延の文字起こしに向けたキャッシュ考慮型のデコードを採用している。英語のみに対応していた nemotron-speech-streaming-en-0.6b を多言語へ拡張したモデルにあたる。
2026-07-24 · FRIDAY · 15:44 PDT
1poolside/Laguna-S-2.1-NVFP489,186 DOWNLOADS · 129 LIKESLaguna S 2.1 is poolside's 117.6B-parameter mixture-of-experts model, with about 8.5B parameters active per token, built for agentic coding and long-horizon tasks; this is its NVFP4-quantized release for efficient inference. It is available via OpenRouter and the Vercel AI Gateway. The release has 129 likes on Hugging Face.Laguna S 2.1 是 poolside 推出的 1176 亿参数混合专家(MoE)模型,每个 token 约激活 85 亿参数,面向 agentic 编程与长周期任务;这是其用于高效推理的 NVFP4 量化版本。它可通过 OpenRouter 和 Vercel AI Gateway 使用。该发布在 Hugging Face 获得 129 个赞。Laguna S 2.1はpoolsideの1176億パラメータのMixture-of-Expertsモデルで、トークンあたり約85億パラメータを活性化し、エージェント型のコーディングや長期タスク向けに設計されている。本リリースは効率的な推論のためのNVFP4量子化版だ。OpenRouterやVercel AI Gatewayから利用でき、Hugging Faceで129のいいねを得ている。
2fdtn-ai/antares-1b4,266 DOWNLOADS · 144 LIKESAntares-1b is a compact 1-billion-parameter text-generation model from fdtn-ai. Despite modest downloads, it is trending on Hugging Face with 144 likes, pointing to community interest in a small new base model. Public documentation is minimal at release.Antares-1b 是 fdtn-ai 推出的紧凑型 10 亿参数文本生成模型。尽管下载量不高,它却在 Hugging Face 上以 144 个赞进入热门,显示社区对这一新的小型基础模型的兴趣。发布时公开文档很少。Antares-1bはfdtn-aiによるコンパクトな10億パラメータのテキスト生成モデルだ。ダウンロード数は控えめながら、Hugging Faceで144のいいねを集めて話題になっており、新しい小型ベースモデルへのコミュニティの関心をうかがわせる。公開ドキュメントは公開時点でごくわずかだ。
3Kwaipilot/KAT-Coder-V2.5-Dev396 DOWNLOADS · 117 LIKESKAT-Coder-V2.5-Dev is a post-trained coding model from Kwaipilot, released as Hugging Face Transformers weights and compatible with vLLM and SGLang for serving. It targets software-development tasks. The early release has 117 likes on Hugging Face.KAT-Coder-V2.5-Dev 是 Kwaipilot 推出的经过后训练的编程模型,以 Hugging Face Transformers 权重发布,并兼容 vLLM 和 SGLang 部署。它面向软件开发任务。该早期发布在 Hugging Face 获得 117 个赞。KAT-Coder-V2.5-DevはKwaipilotによる事後学習済みのコーディングモデルで、Hugging Face Transformers形式の重みとして公開され、vLLMやSGLangでの提供に対応する。ソフトウェア開発タスクを対象とする。早期の公開でHugging Faceで117のいいねを得ている。
4openbmb/MiniCPM-RobotTrack349 DOWNLOADS · 122 LIKESMiniCPM-RobotTrack is a compact vision-language-action model from OpenBMB, built on MiniCPM4-0.5B for embodied target tracking - following a specified object as a robot moves. It extends the small MiniCPM family toward robotics. The release has 122 likes on Hugging Face.MiniCPM-RobotTrack 是面壁智能(OpenBMB)推出的紧凑型视觉-语言-动作模型,基于 MiniCPM4-0.5B 构建,用于具身目标跟踪——在机器人移动时跟随指定物体。它把小型 MiniCPM 家族拓展到机器人领域。该发布在 Hugging Face 获得 122 个赞。MiniCPM-RobotTrackはOpenBMBによるコンパクトな視覚-言語-行動モデルで、MiniCPM4-0.5Bを基盤に、身体化された対象追跡(ロボットが移動しながら指定した物体を追う)向けに構築された。小型のMiniCPMファミリーをロボティクスへ広げる。Hugging Faceで122のいいねを得ている。
5nvidia/Cosmos3-Edge30,303 DOWNLOADS · 110 LIKESCosmos3-Edge is part of Nvidia's Cosmos 3 family of omnimodal world models for physical AI, packaged for edge deployment. World models like these aim to let robots and autonomous systems predict physical dynamics. It has 110 likes on Hugging Face.Cosmos3-Edge 属于英伟达 Cosmos 3 系列——面向物理 AI 的全模态世界模型,并针对边缘部署做了打包。这类世界模型意在让机器人和自主系统预测物理动态。它在 Hugging Face 获得 110 个赞。Cosmos3-Edgeは、フィジカルAI向けの全モーダルな世界モデル群であるエヌビディアのCosmos 3ファミリーの一部で、エッジ展開向けにパッケージ化されている。こうした世界モデルは、ロボットや自律システムが物理的な動態を予測できるようにすることを狙う。Hugging Faceで110のいいねを得ている。
2026-07-23 · THURSDAY · 23:34 PDT
1microsoft/Mage-Flow411 DOWNLOADS · 193 LIKESMage-Flow is Microsoft's compact 4B-parameter generative stack for text-to-image generation and instruction-based image editing, released alongside a companion paper. Rather than scaling to tens of billions of parameters, it targets efficient synthesis at native resolution. The Hugging Face release drew 193 likes.Mage-Flow 是微软推出的紧凑型 40 亿参数生成栈,用于文本生成图像以及基于指令的图像编辑,并配有同名论文。它不追求数百亿参数规模,而是主打原生分辨率下的高效合成。该 Hugging Face 发布获得 193 个赞。Mage-Flowはマイクロソフトのコンパクトな40億パラメータ生成スタックで、テキストからの画像生成と指示ベースの画像編集に対応し、同名の論文も公開された。数百億規模へ拡大せず、ネイティブ解像度での効率的な生成を狙う。Hugging Faceでの公開は193のいいねを集めた。
2openbmb/MiniCPM-RobotManip408 DOWNLOADS · 165 LIKESMiniCPM-RobotManip is a 1.5B vision-language-action model from OpenBMB built for on-device robotic manipulation. It ships as a single generalist policy meant to drive a robot from perception and instructions rather than task-specific models. The release gathered 165 likes on Hugging Face.MiniCPM-RobotManip 是面壁智能(OpenBMB)推出的 15 亿参数视觉-语言-动作模型,面向端侧机器人操作。它以单一通用策略发布,意在直接从感知和指令驱动机器人,而非依赖任务专用模型。该发布在 Hugging Face 获得 165 个赞。MiniCPM-RobotManipはOpenBMBによる15億パラメータの視覚-言語-行動モデルで、オンデバイスのロボット操作向けだ。タスク専用モデルではなく、知覚と指示からロボットを駆動する単一の汎用ポリシーとして公開された。Hugging Faceで165のいいねを獲得した。
3poolside/Laguna-S-2.1-GGUF25,360 DOWNLOADS · 116 LIKESPoolside published GGUF builds of its Laguna S 2.1 model for local inference with llama.cpp, bundled with a DFlash speculative-decoding draft model for faster generation. The quantized release makes poolside's coding-oriented model runnable on consumer hardware. It has 116 likes on Hugging Face.Poolside 发布了其 Laguna S 2.1 模型的 GGUF 版本,便于用 llama.cpp 做本地推理,并附带用于加速生成的 DFlash 推测解码草稿模型。量化发布让 poolside 这款偏编程的模型能在消费级硬件上运行。它在 Hugging Face 获得 116 个赞。Poolsideは自社のLaguna S 2.1モデルのGGUF版を公開し、llama.cppでのローカル推論に対応させ、生成を高速化するDFlash投機的デコード用のドラフトモデルも同梱した。量子化版により、poolsideのコーディング志向モデルが一般的なハードウェアで動く。Hugging Faceで116のいいねを得ている。
4bottlecapai/ThinkingCap-Qwen3.6-27B25,231 DOWNLOADS · 530 LIKESThinkingCap is a finetune of Qwen3.6-27B that reaches similar capability while using about 50% fewer reasoning tokens on average, and over 90% fewer in the best cases. It targets the cost of long chain-of-thought by trimming how much the model thinks before answering. The release drew 530 likes on Hugging Face.ThinkingCap 是对 Qwen3.6-27B 的微调版本,在保持相近能力的同时,平均减少约 50% 的推理 token,最好情况下减少 90% 以上。它针对长链式思考的成本,压缩模型在作答前“思考”的用量。该发布在 Hugging Face 获得 530 个赞。ThinkingCapはQwen3.6-27Bのファインチューン版で、能力を同等に保ちつつ推論トークンを平均で約50%、最良のケースでは90%以上削減する。回答前にモデルが「考える」量を絞ることで、長い思考連鎖のコストに対処する。公開は530のいいねを集めた。
5moonshotai/Kimi-K2.7-Code766,522 DOWNLOADS · 1,255 LIKESKimi K2.7 Code is Moonshot AI's coding-focused agentic model, built on Kimi K2.6 with gains on real-world, long-horizon software tasks. It emphasizes end-to-end completion of complex engineering workflows while improving token efficiency. With over 766,000 downloads and 1,255 likes, it was the day's most-adopted release.Kimi K2.7 Code 是月之暗面(Moonshot AI)偏编程的 agentic 模型,基于 Kimi K2.6 构建,在真实世界的长周期软件任务上有所提升。它强调端到端完成复杂工程流程,同时改善 token 效率。凭借超过 76.6 万次下载和 1255 个赞,它是当日采用最广的发布。Kimi K2.7 CodeはMoonshot AIのコーディング志向のエージェント型モデルで、Kimi K2.6を基盤に、現実の長期的なソフトウェアタスクで性能を高めた。複雑なエンジニアリング工程のエンドツーエンドな完遂を重視しつつ、トークン効率も改善する。76万超のダウンロードと1255のいいねで、当日最も採用された公開だった。
2026-07-22 · WEDNESDAY
1poolside/Laguna-S-2.13,056 DOWNLOADS · 411 LIKESPoolside's Laguna S 2.1 is a code-focused generation model released with a blog post and hosted access on OpenRouter and Vercel's AI Gateway. It is the most-trending new model on the Hub this cycle.Poolside 的 Laguna S 2.1 是一款面向代码的生成模型,随发布博客上线,并在 OpenRouter 和 Vercel AI Gateway 提供托管访问。它是本轮 Hub 上最热门的新模型。PoolsideのLaguna S 2.1は、ブログ記事とともに公開されOpenRouterやVercel AI Gatewayでホスト提供されるコード特化の生成モデル。今サイクルでHub最注目の新モデルだ。
2upstage/Solar-Open2-250B0 DOWNLOADS · 337 LIKESUpstage released Solar Open 2, a 250B-parameter (A15B active) open model - a sizable open-weights entry from a Korean lab, drawing strong early likes on the Hub.Upstage 发布 Solar Open 2,一个 250B 参数(激活 15B)的开放模型——来自韩国实验室的重量级开放权重成果,在 Hub 上早早收获大量点赞。UpstageはSolar Open 2を公開。250Bパラメータ(アクティブ15B)のオープンモデルで、韓国のラボによる大型のオープンウェイト成果としてHubで早くも多くの「いいね」を集める。
3Nanbeige/Nanbeige4.2-3B0 DOWNLOADS · 245 LIKESNanbeige4.2-3B is a compact 3B agentic model built for tool use and multi-step tasks, part of the trend of small models tuned specifically to drive agents.Nanbeige4.2-3B 是一个紧凑的 3B 智能体模型,专为工具调用和多步任务打造,属于"小模型专门调优来驱动智能体"这一趋势的一部分。Nanbeige4.2-3Bは、ツール利用と多段タスク向けに作られたコンパクトな3Bのエージェント型モデルで、エージェント駆動に特化して調整された小型モデルの潮流の一つ。
4conradlocke/krea2-identity-edit0 DOWNLOADS · 500 LIKESKrea 2 Identity Edit is a community image model for identity-preserving edits - changing a scene while keeping a face consistent. It is trending on likes despite being freshly posted.Krea 2 Identity Edit 是一个社区图像模型,用于保持身份一致的编辑——在更换场景的同时保持人脸一致。尽管刚发布,已凭点赞登上趋势。Krea 2 Identity Editは、同一人物の顔を保ったまま場面を変えるアイデンティティ保持編集向けのコミュニティ画像モデル。公開直後ながら「いいね」でトレンド入りしている。
5Motif-Technologies/Motif-3-Beta125 DOWNLOADS · 162 LIKESMotif-3-Beta is a preview checkpoint of Motif Technologies' next model, shared ahead of the final release for early testing - an increasingly common way labs gather community feedback on the Hub.Motif-3-Beta 是 Motif Technologies 下一代模型的预览检查点,在正式发布前放出供早期测试——这是各实验室在 Hub 上收集社区反馈越来越常见的方式。Motif-3-Betaは、Motif Technologiesの次期モデルのプレビュー版チェックポイントで、正式リリース前に早期テスト用として公開された。ラボがHubでコミュニティのフィードバックを集める、ますます一般的な手法だ。
2026-07-21 · TUESDAY
1prism-ml/Ternary-Bonsai-27B-gguf432,196 DOWNLOADS · 889 LIKESPrism ML's ternary-weight Bonsai 27B in GGUF form: full 27B-class reasoning at a fraction of the memory, per the whitepaper. The most-downloaded build of the family this week (an MLX 1-bit variant also trends).Prism ML 的三值权重 Bonsai 27B(GGUF 版):按白皮书说法,以极低内存实现完整 27B 级推理。本周该家族下载最多的版本(另有 MLX 1-bit 变体同时在榜)。Prism MLの三値重みBonsai 27B(GGUF版)。ホワイトペーパーによれば、わずかなメモリで27Bクラスの推論能力をフルに発揮する。今週このファミリーで最もダウンロードされたビルド(MLX 1bit版もトレンド入り)。
2baidu/Unlimited-OCR2,237,351 DOWNLOADS · 2,580 LIKESBaidu's OCR model pitched as one-shot long-horizon parsing: whole documents in a single pass instead of page-by-page pipelines. 2.2M downloads with day-one ms-swift community integration.百度的 OCR 模型,主打一次性长程解析:整份文档一遍过,而不是逐页流水线。220 万下载,发布当天即获 ms-swift 社区集成。BaiduのOCRモデル。ページ単位のパイプラインではなく、文書全体を一度に解析する「ワンショット長距離パース」を掲げる。220万ダウンロード、公開初日からms-swiftコミュニティ統合。
3DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF62,842 DOWNLOADS · 221 LIKESA community Qwen3.6-27B finetune whose model card claims arc-c scores above 700 in 8-bit. Benchmark numbers on community cards are self-reported - read them as claims, not results.社区微调的 Qwen3.6-27B,模型卡宣称 8-bit 下 arc-c 超过 700 分。社区模型卡上的基准数字均为自报,应视作'声明'而非'结果'。コミュニティによるQwen3.6-27Bファインチューン。モデルカードは8bitでarc-c 700超を主張する。コミュニティ製カードのベンチマーク数値は自己申告であり、結果ではなく主張として読むべきだ。
4empero-ai/Qwythos-9B-Claude-Mythos-5-1M-GGUF2,133,420 DOWNLOADS · 2,383 LIKESGGUF quantizations of Empero's Qwythos-9B with a 1M-token context window. Over 2.1M downloads puts it among the Hub's most-pulled models this week.Empero 的 Qwythos-9B(100 万 token 上下文)的 GGUF 量化版。超 210 万下载,位列本周 Hub 拉取量最高的模型之一。EmperoのQwythos-9B(100万トークンコンテキスト)のGGUF量子化版。210万超のダウンロードで、今週Hubで最も取得されたモデルの一つ。
5HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive1,997,690 DOWNLOADS · 2,960 LIKESAn uncensored community build of Qwen3.6-35B-A3B nearing 2M downloads. Sustained demand for unrestricted local models remains one of the Hub's clearest usage signals.Qwen3.6-35B-A3B 的社区去限制版本,下载量接近 200 万。对无限制本地模型的持续需求,是 Hub 上最清晰的使用信号之一。Qwen3.6-35B-A3Bのコミュニティ製アンセンサード版で、ダウンロードは200万に迫る。制限のないローカルモデルへの根強い需要は、Hubで最も明確な利用シグナルの一つだ。
2026-07-17 · FRIDAY
1zai-org/GLM-5.2534,698 DOWNLOADS · 4,078 LIKESZhipu's flagship open model line updated to 5.2, one of the most-liked releases on the Hub this cycle, shipping with a technical report and an active community.智谱的旗舰开放模型系列更新至 5.2,是本周期 Hub 上获赞最多的发布之一,附带完整技术报告和活跃社区。Zhipuのフラッグシップ・オープンモデル系列が5.2に更新。今サイクルのHubで最も「いいね」を集めたリリースの一つで、技術レポートと活発なコミュニティを伴う。
2thinkingmachines/Inkling7,870 DOWNLOADS · 973 LIKESThinking Machines' image-text-to-text model, published in BF16 and NVFP4 builds with Tinker Cookbook integration and full documentation.Thinking Machines 的图文到文本模型,提供 BF16 和 NVFP4 版本,并集成 Tinker Cookbook 与完整文档。Thinking Machinesの画像テキスト→テキストモデル。BF16とNVFP4ビルドで公開され、Tinker Cookbook統合と完全なドキュメントが付属する。
3prism-ml/Bonsai-27B-gguf1,045,182 DOWNLOADS · 403 LIKESA 1-bit 27B model in GGUF form promising full 27B-class reasoning at a fraction of the memory, per Prism ML's whitepaper. Past a million downloads.GGUF 格式的 1-bit 27B 模型,按 Prism ML 白皮书的说法以极低内存实现完整 27B 级推理能力。下载量已超百万。GGUF形式の1ビット27Bモデル。Prism MLのホワイトペーパーによれば、わずかなメモリで27Bクラスの推論能力を発揮する。ダウンロードは100万回超。
4OpenMOSS-Team/MOSS-Transcribe-Diarize83,160 DOWNLOADS · 250 LIKESAn end-to-end 0.9B audio model that transcribes and diarizes in a single pass - who spoke and what they said, without a pipeline of separate tools.0.9B 的端到端音频模型,一次完成转写与说话人分离——谁说了什么一步到位,无需拼接多个工具。0.9Bのエンドツーエンド音声モデル。文字起こしと話者分離を一度に行い、「誰が何を話したか」を複数ツールの組み合わせなしで出力する。
5ATH-MaaS/OvisOCR210,795 DOWNLOADS · 153 LIKESAn OCR-focused vision model released with a technical report and online demo. Document-to-structured-text is the workhorse task behind most RAG pipelines.专注 OCR 的视觉模型,随附技术报告和在线演示。文档转结构化文本是大多数 RAG 流水线背后的基础工作。OCRに特化したビジョンモデルで、技術レポートとオンラインデモ付き。文書から構造化テキストへの変換は、多くのRAGパイプラインを支える基盤タスクだ。