HomeTags#mlx

Tag timeline

#mlx14 total

同じキーワードで束ねられた更新を確認できます。カテゴリをまたいだ関連ニュースや実装トピックの追跡に使えます。

Total14#mlx の全掲載記事All listed entries tagged #mlx
Showing14このページの表示件数Entries on this page
Page1/1静的ページ位置Static page position
Updated公開index snapshotPublished index snapshot

Entriespage 1/1 · 14 total

Sat, Aug 151 entries
新規収集INDEXED公式OfficialLocal Models·Ollama Releases

v0.32.12: qwen3.8に「renderer」とMLXインポートのサポートを追加Ollama Releases v0.32.12

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約Ollama v0.32.12では、Qwen3.8モデル向けに専用レンダラーを追加し、safetensorsインポート時にchatテンプレートのマーカーを検出してMLXインポートにも対応した。思考・ツール・継続などのパース処理も強化されている。

AI SUMMARYOllama v0.32.12 adds a dedicated qwen3.8 renderer and MLX import support by detecting reasoning-effort and preserved-thinking markers in the chat template during safetensors import, improving handling of thinking, tools, and continuation parsing.

Ollama Releases v0.32.12media
Fri, Aug 141 entries
新規収集INDEXED公式OfficialLocal Models·Ollama Releases

Ollama v0.32.10 リリースOllama Releases v0.32.10

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約repeat_penaltyのデフォルト値が1.1から1.0(無効)に変更され、他エンジンとの互換性向上と投機的デコードの高速化を実現。NVFP4 MLXモデルのプリフィル速度も改善された。

AI SUMMARYOllama v0.32.10 changes the default repeat_penalty from 1.1 to 1.0 (off) to match other engines and speed up speculative decoding, while also improving prefill performance for NVFP4 MLX models with a global scale.

Ollama Releases v0.32.10media
Thu, Aug 131 entries
新規収集INDEXED公式OfficialLocal Models·Ollama Releases

Ollama v0.32.10-rc1 リリースv0.32.10-rc1: mlx: avoid pulling MLX models when MLX is missing (#17710)

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約repeat_penaltyのデフォルト値が1.1から1.0(無効)に変更され、他エンジンとの互換性向上と投機的デコードの高速化が図られた。NVFP4 MLXモデルのプリフィル速度も改善されている。

AI SUMMARYOllama v0.32.10 changes the default repeat_penalty from 1.1 to 1.0 (off) to match other engines and speed up speculative decoding, while also improving prefill performance for NVFP4 MLX models.

v0.32.10-rc1: mlx: avoid pulling MLX models when MLX is missing (#17710)media
Tue, Aug 111 entries
新規収集INDEXED公式OfficialLocal Models·Ollama Releases

Ollama v0.32.8 リリースOllama Releases v0.32.8

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約Ollama v0.32.8 では、Muse Glimmerが全プラットフォームで利用可能になり、Claude CodeやCodexなどのコーディングエージェントや長期動作するパーソナルアシスタントの用途に対応する。

AI SUMMARYOllama v0.32.8 brings Muse Glimmer to all platforms, enabling coding agent applications like Claude Code and Codex as well as long-running personal assistants, powered by Ollama's MLX engine.

Ollama Releases v0.32.8media
Mon, Aug 101 entries
公式OfficialLocal Models·Ollama Releases

Ollama v0.32.7 リリースOllama Releases v0.32.7

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約Ollama v0.32.7がリリースされ、Apple SiliconのMLXエンジン経由でMeta製モデル「Muse Glimmer」の初期サポートが追加された。NVIDIA・AMDなど他プラットフォーム向けの最適化は近日提供予定。

AI SUMMARYOllama v0.32.7 adds initial support for Meta's Muse Glimmer model via the MLX engine on Apple Silicon, with broader NVIDIA, AMD, and other platform support coming soon.

Ollama Releases v0.32.7media
Thu, Aug 61 entries
公式OfficialLocal Models·Ollama Releases

Ollama v0.32.6 リリースOllama Releases v0.32.6

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約Qwen3.5がApple GPU上でMLXエンジンのMTPヘッドによる投機的デコードにより高速化され、OpenAI互換ストリーミング形式も修正された。

AI SUMMARYOllama v0.32.6 speeds up Qwen3.5 on Apple GPUs via automatic speculative decoding with the MLX engine, and fixes /v1/chat/completions streaming to match OpenAI's wire format.

Ollama Releases v0.32.6media
Wed, Aug 51 entries
コミュニティCommunityLocal Models·Simon Willison's Weblog

PipeNetwork/minimax-h3-mlx:MLX向けMiniMax-H3ローカル実行ガイドPipeNetwork/minimax-h3-mlx

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約MiniMaxがテキスト・画像・音声・動画を扱うマルチモーダルモデル「MiniMax-H3」を公開し、PipeNetworkがApple SiliconのMLXフレームワーク上でローカル実行できる実装を提供した。

AI SUMMARYMiniMax released MiniMax-H3, an omni-modal model supporting text, image, audio, and video generation including 15-second clips, and PipeNetwork published an MLX-based implementation enabling local inference on Apple Silicon.

PipeNetwork/minimax-h3-mlxog
Sun, Aug 21 entries
コミュニティCommunityLocal Models·Zenn LLM

Ollama 0.30.8はMLXランナーを内蔵するがGGUFは通らない — M1 Max 64GB実測Ollama 0.30.8 ships with an integrated MLX runner for Apple Silicon, but…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約Ollama 0.30.8にMLXバックエンドが統合されたが、バイナリ解析とログ突合の結果、通常の`ollama pull`で取得するGGUFモデルはMLXランナーを経由しないことが判明した。速度改善の恩恵を受けるにはモデル形式の確認が必要となる。

AI SUMMARYOllama 0.30.8 ships with an integrated MLX runner for Apple Silicon, but hands-on investigation on an M1 Max 64GB showed that standard GGUF models pulled via `ollama pull` do not go through the MLX path, meaning users cannot assume a speed gain without verifying the active backend.

Ollama 0.30.8はMLXランナーを内蔵するがGGUFは通らない — M1 Max 64GB実測og
Wed, Jul 291 entries
コミュニティCommunityLocal Models·Zenn LLM

公開MLX変換は本当に動くか — 使えない変換を実測で見分ける方法Even models published on Hugging Face as MLX conversions can be broken — one…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約Hugging Faceに「MLX変換済み」として公開されているモデルでも、ロード不能や全文字化けといった致命的な不具合を抱える例があり、著者がBaiduのOCRモデルを題材に既存変換2種を実測して問題を明らかにした。重みファイルが生成できても正常動作するとは限らず、実測による検証が不可欠だと示している。

AI SUMMARYEven models published on Hugging Face as MLX conversions can be broken — one failing to load and another producing garbled output — as the author discovered when benchmarking two existing conversions of Baidu's Unlimited-OCR (3.3B, MIT). The article argues that generating weight files does not guarantee a working model, and only empirical testing can confirm usability.

Sun, Jul 261 entries
コミュニティCommunityLocal Models·Zenn LLM

日本語OCRモデル Sarashina2.2-OCR を MLX へ移植する実装記録This article documents the process of porting the Japanese OCR model…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約Sarashina2.2-OCRをApple Silicon向けMLXフレームワークへ移植する際、モデルカードに記載されていない実装の詳細を調査・解決した過程をまとめた記事。ローカル環境で高精度な日本語OCRを動かしたい開発者にとって実践的な参考資料となる。

AI SUMMARYThis article documents the process of porting the Japanese OCR model Sarashina2.2-OCR to the MLX framework for Apple Silicon, uncovering implementation details absent from the official model card. It serves as a practical guide for developers aiming to run high-accuracy Japanese OCR locally.

Sat, Jul 181 entries
コミュニティCommunityLocal Models·Zenn AI

ローカルLLM study1-a: gemma4 e2b/e4b の MLX 版はどれだけ速いかThis article benchmarks gemma4 e2b/e4b models running via the MLX framework on…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約Apple Silicon 向け MLX フレームワークで動作する gemma4 の e2b/e4b モデルの推論速度を実測・比較した記事。ローカル環境での実用性を判断する上で参考になるベンチマーク結果を提供している。

AI SUMMARYThis article benchmarks gemma4 e2b/e4b models running via the MLX framework on Apple Silicon, measuring real-world inference speed to assess local deployment viability.

Wed, Jul 11 entries
公式OfficialLocal Models·Ollama Releases

v0.31.1: MLX バックエンドの Gemma4 MoE ローディングコードを改善Ollama Releases v0.31.1

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約Ollama v0.31.1 では、MLX バックエンドにおける Gemma4 の MoE モデルのローディングコードが整理・強化された。Apple Silicon 環境での Gemma4 MoE モデルの読み込み安定性が向上する。

AI SUMMARYOllama v0.31.1 refines the MLX backend's Gemma4 mixture-of-experts model loading code, improving stability and reliability when running Gemma4 MoE models on Apple Silicon.

Ollama Releases v0.31.1media
Thu, Jun 181 entries
公式OfficialLocal Models·Ollama Releases

Ollama v0.30.10 リリースOllama Releases v0.30.10

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約Apple Silicon上でCommand AおよびNorthファミリーのモデルがMLXエンジンで動作可能になり、内部のllama.cppエンジンをbuild 9672へ更新、MLXビルド成果物の修正も含む小規模アップデート。

AI SUMMARYOllama v0.30.10 enables Command A and North family models to run on Apple Silicon via the MLX engine, updates the underlying llama.cpp engine to build 9672, and fixes MLX build artifacts.

Ollama Releases v0.30.10media
Sat, Jun 131 entries
公式OfficialLocal Models·Ollama Releases

Ollama v0.30.8 リリースOllama Releases v0.30.8

重要度 MediumMedium priority公式リリース · Local LLM / Open Modelsofficial release · Local LLM / Open Models

AI要約Ollama v0.30.8がリリースされ、起動時のプロバイダー誤選択を修正。プロンプトキャッシュをコンテキストシフトから分離してKVキャッシュの再利用を改善し、MLX推論の安定性も向上した。

AI SUMMARYOllama v0.30.8 fixes incorrect provider selection at launch, improves prompt caching by decoupling it from context shift for better KV cache reuse, and delivers more stable MLX inference.

Ollama Releases v0.30.8media