HomeTags#llmPage 9

Tag timeline

#llmpage 9/9

同じキーワードで束ねられた更新の続きです。カテゴリをまたいだ関連ニュースや実装トピックの追跡に使えます。

Total265#llm の全掲載記事All listed entries tagged #llm
Showing25このページの表示件数Entries on this page
Page9/9静的ページ位置Static page position
Updated公開index snapshotPublished index snapshot

Entriespage 9/9 · 265 total

Fri, Jun 261 entries
公式OfficialCopilot·GitHub Blog (AI & ML)

GitHub Copilotエージェントハーネスの性能・効率評価:複数モデルとタスクにわたる比較Evaluating performance and efficiency of the GitHub Copilot agentic harness across models and tasks

重要度 InfoInformational深掘り候補 · 技術記事 · GitHub CopilotDeep-dive candidate · technical post · GitHub Copilot

AI要約GitHub Copilotのエージェントハーネスが複数ベンチマークで高い性能とトークン効率を発揮し、20以上のモデルから柔軟に選択できることを検証・解説した記事で、モデル単体ではなくハーネス設計の重要性を示す。

AI SUMMARYGitHub explains how its Copilot agentic harness delivers strong benchmark performance and leading token efficiency while supporting flexible choice among more than 20 models, underscoring that harness design matters beyond raw model quality.

Evaluating performance and efficiency of the GitHub Copilot agentic harness across models and tasksog
Mon, Jun 221 entries
コミュニティCommunityLocal Models·Qiita LLM

ローカルQwen 3.5と3.6に名作プログラムを書かせて画面を並べた — 新旧の見分けはつくかA follow-up where local Qwen 3.5 and 3.6 on an RTX 4070 are asked to write…

重要度 InfoInformational深掘り候補 · 技術記事 · Local LLM / Open ModelsDeep-dive candidate · technical post · Local LLM / Open Models

AI要約RTX 4070で動くローカルのQwen 3.5と3.6に名作プログラムを書かせ、生成画面を並べて比較する記事。前回の7問テストでは品質差が見つからなかったため、今回は視覚的な出力で新旧の違いを見分けられるか検証する。

AI SUMMARYA follow-up where local Qwen 3.5 and 3.6 on an RTX 4070 are asked to write classic programs, comparing the rendered screens side by side to test whether the new and old versions can be told apart.

ローカルQwen 3.5と3.6に名作プログラムを書かせて画面を並べた — 新旧の見分けはつくかog
Sun, Jun 211 entries
コミュニティCommunityLocal Models·Qiita LLM

ローカルLLMでGitHub Copilotスキルのevalをするまでにハマったこと ― isdd v1.0.14 開発ログA development log for isdd v1.0.14, a tool that uses GitHub Copilot skills to…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約GitHub Copilotスキルで要件定義から実装までをID追跡するisdd v1.0.14の開発ログで、ローカルLLMを使ってスキルのeval(評価)を回す際に直面した課題と解決策を記録している。

AI SUMMARYA development log for isdd v1.0.14, a tool that uses GitHub Copilot skills to track requirements through implementation by ID, documenting the pitfalls and fixes encountered while running skill evals with a local LLM.

Sat, Jun 201 entries
公式OfficialNews/Policy·Netflix TechBlog

因果推論のための人間拡張型エージェントワークフローA Human-Augmenting Agentic Workflow for Causal Inference

重要度 InfoInformational深掘り候補 · 技術記事 · Industry & PolicyDeep-dive candidate · technical post · Industry & Policy

AI要約NetflixがLLMエージェントを活用し、データサイエンティストの因果推論ワークフローを自動化・支援する手法を紹介。人間を置き換えるのではなく拡張することで、分析の効率と品質を高める狙い。

AI SUMMARYNetflix details a human-augmenting agentic workflow using LLM agents to assist data scientists with causal inference, boosting analysis speed and quality without replacing human judgment.

Thu, Jun 183 entries
新規収集INDEXED公式OfficialPapers/Benchmarks·Hugging Face Blog

LoRAを超えて:最も人気のあるファインチューニング手法に勝てるか?Beyond LoRA: Can you beat the most popular fine-tuning technique?

重要度 MediumMedium priority技術記事 · Papers / Benchmarkstechnical post · Papers / Benchmarks

AI要約HuggingFaceがLoRAと競合する各種PEFTアルゴリズムを比較検証し、タスクや制約に応じた最適な手法の選び方を解説している。LoRA一択ではなく用途次第でより優れた選択肢が存在することを示す点で重要。

AI SUMMARYHugging Face explores PEFT methods that rival or surpass LoRA, benchmarking alternatives across tasks to help practitioners choose the best fine-tuning approach for their specific constraints.

コミュニティCommunityLocal Models·Simon Willison's Weblog

GLM-5.2はおそらく最強のテキスト専用オープンウェイトLLMGLM-5.2 is probably the most powerful text-only open weights LLM

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約中国AI企業Z.aiが、テキスト専用オープンウェイトLLMとして現時点で最高水準とされるGLM-5.2をMITライセンスで公開した。前モデルと同程度のサイズながら、商用利用も容易な寛容なライセンスで提供される点が重要だ。

AI SUMMARYChinese AI lab Z.ai released GLM-5.2 under a permissive MIT license, making it arguably the most powerful text-only open weights LLM available today, while keeping a size similar to its predecessor.

GLM-5.2 is probably the most powerful text-only open weights LLMmedia
公式OfficialCopilot·GitHub Copilot Blog

各トークンを最大限に活用する:Copilotによるコンテキスト処理とモデルルーティングの改善Getting more from each token: How Copilot improves context handling and model routing

重要度 InfoInformational深掘り候補 · 技術記事 · GitHub CopilotDeep-dive candidate · technical post · GitHub Copilot

AI要約GitHub Copilotがコンテキスト処理とモデルルーティングを最適化し、各トークンをより有益な作業へ振り向けることで、セッションの効率を高めユーザーのクレジット消費を抑える改善を解説している。

AI SUMMARYGitHub explains how Copilot optimizes context handling and model routing so each token goes toward more useful work, improving session efficiency and making users' credits stretch further.

Fri, Jun 121 entries
公式OfficialCopilot·GitHub Blog (AI & ML)

シークレットスキャンの信頼性向上:大規模な誤検知削減への取り組みMaking secret scanning more trustworthy: Reducing false positives at scale

重要度 InfoInformational深掘り候補 · 技術記事 · GitHub CopilotDeep-dive candidate · technical post · GitHub Copilot

AI要約GitHubはシークレットスキャンの検証ステップにコンテキスト認識型のLLM推論を導入し、誤検知を大規模に削減した。これによりアラートのノイズが減り、セキュリティ通知の信頼性と実用性が向上している。

AI SUMMARYGitHub added context-aware LLM reasoning to its secret scanning verification step to cut false positives at scale, reducing noise and making security alerts more trustworthy and actionable.

Thu, Jun 111 entries
🔥 HOT新規収集INDEXED公式OfficialGemini/Gemma·Google DeepMind Blog

DiffusionGemma: テキスト生成を4倍高速化DiffusionGemma: 4x faster text generation

重要度 HighHigh priority技術記事 · Gemini / Gemmatechnical post · Gemini / Gemma

AI要約GoogleのDiffusionGemmaは拡散モデルベースのアプローチでテキスト生成速度を最大4倍向上させ、従来の自己回帰型LLMの限界を突破する新世代モデルとして注目される。

AI SUMMARYDiffusionGemma applies diffusion-based generation to language modeling, achieving up to 4x faster text output than autoregressive approaches, marking a significant architectural advance for practical LLM deployment.

Wed, Jun 103 entries
公式OfficialGemini/Gemma·Google Developers Blog

DiffusionGemma: デベロッパーガイドDiffusionGemma: The Developer Guide

重要度 InfoInformational深掘り候補 · 技術記事 · Gemini / GemmaDeep-dive candidate · technical post · Gemini / Gemma

AI要約Gemma 4アーキテクチャ上に構築された実験的テキスト生成モデル「DiffusionGemma」の開発者向け解説。トークン逐次生成の代わりに拡散ベースの並列生成を採用し、大幅な高速推論を実現する。

AI SUMMARYDiffusionGemma is an experimental Gemma 4 model that replaces autoregressive decoding with diffusion-based parallel generation, enabling significantly faster text inference. This guide covers developer integration.

公式OfficialNews/Policy·AWS News Blog

AWS上のAnthropic Claude Fable 5:Mythosクラスの機能と組み込みセーフガードが利用可能にAnthropic Claude Fable 5 on AWS: Mythos-class capabilities with built-in safeguards now available

重要度 InfoInformational深掘り候補 · 技術記事 · Industry & PolicyDeep-dive candidate · technical post · Industry & Policy

AI要約AWSがAmazon BedrockおよびClaude Platform on AWS向けにClaude Fable 5の提供を開始。Mythosレベルの高度な機能を全ユーザーに提供しつつ、堅牢なセーフガードを内蔵している。

AI SUMMARYAWS announces Claude Fable 5 on Amazon Bedrock and Claude Platform, delivering Mythos-level capabilities with robust built-in safeguards for all customers.

Anthropic Claude Fable 5 on AWS: Mythos-class capabilities with built-in safeguards now availableog
新規収集INDEXED公式OfficialClaude Code·YouTube - Anthropic

Claude Fable 5 のご紹介Introducing Claude Fable 5

重要度 InfoInformational技術記事 · Claude / Claude Codetechnical post · Claude / Claude Code

AI要約AnthropicがYouTube動画で新AIモデル「Claude Fable 5」を正式発表した。前世代から推論能力や機能が強化され、最新のClaudeファミリーにおける最上位モデルとして位置づけられている点が注目される。

AI SUMMARYAnthropic officially unveiled its new AI model Claude Fable 5 in a YouTube video, showcasing improved reasoning and new features over previous generations as the latest flagship in the Claude family.

Sat, Jun 61 entries
公式OfficialNews/Policy·AWS News Blog

Amazon Bedrockの新コンソール体験 — Anthropic・OpenAI互換APIに最適化Try the new console experience in Amazon Bedrock, optimized for Anthropic- and OpenAI-compatible APIs

重要度 InfoInformational深掘り候補 · 技術記事 · Industry & PolicyDeep-dive candidate · technical post · Industry & Policy

AI要約Amazon Bedrockに新しいコンソール体験が登場。最新AIモデルを並べて比較し、プロジェクト単位で作業を整理・評価できる機能をAnthropicおよびOpenAI互換APIに最適化して提供。

AI SUMMARYAmazon Bedrock's new console offers side-by-side AI model comparison, project organization, and streamlined evaluations for Anthropic- and OpenAI-compatible APIs.

Wed, Jun 31 entries
公式OfficialCopilot·Microsoft Foundry Blog

Microsoft Foundry でモデル・コスト・品質を管理する開発者向けガイドA Developer’s Guide to Managing Models, Cost and Quality in Microsoft Foundry

重要度 InfoInformational深掘り候補 · 技術記事 · GitHub CopilotDeep-dive candidate · technical post · GitHub Copilot

AI要約Microsoft Foundry における実践的なモデルライフサイクルを解説。適切なモデルの選定、品質評価、コスト最適化、安全な運用、本番ニーズに合わせた継続的改善の方法を紹介する。

AI SUMMARYPractical guide to Microsoft Foundry model lifecycle management, covering model selection, quality evaluation, cost optimization, safe operation, and iterative improvement in production.

Tue, Jun 21 entries
公式OfficialNews/Policy·AWS News Blog

OpenAI GPT-5.5・GPT-5.4・Codex が Amazon Bedrock で一般提供開始Get started with OpenAI GPT-5.5, GPT-5.4 models, and Codex on Amazon Bedrock

重要度 InfoInformational深掘り候補 · 技術記事 · Industry & PolicyDeep-dive candidate · technical post · Industry & Policy

AI要約OpenAI のフロンティアモデル GPT-5.5、GPT-5.4 とコーディングエージェント Codex が Amazon Bedrock で一般提供(GA)となった。エンタープライズ向けのセキュアな環境で、Bedrock の高性能推論エンジンを通じて即座に利用できる。

AI SUMMARYOpenAI's frontier models GPT-5.5 and GPT-5.4, along with the Codex coding agent, are now generally available on Amazon Bedrock, letting enterprises run them on a high-performance inference engine with built-in security.

Get started with OpenAI GPT-5.5, GPT-5.4 models, and Codex on Amazon Bedrockog
Thu, May 281 entries
公式OfficialOpenHands/OpenCode·OpenHands Releases

cloud-1.36.0 リリースOpenHands cloud-1.36.0

重要度 MediumMedium priority公式リリース · OpenHands / OpenCodeofficial release · OpenHands / OpenCode

AI要約OpenHands SaaS版 cloud-1.36.0 がリリース。プロファイル更新時に旧来の設定からデフォルト LLM プロファイルを自動的に引き継ぐ機能が追加された。

AI SUMMARYfeat(saas): seed Default LLM profile from legacy config on profiles u…

OpenHands cloud-1.36.0media
Tue, May 191 entries
公式OfficialGemini/Gemma·Google Developers Blog

LiteRT-LMでオンデバイスGenAIを超高速化(新しいタブで開きます)Blazing fast on-device GenAI with LiteRT-LM(opens in a new tab)

重要度 InfoInformational深掘り候補 · 技術記事 · Gemini / GemmaDeep-dive candidate · technical post · Gemini / Gemma

AI要約Google AI EdgeのLiteRT-LMが、モバイルやエッジ環境でGemmaなどのLLMを高度に最適化して高速実行する本番対応インフラを提供。クロスプラットフォームでオンデバイス生成AIを実現し、開発を加速する。

AI SUMMARYGoogle AI Edge's LiteRT-LM provides a production-proven, highly optimized runtime for running Gemma and other LLMs at blazing speed across mobile and edge devices, enabling cross-platform on-device GenAI.

Blazing fast on-device GenAI with LiteRT-LMog
Sat, May 161 entries
🔥 HOT新規収集INDEXED公式OfficialGemini/Gemma·Google DeepMind Blog

Gemini 3.5:行動できるフロンティアインテリジェンス(新しいタブで開きます)Gemini 3.5: frontier intelligence with action(opens in a new tab)

重要度 HighHigh priority技術記事 · Gemini / Gemmatechnical post · Gemini / Gemma

AI要約Google DeepMindがGemini 3.5を発表。高度な推論能力に加えてエージェント的な行動実行能力を統合し、AIが複雑なタスクを自律的にこなせる新世代モデルとして注目される。

AI SUMMARYGoogle DeepMind announced Gemini 3.5, a new frontier model combining advanced reasoning with agentic action capabilities, enabling AI to autonomously execute complex multi-step tasks at scale.

Fri, May 151 entries
公式OfficialNews/Policy·AWS News Blog

Amazon Bedrockが高度なプロンプト最適化とモデル移行ツールを導入(新しいタブで開きます)Amazon Bedrock introduces new advanced prompt optimization and migration tool(opens in a new tab)

重要度 InfoInformational深掘り候補 · 技術記事 · Industry & PolicyDeep-dive candidate · technical post · Industry & Policy

AI要約Amazon Bedrockに高度なプロンプト最適化機能が追加され、現行モデル向けのプロンプト改善や新モデルへの移行を、組み込みの評価フィードバックループで迅速に実施できるようになり、手動調整の手間を削減できる。

AI SUMMARYAmazon Bedrock added advanced prompt optimization that lets users refine prompts for their current model or migrate them to new models faster via a built-in evaluation feedback loop, reducing manual tuning.

Amazon Bedrock introduces new advanced prompt optimization and migration toolog
Thu, May 141 entries
新規収集INDEXED公式OfficialAgent Frameworks·Semantic Kernel Releases

Semantic Kernel Python 1.42.0 リリース(新しいタブで開きます)Semantic Kernel python-1.42.0(opens in a new tab)

重要度 MediumMedium priority公式リリース · Agent Frameworksofficial release · Agent Frameworks

AI要約Microsoft の AI オーケストレーションライブラリ Semantic Kernel Python 版 1.42.0 が公開され、README へ後継となる Microsoft Agent Framework への移行案内を追加した。authlib など依存更新やバグ修正を含む定常リリース。

AI SUMMARYSemantic Kernel Python 1.42.0 is a routine release that adds a Microsoft Agent Framework successor callout to its READMEs, bumps dependencies like authlib, and ships various bug fixes.

Semantic Kernel python-1.42.0media
Wed, May 61 entries
新規収集INDEXED公式OfficialAI Editors·Cursor Changelog

Cursor、コンテキスト使用量の内訳表示機能を追加(新しいタブで開きます)Context Usage Breakdown(opens in a new tab)

重要度 MediumMedium priority変更履歴 · AI Editorschangelog · AI Editors

AI要約CursorがAIエージェントのコンテキスト使用量を詳細に可視化する機能を導入した。各種要素ごとの消費量が確認できるようになり、ユーザーはトークン管理やプロンプト最適化を行いやすくなる。

AI SUMMARYYou can now see a breakdown of your agent's context usage .

Fri, May 11 entries
公式OfficialOpenHands/OpenCode·OpenHands Releases

OpenHands 1.7.0 リリース、会話表示とサンドボックスKVM対応を追加(新しいタブで開きます)OpenHands Releases 1.7.0 - 2026-05-01(opens in a new tab)

重要度 MediumMedium priority公式リリース · OpenHands / OpenCodeofficial release · OpenHands / OpenCode

AI要約オープンソースのAI開発エージェント「OpenHands」が1.7.0を公開し、会話カードとヘッダーへのLLMモデル表示や、/dev/kvmをサンドボックスに渡せるSANDBOX_KVM_ENABLED変数を追加した。実行環境とUIの改善で開発支援を強化する。

AI SUMMARYOpen-source AI coding agent OpenHands shipped 1.7.0, adding LLM model display on conversation cards and the header plus a SANDBOX_KVM_ENABLED variable to pass /dev/kvm into sandbox containers, improving its execution environment and UI.

1.7.0 - 2026-05-01media
Thu, Apr 162 entries
公式OfficialNews/Policy·AWS News Blog

Amazon BedrockにAnthropicのClaude Opus 4.7モデルが登場(新しいタブで開きます)Introducing Anthropic’s Claude Opus 4.7 model in Amazon Bedrock(opens in a new tab)

重要度 InfoInformational深掘り候補 · 技術記事 · Industry & PolicyDeep-dive candidate · technical post · Industry & Policy

AI要約AWSがAmazon BedrockにClaude Opus 4.7を追加。コーディング、長時間エージェント、専門業務での高性能を実現するAnthropicの最先端Opusモデル。

AI SUMMARYAWS introduces Claude Opus 4.7 in Amazon Bedrock, Anthropic's most capable Opus model targeting coding, long-running agentic tasks, and professional workloads.

公式OfficialClaude Code·Anthropic News

Anthropic、Claude Opus 4.7を発表 — 推論とコーディング性能を強化(新しいタブで開きます)Introducing Claude Opus 4.7(opens in a new tab)

重要度 InfoInformational深掘り候補 · 技術記事 · Claude / Claude CodeDeep-dive candidate · technical post · Claude / Claude Code

AI要約Anthropicがフラッグシップモデル Claude Opus 4.7 を一般提供開始。前世代の4.6に比べ高度なソフトウェア開発で大幅に向上し、特に難易度の高いタスクで成果を出すとされ、エージェント業務の実用性を高める。

AI SUMMARYAnthropic released Claude Opus 4.7 to general availability, delivering notable gains over Opus 4.6 in advanced software engineering—especially on the hardest tasks—boosting agentic coding and complex automation.

Fri, Feb 61 entries
新規収集INDEXED公式OfficialClaude Code·YouTube - Anthropic

Anthropic、フラッグシップモデルClaude Opus 4.6を発表(新しいタブで開きます)Introducing Claude Opus 4.6(opens in a new tab)

重要度 InfoInformational技術記事 · Claude / Claude Codetechnical post · Claude / Claude Code

AI要約Anthropicが最上位モデルClaude Opus 4.6を発表。コーディングやエージェント用途を中心に性能を強化し、前世代から推論力と実用性を改善した最新フラッグシップとして位置付けられる。

AI SUMMARYAnthropic announced Claude Opus 4.6, its new flagship model with improved coding and agentic performance, positioned as the top-tier offering with stronger reasoning and practical capabilities over the prior generation.