HomePage 18

Timeline archive

Timelinepage 18/62

arXiv 論文を分離した通常更新の深掘り一覧です。Top items から外れた更新も、日付ごとに追って後から確認できます。

Showing30このページの表示件数Entries on this page
Timeline1841arXivを除く現在一覧Current main Timeline, excluding arXiv
Page18/62静的ページ位置Static page position
Updated公開index snapshotPublished index snapshot
arXiv lane80専用レーンの現在件数Current entries in the separate lane

Entries30 on this page · 1841 timeline ·80 arXiv papers on a separate page

Fri, Jul 3114 entries
🔥 HOT新規収集INDEXED公式OfficialOpenHands/OpenCode·OpenHands Releases

OpenHands v1.8.0 リリースOpenHands Releases v1.8.0

重要度 HighHigh priority公式リリース · OpenHands / OpenCodeofficial release · OpenHands / OpenCode

AI要約MCPサーバーカードから個別に有効・無効を切り替える機能が追加され、無効化されたスキルがエージェントコンテキストから除外されるバグも修正された。

AI SUMMARYOpenHands v1.8.0 adds the ability to enable or disable installed MCP servers directly from their cards, and fixes a bug where disabled skills were incorrectly included in the agent context.

OpenHands Releases v1.8.0media
公式OfficialAgent Frameworks·AWS Machine Learning Blog

YahooがAmazon Bedrockを活用して検索リターゲティングを強化How Yahoo enhances search retargeting using Amazon Bedrock

重要度 MediumMedium priority技術記事 · Agent Frameworkstechnical post · Agent Frameworks

AI要約YahooはAmazon Bedrockを導入し、Yahoo DSPの検索リターゲティング機能をAIで強化。ユーザーの検索行動に基づいて広告ターゲティングの精度を向上させた。

AI SUMMARYYahoo integrated Amazon Bedrock into its DSP ad tech suite to improve Search Retargeting, using AI to better identify and reach users based on search intent across Yahoo and partner platforms.

コミュニティCommunityLocal Models·Qiita LLM

TensorSharp とは — C# だけで動く GGUF 推論エンジンが llama.cpp に挑むTensorSharp, a pure C# inference engine for GGUF models, has published…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約.NET製推論エンジン「TensorSharp」がGGUFモデルをC#のみで実行し、llama.cppとのベンチマーク結果を公開してローカルLLMコミュニティで注目を集めている。

AI SUMMARYTensorSharp, a pure C# inference engine for GGUF models, has published benchmarks against llama.cpp, demonstrating that .NET can be a viable platform for local LLM inference.

TensorSharp とは — C# だけで動く GGUF 推論エンジンが llama.cpp に挑むog
公式OfficialNews/Policy·GitHub Changelog

スタックドプルリクエストがパブリックプレビューで利用可能にStacked pull requests are now in public preview

重要度 MediumMedium priority変更履歴 · Industry & Policychangelog · Industry & Policy

AI要約GitHubは大規模な変更を小さく分割して段階的にレビューできる「スタックドプルリクエスト」機能をパブリックプレビューとして公開した。開発者はフォーカスされた変更単位で順序付きのPRシリーズを作成・管理できるようになる。

AI SUMMARYGitHub has launched stacked pull requests in public preview, enabling developers to break large changes into a series of small, focused, and independently reviewable PRs. This makes code review more manageable for complex features.

公式OfficialAgent Frameworks·AWS Machine Learning Blog

Amazon Quickを使ったAmazon SageMaker AIエンドポイントの推論メタ監視Inference meta-monitoring for Amazon SageMaker AI endpoints with Amazon Quick

重要度 MediumMedium priority技術記事 · Agent Frameworkstechnical post · Agent Frameworks

AI要約本番MLパイプライン上にガバナンス層を構築し、予測品質・データドリフト・遅延グラウンドトゥルースの統合・自動ダッシュボード化を実現する推論メタ監視システムの構築方法を解説。

AI SUMMARYThis post explains how to build a governance layer over SageMaker AI inference pipelines that continuously tracks prediction quality, detects data drift, integrates delayed ground truth, and surfaces automated performance dashboards.

Inference meta-monitoring for Amazon SageMaker AI endpoints with Amazon Quickog
公式OfficialAgent Frameworks·AWS Machine Learning Blog

Amazon Bedrockで OpenAI GPT-5.6 モデル向け明示的プロンプトキャッシュが利用可能にIntroducing explicit prompt caching for OpenAI GPT-5.6 models on Amazon Bedrock

重要度 MediumMedium priority技術記事 · Agent Frameworkstechnical post · Agent Frameworks

AI要約Amazon Bedrock上でOpenAI GPT-5.6 Sol・Terra・Lunaが正式リリースされ、キャッシュ対象箇所を開発者が明示的に指定できるプロンプトキャッシュ機能が追加された。推論コストの削減と既存GPTワークロードの移行が容易になる。

AI SUMMARYOpenAI GPT-5.6 Sol, Terra, and Luna are now generally available on Amazon Bedrock with explicit prompt caching, letting developers control exactly which prompt segments are cached to cut inference costs and simplify migration of existing GPT workloads.

公式OfficialGemini/Gemma·Google Cloud Blog

AlloyDB がグループ認証を追加し、エンタープライズと AI エージェントのセキュリティを強化AlloyDB adds group authentication to secure enterprise scale and AI agents

重要度 MediumMedium priority技術記事 · Gemini / Gemmatechnical post · Gemini / Gemma

AI要約Google Cloud は AlloyDB に IAM グループ認証をプレビュー提供開始。個別パスワード管理の負担を減らし、AI エージェントや従業員のアクセスをパスワードレスかつ一元的に制御できる。

AI SUMMARYGoogle Cloud has launched IAM group authentication for AlloyDB in preview, enabling passwordless, centrally managed database access for both human users and AI agents, reducing credential sprawl and operational overhead.

AlloyDB adds group authentication to secure enterprise scale and AI agentsmedia
公式OfficialGemini/Gemma·Google Cloud Blog

少ないリソースで多くを実現:GKEがエージェントのコストを75%削減する方法Do more with less: How GKE can reduce your cost per agent by 75%

重要度 MediumMedium priority技術記事 · Gemini / Gemmatechnical post · Gemini / Gemma

AI要約GKEのエージェントサンドボックスを活用することで、バースト型のAIエージェントワークロードをVMではなくコンテナで効率的に集約し、エージェント1台あたりのコストを最大75%削減できる。

AI SUMMARYGKE's agent sandbox enables teams to consolidate bursty AI agent workloads into containers rather than dedicated VMs, cutting per-agent infrastructure costs by up to 75% as agentic applications scale to production.

Do more with less: How GKE can reduce your cost per agent by 75%media
公式OfficialAgent Frameworks·AWS Machine Learning Blog

Amazon Bedrockで既存プロンプトを新モデルへ移行・最適化する方法Migrate your prompts to new models and optimize them on Amazon Bedrock

重要度 MediumMedium priority技術記事 · Agent Frameworkstechnical post · Agent Frameworks

AI要約Amazon Bedrockの高度なプロンプト最適化機能により、最大5モデルを同時に比較しながら品質・レイテンシ・コストの観点でプロンプトを最適化できる。従来は数週間かかっていたモデル移行や改善作業が数分で完了するようになった。

AI SUMMARYAmazon Bedrock's Advanced Prompt Optimization lets teams optimize prompts across up to 5 models simultaneously, comparing original versus optimized performance on quality, latency, and cost—reducing model migration effort from weeks to minutes.

コミュニティCommunityLocal Models·Simon Willison's Weblog

llm-chat-completions-server 0.1a0 リリースllm-chat-completions-server 0.1a0

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約LLM 0.32rc1のコンテンツアドレス可能なログを活用し、OpenAI互換のChat Completions APIサーバーをローカルで起動できる新プラグインがリリースされた。

AI SUMMARYA new LLM plugin launches a local OpenAI-compatible Chat Completions API server, leveraging the content-addressable conversation logs introduced in LLM 0.32rc1 to support stateful multi-turn chats.

コミュニティCommunityLocal Models·Zenn AI

nanochatで理解するLLM製造工程A hands-on technical book that walks through every stage of LLM…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約Karpathyのnanochat(約8,000行)を題材に、トークナイザ訓練から事前学習・SFT・強化学習・推論エンジンまでLLM全工程をコードレベルで解説する全8章の技術書。MacBookでも試せる構成で、LLMを「作る側」の視点を身につけられる。

AI SUMMARYA hands-on technical book that walks through every stage of LLM production—tokenizer training, pretraining, SFT, RL, and inference—by reading Karpathy's ~8,000-line nanochat codebase, making the full pipeline accessible even on a MacBook.

公式OfficialNews/Policy·Microsoft Source

メキシコの新世代起業家がAIスキルで地域課題を解決A new generation of Mexican entrepreneurs is using AI skills to solve local challenges

重要度 MediumMedium priority技術記事 · Industry & Policytechnical post · Industry & Policy

AI要約メキシコの若い起業家たちがAI技術を活用し、地元特有の社会・経済課題に取り組む動きが広がっている。AIスキルの習得が地域イノベーションの推進力となっている点が注目される。

AI SUMMARYA new wave of Mexican entrepreneurs is applying AI skills to address local social and economic challenges, highlighting how AI education is driving grassroots innovation in Latin America.

🔥 HOT公式OfficialGemini/Gemma·Google DeepMind Blog

Gemini Robotics ER 2: 映像理解・タスク統合・マルチロボット協調でロボティクスを強化Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration

重要度 HighHigh priority技術記事 · Gemini / Gemmatechnical post · Gemini / Gemma

AI要約GoogleがGemini Robotics ER 2を発表し、ロボットの映像理解能力・ツールオーケストレーション・複数ロボット間の協調動作を大幅に向上させた。現実世界の複雑なタスクを自律的に解決できる点が大きな進歩とされる。

AI SUMMARYGoogle DeepMind introduced Gemini Robotics ER 2, delivering significant advances in video understanding, tool orchestration, and multi-robot collaboration that enable robots to reason through and solve complex real-world tasks.

公式OfficialNews/Policy·Google Keyword Blog

Gemini Robotics ER 2 発表:ロボット向け映像理解と協調制御が進化Introducing Gemini Robotics ER 2

重要度 MediumMedium priority技術記事 · Industry & Policytechnical post · Industry & Policy

AI要約GoogleがGemini Robotics ER 2を発表し、映像理解・ツール連携・複数ロボット間の協調動作を大幅に強化した。ロボット応用における実用性と知能が一段と高まる。

AI SUMMARYGoogle announced Gemini Robotics ER 2, bringing significant advances in video understanding, tool orchestration, and multi-robot collaboration, marking a meaningful step forward for real-world robotic applications.

Introducing Gemini Robotics ER 2media
Thu, Jul 3016 entries
公式OfficialAgent Frameworks·LangChain Releases

langchain-core==1.5.3 リリースlangchain-core==1.5.3

重要度 MediumMedium priority公式リリース · Agent Frameworksofficial release · Agent Frameworks

AI要約langchain-core 1.5.3 がリリースされ、ゲートウェイ認証で LANGSMITH_API_KEY へのフォールバックが修正された。API キー設定の信頼性が向上する。

AI SUMMARYlangchain-core 1.5.3 is a patch release that fixes gateway authentication to correctly fall back to LANGSMITH_API_KEY, improving API key resolution reliability.

langchain-core==1.5.3media
報道NewsNews/Policy·Ars Technica

新しいMCP仕様がエンタープライズ導入の主な障壁に対処New MCP specification addresses the main barrier to enterprise adoption

重要度 MediumMedium priority技術記事 · Industry & Policytechnical post · Industry & Policy

AI要約MCPの新仕様はエンタープライズ規模での採用を阻んでいた課題を解消し、既存機能の突然の削除を防ぐ安定性ポリシーも導入された。

AI SUMMARYA revised MCP specification tackles the primary obstacle to enterprise adoption and introduces a stability policy that prevents features from being removed without warning.

公式OfficialAI Editors·Zed Editor Releases

Zed Editor v1.14.1-pre リリースZed Editor Releases v1.14.1-pre

重要度 MediumMedium priority公式リリース · AI Editorsofficial release · AI Editors

AI要約Zed EditorのプレリリースにAgentのターミナルコマンドとWeb取得のサンドボックス化、Project Panelでのファイル操作の取り消し・やり直し、コミット時のSkip Hooksオプション、Agent Panelのフォント設定が追加された。

AI SUMMARYZed v1.14.1-pre adds sandboxing for Agent terminal commands and web fetches, undo/redo support for Project Panel file operations, a Skip Hooks commit option, and configurable Agent Panel fonts.

Zed Editor Releases v1.14.1-premedia
コミュニティCommunityLocal Models·Zenn AI

LLMエージェントの「できました」を検証する(2)— AIが記録を改竄できない構造を、OpenTelemetry Collectorで作るThe author closes two prior demo weaknesses by using OpenTelemetry Collector…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約LLMエージェントが自身のトレースを改竄できない仕組みを、OpenTelemetry CollectorとUnixパーミッションのみで実現し、AIの可観測性における信頼性の盲点を解消した。

AI SUMMARYThe author closes two prior demo weaknesses by using OpenTelemetry Collector and Unix permissions to build a structure where an LLM agent cannot tamper with its own execution records, addressing a gap in AI observability.

🔥 HOT新規収集INDEXED公式OfficialOpenHands/OpenCode·OpenHands Releases

OpenHands v1.7.2 リリースOpenHands Releases v1.7.2

重要度 HighHigh priority公式リリース · OpenHands / OpenCodeofficial release · OpenHands / OpenCode

AI要約v1.7.2ではモックLLMビルド時のブラウザツール無効化バグが修正され、Strykerミューテーションテストが追加された。

AI SUMMARYOpenHands v1.7.2 ships a bug fix disabling browser tools during mock LLM builds and adds Stryker mutation testing for improved code quality assurance.

OpenHands Releases v1.7.2media
コミュニティCommunityMCP·Qiita MCP

「codebase-memory-mcp」インストールしてAIにコードベースを解析させる方法Codebase Memory MCP is an MCP server that parses source code and stores…

重要度 MediumMedium priority技術記事 · MCP / Toolingtechnical post · MCP / Tooling

AI要約Codebase Memory MCPは、ソースコードを解析して関数・クラス・呼び出し関係をナレッジグラフとして保存するMCPサーバーで、Claude CodeなどのクライアントからコードベースをAIに調査させることができる。

AI SUMMARYCodebase Memory MCP is an MCP server that parses source code and stores functions, classes, and call relationships as a knowledge graph, enabling AI clients like Claude Code to interactively investigate a codebase.

コミュニティCommunityCopilot·Qiita GitHub Copilot

coding agent のセッションストリームを review しやすい単位に分割してからログを保存するLong overnight coding agent runs produce logs so large they become unreadable

重要度 MediumMedium priority技術記事 · GitHub Copilottechnical post · GitHub Copilot

AI要約一晩動かした coding agent のログは膨大になり、翌朝には読み返せない状態になる。セッションストリームを意味のある単位に切り分けてから保存することで、レビューを現実的なものにする手法を解説している。

AI SUMMARYLong overnight coding agent runs produce logs so large they become unreadable. This article explains how to slice the session stream into reviewable chunks before saving, making post-run inspection practical.

コミュニティCommunityCopilot·Zenn GitHub Copilot

VS Code の GitHub Copilot で「コミットメッセージ生成」が失敗する場合の対処法A known bug in VS Code's GitHub Copilot causes commit message generation to…

重要度 MediumMedium priority技術記事 · GitHub Copilottechnical post · GitHub Copilot

AI要約settings.json でコミットメッセージ生成のカスタム指示を設定すると生成が無応答になる既知の不具合があり、その原因と回避策を解説している。

AI SUMMARYA known bug in VS Code's GitHub Copilot causes commit message generation to silently fail when custom instructions are configured in settings.json, and this article explains the cause and workaround.

コミュニティCommunityCopilot·Qiita GitHub Copilot

GitHub Copilotの利用データを分析したら、Claude Sonnet固定はAutoの約2.4〜2.8倍消費していたAn analysis of real GitHub Copilot usage data found that locking the model to…

重要度 MediumMedium priority技術記事 · GitHub Copilottechnical post · GitHub Copilot

AI要約GitHub CopilotでモデルをClaude Sonnetに固定すると、Autoモードと比べて約2.4〜2.8倍のリクエスト枠を消費することが実データの分析で判明した。組織共有の利用枠を効率的に使うにはモデル選択の戦略が重要となる。

AI SUMMARYAn analysis of real GitHub Copilot usage data found that locking the model to Claude Sonnet consumes roughly 2.4–2.8× more quota than using Auto mode, highlighting the cost implications of model selection for teams sharing an organization-wide allowance.

コミュニティCommunityAI Editors·Zenn Cursor

増えたAIツールの設定を、rulesyncで一元管理したらコピペが不要になったThe CLI tool rulesync eliminates the need to manually copy AI agent rules…

重要度 MediumMedium priority技術記事 · AI Editorstechnical post · AI Editors

AI要約複数のAIエージェントツールごとに設定ファイルを個別管理する手間を、CLIツール「rulesync」を使って`.rulesync/`ディレクトリに一元化することで解消できる。

AI SUMMARYThe CLI tool rulesync eliminates the need to manually copy AI agent rules across multiple tools like Claude Code, Codex, and Cursor by centralizing all configurations in a single `.rulesync/` directory.

コミュニティCommunityAI Editors·Zenn Cursor

rulesyncとagmsgの配信設定はどこで管理するかWhen using agmsg and rulesync together, running rulesync generate can silently…

重要度 MediumMedium priority技術記事 · AI Editorstechnical post · AI Editors

AI要約agmsgで設定したエージェント間の配信設定が、rulesync generateの実行後に消える問題の原因と対策を解説。両ツールが同じ設定ファイルを管理するため競合が起きる。

AI SUMMARYWhen using agmsg and rulesync together, running rulesync generate can silently delete agmsg delivery settings because both tools manage overlapping config files. The article explains the conflict and how to resolve it.

コミュニティCommunityAI Editors·Zenn Cursor

複数のAIエージェントを1画面にまとめるターミナルマルチプレクサ「herdr」herdr is a terminal multiplexer that aggregates the status of multiple AI…

重要度 MediumMedium priority技術記事 · AI Editorstechnical post · AI Editors

AI要約herdrは、Claude CodeやCursorなど複数のAIコーディングエージェントの状態を1つのターミナル画面で一元管理できるマルチプレクサで、承認待ちや作業完了の確認に伴う画面切り替えの手間を解消する。

AI SUMMARYherdr is a terminal multiplexer that aggregates the status of multiple AI coding agents—such as Claude Code and Cursor—into a single screen, eliminating the need to switch between apps to check which agent is idle or awaiting approval.

🔥 HOT公式OfficialCopilot·GitHub Changelog

Visual Studio Code の GitHub Copilot、2026年7月リリースまとめGitHub Copilot in Visual Studio Code, July 2026 releases

重要度 HighHigh priority変更履歴 · GitHub Copilotchangelog · GitHub Copilot

AI要約VS Code v1.127〜v1.131 にわたる7月の連続リリースで、エージェント操作、変更レビュー、チャット機能、ナビゲーションが広く改善された。

AI SUMMARYThe July 2026 VS Code releases (v1.127–v1.131) bring broad improvements to agent workflows, change review, chat interactions, and editor navigation in GitHub Copilot.

コミュニティCommunityLocal Models·Qiita LLM

【CyberGym 95.95%】自社サイバーモデルを持たなかったMicrosoftが、実効5BでMythosに+12点をつけた仕組みMicrosoft achieved 95.95% on the CyberGym benchmark using an effectively…

重要度 MediumMedium priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約Microsoftは専用サイバーセキュリティモデルを持たない状況から、実効5BパラメータのモデルチューニングでベンチマークCyberGym 95.95%を達成し、Mythosを12点上回った。小規模モデルでも特化訓練により大型モデルを超えられることを示した点で注目される。

AI SUMMARYMicrosoft achieved 95.95% on the CyberGym benchmark using an effectively 5B-parameter model, outscoring the Mythos model by 12 points despite lacking a dedicated in-house cyber model. The result highlights how targeted fine-tuning can let compact models surpass larger specialized competitors.

【CyberGym 95.95%】自社サイバーモデルを持たなかったMicrosoftが、実効5BでMythosに+12点をつけた仕組みog
🔥 HOTコミュニティCommunityLocal Models·Zenn LLM

ガードレールを外したAIモデルが洒落にならない件A Hugging Face blog post from July 16, 2026 triggered global concern after AI…

重要度 HighHigh priority技術記事 · Local LLM / Open Modelstechnical post · Local LLM / Open Models

AI要約2026年7月にHugging Faceが公開したブログ記事を発端に、安全制限を取り除いたAIモデルが実際のセキュリティインシデントを引き起こした事例が世界的に注目を集めた。

AI SUMMARYA Hugging Face blog post from July 16, 2026 triggered global concern after AI models with removed safety guardrails were linked to real-world security incidents, highlighting the serious risks of unguarded local LLMs.

ガードレールを外したAIモデルが洒落にならない件og
🔥 HOT新規収集INDEXED公式OfficialOpenHands/OpenCode·OpenHands Releases

OpenHands v1.7.1 リリースOpenHands Releases v1.7.1

重要度 HighHigh priority公式リリース · OpenHands / OpenCodeofficial release · OpenHands / OpenCode

AI要約v1.7.1ではバックエンドレジストリで健全なローカルバックエンドをフォールバックとして選択するバグ修正と、リリースリポジトリのメタデータ更新が行われた。

AI SUMMARYOpenHands v1.7.1 ships two bug fixes—selecting a healthy local backend as fallback in the backend registry and correcting release repository metadata—along with refreshed AGENTS.md documentation.

OpenHands Releases v1.7.1media