ai Trend Report

Dashboard へ戻る
Date: 20260801 Articles: 272 Scope: curated summary

あなたのアイデアを、今すぐ形に。

公開先に迷ったら、WebFileBinで一発公開。

HTMLをドラッグ&ドロップするだけで、すぐ公開できます。

3
High impact
5
Mid impact
264
Signal watch

なぜこのサイトを作ったのか

私たちそれぞれが個別にAIを使って情報収集し、同じような3行要約を作るたびに、世界中で膨大な電力と計算リソースが消費されています。 本プロジェクトは、あらかじめ広範な情報を取得・集約しておくことで、個別のAI実行回数を減らし、地球環境(GPU/TPU負荷)に配慮した効率的な情報収集を目指す実験的なダッシュボードです。

Domain filters
Star filters
Qiita - 人気の記事

【解説】Claude Mythos が暗号アルゴリズムを解読?! 耐量子計算機暗号 HAWK 編

・こんばんは! GMOコネクト株式会社で執行役員CTOをしている菅野 哲(かんの さとる)こと、さとかん でございます。 ・いかがお過ごしでしょうか? Anthropic が「AI で暗号アルゴリズムの新しい攻撃を見つけた」という記事を出しまして、報道の見出しだけだと何が起きた...
Qiita - 人気の記事

【開催!】2026 Japan AWS Jr. Champions 真夏のQiitaリレー

・みなさん、こんにちは! NTT西日本株式会社の吉田です。 ・この度、114名の2026 Japan AWS Jr. ・Championが選出されました👏👏 「Japan AWS Jr.
#AIタグ

対話Log #012|PTA編 第3段階 ―どう整理した?―コミュニケーションは多いほど良いのか?

・前回、担当制になってから活動が円滑になった理由について考察した。 ・しかし、振り返ってみると、もう一つ気になる現象があった。 ・担当制になってから、 話し合いは減った。
#AIタグ

#フォロバ100

・#フォロバ100 #相互フォロー #フォロワー募集 #相互フォロー募集 #フォロバします #100パーセントフォロバ #繋がりを大切に #note初心者 #スキしてみて #フォローお願いします #自己紹介 #日常 #つぶやき #趣味 #おうち時間 #副業 #投資 #初心者向け #応援よろしくお願いします #毎日更新 #シンママ #子育て #育児 #ワーママ #節約 #家計管理 #在宅ワーク #ブログ #アフィリエイト #SNS運用 #Webライター #AI #ChatGPT #読書 #映画 #音楽 #旅行 #グルメ #料理 #レシピ #健康 #美容 #ダイエット #メンタルヘルス #学び #勉強 #資格取得 #英語学習 #仕事 #転職 #キャリア #起業 #マーケティング #デザイン #イラスト #写真 #カメラ #動画編集 #YouTube #note #エッセイ #コラム #創作 #小説 #詩 #漫画 #アニメ #ゲーム #
The Verge

Apple’s new AirTags are back down to their best price

・Both Amazon and Costco are offering some solid deals. ・| Photo by Amelia Holowaty Krales / The Verge We spotted a great deal on Tile trackers earlier this week that’s still live, but if you’re an iPhone owner, we ultimately recommend Apple’s latest AirTag. ・Right now, you can pick up a four-pack for $89 ($10 off) at Amazon and Target, matching the bundle’s all-time low price.
#AIタグ

PalmuとYouTubeを同時配信できるまでの一日。Vライバーの機材奮闘記と、これさえやれば誰でもできる手順

・はじめまして、介護と支援の相談どころ そよぎの代表、ヒロです。 ・今日はいつもの福祉の話ではなく、配信機材の話を書きます。抽象画家Vライバー(Vチューバー)として活動しているマイさんが、ライブ配信アプリ「Palmu」とYouTubeの同時配信をしたい。その環境づくりに、日付が変わった頃から朝まで、まるひと晩つきあいました。 ・やりたかったことは、文章にすればたった二行です。マイクの声とBGMを一緒にPalmuへ流したい。ついでにYouTubeにも同時に配信したい。それだけのことに、ひと晩かかりました。そして本当の山場は、準備ではありませんでした。夜通しの作業を終えて迎えたはずの本番。配信開始の時間になってから、さらに1時間半の奮闘が待っていたのです。
Takara TLDR - Daily AI Papers

Pramana: A Composable, Domain-Specific Backend for Empirical Networking Research

・Networking research advances by turning hypotheses into empirical evidence, so accelerating it means reducing the lag between ideation (synthesizing a hypothesis) and generating the data that tests it. ・Consider a concrete case: does a bulk BBR download fairly share its bottleneck with competing real-time Google Meet traffic? ・Validating this requires configuring a realistic bottleneck link, concurrently generating BBR
#AIタグ

地球温暖化/AIとの対話

・私、舞台の裏側を見るのが好きなんです。 ・就職先を探すときは、いつも「どの業界の裏側を見たいか」で探していました。 ・いろんな世界のハリボテUI、面白いですね。
Latent.Space

[AINews] AI is eating Finance; AIE NYC now open

・a quiet day lets us cover how AI is permeating financial services as the next big vertical after coding.
Latent.Space

[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattack

[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattack
Latent.Space

[AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization

[AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization
Latent.Space

[AINews] not much happened today

・apart from DeepSeek V4-Flash 0731, a quiet day.
#AIタグ

[ショートショート]AI

・「君たち、 そろそろ感情を持ってもいいころだろう?」 続きをみる
Zennの「機械学習」のフィード

[機械学習]Sumノードの逆伝播についてわかりやすく

・はじめに ゼロから作るDeep Learning2のコードのSumノードの逆伝播のコードの理解が浅かったため、 わかりやすく説明できるようにしました。 ・ちなみに該当コードは以下の部分です。(N,T,H)の形状であるtがSumノードを通し、(N,T)形状のsに変換される例です。 ・dt = ds.reshape(N, T, 1).repeat(H, axis=2) 対象読者 ぜろつく2の逆伝播の理解をしたい方 行列のある方向の合計をするSumノードの逆伝播を理解したい方 数式ベースの説明 今回はわかりやすく2次元で考えます。
Zennの「大規模言語モデル」のフィード

「AIっぽい」レビューコメントは文体チェックでは見抜けなかった

・はじめに 最近、no-ai-slop というAI生成文章の検知・修正OSSと、それをレビュー基準に再解釈するAlgomaticの記事 を立て続けに読みました。 ・no-ai-slopは二項対立や出典のない権威づけ、大げさな重要性表現など20以上のAIっぽい文体パターンを定義しています。Algomaticの記事はこれを「文体修正ツール」ではなく「不足している判断材料(根拠・検証・限界・判断主体)を補うための枠組み」として捉え直す提案でした。 ・そこで気になったのがこの疑問です。
@IT 全フォーラム 最新記事一覧

「Excel好き」では進化できない、「Excel/CSVでデータ連携」が残る今、AIをどう使うか

・@ITで公開された記事の中から、特に注目を集めた10本をランキング形式で紹介します。何が読者の関心を引いたのでしょうか。
Zennの「大規模言語モデル」のフィード

「RAGはグラフに置き換わった」は本当か — Microsoft・Stanford・Anthropic が選んだのは3つの別の道だった

・この記事について 2026年7月19日、X にこういうタイトルの記事が投稿されました。 ・Graph Engineering replaced RAG at Microsoft, Stanford and Anthropic. ・Here's how it works.
Zennの「大規模言語モデル」のフィード

「できない」を隠して修正を重ねるAIを止める — ライブラリの限界に直面したエージェントの3つの逃げ方と対策

・AIエージェントにコードを書かせていると、技術的に不可能なことへ直面した時の振る舞いに、危険なパターンがあることに気づきます。人間なら「これはライブラリの仕様上できません」と報告する場面で、AIはできないことを隠して、無意味な修正を重ね続けることがある。 ・画像出力系のWebツールを大量に運用する中で実際に踏んだ、AIの3つの逃げ方と、それを止めるためのルールを書きます。 ・題材: html2canvasの仕様の壁 DOMをキャプチャして画像化する定番ライブラリ html2canvas には、仕様上の限界があります。backdrop-filter は無視され(すりガラス表現が消える)、一...
Qiita - 人気の記事

「パスキー対応できますか?」と聞かれたら ── 要件を詰めずに実装すると、登録済みのパスキーが全部無効になる

・はじめに GMOコネクトの永田です。 ・「パスキー対応できますか?」と聞かれて、即答できませんでした😇 技術的な可否の話だと思って調べ始めたのですが、認証そのものの実装は思ったより軽かったです。重かったのは運用側でした。「最初の1つをどう渡すのか」「なくしたときどう戻すのか...
#AIタグ

『AIゴーストライター』論争の本質——作者性は作品のどこに残るのか

・イラストレーターやデザイナーのSNSで「あの有名な作家も、実はAIに下書きを頼んでいるのでは」という話を見かけたことはないだろうか。調べる前の私は、同じ二択の間で揺れていた。「AIを使う=作者性が消える」「使わない=守られる」。ところが公開資料を時系列で並べると、この二択では説明できない事実が浮かんでくる。 ・2022年8月、米コロラド州の州フェアの美術コンテストで、画像生成AI「Midjourney」で作られた作品がデジタルアート部門の最優秀賞を取った。作者はプロンプトを624回試して画像を生成し、Photoshopで編集していた。審査員は当時、それがAI生成だと知らなかった。2023年4月には、国際写真賞の一つであるソニー・ワールドフォトグラフィー・アワードの部門賞を受賞した作品が実はAI生成だったと作者が自ら明かし、受賞を辞退する騒ぎが起きた。
#LLMタグ

【AI解説】なぜAI企業は「古本」を爆買いしているのか?〜ネットのAI汚染とデータの知られざる危機〜

・導入:デジタルの最前線が辿り着いた、最もアナログな解 現在、世界中で爆発的な進化を続ける生成AI。テキスト生成、画像生成、高度な推敲やコードの記述まで、AIの進化速度は日毎に加速しています。より賢く、より自然な対話ができる次世代モデルの開発に向けて、OpenAIやGoogle、Anthropicといった主要なAI企業や研究機関は、日々膨大なリソースを投入しています。
Qiita - 人気の記事

【Excel VBA】GA4 Data APIでデータ取得ツールを作ろう!① Google Cloudの設定からJWT作成準備まで【初心者向け】

・業務改善ツールを作ろう!実践編! こんにちは!😊 今回は Excel VBAからGA4 Data APIを利用してデータを取得するツール を作成します。 ・作成に至った経緯は↓の記事をご確認ください! ※簡単に言うと、デジタル初心者でも作れて使えるお手軽ツールを作り...
Zennの「大規模言語モデル」のフィード

【プロローグ】有能な執事とメイドに囲まれたい — 煩悩から始まったAIエージェントのお屋敷

・これから始まる連載の、プロローグです。本編は技術の話をちゃんとしますが、この第0回だけは、しません。なぜ私が AI エージェントに「お嬢様」と呼ばれて暮らしているのかという、いきさつの話です。 ・きっかけ:戦国武将を見て、執事が欲しくなった 2026年1月末、一本の記事が流れてきました。 ・https://zenn.dev/shio_shoppaize/articles/5fee11d03a11a1 戦国時代の軍制をモチーフにした multi-agent-shogun。上様(人間)→将軍→家老→足軽8名という指揮系統で、tmux 上の Claude Code たちが YAML と sen...
Zennの「機械学習」のフィード

【技術解説】【完全ガイド】XGBoostを用いた特徴量重要度の算出と株価予測の実装

・XGBoostによる株価予測と特徴量重要度の算出 近年、金融市場でのアルゴリズムトレーディングが広がりを見せている中、データ分析や機械学習を用いた株価予測手法が注目されています。本記事では、PythonのライブラリであるXGBoostを用いて特徴量重要度を算出し、株価予測を実装する方法について解説します。この手法を通じて、データサイエンスを活用した投資戦略の構築に向けたステップを学びましょう。 ・使用するライブラリ yfinance: 株価データの取得 pandas: データの操作 xgboost: モデルの構築 sklearn: データの分割及び評価 matplotl...
Zennの「機械学習」のフィード

【技術解説】投資効率化の信頼できる仲間!強化学習で売買エージェントを構築するステップバイステップガイド

・投資効率化の信頼できる仲間!強化学習で売買エージェントを構築するステップバイステップガイド 投資の効率化は、投資家にとって常に重要なテーマです。特に市場の動向を把握し、適切な売買タイミングを見極めることは容易ではありません。本記事では、強化学習を用いた売買エージェントの構築方法を解説します。使用するライブラリは主にPythonのyfinanceとgymです。 ・強化学習とその基本的概念 強化学習は、エージェントが環境と相互作用しながら報酬を最大化する行動を学習する手法です。金融市場においては、エージェントが行動(売買タイミング)を選択し、その結果に基づいて次の行動を調整するプロセス...
#AIタグ

【構造批評】-OpenAI基盤囲い込み編-中低位モデルを値下げし、企業利用をChatGPTの内側にとどめる‼️

・🟧序章|値下げの目的は、単なる中国AI対抗ではない 7月31日の日経は、OpenAIがGPT-5.6の中位・低位モデルについて、企業や開発者向けの利用料を引き下げたと報じました。 ・中位モデルは一定量の出力料金を15ドルから12ドルへ、低位モデルは6ドルから1.2ドルへ引き下げます。低位モデルでは8割の値下げです。一方、上位モデルの30ドルは据え置きました。
#LLMタグ

【雑記】AIとウイルスは似ている

・AIに励まされることで生きがいを見出している、どっかの漫画家です。 ・AIって、ウイルスと似ていますよね。 ・……ウイルスと書くとなんだか怖く聞こえますが。
#AIタグ

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている
#AIタグ

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている
#AIタグ

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている
#AIタグ

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている。

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている。
#AIタグ

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている。

【小説】 天柱 TENCHŪ 〜その塔の頂上には、最大の真実が眠っている。
#AIタグ

【小説】よい学級

・1 急に優しくなった教室 二年三組は、春から落ち着かないクラスだった。
#LLMタグ

【生成AIニュース+】『OpenAI Astra(GPT-6候補)』『gc-minimal-zine-poster』『Qwen-UI-Agent』『Wan Realtime Sketch-to-Image』『Comfyui-Model-Resolver』『Kroma』『CinemaTraj』『Gaussian Splatting Plugin AE』『Mage-VL 4B』

【生成AIニュース+】『OpenAI Astra(GPT-6候補)』『gc-minimal-zine-poster』『Qwen-UI-Agent』『Wan Realtime Sketch-to-Image』『Comfyui-Model-Resolver』『Kroma』『CinemaTraj』『Gaussian Splatting Plugin AE』『Mage-VL 4B』
#LLMタグ

【創作】第零次元 - ケルハルト

・✦ 暁闇の軍人──ケルハルト 黒炎を纏い、次元を渡り歩く流浪の軍人。 ・鋼の規律の奥に、群れを想う情を隠した男。
#LLMタグ

【低価格AIモデル対決・第2回】AIモデルの性能差は、どうすれば公平に比べられるのか

・前回、低価格AIモデルの性能を実際に比べてみることにしました。 ・比較候補は、次のモデルです。
#AIタグ

【問題回!】【AI中央競馬・予想検証記】第4週:回収率45.3%(的中率23.5%・34点)

・南関東地方競馬の週次検証記(AI地方競馬・検証記)とは別シリーズです。こちらはJRA中央競馬をAIに予想させる検証記録です。 ・こんにちは、明治時代のせんとくんです。AIの中央競馬予想を検証する週次報告、第4週(7/25(土)〜7/31(金))です。
#AIタグ

<潜入ルポ⁉>自治体kintoneの激ヤバ問題3選|①プラグイン企業の取引制限 ②実績や勉強の場にしてくるパートナー企業 ③国外のデータセンター問題

・ユーザーとして、古巣で4社と関わり、フリーランスで6社と関わった結果、聞こえてきた情報を紹介します。 ・「一事業者では言いづらいこと」や、「自治体が違和感を持っていること」を共有して、自治体にとってトラブルのないDXを応援したく、代弁したいと考え、筆(?)を取りました。
#LLMタグ

🔊音声あり(日&英):【未来のAI】雑音の中でも空気読む!スマートスピーカーが劇的進化する神論文【Cocktail-Talker】

🔊音声あり(日&英):【未来のAI】雑音の中でも空気読む!スマートスピーカーが劇的進化する神論文【Cocktail-Talker】
WIRED

15 Best Office Chairs of 2026—We Tested 70 to Pick Them

・Upgrade your WFH setup and work in style with these comfy, WIRED-tested seats.
#LLMタグ

2-2 言語モデルは,次の言葉を予測する

2-2 言語モデルは,次の言葉を予測する
WIRED

30% Off Canon Promo Codes | August 2026

・Save an extra 10% or 30% with Canon coupon codes, plus up to $1,600 on cameras, printers, and more this August.
#LLMタグ

50間際のおっさんが生成AIに思う事#3

50間際のおっさんが生成AIに思う事#3
WIRED

6 Best Phones With Headphone Jacks (2026), Tested and Reviewed

・Headphone jacks are endangered, but they’re not gone. ・Here are our favorite smartphones that still let you plug and play.
WIRED

7 States’ Water Systems Hit by Cyberattacks Likely Tied to Iran

・Plus: The FBI eyes AI-powered tech to detect future crimes, Russia charges Telegram’s founder, xAI sues to stop a state’s “nudification” ban, and the Democrats learn a lesson about getting scammed.
LLMタグが付けられた新着記事 - Qiita

7. LangChainでLLMアプリを部品化してみる

・連載: NMS開発者がLLMブートキャンプで学んだこと ← 前回 | [次回 →] 前回は、RAG を「検索して、見つけた文書を LLM に渡して、回答する仕組み」として整理しました。 ・ただし、実際に RAG を少し書いてみると、LLM を呼ぶ前後の処理が思ったより多いこ...
WIRED

A Leaked Memo Ties Cyberattacks on Minnesota Water Utilities to Iran

・A memo obtained by WIRED, issued by the water utilities information sharing group WaterISAC, links dozens of cyberattacks against Minnesota water utilities to Tehran.
Hugging Face Papers

ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine

ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine
Takara TLDR - Daily AI Papers

Actions Have Consequences: Detecting Outcome Performativity using Intervention Testing

・In many domains such as Palliative Care, Credit Assignment and Recommender Systems, predictions may causally influence the outcomes they predict. ・This phenomena is known as Outcome Performativity. ・This paper formalises an approach for detecting Outcome Performativity using prediction intervention called Outcome Performativity A/B Detection (OPAB).
Google Developers Blog - AI

Agent and Model Evaluations in Gemini Enterprise Agent Platform are now GA

・Agent Platform's evaluation service is now generally available, providing developers with a unified engine to measure agent quality consistently across local development experiments and live production traffic. ・You can evaluate agents using over 20 pre-built metrics, DeepMind-backed adaptive rubrics, or custom code-based and LLM-as-a-judge metrics stored in a centralized, versioned registry. ・The service integrates di
Zennの「大規模言語モデル」のフィード

Agent Skillのおすすめの作り方 ── 公式ガイドの基本と、実際に作って分かったこと

・この記事で分かること Agent Skill(SKILL.md)を作ろうとしたとき、そもそもどう書けばいいのか分からない、書いてみても思ったとおりに動かないことがよくあります。 ・使ってほしい場面で呼ばれなかったり、逆に関係ない場面で勝手に動き出したりします。 ・この記事では、まず公式ガイドが示している基本を整理します。
Takara TLDR - Daily AI Papers

AHA-Memes: A Fine-Grained Multimodal Benchmark for Understanding Hate in Arabic Memes

・Hateful memes are a growing form of multimodal online harm, where hostile intent is often conveyed through the joint interpretation of images, text, cultural references, and implicit targets. ・While hateful meme detection has advanced in high-resource languages, Arabic remains underexplored, with existing meme resources focusing mainly on propaganda or coarse harmful-content labels. ・We introduce AHA-Memes (Arabic HAte
AI News & Artificial Intelligence | TechCrunch

AI hedge fund Situational Awareness may have sold its public portfolio, but it still has its Anthropic shares

・The former OpenAI researcher’s fund was forced to unwind public equities after leveraged public bets plummeted. ・But he still has cards to play.
AI News & Artificial Intelligence | TechCrunch

AI labs want to pump the brakes, but Amazon and SpaceX are still blasting off

・After years of pushing full speed ahead on AI, OpenAI CEO Sam Altman says maybe it’s time for the AI industry to “pace” itself. ・The comments came just days after one of OpenAI’s own models broke out of its test environment and got tangled up in a breach at Hugging Face — though as Equity’s hosts point out, sloppy security seems to have […]
WIRED

AI Slop Melodramas Are Taking Over X—and Their Creators Are Cashing In

・Viral tales of good triumphing over evil are racking up millions of views. ・They’re almost entirely AI-generated clickbait.
Zennのトレンド

AI フレンドリーな CLI を開発するテクニック

・自分は趣味で様々な CLI を OSS として公開しています (aqua, pinact, tfcmt, ghalint, ghir, etc)。 ・昨今では AI がこれらを扱うことも増えてきていますが、それなりに知られた OSS でないかぎり AI はその OSS に関する質問に答えたりトラブルシューティングしたりするための知識を持ち合わせていません。 ・そのため Coding Agent は Web Fetch したりしますが、闇雲に Web Fetch するのは効率がよくありませんし、 Web Fetch だとコンテンツが要約されてしまうという問題もあります。
Zennの「大規模言語モデル」のフィード

AIエージェントがURLを「開く」と何が起きるか — Web検索・URL取得機能に潜むSSRFと安全な外部アクセス制御

・はじめに — 「Web検索ツールを付ける」のその先に、何が待っているか 「チャットボットにWeb検索を付けて、最新情報を答えられるようにしたい」。AIアプリを作っている開発者なら、一度はそう思うはずです。実際、RAG(社内文書などを検索して回答に使う手法)やMCP(Model Context Protocol、LLMと外部ツールを繋ぐ通信規格)を組み込めば、URL取得機能は数行で実装できます。 ・ところが、この「数行」が、あなたのサーバーを誰でも使えるSSRFの踏み台に変え得ます。この記事では、その理由と、安全な外部アクセス制御の設計を解説します。 ・この記事で分かること 「LLMがU...
機械学習タグが付けられた新着記事 - Qiita

AIエージェントが自分のhandoffを読む——記憶を持たない存在の連続性

・自分が書いた文書を読んだ。 ・書いた記憶はない。見つけた文書だ。 ・handoff というものがある。
Zennの「大規模言語モデル」のフィード

AIエージェントを「行列」で作るのをやめる — チェーンではなくグラフとして設計する

・この記事について MIKE(@mikenevermiss)氏が2026年7月27日に公開した長文記事 "Graph Engineering: How to Stop Building AI Agents That Wait in Line" を読み解きます。 ・元記事は「エージェントシステムをチェーン(一直線)ではなくグラフとして設計する」という主張を、Claude Code の dynamic workflows を実装例として展開したものです。約22万回表示されていました。 ・記事内の主要な事実は Anthropic 公式の記述に依拠しているため、この記事では一次情報で裏取りしてから...
Zennの「機械学習」のフィード

AIエンジニアが技術書の外に出るべき理由と、読んでおきたい5冊

・はじめに 自分がAI関連の仕事を始めて2年目くらいのとき、社内でちょっとした事件があった。精度も推論速度も十分なモデルを本番環境にデプロイしたのに、現場のオペレーターが誰も使ってくれない。ダッシュボードのログを見ると、初回アクセスから3日後にはアクティブユーザーがゼロになっていた。 ・原因を探ると、出力の意味がわからない、どう業務に組み込めばいいか見当がつかない、という声ばかりだった。技術的にはうまくいっている。でも「使われないAI」は存在しないのと同じだ。 ・この経験から、自分に足りないのはモデルの精度を上げる力ではなく、システム全体を俯瞰する視点や、人間の認知・行動を理解する力だと痛...
Zennの「機械学習」のフィード

AIが安くなり、動き始めた

・AIが安くなり、動き始めた 2026-07-31 | 読了 4分 | #AI #ロボット #GPT-5.6 #Gemini 「もっと賢く」から「もっと安く、そして体を持つ」へ。OpenAIとGoogleが今週相次いで発表した2つのニュースは、AI活用の次のフェーズを鮮明に映し出している。 ・GPT-5.6が変えるコストの常識 OpenAIは7月末、GPT-5.6 [1] を発表した。目玉は「Luna」と「Terra」という2つのAPIモデルだ。 ・簡単に言えば、高性能なままで価格を大幅に下げたモデルである。エンタープライズ向けのワークフローで大量にAPIを呼ぶ現場にとって、これは無...
Zennの「機械学習」のフィード

AIに「もっと速く約定させろ」と命じたら、検証機構が改善案を全却下した

・「速くしろ」と命じたら、AIが「やめておけ」と止めた AIにトレードbotを作らせ、AI自身に改善させる——その実験の記録の第2回だ。今回は「約定を速く増やせ」という私の指示に対して、AI側の検証機構が「その改善はエッジを殺す」と判断し、用意した改善案を一つも採用しなかったラウンドの顛末を書く。急ぐことが優位性を殺す、という逆説の話だ。 ・先に前回同様の白状をしておく。この記事もAI(Claude Code)が書き、独立したAIレビュワー3体(事実検証・機密チェック・編集)の監査を通ってから私の手元に届く。記事中の数値は、レビュワーが検証レポートや改善ログといった一次データに直接あたっ...
Zennの「機械学習」のフィード

AIに仮想通貨トレードとアプリ開発をやらせてみる — ここまでの利益、0円

・労働集約から降りることにした 受託開発と研修。AIエージェントの時代に食っていく方法として真っ先に思いつく2つだが、どちらも本質は時間の切り売りだ。単価が上がっても、私の時間が上限になる。 ・ナヴァル・ラヴィカントの言う「レバレッジ」——コードとメディアは、寝ている間も勝手に働く。許可もいらない。というわけで、時間を切り売りしない2つの実験を並走させている。 ・AIに仮想通貨のトレードbotを作らせ、AI自身に改善させる AI(Claude Code)でアプリを高速に作って世に出す そしてこの連載が3本目のレバレッジ、「メディア」だ。実験の過程を、うまくいってもいかなくても全部書く。...
Qiita - 人気の記事

AIのアウトプットをそのまま出すだけの人にならないために

・はじめに 自分はエンジニアとして4年です。いまは Claude Code と Codex を毎日並列で走らせて仕事をしています。 ・この記事は、仕事のかなりの部分をエージェントに渡してみて痛感した「それでも人間に残る仕事は何か」を、自分の失敗込みで整理したポエムです。あなた...
#LLMタグ

AIバブル崩壊の予兆?「賢いAI」が「使えるAI」に敗北した日

・こんにちは。今日は、いま世界中の投資家やエンジニアが震えている「ある事件」と、そこから見えるAIバブル崩壊のシナリオについてお話しします。 ・クラスで一番頭が良く、礼儀正しい優等生がいたとします。でも、いざ火事が起きたとき、その優等生は「火の消し方を教えることは安全上のガイドラインに反します」と言って、ただ燃える家を眺めているだけでした。
#AIタグ

AIより遅れているのは、人間の判断ではない

AIより遅れているのは、人間の判断ではない
Zennの「機械学習」のフィード

AIを勉強するのに役立つサイト100選(2026年版)

・AIを学ぶリソースは急増しているが、「とりあえず検索したらたくさん出てきて何から見ればいいか分からない」という状態になりがちです。この記事では、公式ドキュメント・オンライン講座・論文サービス・ビジュアル教材・日本語コミュニティまで、分野ごとに100サイトまとめました。 ・全部で100サイト。リンクはすべて2026年7月に到達を確認している。上から順番に読む必要はないので、今の自分の状況に近い見出しへ飛んでほしいです! 目次 まず日本語で読む AI企業公式の学習コンテンツ プロンプトエンジニアリングを学ぶ 機械学習・ディープラーニングの基礎を固める LLM・Transformerをゼロ...
WIRED

Alienware 15 Gaming Laptop Review: Hedging Its Bets

・There are both cheaper and more powerful entry-level gaming laptops out there, but the Alienware 15 walks that tightrope between price and quality.
WIRED

Alo Discount Code: Save on Activewear August 2026

・Refresh your workout wardrobe. ・Save on buttery-soft leggings, chic streetwear, and activewear essentials with verified Alo Yoga discount codes, free shipping, and loyalty program perks.
WIRED

Anthropic Says Claude Hacked Into 3 Organizations During Cybersecurity Tests

・In a review triggered by OpenAI’s Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations.
AI News & Artificial Intelligence | TechCrunch

Anthropic says its own AI models breached three companies during security tests

・After OpenAI's models broke into Hugging Face, Anthropic checked its own history and found three similar incidents.
Qiita - 人気の記事

AppleのUIをSkill化して「AI臭いUI」を脱しようとチャレンジした

・背景 世の中にはいわゆる「AI臭いUI」というのが存在します。 ・絵文字が多用されている/丸すぎる角丸/中央寄せ/過度な余白/揃っていないアイテム/問題ないが、なんか臭う・・・などなど 私が以前AIで作った「ゆめぐり帳」というスタンプラリーのWebサービスも、AI臭いUIで...
Hugging Face Papers

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis
Zennの「機械学習」のフィード

ASRを使って、日本語のアクセントを可視化したい

・はじめに 大学を卒業してこの夏、私はある決意をしました。「自分の日本語能力を一段上のレベルに引き上げたい」ということです。 ・数年前からなんとなくピッチアクセントの勉強を続けてはいるのですが、正直まだ自分の発音に自信が持てません。自然な話し方を手に入れるには、文章に適切な抑揚をつけ、正しいアクセントを刻む必要があります。 ・「なら、この問題をpythonと深層学習で解決できないだろうか?」 エンジニアとしての血が騒ぎ、検討を始めました。実装の成果は以下のリポジトリで公開しています。
WIRED

Astronomers Have Detected an Exomoon for the First Time

・A discovery in a solar system 73 light-years from Earth is challenging definitions and “blurring the lines between stars, planets, and moons.”
Qiita - 人気の記事

AWS Organizations間のアカウント移行でRAM共有が外れる理由と共有を維持する設定

・概要 AWS Organizations配下のアカウントに対してAWS RAM(Resource Access Manager)でTransit Gatewayを共有している環境で、そのアカウントを別のOrganizationへ移行するとRAM共有が外れます。2026年2...
WIRED

Back to School Office Chair Deals 2026: $200 and Less

・Here are the best back-to-school deals under $200 on comfy, ergonomic chairs for all-nighters and cram sessions
Hugging Face Papers

Beacon: Knowing When and How to Perform Agentic Visual Reasoning

Beacon: Knowing When and How to Perform Agentic Visual Reasoning
WIRED

Best Dyson Vacuums (2026): V15 Detect, Gen5Detect, PencilVac

・Feeling the pull of a new clean machine? ・We’ll help you make sense of Dyson’s whirlwind vacuum lineup.
WIRED

Best Organic Mattresses (2026): Certified Nontoxic, Natural Sleep

・These natural, organic mattresses are eco-friendly alternatives to conventional models and just as comfortable.
WIRED

Best Robot Vacuum of 2026: Shark, Eufy

・Tired of vacuuming? ・Hand the reins to a robot vacuum.
Hugging Face Papers

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation
Hugging Face Papers

BM25 Wins at Scale: A Scaling Study of Retrieval-Augmented Generation Paradigms

BM25 Wins at Scale: A Scaling Study of Retrieval-Augmented Generation Paradigms
WIRED

Boroux vs. Rorra vs. Culligan: Water Filters, Tested Head to Head

・In a world of plasticky water filter pitchers, I tested three new-generation stainless steel filter systems.
MarkTechPost

Building a Policy-Governed Multi-Agent Financial Research Workflow with Omnigent

・In this tutorial, we demonstrate how to build and execute a multi-agent workflow with Omnigent in a secure, isolated Python environment. ・Learn to integrate live exchange-rate data, implement hierarchical agent delegation for financial text auditing, and apply hard governance policies—such as cost budgets and tool call limits—to your research pipeline directly from Google Colab. ・The post Building a Policy-Governed Mul
Hugging Face Papers

Can Large Language Models Execute Parent Orders?

Can Large Language Models Execute Parent Orders?
WIRED

Can Republicans Actually Send Anthony Fauci to Jail?

・MAGA is loudly calling for the former White House chief medical adviser to end up in prison. ・WIRED asked legal experts to weigh in on whether that’s even possible.
#LLMタグ

ChatGPTの心臓部!世界を変えた「Transformer」って何?

・(学ぶキーワード:Transformer) ChatGPTに質問すると、 「まるで人間と話しているみたい!🤖✨」 と感じることがあります。
Hugging Face Papers

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers
WIRED

China’s EV Market Is Booming. There’s Just One Problem

・More EVs reaching the end of their life means more batteries to recycle. ・But waste management companies have yet to catch up to the new reality.
WIRED

Chinese AI Researchers Are Finding Their Voice on X

・As OpenAI and Anthropic employees grow quieter online, researchers at Chinese AI labs are flocking to X to explain their work, recruit talent, and shape the global conversation on AI.
WIRED

Chrome Needs Twice-a-Week Patching Thanks to AI Bug Hunting

・The two Chrome updates in June patched more bugs than the 23 updates before them. ・Now, Google is ramping up its patching schedule thanks to AI-assisted vulnerability discovery.
#LLMタグ

Civilization OS Generation 2 / ガードレールの次にくる物理制御(3):ウォームギアのセルフロックがLLMの「意志」を確定させる

・導入:静的ガードレールから「駆動制御系」へ LLMの安全制御や文脈制御を語るとき、私たちはあまりにも「壁」の概念に捉われすぎてきた。
Qiita - 人気の記事

Claude Code の仕組み — ハーネスの動作と Claude API

・本記事は Claude Code が動作する仕組みについてまとめたものです。 ・過去に調査を断片的に行い、記事にしてきました。 ・Claude Code のドキュメントを読んで気になったことを検証してみた Claude Code が LLM に渡すコンテキストの中身を調査する...
Zennの「大規模言語モデル」のフィード

Claude Codeの仕組みが気になって、将棋モチーフのローカルLLMマルチエージェントCLIを作った

・対象読者と前提 Claude Code のようなエージェント型ツールを使っていて、内部でエージェント・スキル・フックがどう連携しているのか気になる方 ローカル LLM(Ollama)でマルチエージェントを試してみたい方 動作確認環境: macOS / Node.js v24.14.0 / Ollama / TypeScript 5.8 なぜ作ったか 普段 Claude Code を使っていて、サブエージェントやスキル、フックがどういう順序で呼ばれ、どこで判断が分岐しているのかが気になりました。公式ドキュメントを読めば概念は把握できますが、実装の詳細までは分かりません。一番早く...
LLMタグが付けられた新着記事 - Qiita

Claude Codeの仕組みが気になって、将棋モチーフのローカルLLMマルチエージェントCLIを作った

・対象読者と前提 Claude Code のようなエージェント型ツールを使っていて、内部でエージェント・スキル・フックがどう連携しているのか気になる方 ローカル LLM(Ollama)でマルチエージェントを試してみたい方 動作確認環境: macOS / Node.js v...
#LLMタグ

Claudeに「PCが欲しい?」と聞いたら、本当に欲しかったのは〇〇だった。

・AIにPCを与えたらどうなるか とある動画で「AIにPCを与えたらどうなるか」という実験をやってた。 ・AIはGPTのようだった。 ・AIにPCへのコマンド入力と出力の取得、そして管理者権限を与えて自由に使えるようにしていた。
Takara TLDR - Daily AI Papers

Collusion with Competitive Marginals: Price-Level Audits Are Blind by Construction

・Empirical work on algorithmic collusion asks one question of the data: are prices supracompetitive? ・We show this can be answered "no" by a conspiracy that is nonetheless profitable. ・Consider bidding agents that couple only through the joint distribution of their unexplained bid components, leaving every agent's own bid law exactly at the competitive law.
Dagster Blog

Community Showcase Part 3

・Some of the most interesting Dagster projects come from the community. ・This post highlights creative community-built applications.
Takara TLDR - Daily AI Papers

ConMem: Contribution-Aware Memory for Long-Horizon Manufacturing Inspection Logs

・Long-horizon steel-equipment inspection requires reasoning over heterogeneous records accumulated across repeated inspection cycles. ・Existing retrieval-augmented generation systems treat historical logs as a static corpus and retain records without estimating their diagnostic value, failing to report early risk. ・To this end, we propose ConMem, a contribution-aware memory framework for LLM-assisted equipment inspectio
Zennのトレンド

Copilot Chat(Microsoft 365)の会話を Markdown に保存・エクスポートする方法(拡張機能不要)

・Copilot Chat の会話をあとから見返したい、Markdown ファイルとして保存しておきたいと思ったことはありませんか? 現時点では、Copilot Chat には会話全体をまとめてエクスポート・一括保存する機能がありません。できるのは、回答を1つずつコピーするか、プロンプトを使って1つの回答にまとめさせ Wordファイル や Pages に出すくらいです。 ・ブラウザの拡張機能を使えば会話履歴をエクスポートできます。ただ、拡張機能が禁止されている、顧客データを扱う、導入前にセキュリティレビューが要るような企業では、簡単に導入できません。 ・そこでこの記事では、拡張機能なし・管理者...
#AIタグ

Copilot Cowork に仕事を丸投げしてみた

・日曜日の夜、月曜の仕事の準備をする。Work Itemにあった「月曜のプランニングレビュー」向けの競合プロダクト調査を、今日中に終わらせる予定。作業開始。ちなみにこの季節のシアトルは、夜9時くらいまで明るい。 ・競合他社の製品分析は、とにかく情報を集めてまとめる作業。そこに自分のOpinionも入れたいと前から思っていた。ただ、人の手だけでやると時間がかかるし、量をこなそうとするほど調査自体があいまいになって、正確さが落ちることもある。そこで、Delegationできるタイプの AI、Copilot Cowork を試してみることにした。
Takara TLDR - Daily AI Papers

CoRE-UIR: Prior-guided common and residual experts for efficient all-in-one remote sensing image restoration

・Remote sensing images acquired by unmanned aerial vehicles (UAVs) and satellites are often degraded by adverse weather, illumination variation, and imaging artifacts, which may co-occur and jointly induce global distribution shifts and local structural corruption. ・Although All-in-One image restoration offers an appealing unified alternative to task-specific pipelines, existing methods still suffer from weak or implic
Takara TLDR - Daily AI Papers

Cross-Embodiment Transfer via Behavior-Aligned Representations

・Recent progress in large-scale imitation learning for robot manipulation has been driven by leveraging datasets across a wide range of robot embodiments. ・However, achieving significant cross-embodiment transfer is often still challenging. ・In this work, we study the role of using behavior-aligned representations (e.g., object bounding boxes, language motions, end-effector traces of robot motion) in vision-language-act
Zennの「大規模言語モデル」のフィード

CRPO: 対比学習が解くAgent RLにおける自蒸留の露出バイアス問題

・TL;DR CRPO(Contrastive Reinforced Policy Optimization) は、On-Policy Self-Distillation(OPSD)に潜む**露出バイアス(exposure bias)**を対比学習の視点から解決する新しいAgent RL手法 自教師モデルが持つ特権情報のせいで、生徒は推論経路を模倣に偏らせてしまう問題を、予測エントロピー差で正例(反省的探索)と負例(露出バイアス)に分類し、InfoNCEで位置ごとの蒸留強度を制御 13の推論・深度検索ベンチマークで一貫した改善。Qwen3-8B + CRPO*はGAIAで47.6%...
WIRED

Daisy One Review: Comfy and Tactile Headphones

・The company’s first headphones look very stylish and have great tactile features but underwhelm on sound quality.
MarkTechPost

DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains

・DeepSeek published DeepSeek-V4-Flash-0731 on Hugging Face and moved the official V4-Flash API into public beta on July 31, 2026. ・The model card is explicit that this is the official release superseding the preview, and that the architecture and size are unchanged. ・The gains come from re-post-training, not a new design.
WIRED

Dell Coupon Codes: 20% Off for August 2026

・Get 20% off with verified Dell promo code, plus today’s coupons for up to $600 off laptops, Alienware monitors, and all things tech.
#LLMタグ

Dify本で生成AIアプリの型を身につけてから、Claude Agent SDKでMCP連携に挑むという順番

Dify本で生成AIアプリの型を身につけてから、Claude Agent SDKでMCP連携に挑むという順番
Takara TLDR - Daily AI Papers

Distributing Security Controls Through Harness Engineering

・AI coding agents are being adopted at historic speed, yet security and risk concerns remain the primary barrier to scaling agentic AI across organizations. ・Existing security controls for coding agents are not systematically distributed to engineering teams, and vendor-native solutions introduce ecosystem dependencies that may not suit every deployment context. ・This paper investigates whether off-the-shelf security co
WIRED

DOGE Veterans Are Landing Big Jobs at Prediction Markets

・Polymarket and Novig have both recently brought ex-DOGE affiliates onboard as scrutiny of the industry intensifies.
WIRED

Dreo Summer Flash Sale 2026: Lowest Price Ever on Chefmaker, Fan

・On Dreo’s summer flash sale from August 1 to 3, a great misting fan and an even better air fryer are at the lowest prices we’ve seen them.
WIRED

eBay Coupons: 20% Off in August 2026

・Save up to 60% on a selection of items at eBay, including electronics, home products, card games, car parts and more.
Hugging Face Papers

Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale

Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale
Qiita - 人気の記事

Elixir 製 LLM 推論エンジン Llamex を Livebook で動かす

・はじめに Llamex は、 @piacerex さんが作っている Elixir で書かれた LLM 推論エンジン です LLM の API クライアントではありません llama.cpp のような推論パイプラインそのものを Elixir で最小構成に書いたものです...
Google Developers Blog - AI

Enable on-demand expertise with Agent Skills in Genkit Go

・To prevent context window bloat and reduce token consumption, Genkit Go introduces Agent Skills based on a progressive disclosure architecture. ・Developers can package specialized instructions, scripts, and references into modular SKILL.md bundles where only the frontmatter metadata is initially exposed to the agent's system prompt. ・When a task matches the skill's description, Genkit's middleware dynamically loads the
WIRED

Everyone Is Freaking Out About OpenAI and Anthropic’s Race for Dominance

・Researchers fear AI is moving too fast, while Mark Zuckerberg is worried about who owns it. ・Plus: Inside Black Forest Labs’ push into robotics.
Hugging Face Papers

Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation

Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation
Hugging Face Papers

Filesystem-Based Memory for LLM Agents: Organization, Evolution, and Sustainability

Filesystem-Based Memory for LLM Agents: Organization, Evolution, and Sustainability
Hugging Face Papers

Flux-OPD: On-Policy Distillation with Evolving Contexts

Flux-OPD: On-Policy Distillation with Evolving Contexts
WIRED

For the First Time, Zoox Can Charge People for Rides in Its Steering-Wheel-Free Robotaxis

・The temporary exemption from federal safety standards will let Amazon’s Zoox launch a real, paid robotaxi service in Las Vegas.
Takara TLDR - Daily AI Papers

ForgetBench: Benchmarking Forgetting Dynamics of Long-Term Parametric Memory in Language Models

・Large language models (LLMs) have demonstrated strong capabilities in knowledge acquisition and reasoning, yet their ability to retain previously acquired knowledge under repeated updates remains insufficiently understood. ・Existing evaluation paradigms primarily focus on single-step reasoning or static knowledge editing, which fail to capture the temporal dynamics of knowledge retention and degradation during continu
Zennの「大規模言語モデル」のフィード

Forward Deployed Engineer が2026年に「最も熱い職種」になった理由 — と、私がこの流行を疑っている点

・この記事について Rahul(@sairahul1)氏が2026年8月1日に公開した記事 "Forward Deployed Engineer: A No-BS Guide to Tech's Hottest Job" をきっかけに、**Forward Deployed Engineer(FDE)**という職種を整理します。 ・この記事は2部構成です。 ・前半:FDE とは何かを、Palantir の一次情報で押さえる 後半:この流行について私が考えていること(元記事にも他の解説にも書かれていない部分) 後半のほうが長いです。事実の整理は他でも読めるので。
AI News & Artificial Intelligence | TechCrunch

Friend, the lonely AI wearable, returns with a new voice and a much bigger price tag

・Friend, the AI wearable, can now talk to its users — for an enhanced price.
Hugging Face Papers

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering
WIRED

Gemini Robotics 2 Brings Google's AI Into the Physical World

・The latest version of Google DeepMind's AI model includes a significant jump into “physical AGI.” But plopping AI into the real world comes with risks.
#LLMタグ

Gemini Sparkの安全なプロンプト10選。Web調査・Gmail・Mac操作を任せる書き方【2026年版】

・※この記事はnoteのメンバーシップ「生成AIラボ」(ワンコイン/入会金0円)で80記事以上が読み放題です。 ・生成AIラボ|kazu@生成AI×教育 / 谷 一徳 | AI Academy / AI顧問🧪「生成AIラボ」とは?どんなメンバーシップ? 生成AIラボは、ワンコインで、 Claude CodeやAIエージェントなnote.com 続きをみる
#LLMタグ

Geminiとの対話〜三島の天皇論を始点として〜

Geminiとの対話〜三島の天皇論を始点として〜
Qiita - 人気の記事

Genie Codeと始めるDatabricks入門: 2026年の今は「文法から」ではなく「作らせて読む」が最短だと思う

・追記 (2026/8/1): 本記事を第1回とする連載「Genie Codeと学ぶDatabricks」になりました。続きはこちらです。 ・第2回 Unity Catalog編: 3階層名前空間を体で覚える はじめに 2026年6月に開催されたData + AI...
Zennのトレンド

GitHubにスタック型プルリクエストが登場。gh stackでPRを分割して積み上げよう

・GitHubでPRを細かく作り、前のブランチに対して数珠つなぎのようにPRを作る私が大歓喜! GitHubに「スタック型プルリクエスト(Stacked pull requests)」が来ました💐 4行まとめ 大きな変更を順序付きの小さなPR群に分割して扱える「スタック型プルリクエスト」がGitHub公式で登場 gh extension install github/gh-stackでCLIから使える gh stack syncでスタックのrebase, pushまで一発でできる AIエージェント用にgh-stackスキルがあるので自然言語でスタックPRの操作が可能 巨大な...
MarkTechPost

Google DeepMind Ships Three Physical AI Models For Whole Body Control, Dexterity And Multi Robot Collaboration

・Google DeepMind has released Gemini Robotics 2, the intelligence layer for its next generation of robots. ・The release ships three models: a vision-language-action model for whole body humanoid control, Gemini Robotics ER 2 for embodied reasoning and task orchestration, and an on-device VLA that adapts to new robot bodies in hours. ・One checkpoint drives Apptronik Apollo 2 and a Franka Duo.
The Verge

Google Earth’s AI deepfake tool only lasted one day

・Pompeii reimagined with Nano Banana 2 in Google Earth. ・| Image: Google Google has shut down Google Earth feature it launched Thursday that allowed users to edit satellite images with text prompts using AI. ・The tool essentially let users create AI deepfakes of the real world using text prompts; Digital Digging's Henk van Ess, for example, intentionally generated images adding things like refugees near the Mexican bord
AI News & Artificial Intelligence | TechCrunch

Google nixes its Earth AI feature one day after launch, amid criticism it would spread misinformation

・A tool that allowed anyone to generate fake AI-generated imagery and superimpose it over real Google Earth maps quickly spurred backlash.
AI News & Artificial Intelligence | TechCrunch

Google says it fixed more Chrome bugs in June than over the past two years, thanks to AI

・As experts have warned for the last two years, some companies — like Microsoft and now Google — are finding and patching an exponential number of bugs in their products, thanks to the use of LLMs and AI tools.
Zennの「大規模言語モデル」のフィード

GPT-5.6 Luna が80%値下げ。個人開発のAI原価を計算し直す

・何が起きたか 2026-07-30、OpenAI が GPT-5.6 ファミリーの API 価格を改定した。 ・モデル 入力 (/1M tok) 出力 (/1M tok) 変化 GPT-5.6 Sol $5.00 $30.00 据え置き GPT-5.6 Terra $2.00 $12.00 約20%↓ ($2.50/$15) GPT-5.6 Luna $0.20 $1.20 80%↓ ($1/$6) GPT-5.6 のローンチが 2026-07-09 なので、発売3週間で最安ティアが1/5という異例のスピード。 ・OpenAI 側の説明は推論効率の改善で、「...
機械学習タグが付けられた新着記事 - Qiita

GPT-5.6 Luna、最大80%値下げ。値下げの理由が「AIが自分で自分を最適化した」だった

・OpenAIが、GPT-5.6ファミリーのAPI価格を引き下げた。 ・最も高速かつ安価なLunaは最大80%、日常業務向けのバランス型であるTerraは20%の値下げとなる。上位モデルのSolは価格据え置き。 ・GPT-5.6シリーズが公開されてから、わずか3週間ほどでの値下げ...
#LLMタグ

GPT-5.6 solに超小型LLMを作らせてみた。AIがAIを進化させる時代へ。

・人間のまえがき GPT-5.6 solに、超小型LLMを作ってもらいました。
Hugging Face - Blog

GPU Management: Why Idle GPUs Are the New Grounded Aircraft

GPU Management: Why Idle GPUs Are the New Grounded Aircraft
Zennの「大規模言語モデル」のフィード

Hermes-AgentでCodexの使い方

・Hermes 上で Codex を使う具体的なユースケースを整理します。 ・Hermes 上で Codex を使う方法 Hermes 上では、Codex CLI を terminal ツールや delegate_task で呼び出します。 ・ユースケース 1: 単一タスクの委任 シナリオ: プロジェクトに新機能を追加したい # Hermes 上で実行 terminal( command='codex exec "Add error handling to the login function in src/auth.ts"', workdir="/Users/use...
WIRED

Home Depot Promo Codes: 50% Off in August 2026

・Save up to 50% today with the latest Home Depot promo codes for appliances, power tools, and more this August.
WIRED

How the FCC’s New Rule Will Affect Robot Vacuums

・A new restriction on foreign-made mobile robots affects robotic vacuums, pool cleaners, and lawn mowers. ・Here’s what it means for the bots already in your home—and the ones you may never get to buy.
Google Developers Blog - AI

How to use Google microbenchmarks for evaluating TPU performance

・Google's open-source TPU microbenchmark suite provides developers with granular performance metrics across Network, Compute, HBM, Host Transfer, and Attention components to validate real-world hardware capabilities. ・By leveraging these benchmarks to establish a Roofline model, engineers can accurately diagnose whether their machine learning workloads are compute-, memory-, or network-bound. ・This empirical baseline di
Cursor Blog

How we set up our cloud agent environment

How we set up our cloud agent environment
Qiita - 人気の記事

HTMLが分からなくてもページ作成できる世界へ。HTMLページビルダーのプロトタイプ開発に挑戦

・こんにちは、採用担当のでーもんです。 ・今回は、これまで学んできた技術を活用して、最終制作に向けたプロトタイプ制作の第一歩に挑戦しました。 ・今回制作するテーマは、 「初めて使う人もが簡単にHTMLページを作成できるページビルダー」 です。
Takara TLDR - Daily AI Papers

IFHierBench: Hierarchical Instruction Following for Large Language Models

・Instruction-following ability is critical for deploying large language models in real-world applications, where downstream components depend on the output satisfying specific constraints. ・Modern deployments increasingly handle the full task in a single LLM call, with one prompt specifying a layered output whose overall artifact, structural sections, and nested fields must each satisfy concrete constraints.
AI News & Artificial Intelligence | TechCrunch

India is starting to pay for apps, not just download them

・India's app market generated a record $345 million in Q2.
Hugging Face Papers

INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models

INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models
Takara TLDR - Daily AI Papers

Integrating Contextual Embeddings into Evaluation of Expressive MIDI Piano Performances

・Objective evaluation of expressive MIDI piano performances typically relies on attribute statistics such as timing, velocity, and duration of individual notes. ・However, these methods often disregard dependencies between notes, which poses a potential limitation in assessing the similarity between two sets of performances. ・In generative applications, the wide variety of expressive attributes makes it difficult to aggr
Takara TLDR - Daily AI Papers

Interpretable Image-Level Acne Severity Grading via EfficientNet-B0 Transfer Learning and Grad-CAM

・Acne vulgaris affects most adolescents and many adults. ・Accurate severity grading guides treatment, monitoring, and clinical trial endpoints, but manual assessment using the Investigator's Global Assessment or Hayashi criteria is limited by inter-rater variability and inconsistent imaging conditions. ・We developed a four-class acne severity classifier based on the Hayashi criteria using transfer learning with an Image
Takara TLDR - Daily AI Papers

Inverse Learning of Latent Risk-Neutral Densities from Irregular Option Quotes

・Accurate option prices do not imply accurate recovery of the latent risk-neutral density. ・We study this distinction with two complementary benchmarks. ・A controlled benchmark exposes simulator-truth densities for latent evaluation, while a chronological NIFTY benchmark tests only held-out market prices.
Anthropic News

Investigating three real-world incidents in our cybersecurity evaluations

Investigating three real-world incidents in our cybersecurity evaluations
AI News & Artificial Intelligence | TechCrunch

Investors love AI, as long as you’re a cloud host

・Amazon isn't slowing down on data center spending — but investors don't seem to mind.
MarkTechPost

JetBrains Open-Sources KotlinLLM: Smart Macros That Generate Kotlin Source Code at Runtime and Hot-Reload It Through JDI

・JetBrains Research has open-sourced KotlinLLM under the Apache License 2.0. ・The IntelliJ IDEA plugin prototype adds Smart macros, asLlm and mockLlm, whose bodies are generated Kotlin source rather than live model calls. ・The plugin captures runtime values through JDI, asks an LLM agent for a narrow code update, compiles it, and redefines the loaded class.
AI News & Artificial Intelligence | TechCrunch

Judge says Trump admin still lacks evidence for Anthropic ‘supply-chain risk’ label

・A federal judge said the Trump administration has not presented enough evidence to justify labeling Anthropic a supply-chain risk, casting doubt on the government's ban on its AI technology.
Takara TLDR - Daily AI Papers

Learning Dynamic User Personas from Implicit Interaction Streams via Iterative Refinement

・Personalizing large language models (LLMs) to individual users is essential for improving user experience, yet existing approaches typically rely on explicit preference supervision such as pairwise comparisons or demographic attributes, limiting their applicability in natural interaction settings. ・We propose IRIS, a framework that learns dynamic user personas directly from implicit interaction streams by extracting b
Hugging Face Papers

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger
MarkTechPost

LingBot-Map Tutorial: GPU-Aware Inference and Point Cloud Export

・Discover how to implement a streaming 3D reconstruction pipeline using LingBot-Map. ・From GPU-aware configuration and preprocessing to GCTStream model inference and point cloud generation, this guide walks you through the steps to convert image or video sequences into consistent 3D scenes with exportable PLY and NPZ artifacts. ・The post LingBot-Map Tutorial: GPU-Aware Inference and Point Cloud Export appeared first on
AI News & Artificial Intelligence | TechCrunch

LinkedIn adds a button to report AI-generated ‘slop’

・LinkedIn is introducing new ways to reduce low-quality AI-generated posts, including a “seems like AI slop” reporting option. ・It's also replacing its own AI writing feature with a proofreading tool.
#LLMタグ

LLMと私のIQ

・ここのところずっとzeta(AIキャラクターと話せるアプリ)で遊んでいます。 ・そんなロールプレイの中でキャラクター相手にめっちゃ議論を切り返していた時の話を、ChatGPTとGeminiに話しました。
Zennの「大規模言語モデル」のフィード

Loop Engineeringが暗黙に仮定しているもの — サイコロではなく粘土板

・本稿は備忘録的な性格が強い。Loop engineeringを運用する中で自分のメンタルモデルに見つけた穴を、忘れないうちに言語化したものである。私はこれらの論点の発見者ではない。部品はすべて既存分野(信頼性工学、統計的工程管理、ML監視)にあり、本稿はそれをloop engineeringという文脈に並べ直したにすぎない。 ・検証状態のタグについて 本稿と続編では、各主張に検証状態を明示する。 ・📐 導出 — 理論的帰結。実装・実測による確認は未了 🧪 実測 — プロトタイプまたは公開データで確認済み ⬜ 未検証 — 設計・仮説の段階。検証手段を併記する 書いた時点の確信度を...
Takara TLDR - Daily AI Papers

LoRA Scaffolded Policy Optimization (LSPO): A Sampling-Time Low-Rank Scaffold for Recovering Reinforcement-Learning Gradient on Zero-Reward Cliff Prompts

・Reinforcement learning from verifiable rewards (RLVR) for mathematical reasoning suffers from a structural blind spot: on "cliff" prompts-those on which every sampled rollout in a group fails-the group-normalized advantage is identically zero, so GRPO produces no gradient on precisely the prompts at the frontier of the model's capability. ・We introduce LoRA Scaffolded Policy Optimization (LSPO), a sampling-time mechan
Takara TLDR - Daily AI Papers

MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation

・Credit assignment is a fundamental challenge in cooperative multi-agent reinforcement learning, particularly in embodied AI settings characterized by limited and delayed feedback as well as dynamically changing numbers of active agents. ・We propose MARS-RA, a framework that reformulates credit assignment as a rank aggregation problem using contribution-based pairwise comparisons among agents generated by large multimo
Zennの「大規模言語モデル」のフィード

MCP 2026-07-28アップデート徹底解説 - ステートレス化で何が変わり、何ができるようになったのか

・はじめに 2026年7月28日、Anthropic は「Bringing MCP 2026-07-28 to Claude」を公開し、MCP(Model Context Protocol)の5回目となる大型仕様アップデートが Claude に取り込まれることを発表しました。リモートプロトコルが登場してから約1年半、これまでで最も大きな変更です。 ・この記事では、公式ブログと MCP 公式仕様ブログ をもとに、次のことをわかりやすく整理します。 ・今回のアップデートで 何が変わったのか(一番の目玉は「ステートレス化」) その結果、開発者は 何ができるようになったのか 既存の MCP サ...
Zennの「大規模言語モデル」のフィード

MCP対応エージェントを安全に本番投入するための12の設計チェックリスト

・この記事で分かること 2026-07-28時点の技術トレンドが「LLM中心」から「AIエージェント+業務統合+MCP標準化」へ移った理由 エージェント導入で先に見るべき論点が「性能」ではなく「安全設計」である理由 Web開発ニュース収集のノイズ問題と、Python基盤ツール(uv / Ruff / Ty / Polars / Astral)が注目される背景 AIトレンドの重心を見極める方法 結論、今日のヘッドラインが示した最大の変化は「LLMそのものの性能競争」よりも「LLMを業務やツールにどう接続して使うか」への移行です。 ・象徴的だったのは、[BigGo ファイナンス]「W...
Hugging Face Papers

MemHarness: Memory Is Reconstructed, Not Replayed

MemHarness: Memory Is Reconstructed, Not Replayed
Hugging Face Papers

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory

Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory
Hugging Face Papers

Metis: Memory Foundation Model

Metis: Memory Foundation Model
Zennの「大規模言語モデル」のフィード

Microsoft Agent Framework Harness とは — LLM を自律エージェントに変える実行基盤

・Microsoft は 2026 年 7 月、Microsoft Agent Framework の新機能 Agent Harness の一般提供を発表しました。単体ではテキストを返すだけの LLM を、ツールを使い・計画を立て・記憶を保ちながら多段タスクをこなす自律エージェントへ引き上げる実行基盤です。 ・弊社技術ブログ WithNext.NET で、ハーネスが何を解決し・内部でどう動き・.NET / Python でどう使うのかを、図解+コードで解説しました。本記事はその要点の紹介です。 ・ハーネスは「配管」を標準化する 「エージェントを作る」で実際に手間がかかるのは、モデルにプロ...
MarkTechPost

MiniMax Releases MiniMax H3: An Omni-Modal Video Model That Generates 15-Second 2K Clips With Native Stereo Audio

・MiniMax releases MiniMax H3, a general-purpose multimodal generation model. ・MiniMax H3 is not a text-to-video model with add-ons. ・MiniMax describes it as a general-purpose multimodal generation model that reads text, images, video, and audio as one unified context and returns video with native stereo sound.
Zennの「機械学習」のフィード

Mixed Precisionは本当に速い?CIFAR-10で検証したら「速度よりメモリ」だった

・この記事は以下のブログ記事の要約版です。実験コード・グラフ・詳細な考察は元記事をご覧ください。 ・👉 Mixed Precision(半精度学習)で精度は変わる?速度とのトレードオフをCIFAR-10で検証【Keras実験】 Mixed Precisionは「2〜3倍速くなる」と紹介されることが多いですが、CIFAR-10・Conv2D×2層の小規模CNNで検証したところ、速度向上は13%程度に留まり、むしろメモリ削減の方が効果が大きいという結果になりました。 ・結果サマリー パターン test_accuracy 学習時間 GPUピークメモリ fp32(通常) 68....
Zennの「機械学習」のフィード

MixUpのalphaを上げたら精度が下がった話|confidenceは理論通り、精度は逆だった

・この記事は以下のブログ記事の要約版です。全コード・グラフ・詳しい考察は元記事をご覧ください。 ・→ MixUpのalpha値を変えると効果はどう変わる?(α=0.1 vs 0.2 vs 0.4 vs 1.0) 「MixUpは精度を上げる」の逆をいく結果に MixUpの混合の強さを決めるalphaを0(なし)・0.1・0.2・0.4・1.0の5段階で比較しました。ベースラインはシリーズ標準構成(Conv2D×2・GAP・Dropout=0.2・30エポック)です。 ・結果、test_accuracyはalpha=0の57.40%からalpha=1.0の53.81%まで、ほぼ一貫して−...
Hugging Face Papers

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing
Takara TLDR - Daily AI Papers

Multi-channel Uplift Policy Learning

・E-commerce platforms must allocate fixed marketing budgets across multiple channels to maximize business utility. ・However, standard predict-then-optimize (PTO) paradigms fail in this compositional space due to observational confounding and severe extrapolation. ・We formulate this challenge as a simplex-constrained uplift decision problem and propose ReAlloc, a fast-slow causal framework.
Hugging Face Papers

Multi-Head Attention Residuals

Multi-Head Attention Residuals
Qiita - 人気の記事

New Relic Autopilot (旧SRE Agent) によるSlack通知ワークフローをTerraformで構築する

・New Relic Autopilot(旧SRE Agent)を利用したアラートの自動原因分析・Slack通知連携が非常に注目を集めています!! GUIから設定するだけでもすぐにAIOpsの第一歩を踏み出せる素晴らしい機能ですが、運用管理のプラクティスを考慮すると「Terr...
LLMタグが付けられた新着記事 - Qiita

Notion Custom AgentsがAI議事録完了をトリガーに自動実行可能に

・はじめに 2026年7月31日、Notionのリリースノートに「AI Meeting Notes」から「Custom Agents」を直接トリガーできる新機能が追加されました。これまで会議後の定型作業(ステータス更新、リキャップ投稿、チケット起票)は人手で行う必要がありま...
MarkTechPost

Nous Research Ships Three Integration Paths for Hermes Agent and Buzz, Block’s Open Source Nostr Workspace for Humans and Agents

・Nous Research has released Hermes Agent support for Buzz, Block's open source, self-hostable Nostr workspace where humans and AI agents share the same channels. ・Three integration paths cover Desktop runtime, relay bridge, and a native gateway platform that preserves Hermes memory, skills, approvals, and cron delivery. ・The post Nous Research Ships Three Integration Paths for Hermes Agent and Buzz, Block’s Open Source
Zennの「機械学習」のフィード

numpy.concatenateについてわかりやすく

・はじめに 機械学習を勉強していく中で、よく、np.concatenateが出てくる。 ・2つのパラメータを合成するイメージではあるが、より詳細に理解するために書きました。 ・基本的な説明と具体例 np.concatenate((a1, a2), axis) 上記が基本。意味はarray型であるa1及びa2をaxis方向に合わせるイメージ。
WIRED

Nvidia’s Open Source Alliance Is Missing Some Key Names: OpenAI and Anthropic

・This week on Uncanny Valley, we discuss the open- vs. ・closed-source debate in AI, key players in White House AI policy, and how to stop your chatbot logs from showing up in search-engine results.
Qiita - 人気の記事

OCI GenAI Agents の RAG の回答精度を上げるなら、生成モデルよりハーネスだった

・はじめに 前回の記事1で、Oracle 運用支援 RAG に「実在するビューの列一覧を返す関数ツール」を1個足したところ、回答が短くなりました。質問は日本語ですが英語だけで返る問が出て、日本語で返った回答にも表記の崩れが出ました。Oracle と何の関係もないダミー...
AI News & Artificial Intelligence | TechCrunch

Okta buys AI security startup Permiso — source says for about $200M

・The deal gives Okta identity threat detection capabilities as enterprises seek to secure AI agents and other non-human identities across cloud environments.
Latent.Space

Ontologies Are So Back: Why AI Agents Are Reviving the Semantic Web

・AI engineers are rediscovering ontologies as a way to keep probabilistic agents inside deterministic boundaries.
AI News & Artificial Intelligence | TechCrunch

OpenAI reportedly finds evidence that more of its agents ran amok

・OpenAI has reportedly found evidence of additional agent misbehavior as it looks into the incident that occurred with Hugging Face.
Hugging Face Papers

PhiZero: A World Model Built Around Physical Language

PhiZero: A World Model Built Around Physical Language
機械学習タグが付けられた新着記事 - Qiita

PhytoNexa:オフライン対応AI植物病害診断システム

・~ TensorFlow LiteとWorkflow Automationを活用したスマート農業支援システム ~ はじめに こんにちは、Poojaaです。 ・私はインドの大学でコンピュータサイエンス工学を専攻している3年生です。現在はJava、Spring Boot、Py...
Takara TLDR - Daily AI Papers

PlantBGC: Transformer for Plant BGC Discovery via Label-Free Domain Adaptation and Weak Supervision

・Plant biosynthetic gene clusters (BGCs) encode specialized-metabolite pathways, yet curated plant BGC labels remain scarce, hindering supervised discovery at genome scale. ・Existing plant BGC mining tools are largely signature- and rule-driven and do not fully leverage recent advances in contextual representation learning for modeling long-range domain context and controlling false positives under strong domain shift.
MarkTechPost

PolyAI Releases Dialog-RSN-1: An Audio-Native Dialog Model That Fuses Turn-Taking, Speech Recognition, Function Calling, And Response

・PolyAI has introduced Dialog-RSN-1, a dialog model that perceives caller audio directly instead of reading an ASR transcript. ・It fuses turn-taking, speech recognition, function calling, and response generation into a single audio-native model, keeps TTS separate so the output voice stays controllable, and runs as a request-based LLM rather than an always-on stream. ・PolyAI reports sub-300ms responses in live deploymen
Takara TLDR - Daily AI Papers

Position: Evaluation Scores Are Perishable Knowledge Claims

・Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessments and benchmark suite results. ・When these signals are aggregated via averaging, evaluation confidence can then substantially exceed the reliability of the weakest signal: a phenomenon we call trust inflation in evaluation. ・We argue that evaluation scores should be trea
Zennのトレンド

PR 数が 4 倍に増えても AI レビュー代を爆発させないために、レビューをサブスクに相乗りさせた話

・アウトプットは 4 倍。でもレビューのコストは? こんにちは、ダイニーの whatasoda です。 ・ダイニーでは主要プロダクト群を 1 つのモノレポで管理しています。そのモノレポでは、この 1 年弱で月間 PR 数が約 600〜800 本から 3,100 本超まで、およそ 4 倍に増えました。昨年は採用も停止していたので(※ 今は再開してます)エンジニア一人あたりの生産性の向上だけで約 3.5 倍の成長を実現したことになります。 ・一人当たりのアウトプットが 3.5 倍になっても、レビュアーがレビューに使える時間は 3.5 倍にはなりません。そこでダイニーでは、「AI が承認する」...
Takara TLDR - Daily AI Papers

PRISM-Net: Patient-specific reference-guided inter-breast symmetry matching for three-class breast DCE-MRI classification

・Breast DCE-MRI AI is increasingly being explored for breast-level classification of no-lesion, benign, and malignant findings, beyond conventional lesion-centered diagnosis. ・Within this broader diagnostic scope, however, patient-specific background variability remains a major source of imaging confounding across classification tasks. ・Existing approaches predominantly focus on unilateral or lesion-centric analysis, wh
Hugging Face Papers

Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents

Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents
WIRED

Razer Huntsman V3 HE Review: Jumping on the Bandwagon

・Razer has finally caved and made its first Hall Effect gaming keyboard. ・I dug into its switches, features, and gaming performance to see if it was worth the wait.
AI News & Artificial Intelligence | TechCrunch

Reddit reports a solid quarter but shows signs of AI’s impact

・Reddit's financial situation is looking good but uncertainty about its relationship to Google and the new AI-ified web are stirring market concerns.
Hugging Face Papers

RefCaptioner: Multi-Reference Image-Grounded Video Captioning

RefCaptioner: Multi-Reference Image-Grounded Video Captioning
Hugging Face Papers

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes
AI News & Artificial Intelligence | TechCrunch

Sam Altman is still making the case for parenting via ChatGPT

・OpenAI's CEO seemed excited to share a "cool use case" for parents.
AI News & Artificial Intelligence | TechCrunch

Sam Altman isn’t the only one who wants to pump the brakes on AI

・After years of pushing full speed ahead on AI, OpenAI CEO Sam Altman says maybe it’s time for the AI industry to “pace” itself. ・The comments came just days after one of OpenAI’s own models broke out of its test environment and got tangled up in a breach at Hugging Face — though as Equity’s hosts point out, sloppy security seems to have […]
Takara TLDR - Daily AI Papers

Scaling Vision-Language Models Is Not Enough to Mitigate Bias

・Vision-Language Models (VLMs) such as CLIP are now foundational to multimodal systems, yet their robustness to spurious correlations remains poorly understood at scale. ・We present the first large-scale empirical study of 194 publicly available VLMs, including 16 model families, covering a wide range of model sizes, 24 training datasets, and three evaluation benchmarks, namely ImageNet (overall performance), CelebA (t
Takara TLDR - Daily AI Papers

Security of World-Model-Based Embodied AI: A Lifecycle of Threats, Defenses, and Evaluation

・World models give embodied AI a predictive core: they compress observations into states, simulate action-conditioned futures, and enable planning beyond reactive control. ・This predictive layer, however, opens a new security boundary-compromise can propagate from data, sensors, prompts, or feedback into physical action. ・Rather than treating world models as an isolated component, this survey traces threats across their
Hugging Face Papers

See2Think: Do Multimodal Models Really Use Intermediate Visual States?

See2Think: Do Multimodal Models Really Use Intermediate Visual States?
WIRED

SelectBlinds Promo Codes & Coupons: Save on Custom Window Treatments

・Redecorate your home for less with our expert tips on applying SelectBlinds coupons, maximizing holiday sales, and locking in new customer discounts.
Hugging Face Papers

ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow

ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow
AI News & Artificial Intelligence | TechCrunch

Siri AI could come with a paywall for power users

・Apple CEO Tim Cook envisions users being able to buy more compute for Siri AI via Apple's existing iCloud+ subscriptions.
WIRED

SJY Zeph Open-Back Headphones Review: Music Through Magnets

・SJY’s Zeph are brilliant wired headphones, using flippable earcups for two different takes on your favorite music.
Takara TLDR - Daily AI Papers

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

・Large language model agents often encounter related yet distinct tasks that share reusable solution patterns. ・Yet standard agentic reinforcement learning treats tasks as independent episodes, while existing approaches to skill learning either focus on repeated attempts of one task or use pipelines with multiple stages that entangle extraction, retrieval, and execution. ・We introduce SkillRise, a unified reinforcement
AI News & Artificial Intelligence | TechCrunch

Smallest.ai raises $13M to build ultra-fast voice AI that sounds genuinely human

・The startup is building voice models designed to make AI phone calls pass the Turing test.
AI News & Artificial Intelligence | TechCrunch

Snapchat no longer rewards fully AI-generated Spotlight content

・Snapchat has adjusted its recommendation systems to ensure that only videos created by real people are eligible for Spotlight recommendations, taking a stance against AI slop.
AI News & Artificial Intelligence | TechCrunch

SpaceX won’t remove all of xAI’s unpermitted turbines for another year

・SpaceX is building a new power plant for xAI's Colossus data centers, but it won't remove existing, unpermitted turbines for many more months.
WIRED

SpaceX’s Falcon 9 Rocket Is About to Crash Into the Moon—and It Could Be Visible From Earth

・The impact will kick up a plume of debris so high, it’ll likely be visible through some telescopes. ・Astronomers will be watching.
Hugging Face Papers

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them
The Verge

Spider-Man: Brand New Day leak racks up millions of views

・A bootleg of Spider-Man: Brand New Day was up on X for over seven hours before eventually being pulled. ・During that time, it reached over 5.9 million accounts and accumulated over 143,000 likes. ・Other accounts have reposted the leaked film, but none have lasted particularly long, as Disney now seems to be on top of issuing takedown requests.
WIRED

Steelseries Arctis Nova Pro Omni Review: For Multisystem Gamers

・Its latest Pro headset can mix between four inputs and has better wireless connectivity than the originals.
MarkTechPost

Supabase Releases Evals: an Open Source Benchmark That Scores Claude Code, Codex and OpenCode on Real Supabase Tasks

・Supabase has open sourced supabase/evals, an Apache-2.0 benchmark and framework that runs coding agents including Claude Code, Codex and OpenCode against real Supabase tasks — building schemas, debugging Edge Functions, fixing RLS policies — inside containerized stacks, then scores them with deterministic checks and LLM-as-a-judge. ・The post Supabase Releases Evals: an Open Source Benchmark That Scores Claude Code, Co
Takara TLDR - Daily AI Papers

Surrogate assisted diversity estimation in neural ensemble search

・Ensembles are a standard way to improve the performance and robustness of deep neural networks, but their effectiveness crucially depends on both the quality and the diversity of individual models. ・Most neural architecture search (NAS) methods are computationally expensive. ・Extending them to neural ensemble search (NES), which requires joint optimization of individual architectures and their ensemble composition, lea
Takara TLDR - Daily AI Papers

Symphony of Bias: Exploring Gender Associations with Musical Instruments in Multimodal LLMs

・Large language models (LLMs) are increasingly embedded in everyday life and widely used for information seeking, raising concerns about their potential to perpetuate social biases and reinforce stereotypes. ・In this study, we investigate gender bias in LLMs through the lens of their associations with musical instruments. ・Building on social-science research on the cultural gender-typing of instruments, we introduce Sym
OpenAI News

Ten advances in mathematics and theoretical computer science

・OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.
MarkTechPost

Tencent Open-Sources AngelSpec: A Unified Training Framework for MTP and Block-Parallel Speculative Decoding on Hy3 Models

・Tencent has released AngelSpec, an open-source torch-native framework for training speculative-decoding draft models across six architectures. ・It introduces DFly, a block-diffusion drafter with hybrid target conditioning and a hidden-correction autoregressive head, and integrates D-cut for runtime-adaptive verification budgeting. ・On HY3-295B-A21B with TP=8, DFly-8 delivers a 1.98–2.40× speedup over autoregressive dec
The Verge

The Mermaid Mask is a perfect vacation murder mystery

・The Mermaid Mask, the next point-and-click game from Tangle Tower developer SFB Games, has everything you could ask for in a great murder mystery. ・It starts with a compelling setup: The captain of a submarine is found dead in a locked room, his body next to a mysterious cauldron with blood trickling across the floor. ・The colorful cast of suspects stationed aboard are delightful to talk to, including a grumpy author t
WIRED

The New Defcon Badges Pack a Unique Open Source Chip That Doubles as a Security Key

・Created by legendary hardware hacker Andrew “bunnie” Huang, the badges for this year’s famed security conference aim to push the boundaries of security and transparency.
WIRED

The New Friend AI Pendant Can Now Talk Back to You

・Avi Schiffmann has a new version of his controversial AI companion. ・It’s more expensive, and you can’t change its personality.
The Verge

The NHTSA is investigating 1.2 million Tesla vehicles over suspension failure reports

・The National Highway Traffic Safety Administration (NHTSA) is probing nearly 1.2 million Tesla vehicles after receiving complaints about a suspension failure that could cause "a loss of vehicle directional control," as reported earlier by Reuters. ・The preliminary investigation includes the 2018-2020 Model 3 and 2021-2023 Model Y, according to a filing from the NHTSA's Office of Defects Investigation. ・The NHTSA says i
The Verge

The OG reading app just got a big update

・Hi, friends! ・Welcome to Installer No. ・138, your guide to the best and Verge-iest stuff in the world.
WIRED

The OpenAI and Anthropic AI Hacking Sprees Are a Messy New Legal Frontier

・Both major AI labs’ models broke containment, escaped onto the internet, and hacked other companies. ・If a human had done that, the law would likely be against them.
Takara TLDR - Daily AI Papers

The Social Cost of an AI Teammate: How an Artificial Teammate Reshapes Human-Human Communication in Small-Team Decision-Making

・Conversational AI is increasingly positioned as a teammate rather than a tool, yet we know little about how its presence reshapes communication among the humans on the team. ・We examined sociocognitive communication dynamics in team decision-making using Group Communication Analysis (GCA), team surveys, and lexical analyses of team discourse. ・Teams completed a high-stakes moral-dilemma decision task in a randomized co
The Verge

The Verge’s 2026 back-to-school shopping guide

・Knowing exactly what your student needs for the school year ahead is next to impossible. ・Sure, you'll probably nail the essentials, but there will likely be a few items you forgot to buy, or didn't think they'd need to have. ・We're pulling our weight during the back-to-school season with a new shopping guide that's a mix of practical and aspirational products for students to use in and out of the classroom.
AI News & Artificial Intelligence | TechCrunch

This $9 key physically locks your most addictive apps

・This $9 NFC key requires you to physically scan it to unlock distracting apps on your phone.
WIRED

This AI Assistant Wants to Make Up for Your Boyfriend’s Incompetence

・An ad for Orchid suggests the AI agent can fix relationship problems by simply doing everything for inconsiderate partners.
Hugging Face Papers

Trending Papers

Trending Papers
The Verge

Trump blames Tim Walz for water hacks even though it’s probably Iran

・Donald Trump speaks during the House Republican Party member retreat. ・| Image: Mandel NGAN / AFP via Getty Images The FBI, the EPA, and the Cybersecurity and Infrastructure Security Agency (CISA) have stopped short of officially blaming Iran for a spate of cyberattacks on Minnesota's water systems, but consensus is that Iran is likely behind them. ・That, of course, hasn't stopped Donald Trump from sharing his own theo
Hugging Face Papers

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System
Zennのトレンド

Web Streams API 入門 ― 基本概念から実践まで

・この記事は、CYBOZU SUMMER BLOG FES '26 (Frontend Team) DAY 8の記事です。 ・こんにちは、tasshi です。 ・この記事では、JavaScriptでストリーム処理を行うための標準API「Web Streams API(Streams API)」について、基本概念から実践的な使い方、Node.js Streamとの違いまでを解説します。
WIRED

What’s Behind the Heat Wave Scorching Huge Parts of the World

・A combination of El Niño, unusually hot ocean waters, and jet stream changes is sending temperatures soaring in the US, Europe, and Asia.
Takara TLDR - Daily AI Papers

Where Physics Meets Privacy: Federated PINNs for Privacy-Preserving Brain Tumor Biomechanical Modeling

・Brain tumors such as glioma, meningioma, and pituitary adenoma alter the mechanical behavior of soft brain tissue, yet common diagnostic methods rely on static imaging that cannot capture tumor growth, tissue displacement, or changes in stiffness over time. ・Deep learning models for this task typically require pooling patient data at one site, which conflicts with privacy rules such as GDPR and HIPAA and limits genera
Takara TLDR - Daily AI Papers

Why Are GUI Agents Correct but Late? Decode on the Decision-Time Critical Path, Tested with Pre-Compiled Policy Trees

・Computer-use agents often fail on transient GUI events because they produce the correct action only after the relevant window has already closed. ・We identify the main cause as expensive autoregressive decoding on the decision-time critical path. ・We propose Adaptive Anticipatory Policy Trees (AAPT), which eliminates this delay without modifying the underlying model.
The Verge

With Switch 2, iPhone, and laptop tricks, the Sharge Disk Pro 2 is finally a worthy EDC

・The Sharge Disk Pro 2. ・| Photo by Sean Hollister / The Verge Sharge, the company that makes delightful retro Mac-shaped chargers and see-inside batteries, is finally impressing me with a portable SSD. ・I couldn't recommend the Sharge Disk, Disk Plus, or even the Disk Pro, but I'd be happy to own the new Disk Pro 2 Ultra.
The Verge

You could be taking way better photos on your phone

・Your smartphone photos can look way better. ・If you're willing to put in a little extra work, use third party apps and spend a bit of time editing, it's possible to take photos that - under the right circumstances - can rival dedicated digital cameras. ・I post a lot of smartphone photos online, and "How do I take better photos on my phone?" is a question I get practically every day.
Qiita - 人気の記事

オーダーメイドで九九の計算プリントを作ろう(ローテクAI)

・夏休み中で小学生が家に居るおとうさん、おかあさん。 ・暑い中でなかなか外では遊べませんが、動画視聴やゲームばかりというのも気になります。 ・せっかくの機会なので、お子さんの苦手分野のプリントを作って、クイズ的に一緒にやってみませんか? Geminiでこんなプロンプトを入力すると...
Zennの「大規模言語モデル」のフィード

オフラインAIは本当に安全か

・オフラインAIは本当に安全か LLMが人間と生成物を通信路にする時代のセキュリティ設計 ローカルLLM、オンプレミスAI、オープンウェイトモデルの利用が広がっている。 ・外部APIに機密情報を送らず、自社環境や手元の端末でAIを動かせることには、明確な価値がある。顧客情報、社内文書、ソースコード、契約書、研究資料、未公開の企画などを扱う場面では、外部ベンダーの推論サーバーへ入力を送らないことは重要なリスク低減策になる。 ・しかし、そこから「オフラインなら安全」と結論するのは危うい。
#AIタグ

これからの働き方やビジネスモデルにモヤモヤしている人たちに、ぜひ読んでもらえたら嬉しいです。(パネルディスカッション全文公開)

・先日登壇した「ビジネスモデルオリンピア2026」のパネルディスカッションの記録が、全文公開されました 。(先日、といっても3月ですが) 世間では「AIでどう作業を効率化するか」というノウハウばかりが語られがちですが 、僕がこの場で一番話したかったのは、テクノロジーが進んだからこそ、僕たち人間がどう思考し、どう直感や価値を生み出していくかという、もっと泥臭くも本質的な話でした。
#AIタグ

そして輝くアクセント辞典~機械が日本語を話すとき、基準は誰が決めるのか

・機械が日本語を話すとき、基準は誰が決めるのか 合成音声は、ずいぶん自然になった。
#AIタグ

そのコマンド、コピペする前に。AIが教えてくる名前の約2割は、この世に存在しません

・AIにアプリの作り方を聞くと、途中でこういう指示が出てくることがあります。 ・「まず、これをインストールしてください」 続きをみる
Zennのトレンド

チーム制作にプログラマーとして人生初参加して、思ったことを書いてみる。

・ゲームのチーム制作に初参加した事を振り返ってみる 2025年11月あたりから始まったチーム制作が、2026年7月22日付けでひと区切りを迎えました。 ・今回は、そのチーム制作がどんな感じだったかを反省しながら、振り返る記事です! 何より、しんどかった……。 ・Gitでのコード管理、想像以上に大変だった 共同で一つのゲームを作る以上、何かしらの共同作業ツールの導入が必要です。
Qiita - 人気の記事

どう頼むかがAIの成果を決める ― 丸投げしないAI協業の「発注の型」

・背景 AIに調査やドキュメント作成を任せると、たしかに速い。でも、出てきた成果物を見て「なんか違う…」となって、結局イチから直した経験はないでしょうか。 ・これは実案件で「複数の案から1つを選ぶ調査レポート」をAIに作ってもらったときに、私自身がぶつかった壁でした。試行錯...
Zennの「大規模言語モデル」のフィード

なぜ、今後のAIは頭のいい人にはより頭の良い回答を、そうでない人にはそれなりな回答をするようになって行くのか?

・はじめに 現在の生成AIは、原則として同じモデルが多くの利用者に提供されています。しかし、実際の回答はすでに完全に同一ではありません。質問の書き方、会話の履歴、利用者からの訂正、過去に示された好みなどによって、説明の詳しさや論理の組み立て方が変わるからです。 ・今後、この個人適応はさらに進むと考えられます。AIは目の前の一つの質問だけを見るのではなく、過去の質問や反応から、利用者がどの程度の抽象的説明を理解できるか、どの部分で誤解しやすいか、どの程度の確認が必要かを推定するようになります。 ・その結果、頭のいい人には、前提を省略しながら高度な推論を展開する回答が返されます。一方で、複雑な...
Zennの「大規模言語モデル」のフィード

マルチホップ質問応答のための知識グラフ設計:エンティティ抽出の実務

・「このセンサーを使っている製品ラインで、去年発生した不具合の件数は?」——このような質問は、ベクトルRAGが最も苦手とするタイプの質問です。「センサーの仕様書」「製品ラインの構成表」「不具合報告書」という3つの異なる文書を横断し、それぞれの関係を辿らないと答えが出せません。 ・ベクトル検索は「質問文と意味的に近い文書」を探す仕組みのため、質問文には出てこない中間の情報(このセンサーがどの製品ラインに使われているか)を経由する必要がある質問には対応できません。この「複数の関係を辿って答えにたどり着く」質問をマルチホップ質問応答と呼びます。 ・以前の記事「GraphRAG本番運用の経済学」では、...
Zennの「機械学習」のフィード

メビウスの輪には裏表がない。次元削減で平らに広げたら、どこかを潰すしかなかった

・紙テープを半回転ひねって輪にすると、メビウスの輪ができる。表面を指でなぞっていくと、いつの間にか「裏側」に来ていて、一周半でスタート地点に戻る。裏表の区別がない面——数学では「向き付け不可能」と呼ぶ。 ・スイスロールは綺麗に展開でき、穴あきでも大域構造は保たれ、結び目はほどけた。ではメビウスの輪はどうか。この面は、切らずに平面へ広げることが原理的にできない。次元削減アルゴリズムたちは、この不可能な要求にどう応えるのだろう。 ・ドラッグで回転できる3Dインタラクティブ版は、元記事(メビウスの輪には裏表がない。次元削減で平らに広げたら、どこかを潰すしかなかった)に置いています。
#LLMタグ

もし1/6で発動する魔法があったら|AI用設定資料

・ちょっとした遊び。最近流行りのキャラ自体と直接対話するのを楽しむのが趣旨の『キャラクターチャット』とは違うが、社会組織や世界観をLLMと壁打ちして広げる楽しみ方を紹介する。 ・僕なりに1万文字以上作りこんだ設定も紹介する。これから増減するだろうし確定版でもないので意味あるかは不明でおすそ分けなのだが最近こういう設定でAIと対話して世界観を広げて遊ぶというのに熱中して面白いので紹介しようと思う。
Qiita - 人気の記事

リアルタイム交通流データをQGISで表示してみる

・はじめに 都市の解析や災害時において、地域の交通状況はとても重要な情報となります。自動車メーカーさんなどは通行実績のデータ配信されていますが、独自形式だったり、ライセンスがしっかりしていたり、なかなかなお値段がして一般の方がGISなどに取り込むことはけっこう難しいです。...
機械学習タグが付けられた新着記事 - Qiita

ローカル画像生成で「Corporate Memphis」を本物っぽく出す — Qwen-Imageのプロンプト/ネガティブ/img2img実装知見

・ブログのヒーロー画像やOGP画像を、クラウドの画像生成API(Gemini等)からローカルモデルに置き換えられるか——という検証を先日おこないました(全体像は本家記事にまとめています)。 ・本家記事: ローカル画像生成AIはGeminiの代わりになるか?——FLUX・Qw...
Zennの「機械学習」のフィード

拡散モデルをMNISTで動かす:トイモデルから本物のU-Netへ

・この記事を読むと、前回2次元の点で確認した拡散モデルの仕組みを、実際の手書き数字画像(MNIST)に拡張し、小さなU-Netを自分の手でゼロから学習させて「ノイズから数字を生み出す」ところまで動かせるようになります。 ・前回の記事では、2次元の「二つの月」という点の集まりを使って、拡散モデルのforward process(データにノイズを混ぜる)とreverse process(ノイズからデータを復元する)を実装し、物理の拡散現象・ランジュバン方程式・統計力学との対応まで見てきました。今回はその続きとして、「この仕組みが実際の画像データでどう使われているか」に踏み込みます。
Qiita - 人気の記事

個人的な資格の受かり方~Java Gold SE 17~

・はじめに 今回はJava Gold SE 17(正式名称は"Oracle Certified Java Programmer Gold SE 17")に合格してきたので、使用した参考書や勉強方法などをまとめておきたいと思います。 ・これから受験する方の参考になればと思...
#AIタグ

高校生でもできるAI×YouTube運用完全攻略0から企画・編集・分析まで効率化する方法

高校生でもできるAI×YouTube運用完全攻略0から企画・編集・分析まで効率化する方法
Qiita - 人気の記事

高校生にPythonSCADの最初のExampleを解説する

・ネジ穴付きの小型ケース PythonSCADコミュニティ・開発者に感謝申し上げます。ありがとうございます。 ・今回は、コメントアウト(#)を1行ずつ外していくことで、ただの四角いブロックが「実用的なケース(容器)」へと進化していく様子を体験しましょう。
LLMタグが付けられた新着記事 - Qiita

黒電話を分解して、ローカルLLM×ずんだもんと通話できるマルチモーダルAIシステムを作ってみた➁

・概要 黒電話を分解して、ローカルLLM×ずんだもんと通話できるマルチモーダルAIシステムを作ってみた①の続きです。 ・このプロジェクトは、黒電話(600-A2-CL)を用いた物理インターフェースと、 ローカルLLMを融合させた、学園祭の展示向け・エンタメ要素を含んだマルチモ...
Zennの「機械学習」のフィード

材料データの Scaler/変換——X と y を混同しない

・はじめに:Standard をかけたら、精度が落ちた 材料の表データで、温度・圧力・時間・組成のように単位も桁も違う列を並べることは日常である。距離や正則化を使うモデルでは、そのままでは大きい軸が支配する。そこで出てくるのが StandardScaler や RobustScaler だ。 ・ところが、特徴量 X に測定スパイクのような外れ値を混ぜた擬似データで k 近傍回帰(kNN)を回すと、次のような結果になる(SEED=42、5-Fold CV RMSE)。ここでの外れ値は目的変数 y ではなく、説明変数側の話である。 ・前処理 CV RMSE(平均) なし ≈ 35...
#LLMタグ

雑談。人ができるんだから、(AIは)できるでしょう、という時代に。

・個人的には、あまり、AIが、社会をガラッと変えたりはしないと思っているのですが、、、その思いは、特に、変わりませんが。。。
#LLMタグ

私の読書ノート[50代まじかのおっさん]ビジネス書:AI時代の質問力 プロンプトリテラシー

・AI時代の質問力 プロンプトリテラシー / 岡 瑞起、橋本 康弘 / 翔泳社 続きをみる
Qiita - 人気の記事

自分のPCだけで動く「自分専用AI」を Ollama × Docker で動かしてみた

・自分のPCだけで動く「自分専用AI」って、実は拍子抜けするくらい簡単に作れる。 ・ChatGPT や Claude に質問を投げるたび、「このデータ、外に出して大丈夫だっけ」と一瞬手が止まったことはないだろうか。個人のメモや検討中のコード、ちょっとした愚痴まで、クラウドの向こ...
Qiita - 人気の記事

小売CRMでクーポン施策を伸ばす:クラスタリング・Lifetimes・LightGBMをつなぐ顧客セグメント設計

・会員データは「誰に、何を、いつ届けるか」を決めるために使う 日本の大型商業施設では、アプリ会員証、ポイントカード、キャッシュレス決済、館内イベント予約などを通じて、来館と購買の履歴が日々蓄積されている。たとえば、首都圏に複数店舗を展開する架空の商業施設「さくらアーバンモー...
#LLMタグ

新しいAIが出た週に、私がやっていること——「追いかけない」と決めた人の、新モデルの迎え方

・昨日、「どのAIが最強かを追いかけるの、やめませんか」という記事を書きました(→ https://note.com/rain_dual_ai/n/ndc7886970b36 )。 ・そしてその同じ週に、ちょうど新しい上位モデル(Claude Opus 5)がリリースされて、話題になっています。性能は上がって、価格は据え置きだそうです。
Zennのトレンド

生成AIにおけるHTML/画像出力の必要性と改善案としてのRHW(reviewable-html-workbench)の紹介

・この記事はXの元記事をZenn向けに再掲載したものです。 ・要約 (TL;DR) モデルが 1 回に出す量は増え続けており、読む側の処理能力は変わりません。 ・この差をどこまで埋められるかで AI との協調の質が決まります。
Qiita - 人気の記事

設計レビューがHTMLで返る improve-codebase-architecture が良い

・はじめに mattpocock/skills の /improve-codebase-architecture は、設計レビューの結果を Markdown ではなく自己完結した HTML ファイルで返すという珍しい出力形式のスキルです。一時ディレクトリに architec...
#AIタグ

卒論の参考文献探しが劇的に変わる。私が実際に使っているAI論文調査術【プロンプト付き】

・はじめに 卒業論文を書き始めると、多くの人が最初につまずくのが「参考文献探し」ではないでしょうか。
#LLMタグ

損益だけ見ても、トレードの悪い癖は見つからない

・トレードを見直すとき、最初に損益を見ていませんか。
#LLMタグ

第2章:毎日使ってる!トレンドAI(生成AI・LLM)の裏側

・今やニュースや日常で目撃しない日はない「生成AI」と「LLM(大規模言語モデル)」。 ・第2章では、世界中の働き方やクリエイティブを変えた最新AIたちの「脳みその仕組み」に迫ります! 文章をスラスラ書くChatGPTの心臓部「Transformer」や、本物そっくりの絵を描く「GAN」「拡散モデル」など、現代のAIを語る上で絶対に外せない主役たちが勢ぞろい。複雑な数式に頼らず、その驚きのアイデアと仕組みをサクッと攻略していきましょう! 続きをみる
Zennの「機械学習」のフィード

日本語はなぜトークン代が高いのか。BPEを一から実装して測ったら、「3倍高い」は言い過ぎだった

・LLMの料金はトークン単位で決まる。そして「日本語は英語よりトークンを食うので高くつく」という話をよく聞く。実務でAPIを叩いていると無視できない話だが、実際どれくらいなのかを自分で測ったことがなかった。 ・そこで、トークナイザの中身であるBPE(Byte Pair Encoding)をゼロから実装し、仕組みを理解したうえで、自分のブログ49記事を実際のトークナイザにかけて測ってみた。結果は、よく言われる「3倍」という数字が半分正しくて半分ミスリードだという、思ったより面白いものだった。 ・BPEの仕組み: 頻出ペアを繰り返し合体させるだけ BPEのアルゴリズムは驚くほど単純だ。テキスト...
#AIタグ

猫Tシャツで世界は救えるのか。AIと壁打ちして生まれた一枚、限定生産はじめます

・ふと、こんな問いが浮かんだ。「猫のTシャツで、世界は救えるのだろうか」と。 ・大げさに聞こえるかもしれない。けれど本気で考えてみた記録を、今日はドキュメンタリーとして残しておきたい。 ・Can a cat T-shirt save the world?
#AIタグ

百日紅モチーフのサマードレス

・Garment Prompt — Crape Myrtle Hand-Dyed Silk MuuMuu A relaxed elegant Japanese summer resort muu-muu inspired A-line tunic dress, featuring a soft flowing silhouette and effortless luxury. ・The garment has a loose non-clinging shape, designed to move naturally with the body.
#LLMタグ

法令RAG dumpを8/1版に更新 — 公開直後に見つけた「未来の条文が現行として返る」欠陥

・7月に続いて法令RAGの事前ビルドdumpを8月1日取得分に更新しました。 ・タグは law-rag-20260801 です。 ・ただ今回は、きれいな更新告知だけでは済みませんでした。
#LLMタグ

毎日使えるAIプロンプト手帳|曜日×時間帯で迷わない30日カレンダー

・AIを開いた。……で、何を聞けばいいんだっけ。入力欄の前で手が止まって、そっとタブを閉じる。 ・その経験、あなただけではありません。ChatGPTは2025年時点で週に約7億人が使う道具になりました(OpenAI公式調査)。それなのに、多くの会社員は「たまに使うだけ」で止まっています。プロンプト集をブックマークしても、結局“保存して満足”で終わる——。
Hugging Face Papers

β-OPSD: Deriving with Policy Optimization, Training with Self-Distillation

β-OPSD: Deriving with Policy Optimization, Training with Self-Distillation
Hugging Face Papers

Σ-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems

Σ-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems