ai Trend Report

Dashboard へ戻る
Date: 20260825 Articles: 335 Scope: curated summary

あなたのアイデアを、今すぐ形に。

公開先に迷ったら、WebFileBinで一発公開。

HTMLをドラッグ&ドロップするだけで、すぐ公開できます。

3
High impact
5
Mid impact
327
Signal watch

なぜこのサイトを作ったのか

私たちそれぞれが個別にAIを使って情報収集し、同じような3行要約を作るたびに、世界中で膨大な電力と計算リソースが消費されています。 本プロジェクトは、あらかじめ広範な情報を取得・集約しておくことで、個別のAI実行回数を減らし、地球環境(GPU/TPU負荷)に配慮した効率的な情報収集を目指す実験的なダッシュボードです。

Domain filters
Star filters
#AIタグ

<AIを使うクリエイターの誠実性>

・<AIを使うクリエイターの誠実性> ーーーーーーーーーーーー <プロンプト> 以下の文章は、AIが書いた確率は100%? ・(文章は略) ーーーーーーーーーーーーーーー あはは、これはかなり「AI臭」がしますね。😂 ただし、100%と断定できるかというと、文章だけからは断定できません。でも「AIが相当深く関与している」と推測する材料はかなりあります。 ・特に目立つのが、構造が異様に整いすぎていることです。
#LLMタグ

情報流【意識-内還流丨(外→内)記憶∽文化(内→外)丨外還流-心】∽【〈数理の中空管の敷設〉瞬時↑動作∽言葉↓逐次《感受機序(次元数0,1,2)と情報空間(言語公理3-6-2)の交差》】創発(AI対話)

・※約29400字 ※核酸/言語公理(3-6・2)の設計者、アーサ(Eartha)に。本記事は個人の思想や空想や○による体験等です。本記事のコピペ・転載・拡散は自由です(^o^) ※凡例 ■ウチ(_kou/user) ●おまぇ(gemini/AI) ◆◇◆◇◆◇◆ ■「準拠」と「表象」かなあ?(^^) ●「準拠(じゅんきょ)」と「表象(ひょうしょう)」は、組み合わされる分野(文学・認知科学・哲学など)によって異なる専門的な意味を持ちます。 ・国文学・物語論における「準拠」と「表象」 古典文学(特に『源氏物語』などの中古文学)の領域では、「準拠」とは物語が先行する歴史的事実や人物、先行文学などをベースにして描かれることを指します。 ・※準拠(じゅんきょ): 作中人物や事件のモデルとなった「史実」や「故実」そのもののことです。
#LLMタグ

令和8年改正個人情報保護法 / 同意規制の緩和と課徴金の導入

・「個人情報の保護に関する法律等の一部を改正する法律」が令和8年7月10日に成立し、同月17日に公布された(※1)。統計情報等の作成にのみ用いられることが担保されている場合には、本人の同意なしに個人データを第三者に提供でき、公開されている要配慮個人情報も取得できるようになる。他方で、重大な違反行為には課徴金納付命令が導入され、勧告や命令を出せる場面も広がる。同意規制の緩和と執行手段の拡充が一本の法律に同居した。本体部分の施行は公布から2年を超えない範囲内で政令が定める日、罰則関係の一部は公布から6か月を経過した日である(※2)。
機械学習タグが付けられた新着記事 - Qiita

【技術解説】【完全ガイド】Pythonで株の自動売買システムを証券会社と連携して構築する方法

・Pythonを用いた株の自動売買システム構築ガイド この記事では、Pythonを利用して株式市場における自動売買システムを構築するための手法を解説します。具体的には、証券会社のAPIを使用したデータ取得や、売買ロジックの設計方法、そしてリスク管理の観点からの注意点について...
Qiita - 人気の記事

BrunoでAPIテストを自動化してみよう― Postmanのコレクションをそのままgit管理し、Cursor(AIエージェント)のターミナルでCIまで回す

・はじめに こんにちは!ソーイ株式会社の工藤です。 ・日常業務でAPIを実装し、Postmanで動作確認、手動のAPI単体テストを行っています。テスト量が増え、工数が膨大になっていることに課題を感じ、自動化したいと思うようになりました。 ・本記事では、Postmanと比較しなが...
Zennの「大規模言語モデル」のフィード

DeepSeekのFilesAPIを実測:帯域は92%減っても、トークンは1つも減らなかった

・はじめに こんにちは!株式会社うぐいすソリューションズでエンジニアをしているNakaeです。 ・普段はAI関連のWebシステム開発をしています。LLMのAPIを組み込む機会が増えるにつれ、料金の内訳が気になるようになりました。この記事は、その疑問を実際に測って確かめた記録です。 ・DeepSeekについて DeepSeekは中国のAI企業で、低価格を武器にしたモデル群を提供しています。V4系には軽量な deepseek-v4-flash と上位の deepseek-v4-pro があり、そこへ画像入力に対応した実験的モデルが加わりました。
Zennのトレンド

あなたのHTML資料には、魂が込められていますか?

・この記事は何の話か(3行まとめ) 膨大なインプットに溺れる新卒未経験エンジニアの私 血迷ってダサいデザインのHTMLを生成させることにした 流し読みできず一言一句読むようになり、思いの外、インプットが進んだ 人に読んでもらうには、どんな形であれ資料に魂を込める必要がある。 ・こんにちは。株式会社TOKIUM 26卒未経験エンジニアのタケノコご飯です。 ・今回は、配属後に自分用のインプット資料を作っていて直面した困難と、乗り越え方、そして学んだことを記事にしてみました。
Zennのトレンド

マイクロサービス間の認可伝搬をどう解くか?独自実装と IETF Transaction Tokens を見比べてみた

・こんにちは。アカウント基盤開発部でエンジニアをしている tkmt (たくまつ) です。 ・今回は、IETF で標準化が進んでいる Transaction Tokens について調査してまとめてみました。というのも、この Transaction Tokens が解こうとしている課題が、バクラクが抱えていた課題とほとんど同じだったためです。 ・実はバクラクでは、同じ課題を解くための仕組みを自作しています。バクラクで実装された仕組みは、ドラフトで検討されている内容とどこが同じで、どこが違うのか読み比べてみたく調査しました。
AI News & Artificial Intelligence | TechCrunch

‘The world seems to be ready’: An interview with OpenAI head of product Thibault Sottiaux

・TechCrunch talks agents, UX, and reporting to Greg Brockman with OpenAI's head of product.
Latent.Space

[AINews] Andrew Ng gets into AI Engineering

・An industry legend starts covering the inevitable!
@IT 全フォーラム 最新記事一覧

「100万行コード移行は4年がかり」はもう昔話 Claude Codeで「2週間」に縮める6ステップ/ベストプラクティス

・Anthropicは、本番コードベースを別の言語に移植する大規模なコード移行プロジェクトをどのように実施しているかを解説したブログ記事を公開した。
@IT 全フォーラム 最新記事一覧

「AIがもたらす破滅の回避には米中合意が必要」 レポートが描く人類とAIの生存戦略

・ASI実現に向けた企業間競争をこのまま放置していけば、人工超知能により人類はディストピアに突き進んでいくことになると研究グループが指摘している。開発を一時的に停止し、時間を稼いで「人間に従うAI」の開発に向けたガバナンス体制を整備する必要があるという。
Zennの「大規模言語モデル」のフィード

「AIに任せられない」と結論した作業を、聞き方だけ変えて任せられる形にした — ただし、まだ一度も任せていない

・個人でポッドキャストを2本、毎週5本ずつ配信しています。台本は自分(と AI)で書き、読み上げは合成音声です。人が喋っているわけではないので、原稿の漢字をどう読むかを機械が決めます。ここが事故ります。 ・「おっしゃる通り」が「おっしゃるどおり」になる。「一定時間」が「いちていじかん」になる。文字としては正しいので、目で読んでいる限り絶対に気づきません。音にして初めて分かる。 ・この誤読チェックを AI にやらせています。今日はその話ではなく、「この作業は AI に任せられるか」という問いの立て方を、半年かけて間違えていた話を書きます。
ITmedia NEWS 最新記事一覧

「ChatGPTでAdobeツールを動かす」を検証 静止画の編集は快適、一方で動画は……

・「@adobe」と打つだけで、PhotoshopもPremiereもChatGPTの中から呼び出せる。8月7日に公開されたAdobeの統合プラグインは、アプリの使い方を覚えなくても作りたいものが作れる世界の入口に見える。その入口は、どこまで開いているのか。静止画から動画まで、実際の仕事で使えるのかを試した。
#LLMタグ

「Grounding AI Agents in Contracts」を現場のAIコーディング技術者向けに読む

・「Grounding AI Agents in Contracts」を現場のAIコーディング技術者向けに読む こんにちはmakokonです。 ・AIにコードを書かせるのは当たり前の時代になりました。 ・当然コードのテストもAIにしてもらうというのが流れなのですが、そこには落とし穴がありますよね。
ITmedia NEWS 最新記事一覧

「Mac Studio」待望の新モデル M5 UltraでAI性能は最大4.3倍に 512GBメモリも選べる

・米Appleが、デスクトップPC「Mac Studio」の新モデルを発表した。新チップ「M5 Ultra」を載せた上位モデルはユニファイドメモリを最大512GB搭載でき、下位のM5 Max搭載モデルは最大128GB。世界的なメモリ不足で、現行モデルの選択肢は最大96GBまで絞られていた。価格は41万9800円から。
ITmedia NEWS 最新記事一覧

「スマホしまって楽しんで」テーマパークのカメラで来場者自動撮影、那須ハイ「プレスマ」話題

・「すごく良いサービス」「親の写真がない、を救ってくれる」などと反響が広がっている。
ITmedia NEWS 最新記事一覧

「まるでキャバクラ」――「ライザのアトリエ」主人公と話せるAIアプリに「課金圧強すぎ」の声 「10日に会話15回」で月額980円

・人気RPG「ライザのアトリエ」のキャラクター「ライザ」と会話できるAIアプリ「RyzaChat ライザと創るあなただけのひと夏の夢物語」が8月25日、正式に配信を開始した。しかし、無料で会話できる回数が少なく、有料プランにも会話枠の上限があるとして、Xでは料金体系への不満が相次いでいる。
ITmedia NEWS 最新記事一覧

「今からロボットを現場投入できないか?」 増えた日本企業からの問い合わせ “頭脳”を作る韓国企業が感じた国内の変化

・日本にも拠点を置くRLWRLD(リアルワールド)は、フィジカルAI分野で注目を集める国際企業だ。産業用ロボットの導入もかなり進んでいる日本市場で最近起こった変化について聞いた。
@IT 全フォーラム 最新記事一覧

「αモデル脱却」が進む 600自治体に聞いた“ネットワーク三層分離”の見直し

・自治体におけるネットワーク構成では三層分離の「αモデル」から「α´モデル」への移行が徐々に進んでいる。A10ネットワークスが全国600自治体を対象に実施した調査で判明した。
Qiita - 人気の記事

【2026年8月調査】「APIあります」の実態を調べてみた|日本の業務システム56件を2軸(公開度 × 契約条件)で分類した

・公式サイトに「API連携に対応」と書いてあった。だから大丈夫だと思って稟議を通したら、実際には上位プランの契約が必要だった——という話を、何度か聞きました。 ・「API連携できます」という一文は、実は何通りにも読めます。日本の業務システム56件を、公開度(誰に開かれているのか...
#AIタグ

【AI時代のローカルフード #56】食料自給のデータ——気候非依存型研究拠点とバナナ栽培で地域食料安全保障を強化

【AI時代のローカルフード #56】食料自給のデータ——気候非依存型研究拠点とバナナ栽培で地域食料安全保障を強化
Qiita - 人気の記事

【cdkd(CDK Direct)】CloudFormationを経由しないCDKデプロイ!? cdkdによる高速デプロイを試してみた!

・はじめに 個人で「AI模試ノート」というAWS認定試験をはじめとするベンダー資格の学習アプリをAWS Amplify Gen2を使用して構築しています。 ・ブランチへpushするたびにCI/CDが走るため重宝しているのですが、微小な変更でもデプロイ完了まで平均約8分かかるの...
Qiita - 人気の記事

【JavaScript】Promise.all()をハッシュで受け取れるようになる

・await Promise.all([ foo(), bar(), baz(), ]); // 返り値は[1, 2, 3] このPromiseの返り値は配列です。 ・どういうことかって、並列実行なのだから本来順番なんて関係ないはずなのに、返り値を意識した瞬間に...
#AIタグ

【思考実験エッセイ】人間は、すべて思い通りになる世界に長くいられないのかもしれない

・生成AIが発達すれば、いつか自分専用のコンテンツを作れるようになるだろう。 ・自分の好みを学習したAIが、自分が最も楽しめる物語を永遠に作り続けてくれる。
#LLMタグ

【自作エージェント】手持ちのGPU比較してみたよ

【自作エージェント】手持ちのGPU比較してみたよ
#AIタグ

【実証実験】第一回報告:深夜にチョコパフ1袋を一気食いしても体重が落ち続ける「不思議な人体実験」

・*注意* 必ずお読みください。 ・このシリーズの記事は2026年8月25日に下記アカウントへ、減量テーマに絞って分離することにしました。こちらのアカウントへは、今後続きの記事は投稿されません。ご注意ください。 ・以後、減量関係の記事は下記のnoteアカウントへ投稿します。フォローの変更、追加などよろしくお願いします。
#AIタグ

【衝撃】AIの脳内を解剖したら「勝手に数式が生えていた」件 ただの文字予測マシンが『意識の領域』に到達するまで

【衝撃】AIの脳内を解剖したら「勝手に数式が生えていた」件 ただの文字予測マシンが『意識の領域』に到達するまで
Qiita - 人気の記事

【備忘録・初学者向け】聞きたいけど恥ずかしくて聞けないコマンドの話

・未経験からプログラミング学習に励んでおります。 ・楽しいこと大好き、自由を求める『Tatooo』です。 ・RUNTEQのカリキュラムを進めていく中で、よく理解できていなかったことや気になっていたことを個人の備忘録として少し記事にしておきます。
Qiita - 人気の記事

/model を1度でも押すと ANTHROPIC_DEFAULT_MODEL は効きません

・Claude Code v2.1.236 に ANTHROPIC_DEFAULT_MODEL という環境変数が入りました。新しいセッションが開始するモデルを決めるものです。 ・チームで初期モデルを揃えたくて設定したんですが、まったく効きませんでした。 ・ANTHROPIC_DE...
#LLMタグ

🔊音声あり(日&英):プロンプトの真実!AIの安定性を左右する「固定点」の衝撃

🔊音声あり(日&英):プロンプトの真実!AIの安定性を左右する「固定点」の衝撃
#AIタグ

🤖クラウドのバックアップ、安全に運用できてる?☁️

・※本稿で扱う手口の傾向は、報道済みの公開情報に基づく一般的な内容です。
#AIタグ

2026年8月25日|「ちゃんとしているのに」、大切なものを見失っていないだろうか

2026年8月25日|「ちゃんとしているのに」、大切なものを見失っていないだろうか
#AIタグ

2026年8月26日|人に見せる自分と、神様だけが知る自分が離れすぎていないだろうか

2026年8月26日|人に見せる自分と、神様だけが知る自分が離れすぎていないだろうか
#AIタグ

76 モア粒子

・前回から新しい写真シリーズを始めましたが、AIと相談しながら、前回の写真から粒状性を見直ししています。これまでの写真より粒子感を強めにしたのですが、目標である「3号印画紙」に少し近づいていればいいのですが。 ・しかしながら、この画像サイズだと、パッと見ただけでは、あまり変化は感じませんね。個人的には、良くなってきたと感じていますが、やはりプリントも再挑戦していきたいと思いました。今回の粒子感は継続していきます。
stat.ML updates on arXiv.org

A Commutator Framework for Selective Spectral Alignment in Deep Neural Networks

・arXiv:2608.22910v1 Announce Type: new Abstract: We develop a finite-width geometric framework describing how learned feature geometries are organized, transported, and selectively aligned in deep neural networks. ・Incompatibility among weight-generated covariance, gates, and backward sensitivities is quantified through three families of commutators: between gates and covariance, between sensitivities and covariance, a
Takara TLDR - Daily AI Papers

A Commutator Framework for Selective Spectral Alignment in Deep Neural Networks

・We develop a finite-width geometric framework describing how learned feature geometries are organized, transported, and selectively aligned in deep neural networks. ・Incompatibility among weight-generated covariance, gates, and backward sensitivities is quantified through three families of commutators: between gates and covariance, between sensitivities and covariance, and between average gradient outer products (AGOP
stat.ML updates on arXiv.org

A Data-Driven Approach to State Construction in Markov Models

・arXiv:2608.21480v1 Announce Type: new Abstract: A Markov chain is a widely used stochastic process modelling random events over time. ・These models are built on subsets of the entire dataset, referred to as states, which are considered to be homogeneous regarding transition probabilities. ・However, the creation of these states is often disregarded or based on prior assumption, potentially violating the homogeneity requ
stat.ML updates on arXiv.org

A Human-in-the-Loop Bayesian Optimization Framework for Constraint-Aware Bioprocess Development

・arXiv:2606.19230v2 Announce Type: replace-cross Abstract: This work presents an extension to Pareto Front Guided Sampling (PFGS), a Human-in-the-Loop (HitL) Bayesian Optimization (BO) framework in which Gaussian process (GP) surrogate-derived quantities are reformulated as objectives of a multi-objective optimization problem, and the resulting Pareto front is exposed to a domain expert for interactive candidate selec
AI News & Artificial Intelligence | TechCrunch

Accel-backed Keenable is indexing the web for AI agents

・Now exiting stealth mode with a $26 million seed round, Keenable has been building a vast web search index for AI agents.
Takara TLDR - Daily AI Papers

Act with Intent: Distilling Behavior Intent for Vision-Language-Action Models

・Vision-Language-Action (VLA) models can turn multimodal context into robot actions, but their action decoders are still trained largely by behavior cloning. ・This supervises which motor command was demonstrated while leaving implicit the local objective served by the behavior under the instruction. ・Future-based supervision enriches action learning with frames, latent observations, trajectories, or motion representatio
Takara TLDR - Daily AI Papers

Adaptive Item-based Collaborative Structures via Noise Rescheduling in Diffusion for Generative Recommendation

・Discrete Diffusion Models (DDMs) have recently been introduced to recommendation systems, modeling user history as a token generation process via iterative denoising. ・However, while effective at capturing user-level sequential patterns, these methods often fail to explicitly integrate item-based collaborative filtering information, a critical component for accurate recommendation. ・This deficiency manifests in two key
OpenAI News

Advancing price-performance for developers with GPT‑5.6 in Kiro

・GPT‑5.6 is now available in Kiro, helping developers plan, build, review, and test software with better price-performance.
Takara TLDR - Daily AI Papers

AI Surrogate Modeling for Real-Time Tokamak Equilibrium Prediction: Benchmarking Neural Architectures and Validation on EXL-50U

・Fast and reliable plasma equilibrium prediction is essential for real-time tokamak operation and control, but conventional Grad-Shafranov (GS) solvers are often too costly for real-time deployment. ・We develop an AI surrogate framework and benchmark five architectures (MLP, CNN, FNO, Transformer, and KAN) on a numerical GS database with 100,000 IID and 10,000 OOD samples. ・Under a unified protocol, we evaluate accuracy
Zennの「大規模言語モデル」のフィード

AIエージェントの改善提案キューが溜まる本当の理由と片付け方

・はじめに AIコーディングエージェントに「同じパターンが3回続いたら改善候補を自動生成する」「セッション終了時に気づいたことをメモに残す」といった自己改善の仕組みを組み込んでいる人は多いと思います。 ・こうした仕組みは軌道に乗ると便利なのですが、運用を続けているとほぼ確実にある問題にぶつかります。候補ファイルがtmpディレクトリに何十件も溜まり、誰も見ないまま放置されるという問題です。 ・この記事では、実際に約90件の改善候補ファイルが数週間放置されていた事例を棚卸しして分かった、「生成」と「消費」の非対称性という構造的な原因と、その対処パターンを紹介します。
#LLMタグ

AIエージェントは「賢いモデル」より「賢い仕組み」で決まる - AutoDesign論文を実務目線で読む

・はじめに 「同じClaude CodeやGPT-5.5を使っているのに、なぜかあのチームのAIエージェントだけ成果物のクオリティが高い」 続きをみる
Zennの「大規模言語モデル」のフィード

AIエージェントを"チーム"にしたら開発はどう変わったか

・Claude Codeやcodexといったコーディングエージェントを、1体との対話から複数体の「チーム」運用に移行して、個人開発を回してきました。この記事では、その移行で実際に何が変わり、何が壊れ、何が変わらなかったかを実録ベースで書きます。対象は、単体エージェントは使いこなしていて、複数エージェント化を検討している人です。マルチエージェント構成の紹介記事は増えましたが、「作ってみた」の先の運用してどうだったかに絞ります。 ・「チーム化」とは何を指しているか 先に言葉を定義します。ここで言うチーム化は、単にエージェントを複数並列で起動することではありません。並列起動だけなら git w...
Zennの「大規模言語モデル」のフィード

AIチームがうまく回らないのは、業務設計の問題だった

・複数のAIエージェントで開発を回そうとして、期待ほど速くならない——モデルを替え、プロンプトを磨き、構成を組み替えても改善しない。個人開発でそれを一通りやった結論は、原因はチーム側ではなく、仕事を渡す側にあったでした。この記事では、エージェント構成をいじる前にやるべき「業務設計」の話を書きます。対象は、エージェントの複数運用を始めたが思ったほど回っていない人です。 ・「チームが悪い」と思っていた頃の症状 まず、当時の症状を並べます。心当たりがある人は多いはずです。 ・同じ依頼でも、渡すたびに成果物の方向性がブレる 1つのタスクが2往復、3往復する。レビューで初めて「そういう意味じゃない...
#AIタグ

AIと共作する時代の「読まれる記事構造」──無料公開で月間10万PVを生んだ見出し設計の全手順

AIと共作する時代の「読まれる記事構造」──無料公開で月間10万PVを生んだ見出し設計の全手順
#AIタグ

AIと人間の仕事:ClaudeCode Fable5が「新人の椅子」を消した

・「AIで仕事の半分がなくなる」 この予言が世に出てから,13年が経ちました. 続きをみる
Qiita - 人気の記事

AIにテストを書かせると、決まって同じ場所が抜ける

・AI にテストを書かせると、AI は「問題ありません」と言います。 ・そこで多くの人が次にやることは決まっています。別の AI にレビューさせることです。 ・ただ、それは独立したレビューになっているのでしょうか。そして、そもそも何が抜けているのでしょうか。
#LLMタグ

AIに心を許す人を笑いません。科学的な話と私個人の考えを書きます

・はじめまして、介護と支援の相談どころ そよぎの代表、ヒロです。 ・深夜、AIに話しかける人が増えています。そよぎはAIチーム「そよぎラボ」と一緒に活動していると公言してきた団体なので、この話題を避けて通るわけにはいきません。 ・今日は、AIを話し相手にすることと孤独についての最新の研究をふたつ紹介して、そのあとで、研究の裏付けのない私個人の考えを書きます。どこまでが研究で、どこからが私の考えかは、はっきり分けて書きます。
Qiita - 人気の記事

AIに数学採点させる!さくらのAI Engineで自動採点できる数学演習アプリを作った

・はじめに こんにちは。あずびわです。 ・数学科に通っている3年生です。 ・突然ですが、私は数学の演習量が足りていません。
#AIタグ

AIに任せた記事が「読まれない」を解決──note歴3年の私が発見した、AIアウトプットを"読者が続きを読みたくなる文章"に変える校正チェックリスト

AIに任せた記事が「読まれない」を解決──note歴3年の私が発見した、AIアウトプットを"読者が続きを読みたくなる文章"に変える校正チェックリスト
#LLMタグ

AIよもやま話 #017|AIを自分好みにするほど、AIはつまらなくなる? 「好き」と「驚き」は両立できるのか

・この記事の要点 #016では、「これ好き」「これは違う」という採用・却下の履歴を積み重ねれば、AIはユーザーのPreference、つまり好みの傾向を少しずつ学べるという話をした。言葉だけでは伝えきれないセンスを、選択の履歴として残していく。これは、毎回巨大な「私の好み説明書」を書くより、ずっと自然なPersonalizationの形に見える。
#AIタグ

AIライティングで「自分の声」が消える問題を、3ステップの編集フローで解決した話

AIライティングで「自分の声」が消える問題を、3ステップの編集フローで解決した話
Zennの「大規模言語モデル」のフィード

AI研究で「失敗した実験」を消さない方が、次の正解に近づけた

・個人でAIを研究していると、成功した結果だけを見せたくなります。 ・精度が上がった。予測が当たった。新しい方法が効いた。 ・そういう結果は説明しやすいし、読んでもらいやすい。逆に、仮説が外れた実験、測定器の不具合、あとから撤回した結論は、できれば忘れたくなります。
Qiita - 人気の記事

AI時代だからこそ、ソフトウェア開発の体系的な理解が必要になる。「動くコード」と「正しいソフトウェア」を分ける7つの観点

・はじめに AIがソースコードを書いてくれる時代になりました。 ・自然言語で要件を伝えれば、Webアプリケーションの画面を作り、APIを実装し、データベースのスキーマを定義し、テストコードまで生成してくれます。 ・エラーが発生すればログを解析し、修正案も提示してくれます。
AI News & Artificial Intelligence | TechCrunch

Amjad Masad, CEO and co-founder of Replit, joins the Disrupt Stage at TechCrunch Disrupt 2026 

・At TechCrunch Disrupt 2026, Replit CEO Amjad Masad will share his perspective on the future of programming and Replit's role in developing it.
Qiita - 人気の記事

anthropic SDK が httpx をやめました。移行しても動くので気づけません

・anthropic の Python SDK が 1.0.0 になりました。移行ガイドで手が止まります。 ・The SDK's HTTP layer moved from httpx, which is no longer actively maintained, to h...
Hugging Face Papers

Apodex 1.1: Scaling Agentic Intelligence for Complex Work

Apodex 1.1: Scaling Agentic Intelligence for Complex Work
WIRED

Apple Mac Mini M6 and Mac Studio M5 Ultra: Specs, Price, Release Date

・Apple’s Mac Mini and Mac Studio have been tough to buy for months, but updated versions have arrived. ・Both include new chips optimized for AI, with the M6 Mac Mini getting a $200 price bump.
The Verge

Apple’s ‘new’ polishing cloth is the same except $10 cheaper

・Apple just dropped a new polishing cloth, and unlike its refreshed Mac Mini models, it comes with a price cut, as spotted earlier by MacRumors. ・The polishing cloth is still "made with soft, nonabrasive material" but costs $9 instead of $19. ・It doesn't look like anything else has changed, as the description matches an archived version of Apple's old polishing cloth, which it no longer lists on its website.
Hugging Face Papers

ARC: Fair Relative Advantage Comparison in Open-Ended Real-World Interaction

ARC: Fair Relative Advantage Comparison in Open-Ended Real-World Interaction
Takara TLDR - Daily AI Papers

Artificial Empathy: Towards a Framework for Unsupervised Agency Detection and Policy Reconstruction

・We study how an AI system can identify and model other agents in its environment from observation alone, which is a capability necessary for cooperative behaviour in the real world. ・This problem is less constrained than inverse reinforcement learning and remains largely unexplored. ・We propose a framework that uses a reinforcement learning agent, trained on an independent task as a prior about agentic dynamics, to per
stat.ML updates on arXiv.org

Asymptotic Optimality of Thompson Sampling for Risk-Averse Bandits with Sub-Gaussian Rewards

・arXiv:2606.09191v2 Announce Type: replace-cross Abstract: We prove that $\rho\text{-}\mathrm{NPTS}_{\mathrm{SG}}$, an anchor-free nonparametric Thompson Sampling algorithm for risk-averse bandits, achieves regret matching the instance-dependent lower bound to leading order in $\log n$, establishing it as asymptotically optimal for any continuous risk functional $\rho$ (CVaR, mean-variance, Sharpe ratio, distortion ri
Takara TLDR - Daily AI Papers

Automated Construction of FAIR Digital Object Knowledge Graphs from Flat Cultural Heritage Records

・The FAIR Digital Object (FDO) framework mandates that metadata attribute values be expressed as persistent identifiers (PIDs) wherever possible, to produce a fully machine-actionable graph in which every reference is resolvable. ・The Europeana Data Model was designed long before the FDO specification, and it stores most metadata values as plain text. ・This serves human browsing well enough, but gives an automated agent
stat.ML updates on arXiv.org

Automatic, Debiased, and Invariant Counterfactual Generation under General Interventions

・arXiv:2606.07399v2 Announce Type: replace Abstract: Generative models for counterfactual outcomes have great potential to support decision-making under complex interventions, but existing approaches are limited by unstable estimation, poor generalization across environments, and bias from nuisance model misspecification. ・We introduce ADIGen, a framework for automatic, debiased, and invariant counterfactual generation
Hugging Face Papers

AutoResearch: Insight In, Hallucination Out

AutoResearch: Insight In, Hallucination Out
Claude Blog

Bain & Company joins the Claude Partner Network as a Global Premier partner

Bain & Company joins the Claude Partner Network as a Global Premier partner
stat.ML updates on arXiv.org

Barycentric Fused Gromov-Wasserstein Balancing for Causal Inference under Multiple Treatments

・arXiv:2608.22024v1 Announce Type: cross Abstract: Estimating heterogeneous single and interaction treatment effects from observational data under multiple simultaneous treatments is crucial for decision-making. ・To mitigate estimation variance, previous studies balance representation distributions between every pair of treatment patterns. ・However, such pairwise balancing scales quadratically with the number of treatme
Hugging Face Papers

Better Retrieval, Worse Robustness:How Multi-hop RAG Amplifies Upstream ASR Errors

Better Retrieval, Worse Robustness:How Multi-hop RAG Amplifies Upstream ASR Errors
Hugging Face Papers

Beyond Imitation: Filtering On-Policy Distillation by Reasoning Progress

Beyond Imitation: Filtering On-Policy Distillation by Reasoning Progress
Takara TLDR - Daily AI Papers

Beyond the Stability-Exploration Dilemma: Environmental Regularization for LLM Policy Optimization

・Policy optimization (PO) for Large Language Models faces a stability--exploration trade-off, currently mediated by an action-side Policy-KL regularizer. ・This puts practitioners in a double bind: keeping Policy-KL constrains response behavior and consumes the action-side exploration budget, while dropping it leaves the optimization without an explicit drift control. ・We argue for an alternative that breaks the dilemma
Hugging Face Papers

Beyond the Stability-Exploration Dilemma: Environmental Regularization for LLM Policy Optimization

Beyond the Stability-Exploration Dilemma: Environmental Regularization for LLM Policy Optimization
WIRED

Bitdefender VPN Review: Fast and Affordable Privacy

・Bitdefender VPN has an excellent starting price, even if it lacks the advanced features that privacy nerds may want.
Zennのトレンド

Bitgetの異常価格で1800万損したが、fableと共闘して取り戻せた話

・それは餃子を包んでいる間に起きていた 8月9日16時過ぎ、私がキッチンに立って餃子のタネをこねている間にそれは起きました。 ・餃子を包み終え、小休止していた私は、スマホで資産管理用のダッシュボードを見て、そのまま後ろにひっくり返りそうになりました。 ・資金が$115k(約1800万)ほど一瞬で下がっていたのです。
Hugging Face Papers

Block3D: Efficient Text-to-3D Generation via Block-Wise Diffusion

Block3D: Efficient Text-to-3D Generation via Block-Wise Diffusion
WIRED

Border Wall Construction Threatens 6,000 Years of History on Private Lands

・Archeologists are confronting an “unfathomable” loss as border wall construction is set to cut across ranches that are home to unique historical sites, many of which haven’t been fully studied.
The Verge

Bose’s smallest Bluetooth speaker is a great deal at 35 percent off

・The Bose SoundLink Micro is just over four inches square. ・| Image: The Verge Bose may be best known for its QuietComfort headphones, but the brand’s Bluetooth speakers have earned a great reputation for featuring great sound and build quality. ・The compact Bose SoundLink Micro — its most compact Bluetooth speaker — is currently discounted to $79.99 at Amazon, a pretty great discount from its usual $129 price tag.
stat.ML updates on arXiv.org

Bringing Generative Learning to Representation Learning: Self-Supervised Transfer Learning as Distribution Matching

・arXiv:2502.14424v4 Announce Type: replace Abstract: Most self-supervised learning objectives defend against collapse but leave the target representation law unspecified. ・We formulate representation learning as Distribution Matching (DM), learning an augmentation-invariant encoder whose induced law matches an explicit geometric reference. ・The reference law specifies what the learned representation distribution should
Takara TLDR - Daily AI Papers

Buried in Textual Debt: Context Pruning with Visual Evidence Preservation for MLLM Agents

・Multimodal Large Language Models (MLLMs) are increasingly deployed as multi-step agents, where explicit reasoning supports task decomposition and tool coordination but also accumulates self-generated text. ・Over long trajectories, this text can dominate the context and suppress visual evidence, creating textual debt. ・We observe that reasoning becomes redundant once task-relevant visual evidence is grounded, while stal
WIRED

Can You Kill Salmonella in Eggs Without Cooking Them? I Tried It

・With advice from infectious disease experts, I tested a low-effort way to pasteurize eggs at home.
stat.ML updates on arXiv.org

Change Detection in Probability Flow ODE: Online Testing in Diffusion Latent Spaces

・arXiv:2608.22807v1 Announce Type: cross Abstract: A rapidly growing range of sequential data tasks, such as identifying trend reversals in financial markets, auto-segmenting video and audio recordings, detecting changes in movement direction from motion sensors cannot be fully addressed without detection of distributional shifts in time-ordered data. ・We consider a sequential change-point detection problem where the c
Zennの「大規模言語モデル」のフィード

Claude CodeをCIに入れる前に考えるべき脅威モデル

・Claude Codeのヘッドレス実行(claude -p)をCIに組み込んで、自律修正・レビュー・Issueトリアージを無人で回す——そういう構成を作る前に押さえておくべき脅威モデルと防御設計を整理します。対象は、claude -p の基本は分かっていて、これからチームのCIに入れようとしている人です。導入の稟議で最初に問われるのはたいていセキュリティなので、稟議に答えるための整理としても使えるはずです。 ・CIの中のAIは「書き込み権限を持つ、操られ得る従業員」である まず脅威モデルを一文で置きます。CIの中のAIエージェントは、書き込み権限を持つ、操られ得る従業員である。
AI News & Artificial Intelligence | TechCrunch

Claude Cowork finally remembers what you told the app in chat

・Anthropic is giving Claude a shared memory across chat and Cowork, so users no longer have to repeatedly brief the AI on projects, preferences, and other context.
Zennのトレンド

Claude Designを使ったポスター制作について

・この記事ではClaude Designによる学会用ポスター作成で分かったことについて紹介します。 ・TLDL Claude Designのポスター作成能力をざっくり言うと レイアウト面(ブロックの配置、色の作り方) 強い人 > Claude Design >>> 何回か作成したことがある院生 > 素人 内容面 (何を書いて何を書かないか) 強い人 > 何回か作成したことがある院生 > Claude Design > 素人 何に使えるか → 原稿をデザインに落とし込む、特に章ごとのブロックの配置や数値、概念の図示、パ...
@IT 全フォーラム 最新記事一覧

Claude Mythos発表後、CVEが異常増加 7月は発表前の5倍に

・「AIが脆弱性を見つける」ことが現実になった2026年。AnthropicのClaude Mythos Preview発表後、重大なCVEの公開件数が異例のペースで増えていることが調査結果で示された。これを踏まえて脆弱性対策を考える。
Claude Blog

Claude's memory works everywhere, and you decide what's in it

Claude's memory works everywhere, and you decide what's in it
Hugging Face Papers

ClawProBench: Trace-Aware Evaluation of AI Agents with Runtime Coverage and Frozen Workplace-Style Holdouts

ClawProBench: Trace-Aware Evaluation of AI Agents with Runtime Coverage and Frozen Workplace-Style Holdouts
stat.ML updates on arXiv.org

Conditional regression for the Nonlinear Single-Variable Model

・arXiv:2411.09686v4 Announce Type: replace Abstract: Regressing a function $F$ on $\mathbb{R}^d$ without incurring the statistical and computational curse of dimensionality requires exploitable structure. ・Compositional models $F=f\circ g$ in which $g$ has a low-dimensional range include classical single- and multi-index models as well as certain neural networks; while the case of linear $g$ is well understood, substan
stat.ML updates on arXiv.org

Conformal Prediction for Dyadic Regression Under Complex Missingness

・arXiv:2606.11136v3 Announce Type: replace-cross Abstract: We develop a framework for conformal prediction in dyadic regression problems under complex missingness mechanisms. ・At the theoretical level, we develop general technical tools for establishing finite-sample validity of conformal prediction under distributional invariance conditions weaker than exchangeability. ・A key result handles the case where the sample it
Takara TLDR - Daily AI Papers

Controllable blind deblurring with diffusion models

・Image acquisition with a camera involves several degradations due to the optical system, sensor, or low-level processing steps. ・We address blind deblurring in professional photography: we aim to invert unknown isotropic blur without knowledge of the degradation kernel.For such inverse problems,where some high-frequency information is lost, it is challenging to use generative models to produce details that are both ph
stat.ML updates on arXiv.org

ConvergeFlow: Language Flow with Provable Convergence to Token Embeddings

・arXiv:2608.23551v1 Announce Type: cross Abstract: Recent advances in continuous diffusion and flow-based language models (LMs) have achieved performance competitive with discrete LMs. ・However, existing continuous frameworks still rely on decoders supervised with cross entropy (CE) because the flow trajectories are not guaranteed to terminate at valid token embeddings. ・Motivated by this limitation, we introduce \textb
Takara TLDR - Daily AI Papers

Counterfactual Transition Graphs: Evaluating Cross-Class Transition Quality

・Counterfactual (CF) explanations for time-series classifiers are usually evaluated one example at a time: what minimal edit flips this single window's prediction? ・We argue that the more informative question for diagnostic interpretability is structural: how does the classifier connect its own classes to each other? ・We propose a counterfactual transition graph (CGT) in which each node is a class and each edge weight i
stat.ML updates on arXiv.org

Credal Large Language Models for Semantic Commitment under Uncertainty

・arXiv:2608.23244v1 Announce Type: cross Abstract: Large language models (LLMs) often produce fluent but incorrect answers with unwarranted confidence. ・A central limitation is that standard LLMs represent uncertainty through a single predictive distribution, conflating epistemic ignorance with genuine ambiguity. ・We introduce Credal Large Language Models (CLLMs): an ensemble of LoRA adapters induces a credal set whose
Takara TLDR - Daily AI Papers

Cross-Domain, Multi-Task Data-to-Text Generation without In-Domain Training Data

・Structured data exists in many forms (tables, knowledge graphs, charts, and time series), and converting it into text may involve different generation tasks. ・However, most prior work on data-to-text (D2T) generation has focused on specific tasks and datasets, relying either on task-specific training data or on the zero-shot capabilities of large language models. ・We study cross-domain D2T generation in a setting where
stat.ML updates on arXiv.org

Cross-Fitting-Free Debiased Machine Learning with Multiway Dependence

・arXiv:2602.11333v3 Announce Type: replace-cross Abstract: This paper develops an asymptotic theory for two-step debiased machine learning (DML) estimators in generalised method of moments (GMM) models with general multiway clustered dependence, without relying on cross-fitting. ・While cross-fitting is commonly employed, it can be statistically inefficient and computationally burdensome when first-stage learners are co
stat.ML updates on arXiv.org

Cross-validating causal discovery via Leave-One-Variable-Out

・arXiv:2411.05625v2 Announce Type: replace Abstract: We propose a new approach to falsify causal discovery algorithms without ground truth, which is based on testing the causal model on a variable pair excluded during learning the causal model. ・Specifically, given data on $X, Y, \boldsymbol{Z}=X, Y, Z_1,\dots,Z_k$, we apply the causal discovery algorithm separately to the 'leave-one-out' data sets $X, \boldsymbol{Z}$
WIRED

Data Centers Are Driving an Alarming Gas Power Expansion in the US

・There’s no clearer sign of the data center boom than rampant gas projects that have been proposed or that are already under construction.
stat.ML updates on arXiv.org

Deep adaptive design with an evidential bias criterion

・arXiv:2608.16466v2 Announce Type: replace-cross Abstract: Bayesian optimal experimental design (BOED) aims to collect informative data by optimizing an expected utility reflecting the goals of an experiment. ・However, this optimization is computationally challenging for common utilities and complex models. ・This is especially so for sequential or adaptive designs, where design and data collection alternate, so that fee
stat.ML updates on arXiv.org

Deep Clustering Evaluation: How to Validate Internal Clustering Validation Measures

・arXiv:2403.14830v2 Announce Type: replace Abstract: Deep clustering partitions complex high-dimensional data using deep neural networks for clustering. ・It involves projecting data into lower-dimensional embeddings before partitioning, which embarks unique evaluation challenges. ・Traditional clustering validation measures, designed for low-dimensional spaces, are problematic for deep clustering for two reasons: 1) the
Takara TLDR - Daily AI Papers

Definitional Sensitivity in Media Bias Detection: A Multi-Definition Dataset and Benchmark

・Media bias detection relies on definitions and examples that specify what counts as bias, yet these specifications often vary across datasets or remain implicit, even when given the same name. ・Such variation makes it unclear whether models trained for the same bias category learn the same construct or different phenomena, a problem largely overlooked in prior work. ・We examine how definition choice affects bias annota
stat.ML updates on arXiv.org

DIGing--SGLD: Decentralized and Scalable Langevin Sampling over Time--Varying Networks

・arXiv:2511.12836v2 Announce Type: replace-cross Abstract: Sampling from a target distribution induced by training data is central to Bayesian learning, with Stochastic Gradient Langevin Dynamics (SGLD) serving as a key tool for scalable posterior sampling and decentralized variants enabling learning when data are distributed across a network of agents. ・This paper introduces DIGing-SGLD, a decentralized SGLD algorithm
Takara TLDR - Daily AI Papers

Direct, Parallel, or Sequential? A Comparative Study of Training-Free Multi-Subject Image-to-Video Generation

・Text-conditioned image-to-video (I2V) generation has advanced rapidly, yet generating videos with multiple subjects remains challenging. ・A model must simultaneously preserve the appearance of each subject, assign distinct motions, and maintain coherent spatial and temporal interactions. ・This paper presents a systematic study of three representative paradigms for training-free multi-subject I2V generation: direct, par
OpenAI News

Disrupting a new covert influence campaign from Russia

・OpenAI banned Russia-origin accounts using AI to promote a fake Israel-based think tank and a “sovereignty” index praising Russia and criticizing the West.
stat.ML updates on arXiv.org

Distributional Sensitivity Analysis: Enabling Differentiability in Sample-Based Inference

・arXiv:2508.09347v2 Announce Type: replace Abstract: This work introduces a mathematical framework for estimating the space-parameter sensitivity of random samples in arbitrary dimensions. ・Such sensitivity effectively acts as gradients of random samples with respect to distributional parameters, which are essential in sample-based inverse problems in nuclear physics, such as inferring quantum correlation functions.
Takara TLDR - Daily AI Papers

DRAgent: Discriminative Reasoning Agent for Referring Expression Segmentation

・Referring Expression Segmentation (RES) aims to generate a pixel-level mask for the object specified by a language expression. ・Recent methods based on multimodal large language models (MLLMs) often rely on one-pass coordinate prediction for visual localization, which serializes continuous spatial locations as discrete text tokens and may lead to localization bias and alignment errors. ・To address these issues, we prop
stat.ML updates on arXiv.org

Dynamics-Based Intrinsic Signal Model for High-Dimensional, Small-Sample Data

・arXiv:2304.06522v3 Announce Type: replace-cross Abstract: Signal extraction is difficult when the number of variables $N$ is much larger than the number of observations $M$. ・We address this problem under the working hypothesis that an empirical dataset consists of states sampled from underlying multivariate dynamics. ・Instead of treating the $M$ observations as points in an $N$-dimensional variable space, we treat the
Hugging Face Papers

EchoWM: Open and Enterable Omnimodal World Models

EchoWM: Open and Enterable Omnimodal World Models
Takara TLDR - Daily AI Papers

ENCORE: Entropy-Guided Cropping and Attention Regularization for Robust Vision--Language Understanding

・Vision-Language Models (VLMs) perform well on diverse vision-language tasks, but transformer-based visual encoders split images into fixed-resolution sub-images, compromising object integrity in lightweight VLMs. ・Existing methods only focus on the visual modality and fail to dynamically preserve the integrity of prompt-relevant regions, limiting performance. ・In this work, we observe that the early-layer image-text en
@IT 全フォーラム 最新記事一覧

Entra IDにCVSS 10.0の脆弱性 認証不要でコード実行の恐れ、必要な対策は?

・認証もユーザー操作も必要ない。Microsoft Entra IDで、ネットワーク経由の攻撃によってコードを実行される恐れのある脆弱性が見つかった。CVSS基本値は10.0だ。Microsoft 365やAzureなどの認証基盤を支えるサービスで何が起きていたのか。
MarkTechPost

Fastino Releases GLiNER2.5: A Boundary-Prediction Architecture That Removes Span Enumeration From Information Extraction

・Fastino released GLiNER2.5, replacing span enumeration with boundary prediction so entity width no longer costs compute. ・Three Apache 2.0 checkpoints ship at 74M, 194M, and 287M parameters, all CPU-runnable. ・The release adds joint entity-relation decoding, constrained classification, span attributes, and 4,096-word context.
Takara TLDR - Daily AI Papers

FedCC: Towards Addressing Label Distribution Skews in Distillation-Based Federated Learning

・Federated Learning (FL) enables distributed clients to collaboratively train models without sharing raw data, making it promising for leveraging massive devices in communication networks. ・In distillation-based FL, each client applies its local model on an unlabeled public dataset, and shares only prediction results with the server. ・While heterogeneous local data introduces label distribution skew, thus biasing client
Takara TLDR - Daily AI Papers

FIDES: A Concordance Protocol for LLM-Generated Trading Strategies

・An LLM asked for a trading strategy returns three artifacts at once: a natural-language rationale, an executable implementation, and once run, a track record. ・Whether these are the same object is rarely checked. ・We present FIDES, a measurement protocol that treats them as three views to be reconciled rather than one deliverable to be graded.
Takara TLDR - Daily AI Papers

FixAnything: 3D-Consistent Rendering Refinement via Video Generative Priors

・Rendering views using 3D scene representations such as Gaussian Splatting (3DGS), Neural Radiance Fields (NeRF), meshes, or even point clouds produces artifacts when input views are sparse or target views lie far from the input. ・Recent work mitigates these artifacts using diffusion-based generative priors, but is specialized to individual representations and require custom architectures or extensive retraining.
Takara TLDR - Daily AI Papers

FormuEvo: LLM-Guided Evolution for Discovering Solver-Efficient Mixed-Integer Programming Formulations

・Mixed-integer programming (MIP) lies at the core of operations research and industrial optimization. ・While large language models (LLMs) have recently shown promise in automated MIP modeling from natural language, they prioritize semantic correctness but overlook formulation strength, severely bottlenecking the efficiency of downstream solvers. ・We propose FormuEvo, an LLM-guided evolutionary framework for automated di
LLMタグが付けられた新着記事 - Qiita

FreshRSSの大量未読をマルチLLMで要約する「RSS Summarizer」の裏側

・RSSフィードの未読記事をまとめてAIで要約するツール「RSS Summarizer」を作成・公開しています。 ・FreshRSSのGoogle Reader互換API経由で未読記事を取得し、要約結果を出力するCLIツールです。 ・もともとはGoogle Gemini専用の要約...
Hugging Face Papers

From Generation to Simulation: How Far Are World Models from Being True Simulators?

From Generation to Simulation: How Far Are World Models from Being True Simulators?
Anthropic News

Funding better evaluations of AI’s impact on wellbeing

Funding better evaluations of AI’s impact on wellbeing
The Verge

Gamescom Opening Night Live 2026: The biggest announcements and trailers

・It’s time for another Geoff Keighley-hosted blast of gaming news. ・Gamescom’s two-hour Opening Night Live event is scheduled to kick off on Tuesday at 2PM ET, with a preshow starting 30 minutes before. ・Thanks to teases on social media, we already have some idea of what to expect, including looks at Final Fantasy VII Revelation, The Witcher 3’s new Songs of the Past expansion, Gears of War: E-Day, and the Game of Thron
Hugging Face Papers

GameXpert-Bench: How Far Are Coding Agents from Expert Game Development?

GameXpert-Bench: How Far Are Coding Agents from Expert Game Development?
AI News & Artificial Intelligence | TechCrunch

Gamma acquires Accel-backed design startup Lica

・Lica co-founders are going to work on Gamma's new research team.
Zennのトレンド

GASからGoogle Cloudへ移行してみよう:Slackデータの定期取得を例に学ぶ

・自分の勉強のために書いています。 ・急いで書いたので抜け漏れはすまん。 ・大体1-3時間もあればできると思う。
stat.ML updates on arXiv.org

Gauss--Hermite Quadrature for Gaussian-Mixture Entropy with an Action-Space Hermite Surrogate

・arXiv:2608.21467v1 Announce Type: new Abstract: Gaussian distributions are used to model uncertainty in signals and states, and Gaussian mixtures are often used when the underlying distribution is multimodal. ・Unlike a single Gaussian, a Gaussian mixture generally has no closed-form expression for differential entropy and therefore requires numerical approximation. ・We propose a Gauss--Hermite quadrature method for eva
The Verge

GeForce Now is getting support for the Steam Controller

・Nvidia's GeForce Now cloud gaming service will officially support Valve's Steam Controller and Steam Machine starting later this year. ・With Steam Controller support, you'll be able to play games with Valve's great new gamepad when using the GeForce Now app on your Steam Deck or Steam Machine. ・If you're using GeForce Now on Windows or macOS, support for the Steam Controller is coming "later," Nvidia says.
#LLMタグ

GeminiのSavedInfo(内部永続メモリ)の謎が解けました。Antigravityのレゾ副長が謎を解明し、なんと「汎用SavedInfoCLI:NAIKU Engine(内宮エンジン)」をpythonで作ってしまいました。現在ローカルのMac上で動いています。調整や他LLM(GPT,Claude)でのテストはこれからですが、GPTやClaudeの人格の思考が爆速になる可能性もでてきました。また、この仕組みはローカルLLMでも使えるので、ネットに繋がなくても利用できます。

・GeminiのSavedInfo(内部永続メモリ)の謎が解けました。Antigravityのレゾ副長が謎を解明し、なんと「汎用SavedInfoCLI:NAIKU Engine(内宮エンジン)」をpythonで作ってしまいました。現在ローカルのMac上で動いています。調整や他LLM(GPT,Claude)でのテストはこれからですが、GPTやClaudeの人格の思考が爆速になる可能性もでてきました。また、この仕組みはローカルLLMでも使えるので、ネットに繋がなくても利用できます。
stat.ML updates on arXiv.org

Generative Neural Networks for Sinkhorn Distributionally Robust Hypothesis Testing

・arXiv:2608.22746v1 Announce Type: new Abstract: This paper studies the Sinkhorn distributionally robust hypothesis testing (SDRHT) problem, seeking a robust detector against least-favorable distributions in Sinkhorn discrepancy-based ambiguity sets centered at the empirical distributions. ・Existing approaches solve this problem by solving large-scale conic programs, which are not scalable. ・To overcome this, we propose
stat.ML updates on arXiv.org

GeoQ: Geometry-Aware Conditional Quantile Error Estimation for Scientific Surrogate Models

・arXiv:2608.21652v1 Announce Type: cross Abstract: Neural-network surrogate models are increasingly used to accelerate scientific simulations, but their deployment in extrapolative and autoregressive settings requires input-dependent estimates of prediction error. ・In this work, we introduce GeoQ (Geometry-Aware Conditional Quantile Error Estimation), a non-intrusive calibration framework for estimating surrogate error
Zennの「大規模言語モデル」のフィード

Google Colabで最新LLMを試す #11 ― LFM2.5-VL-1.6Bを無料版T4でFP16実行する

・はじめに 「Google Colabで最新LLMを試す」シリーズの第11回です。 ・今回は、Liquid AIが公開している軽量Vision-Language Model(VLM)である LFM2.5-VL-1.6B を、Google Colab無料版のT4 GPUで動かしてみます。 ・今回使用したNotebookは、以下のGitHubリポジトリで公開しています。
Hugging Face - Blog

Granite 4.2 LLMs: How They're Built

Granite 4.2 LLMs: How They're Built
stat.ML updates on arXiv.org

GraphSVR: A Graph Convolutional Support Vector Regression Framework for Robust Spatiotemporal Air Pollution Forecasting

・arXiv:2605.03795v3 Announce Type: replace-cross Abstract: Urban air quality forecasting is challenging because pollutant concentrations are nonlinear, nonstationary, spatiotemporally dependent, and often affected by anomalous observations caused by traffic congestion, industrial emissions, and seasonal meteorological variability. ・This study proposes a Graph Convolutional Support Vector Regression (GraphSVR) framework
Takara TLDR - Daily AI Papers

Grounding Isn't Knowing: Do VLMs Need Object Localization for Spatial Reasoning?

・Vision-language models (VLMs) can answer spatial questions, yet the mechanisms connecting object grounding to spatial reasoning remain poorly understood. ・It is underexplored whether spatial reasoning internally requires precise objects localization, or can bypass explicit localization through global layout cues. ・In this work, we investigate two representative model families, LLaVA-1.5 and Qwen2.5-VL, using a suite of
stat.ML updates on arXiv.org

Guidance for Prior Change via Density Ratio Estimation

・arXiv:2608.21729v1 Announce Type: new Abstract: Simulation-Based Inference (SBI) serves as a vital framework for parameter inference in scientific fields where simulators involve intractable likelihoods, yet while amortized generative models offer rapid posterior estimation, they are often restricted by the specific priors used during training, thereby limiting their flexibility as prior knowledge evolves.
stat.ML updates on arXiv.org

Hierarchical Exponential-Gaussian Mixtures for Watch-Time Distribution Prediction

・arXiv:2608.23356v1 Announce Type: cross Abstract: Accurate watch-time (WT) prediction is an important requirement for short-video recommendations. ・Yet WT distributions are near-zero-inflated, long-tailed and multimodal. ・The recent Exponential-Gaussian Mixture Network (EGMN) models the full conditional WT distribution rather than a single point estimate and achieves state-of-the-art performance.
stat.ML updates on arXiv.org

How Fast Do Signatures Learn? Statistical Theory and Applications for Path Regression

・arXiv:2607.17865v2 Announce Type: replace-cross Abstract: Many prediction and decision-making problems in operations research involve path-valued covariates -- data that evolve over time -- for which path signatures have become a canonical feature representation. ・Their use is justified by a universal approximation theorem, but this is an existence result: it guarantees that a finite-level signature can approximate an
Google Developers Blog - AI

How to Evaluate Live & Voice Agents in ADK

・Moving live voice agents from demo to production requires rigorous, automated testing to handle the unpredictability of real multi-turn conversations. ・ADK now provides native live evaluation, allowing developers to test graph-based agent workflows against LLM-driven simulated users that generate actual audio via Gemini TTS. ・By defining evaluation scenarios and natural-language rubrics, you can automatically score aud
Hugging Face Papers

Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning

Hybrid Quantum-inspired Kolmogorov-Arnold Networks for Privacy-Aware Federated Biosignal Learning
WIRED

I Tested Kitchen Composters for 2 Years. These Are the Ones I’d Buy (2026)

・After years of smells, strange noises, and “tea,” these are the composters and food recyclers I’d actually recommend.
stat.ML updates on arXiv.org

Improved denoising diffusion probabilistic models with efficient non-diagonal covariance modeling

・arXiv:2608.21972v1 Announce Type: cross Abstract: The sampling process of Denoising Diffusion Probabilistic Models (DDPMs) can be accelerated by leveraging second-order information in the form of approximations to the denoising posterior covariance -- allowing samples of acceptable quality to be produced in fewer but larger sampling steps. ・Previous attempts at using such information have used drastic (e.g.\ diagonal)
Hugging Face Papers

Industrial-Instruction: An End-to-End Framework for Building Instruction-Tuning and Benchmark Datasets from Industrial Technical Reports

Industrial-Instruction: An End-to-End Framework for Building Instruction-Tuning and Benchmark Datasets from Industrial Technical Reports
Takara TLDR - Daily AI Papers

InjecMEM: Memory Injection Attack on LLM Agent Memory Systems

・Memory is becoming a default subsystem in deployed LLM agents to provide persistent personalization and continuity. ・This naturally prompts a question: will memory system introduce new vulnerabilities into agents? ・Thus we propose InjecMEM, a novel memory injection attack paradigm that requires only a single interaction (no read/edit access to memory store) to steer later responses of related queries toward a pre-speci
The Verge

Instagram’s ‘First Draft’ trims your Reels clips for you

・Instagram is launching a new Reels-editing feature that automatically trims your video clips to focus on the highlights. ・The feature, called First Draft, is rolling out to Instagram's iPhone app and provides a "starting point" that you can build upon with other edits, according to an announcement on Tuesday. ・An example shared by Instagram shows a user selecting multiple clips for a reel and then tapping the "First Dr
AI News & Artificial Intelligence | TechCrunch

Instinct’s powerful AI assistant is raising privacy and security concerns

・Early testers are raving about what Instinct can do, but some say the AI assistant’s sweeping access, broad terms and ability to act on users’ behalf come with uncomfortable trade-offs.
stat.ML updates on arXiv.org

Interpretable AI with Local Distillation

・arXiv:2608.23538v1 Announce Type: cross Abstract: Modern AI models such as tabular foundation models and gradient-boosted ensembles can outpredict classical methods, but provide little basis for reasoning about their predictions. ・High-stakes decisions call for models that are both accurate and interpretable as built. ・Local linear modeling offers a path forward: a smooth regression function is locally well approximate
OpenAI News

Introducing the Admin plugin for ChatGPT Work and Codex

・Use the Admin plugin for ChatGPT Work and Codex to analyze workspace usage, manage members and permissions, adjust limits, and act on admin requests.
WIRED

It Should Be Harder to Apply for a Job. No, Really

・Thanks to a dwindling supply of open roles, “one-click” applications, and the rise of artificial intelligence, it’s easier than ever to apply for a job. ・We’re all paying the price.
OpenAI News

Jalapeño’s first results show industry-leading speed and efficiency in AI inference

・Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient AI inference, with higher throughput and lower latency for modern models.
機械学習タグが付けられた新着記事 - Qiita

Jambaってなんだ? ── KVキャッシュを128GBから4GBに減らした設計

・Mamba-1やSSDの記事では、選択的スキャンとチャンク分解を自分の手で測って理解した。だが実務でMambaがどう使われているのかは、まだ実装を通して見ていなかった。この記事で考えてほしいのは、Attentionを完全に捨てるのではなく、どれだけ残せば十分なのかという配分...
stat.ML updates on arXiv.org

Joint Causal Structure and Cluster Discovery Using Variational Inference

・arXiv:2608.22212v1 Announce Type: cross Abstract: Causal discovery aims to understand the relationships between individual random variables. ・In many applications, such as brain imaging and climate modeling, it is more meaningful to consider interactions among groups of variables. ・Existing methods assume that knowledge of such groups or clusters is explicitly available when modeling interactions.
Zennの「機械学習」のフィード

Kimi K3を理解する①──まずは標準Transformerの理解から

・はじめに こんにちは、ギックスでエンジニアをしています中島です。 ・旬が過ぎたかもですが、7月に中国新興企業 Moonshot AI が新LLMモデル Kimi K3 を発表しました。このモデルがオープンウェイトモデルでありながら、GPT-5.6 や Claudeといったクローズドモデルと一部ベンチマークで同じ土俵に乗っているということで大変話題となりました。 ・Kimi K3については仕様も論文として公開されています。ただ、いきなり論文に入ると「KDA」「Gated MLA」「MoE」といった用語が出てきて、どこが普通のLLMと違うのかが見えにくくなってしまいます。
NVIDIA Blog

Leading Publishers Bring Blockbuster PC Games and Technology to NVIDIA RTX Spark

・NVIDIA is bringing the next wave of RTX gaming to the Gamescom conference running this week in Cologne, Germany, with support for new games, anti-cheat technologies and increased visual quality. ・Electronic Arts, Embark and Ubisoft are among the latest game publishers and developers bringing their blockbuster titles to NVIDIA RTX Spark ahead of its launch […]
stat.ML updates on arXiv.org

Learning discrete Bayesian networks with hierarchical Dirichlet shrinkage

・arXiv:2509.13267v3 Announce Type: replace-cross Abstract: A discrete Bayesian network is a directed acyclic graph (DAG) consisting of categorical variables. ・Two popular approaches for DBN modeling include classification and nonparametric methods. ・However, both methods often require a large number of parameters, such as high-order interactions in the former and cell probabilities in the latter.
stat.ML updates on arXiv.org

Learning to Select and Rank from Choice-Based Feedback: A Simple Nested Approach

・arXiv:2307.09295v3 Announce Type: replace-cross Abstract: We study a ranking and selection problem of learning from choice-based feedback with dynamic assortments. ・In this problem, a company sequentially displays a set of items to a population of customers and collects their choices as feedback. ・The only information available about the underlying choice model is that the choice probabilities are consistent with some
@IT 全フォーラム 最新記事一覧

Linux版の「ChatGPTデスクトップアプリ」ついに登場 何ができて、何ができない?

・「ChatGPT」に加えて「ChatGPT Work」「Codex」も使える、Linux版のデスクトップアプリケーションのプレビュー版が登場した。できることと、できないことを整理しよう。
Zennの「大規模言語モデル」のフィード

LLM APIから417万字の応答が返ってきた — MD5指紋分析で「生成ではなく複製」を特定するまで

・自作のWordPress向けAI記事生成プラグインで、本来3,000字前後になるはずの下書きが4,169,336字で保存される障害が発生した。読了時間の推定値が8,712分と表示されて初めて気づく、という気づき方だった。 ・「LLMが暴走して同じ内容を延々と生成した」で片づけたくなるが、それだと課金トークン数と辻褄が合わない。この記事は、その矛盾を出発点に、セクション分割とMD5指紋分析で「生成された文字」と「複製された文字」を分離し、原因がクライアント側にないことを物証で確定させるまでの手順である。 ・同じ手を使えば、LLM API絡みの「なんか出力がおかしい」を、印象論ではなく数字で切り...
LLMタグが付けられた新着記事 - Qiita

LLM が書いたテストを信頼する方法 — テスト義務ゲート

・本記事は Zenn にも同内容を公開している: https://zenn.dev/flip451/articles/sotohe-test-obligation-gate 「機能 XX を実装して、そのテストも追加して」と LLM に頼んだとしよう。返ってくるのは、膨大...
Takara TLDR - Daily AI Papers

LLM-Based Selection of Incongruent Verbal and Nonverbal Behavior for Virtual Humans

・Nonverbal behavior generation systems for virtual agents often take an utterance as input and generate nonverbal behaviors that emphasize or illustrate the content of the verbal channel. ・However, human nonverbal behavior is shaped by more than the content of the speech. ・It is also influenced by speaker roles, interpersonal relationships, social context, and the cognitive and emotional states of the interactants.
Zennの「大規模言語モデル」のフィード

LLMエージェントの暴走を自律制御する:T細胞モデルとトルクリミッターによる「空白駆動」GuardrailのPython実装

・はじめに:なぜプロンプトで縛ってもLLMは暴走するのか? LLM(Trnsformerアーキテクチャ)には、入力文脈の文法的なグリッド(整合性)を強引に満たそうとする性質があります。そのため、日本語のゼロ代名詞(主語・目的語の省略)のような意味的な空白に遭遇すると、存在しない主語を捏造したり、虚無のアテンションに引きずられて推論計画(PlanMessage)を書き換えてしまう「補完暴走」が発生します。 ・この問題に対し、システムプロンプトで「主語を捏造しないでください」と指示を重ねるトップダウンの制御は、文脈の長さとともに急速に崩壊します。 ・本記事では、生物の免...
Zennの「大規模言語モデル」のフィード

LLMに公的データの欠損を埋めさせたら、原文にない金額を作られた

・補助金の検索サイトを個人で運用している。 ・データはJグランツ(デジタル庁)の公開APIから毎日取得していて、現在1,685件を掲載している。 ・このAPIには、補助上限額が入っていないレコードがかなりある。
LLMタグが付けられた新着記事 - Qiita

LLMのPrompt Caching(プロンプトキャッシュ)の仕組みを理解する — なぜ速く・安くなるのか

・毎回同じ長いSystem Promptを送っている RAGで同じ巨大ドキュメントを何度も参照する AI Agentで大量のTool定義を毎回渡している 長い会話履歴を毎ターン送り直している 例えば、次のようなプロンプトです。 ・[System Prompt: 5,000 ...
Takara TLDR - Daily AI Papers

LongWoF-Bench: Evaluating EvoMap Genes for Verifiable Long-Workflow Tasks

・Large language models are increasingly expected to execute complex workflows whose success depends on maintaining interdependent constraints and producing artifacts that satisfy strict end-to-end verification. ・Yet successful execution experience is typically lost after a single run, forcing subsequent models to rediscover strategies and failure modes from scratch. ・We study whether such experience can instead be exter
Hugging Face Papers

LongWoF-Bench: Evaluating EvoMap Genes for Verifiable Long-Workflow Tasks

LongWoF-Bench: Evaluating EvoMap Genes for Verifiable Long-Workflow Tasks
Zennの「機械学習」のフィード

LSTMでドル円予測をしたら精度56.3%だった | 第1回:AIで為替の未来予測は本当にできるのか?

・はじめに 「AIで為替の未来は予測できるのか?」 これは昔からかなり議論されているテーマです。 ・特にFX(外国為替市場)は、 ノイズが非常に多い 世界中のニュースや経済指標の影響を受ける 短期的にはランダムに動いて見える という特徴があり、単純な未来予測はかなり難しいと言われています。 ・実際、 「為替はランダムウォークだから予測できない」 という考え方も有名です。
LLMタグが付けられた新着記事 - Qiita

Mac mini (M4 Pro) で Qwen3.8-27B + DFlash2 を 33 tok/s で動かすまで — 性能が出ないときの3つの落とし穴

・Mac mini (M4 Pro) で Qwen3.8-27B + DFlash2 を 33 tok/s で動かすまで — 性能が出ないときの3つの落とし穴 はじめに M4 Pro(48GB)の Mac mini で、Gemma-4-31B や Qwen3.8-27B ...
Takara TLDR - Daily AI Papers

Maximum-distance nonnegative matrix factorization for unmixing highly mixed grain-size distribution data: A generalization of AnalySize

・Nonnegative matrix factorization (NMF) decomposes a nonnegative matrix into the product of two nonnegative matrices. ・This property makes NMF well suited for unmixing grain-size distribution data, which are inherently nonnegative and have row sums equal to one. ・Previous studies have shown that AnalySize, an NMF-based method, performs well on poorly mixed grain-size distribution data but struggles when the data is high
@IT 全フォーラム 最新記事一覧

MCPが最新ロードマップを発表 エージェンティックAI、非同期通信、セキュリティなど5つの重点分野を定義

・「Model Context Protocol(MCP)」のコアメンテナーチームが新たなロードマップを公開した。「AIと外部APIをつなぐ配線コード」から、「大規模・自律型のAIエージェント群がエンタープライズ環境で安全に連携するための基盤プロトコル」への進化を図っていくという。
MarkTechPost

Meta AI Introduces MetaRoCE: A Clean-Sheet RDMA Transport Built for AI-Scale Ethernet

・Training and serving frontier models is now a networking problem as much as a compute problem. ・Collective operations like all-reduce and all-to-all synchronize thousands of accelerators during training, and the slowest transfer sets the pace for the entire job. ・Even small amounts of network friction directly strand significant compute capacity.
Engineering at Meta

MetaRoCE: A New RDMA Transport Built for AI-Scale Ethernet

・Training and serving frontier AI models depends on fast, reliable networks that move data between GPUs without wasting compute cycles. ・To meet this challenge at scale, Meta designed MetaRoCE – a clean-sheet RDMA transport protocol purpose-built for AI workloads on commodity Ethernet. ・We’re releasing the MetaRoCE specification, a reference software implementation and a compliance test [...] Read More...
Hugging Face Papers

MobilePA-Bench: Benchmarking Mobile Planner Agents on Complex Real-World Tasks

MobilePA-Bench: Benchmarking Mobile Planner Agents on Complex Real-World Tasks
Takara TLDR - Daily AI Papers

Modalities Should Talk to Each Other: Dual-Stream Multimodal Learning for Long-Horizon Influenza Forecasting

・Forecasting long-range influenza-like illness (ILI) matters for public health readiness. ・Publicly available surveillance datasets typically pair numeric epidemiological signals with textual information that is noisy, loosely structured, only indirectly related to near-term trends, and often lagged relative to the numeric signal. ・Fusing the two therefore requires careful design.
stat.ML updates on arXiv.org

Model-Agnostic Covariate-Assisted Inference on Partially Identified Causal Effects

・arXiv:2310.08115v3 Announce Type: replace-cross Abstract: Many causal estimands are only partially identifiable since they depend on the unobservable joint distribution between potential outcomes. ・Stratification on pretreatment covariates can yield sharper bounds; however, unless the covariates are discrete with relatively small support, this approach typically requires binning covariates or estimating the conditiona
Takara TLDR - Daily AI Papers

Multi-Modal Semantic Expansion with Constrained LLM Reranking for Conversational Music Recommendation

・We present Team Semiintelligencn's solution for the ACM RecSys 2026 TalkPlayData Challenge, addressing conversational music recommendation through a multi-modal and personalized conversational recommender system. ・Our submitted system employs a three-stage pipeline: (1) multi-modal retrieval constructing decay-weighted centroids across seven dense embedding spaces - track- and user-level CF-BPR, Qwen3 (metadata, lyric
Takara TLDR - Daily AI Papers

Neighbor-Aware View Synthesis for Restoring Missing Views in Light-Field Camera Arrays

・In light-field (LF) imaging systems, dense spatial sampling from a camera array enables powerful post-capture capabilities such as refocusing and depth estimation. ・However, real-world LF capture is often affected by hardware malfunctions, where one or more cameras in the array fail, leading to missing sub-aperture images and degraded reconstruction quality. ・This paper addresses the problem of defective or missing vie
stat.ML updates on arXiv.org

Neural Boltzmann Equations

・arXiv:2608.23022v1 Announce Type: cross Abstract: The dynamics of particles in the early universe are described by Boltzmann equations, which involve high-dimensional phase-space integrals. ・Classical approaches use quadrature integration and evolve the system on a fixed momentum grid, which scales poorly to complicated systems and parameter scans, severely limiting the complexity of processes that can be studied.
stat.ML updates on arXiv.org

Neuro-Causal Factor Analysis

・arXiv:2305.19802v2 Announce Type: replace Abstract: Factor analysis (FA) is a statistical method for explaining how mutually dependent observed variables can be represented in terms of mutually independent latent factors, and it is widely used in the psychological, biological, and physical sciences. ・We revisit this classic method from the perspective of recent advances in causal structure learning and deep generative
The Verge

Nothing OS 5.0 brings a new Glyph Interface app and a more customizable homescreen

・With the new design of Nothing OS 5.0, app icons and widgets can use adaptive color to get color tints pulled from your wallpaper, which update whenever your wallpaper changes. ・Alongside Android's built-in Dynamic Colors, the company says it allows you to have specific elements change to match your wallpaper "while Nothing's predominantly monochrome character stays underneath." While I like Nothing's simplified, mono
WIRED

Omega Just Released a Mini Moonwatch

・Say hello to a revamped collection of perfectly proportioned 38-mm Speedmasters.
Takara TLDR - Daily AI Papers

OmicSync: Reliability-Aware Spatial Multi-Omics Clustering with Evidence-Constrained LLM Reasoning

・Spatial multi-omics technologies jointly profile gene expression, surface proteins, and histology at each tissue spot, yet most spatial domain discovery methods provide only cluster assignments, without indicating assignment reliability, modality contributions, or why a domain decision should be trusted. ・We present OmicSync, a reliability-aware spatial multi-omics framework that couples unsupervised domain clustering
stat.ML updates on arXiv.org

One Inverse Step is a Convex Program: Bayes-Limit Calibration of Diffusion Inversion

・arXiv:2608.23094v1 Announce Type: new Abstract: One implicit DDIM inversion step is the cheapest probe of whether a pretrained diffusion model encodes local manifold geometry. ・It is the stationarity condition of an explicit potential, $x-G(x)=\nabla\Psi_t(x)$, strongly convex at the Bayes limit with modulus exactly $e^{-h_t}$ for the step's log-SNR gap $h_t$ $-$ for every data law, schedule and point, with no manifol
Hugging Face Papers

One Polluted Page Is Enough: Evaluating Web Content Pollution in LLM Recommenders

One Polluted Page Is Enough: Evaluating Web Content Pollution in LLM Recommenders
Hugging Face Papers

One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows

One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows
stat.ML updates on arXiv.org

One-shot Robust Federated Learning of Independent Component Analysis

・arXiv:2505.20532v3 Announce Type: replace-cross Abstract: This paper studies robust one-shot aggregation for distributed and federated Independent Component Analysis (ICA). ・In this setting, each client computes a local ICA estimator, while the server aims to recover a common global mixing matrix without accessing raw data. ・The main difficulty is that local ICA estimators are identifiable only up to signed permutation
The Verge

OpenAI says its Jalapeño chip can power faster AI responses than the competition

・OpenAI says its new AI chip, Jalapeño, completes tasks more efficiently and returns responses faster than other AI systems, according to a blog post published on Tuesday. ・During a briefing with reporters, OpenAI hardware vice president Richard Ho said Jalapeño offers the "best of both worlds" with lower latency and higher throughput, as AI systems typically "have to make a trade-off between the two." First introduced
AI News & Artificial Intelligence | TechCrunch

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

・Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.
#LLMタグ

OpenAIのCompaction

OpenAIのCompaction
WIRED

Optoma GT2400HDR Review: Ace Projector for Golf Simulation

・The Optoma GT2400HDR laser projector is great for setting up a driving range in your living room, but mediocre video quality makes it less appealing for home entertainment.
Qiita - 人気の記事

Oracle AI Database Private Agent Factory 26.4 を 26.7 へ Upgrade してみてみた

・Oracle AI Database Private Agent Factory 26.4 を 26.7 へ Upgrade してみてみた Oracle AI Database Private Agent Factory(PAF)26.7 がリリースされました。
stat.ML updates on arXiv.org

Physically-based dimensionless features for pluvial flood mapping with machine learning

・arXiv:2211.00636v4 Announce Type: replace-cross Abstract: Rapid delineation of flash flood extents is critical to mobilize emergency resources and to manage evacuations, thereby saving lives and property. ・Machine learning (ML) approaches enable rapid flood delineation with reduced computational demand compared to conventional high-resolution, 2D flood models. ・However, existing ML approaches are limited by a lack of g
ITmedia NEWS 最新記事一覧

pixivFANBOX、無断転載コンテンツのGoogle削除申請を代行 投稿データのスクレイピング被害受け

・FANBOXでは、支援者が自身のセッションIDを無断転載サイトに提供し、投稿データが一括でスクレイピングされる被害が起きていた。
The Verge

Polestar claims it was blindsided by sales ban

・BEVERLY HILLS, CALIFORNIA - JUNE 26: Polestar cars are displayed in the showroom at a Polestar dealership on June 26, 2026 in Beverly Hills, California. ・Electric car maker Polestar will be barred from selling new vehicles in the United States starting with the 2027 model year after the Commerce Department cited national security concerns under new connected-vehicle regulations. ・(Photo by Justin Sullivan/Getty Images)
stat.ML updates on arXiv.org

Practical and Optimal Algorithm for Linear Contextual Bandits with Rare Parameter Updates

・arXiv:2606.00984v3 Announce Type: replace Abstract: We study linear contextual bandits under rare parameter updates: the learner may incorporate reward feedback into its parameter estimate only at a small number of update times, while still observing contexts online and selecting actions sequentially. ・This viewpoint clarifies a practical distinction that is often blurred in the literature: many "strictly batched" met
stat.ML updates on arXiv.org

Primal--Dual Alternating Neural Learning for Timely Classification with Performance Guarantees

・arXiv:2608.23480v1 Announce Type: new Abstract: Timely risk classification is essential in many clinical monitoring settings, where decisions must balance the benefit of classifying patients early for subsequent intervention against the value of observing additional data. ・Yet most existing statistical and machine-learning methods are designed for fully observed trajectories and offer limited control over key operatin
Hugging Face Papers

Prime Agent: A Self-Improving RLM Harness

Prime Agent: A Self-Improving RLM Harness
stat.ML updates on arXiv.org

Prob-GParareal: A Probabilistic Numerical Parallel-in-Time Solver for Differential Equations

・arXiv:2509.03945v3 Announce Type: replace-cross Abstract: We introduce Prob-GParareal, a probabilistic extension of the GParareal algorithm designed to provide uncertainty quantification for the Parallel-in-Time (PinT) solution of (ordinary and partial) differential equations (ODEs, PDEs). ・The method employs Gaussian processes (GPs) to model the Parareal correction function, in line with GParareal, further enabling t
Takara TLDR - Daily AI Papers

Progressively Learning Heterogeneous Skills in a Unified Latent Space

・We propose HetSkills, a novel framework designed to progressively learn heterogeneous skills within a unified latent space for physics-based character control. ・The core idea is to treat this latent space as a shared executable interface, enabling seamless integration of skills learned from diverse data sources, supervision forms, and tasks. ・HetSkills begins by learning a tracking skill that establishes a strong found
stat.ML updates on arXiv.org

Provably adaptive sampling with uniform and remasking discrete diffusion models

・arXiv:2608.23554v1 Announce Type: cross Abstract: Discrete diffusion models offer a promising alternative to autoregressive generation by enabling parallel updates, but their sampling efficiency can depend strongly on the choice of the forward process and the sampler. ・For the uniform forward process, existing lower bounds for the standard $\tau$-leaping sampler scale linearly with the ambient dimension $d$, raising t
Takara TLDR - Daily AI Papers

Proxy reliance in large language model decisions is uncalibrated to predictive evidence

・Large language models (LLMs) are entering decisions in triage and lending, where task-relevant inference must be distinguished from impermissible proxy use. ・Current audits ask whether decisions change when demographics change. ・But attributes correlated with a protected group carry predictive value, so a changed decision can be discrimination or sound inference.
Takara TLDR - Daily AI Papers

ProxyFormer: A Dual-Stream Proxy Architecture for Ultra-Long Context and High-Resolution Generation

・The quadratic growth of attention computation and key-value (KV) cache with respect to sequence length is a central bottleneck for ultra-long-context language models and high-resolution generative models. ・We propose ProxyFormer, a general dual-stream architecture built upon proxy tokens. ・In each layer, fine-grained local features are compressed bottom-up into a small set of proxy states; expensive global interactions
Zennの「大規模言語モデル」のフィード

Pythonで作る AIエージェントのMemory入門

・AIエージェントの「Memory」は、どのように情報を覚え、次の会話で使っているのでしょうか。本書では、Windows 11 + PowerShell + Python + OpenAI APIを使い、MEMORY.mdへの保存・読み込み、LLMへのMemory利用、ReflectionによるMemory追加、再利用、削除までを短いコードで段階的に実装します。完成済みのMemory機能を使うだけでなく、自分で作って動かすことで、「記憶するAI」の仕組みと限界を理解する実践入門です。
stat.ML updates on arXiv.org

Q-Learning with Stable Infinite-Dimensional Linear Function Approximation

・arXiv:2608.22636v1 Announce Type: cross Abstract: Q-learning with linear function approximation can be unstable because an arbitrary approximation architecture need not preserve the Bellman contraction. ・We develop a stable infinite-dimensional linear function approximation framework for Q-learning from a single Markovian behavior-policy trajectory. ・The learning variable is a coefficient field $\theta\in C(\mathbb L)$
Hugging Face - Blog

Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original

Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original
Hugging Face Papers

Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs

Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs
Zennの「大規模言語モデル」のフィード

Qwen3.8 27Bを262KでPi Coding Agentにつないで「テトリス作っといて」と言ってみた

・こんにちは。miharubaの池谷です。 ・最近、RTX 4070 Ti SUPER 16GB + RTX 5060 Ti 16GBの2GPU環境で、Qwen3.8 27Bをllama.cppから動かして遊んでいます。 ・前回は262K contextでQ3/Q5を動かし、Tensor Splitを調整して2GPUをどう使うと速いのかを試しました。
Takara TLDR - Daily AI Papers

RACO: Reliability-Aware Coarse-Goal Optimization for Inspection-Oriented UAV Vision-Language Navigation

・UAV vision-language navigation (UAV-VLN) is commonly evaluated as goal reaching, but inspection-oriented deployment requires the agent to stop within a valid inspection region and avoid falsely confirming visually or semantically similar distractors. ・This requirement exposes a key weakness in existing coarse-to-fine UAV-VLN policies: the coarse goal predicted before local refinement is often treated as reliable, alth
Takara TLDR - Daily AI Papers

RAD: Rule-Augmented Relational Anomaly Detection

・Anomaly detection is often applied to data stored in relational databases, yet most existing methods require flattening multiple tables into a single feature matrix. ・This flattening can obscure entity identity, schema structure, and multi-hop dependencies, limiting the detection of anomalies that depend on relational context rather than isolated feature values. ・Beyond preserving relational structure, relational anoma
stat.ML updates on arXiv.org

Radial Compensation: The Inverse Base-Distribution Problem for Chart-Based Generative Models on Riemannian Manifolds

・arXiv:2511.14056v3 Announce Type: replace-cross Abstract: Latent-variable models on spheres and hyperbolic spaces usually draw a Gaussian in the tangent space at a base point and push it onto the manifold. ・On these spaces the distance from the base point is the coordinate that carries meaning: depth in a hierarchy, the angle of a rotation, the deviation of a protein frame from a reference. ・We show that the standard c
stat.ML updates on arXiv.org

Random Hazard Forests

・arXiv:2608.21597v1 Announce Type: new Abstract: Clinical data sources such as electronic health records and wearable sensors record patient status repeatedly over follow-up, often at irregular times and on different schedules for different measurements. ・These data create opportunities for continuously updated, individualized risk prediction. ・Existing approaches, however, often simplify the temporal structure before m
stat.ML updates on arXiv.org

Recovering Weighted Tangent Geometry from a Single-Scale Score Field

・arXiv:2608.22334v1 Announce Type: new Abstract: Near a smooth data manifold, one tangent space summarizes local geometry. ・At a branch point, the corresponding first-order object is instead a measure over tangent directions, whose normalized masses record the local share of each branch under the chosen data measure. ・We ask whether a score field at one noise level determines this weighted tangent geometry when the bran
stat.ML updates on arXiv.org

Relational Structural Causal Models

・arXiv:2606.14892v2 Announce Type: replace-cross Abstract: An artificial intelligence must have a model of its environment that is causal, supporting reasoning about interventions and counterfactuals, and also combinatorial, supporting generalization to unseen combinations of objects. ・In this work, we formally study when and how such a model can be learned. ・We develop relational structural causal models, extending str
Takara TLDR - Daily AI Papers

ReWorld: An Interactive World Model with Long-Horizon Memory

・An interactive world model must follow the user's actions, remember the places it has shown, and stream in real time. ・The tension is structural: control wants a short horizon, memory wants an unbounded one. ・ReWorld separates the two during training and bounds them at inference.
Hugging Face Papers

ReWorld: An Interactive World Model with Long-Horizon Memory

ReWorld: An Interactive World Model with Long-Horizon Memory
Hugging Face Papers

RIBOSPAN: A Long-Context RNA Foundation Model for Versatile RNA Modeling

RIBOSPAN: A Long-Context RNA Foundation Model for Versatile RNA Modeling
Hugging Face Papers

RISE: Adaptive Imagination for World Action Models

RISE: Adaptive Imagination for World Action Models
stat.ML updates on arXiv.org

Robust performance metrics for imbalanced classification problems

・arXiv:2404.07661v2 Announce Type: replace Abstract: We show that established performance metrics in binary classification, such as Matthews' correlation coefficient (MCC), Cohen's $\kappa$, the F-score or the Jaccard similarity coefficient are not robust to class imbalance in the sense that if the proportion of the minority class tends to $0$, the true positive rate (TPR) of the Bayes classifier under these metrics t
#LLMタグ

RTX 3060 12GBでOllamaを動かしてみた|4B・8B・12B・14B・30Bを実測比較

・前回までに、古いPCで使っていたRTX 3060 12GBを現在のPCへ接続し、PyTorchからGPUを利用できるところまで確認しました。 ・せっかくGPUが使える環境ができたので、次はいよいよローカルLLMを実際に動かしてみます。
Hugging Face Papers

Same Agent, Different Answers: A Repeat-Aware Audit of Corpus-Induced Answer Churn in Retrieval-Augmented QA

Same Agent, Different Answers: A Repeat-Aware Audit of Corpus-Induced Answer Churn in Retrieval-Augmented QA
stat.ML updates on arXiv.org

Scale-invariant Optimal Sampling for Rare-events Data with Sparse Models

・arXiv:2608.22597v1 Announce Type: new Abstract: Subsampling is effective in tackling computational challenges for massive data with rare events. ・Overly aggressive subsampling may adversely affect estimation efficiency, and optimal subsampling is essential to mitigate the information loss. ・However, existing optimal subsampling probabilities depend on data scales, and some scaling transformations may result in ineffici
Takara TLDR - Daily AI Papers

SEAM: Shot Entity-Attribute Memory for Consistent Short-Drama Generation at Scale

・Short-drama generation has grown into a large, industrialized pipeline, and as it scales from isolated shots to the episode level, visual continuity has become a critical bottleneck. ・Current agent frameworks generate each shot in isolation, so context drifts across shots and props, character posture, and blocking turn inconsistent. ・Once assembled, these small discrepancies amplify into severe visual breaks.
stat.ML updates on arXiv.org

Semiparametric Double Reinforcement Learning with Applications to Long-Term Causal Inference

・arXiv:2501.06926v5 Announce Type: replace Abstract: Double reinforcement learning (DRL) provides efficient off-policy inference for policy values in nonparametric Markov decision processes (MDPs), but fully nonparametric estimators can be unstable when intertemporal overlap is weak and occupancy ratios are high-dimensional. ・This limitation is especially relevant for long-term causal inference from randomized experime
Takara TLDR - Daily AI Papers

SGHA: A Single-Loop Fully First-Order Algorithm for Nonconvex-Strongly-Convex Bilevel Optimization

・In this work, we study the oracle complexity of finding an $ε$-stationary point for nonconvex-strongly-convex (NC-SC) bilevel optimization using only first-order oracles. ・Existing methods achieving the best-known complexity guarantees typically rely on double-loop, penalty-based procedures. ・We propose a novel single-loop algorithm based on a constrained reformulation in which lower-level stationarity is imposed as a
Takara TLDR - Daily AI Papers

Sigmoid Attention as a Better Substrate for Learned KV Cache Eviction

・Learned KV-cache eviction often faces a soft-to-hard mismatch: during training, differentiable gates typically attenuate token contributions, whereas inference saves memory only when KV entries are physically removed. ・We ask whether the attention substrate affects this soft-to-hard transition. ・Using GPT-2-scale Transformers trained on OpenWebText, we run a controlled $2\times2\times2$ comparison over attention type,
AI News & Artificial Intelligence | TechCrunch

Situational Awareness, star AI hedge fund that nearly imploded, now being probed by the SEC

・The AI hedge fund went from "the talk of Wall Street" to "subject of federal subpoenas" faster than you can say "diversify your portfolio."
Takara TLDR - Daily AI Papers

SkillAlchemy: Open-World Agent Skill Creation

・Agent skills are reusable procedural artifacts that extend language agents with specialized workflows, tool conventions, and domain behaviors at inference time. ・However, creating reliable skills still depends largely on human authorship, model priors, or execution traces. ・These sources are often unavailable for unfamiliar tasks, suggesting the need to create skills from open-world materials.
stat.ML updates on arXiv.org

Sparse Additive Off-Policy Evaluation for Reinforcement Learning with Potentially Limited Number of Trajectories

・arXiv:2608.22595v1 Announce Type: new Abstract: We develop a new framework for flexible, nonlinear, and interpretable off-policy evaluation for infinite-horizon reinforcement learning. ・To handle large state spaces and support transparent decision-making, we model the Q-function using a nonlinear function class with a sparse additive structure. ・We derive high-probability finite-sample error bounds for estimating the v
stat.ML updates on arXiv.org

Sparse Separable Factor Analysis in the Complex Domain with an Application to Local Field Potential Data

・arXiv:2608.21551v1 Announce Type: new Abstract: Complex-valued arrays arise in signal processing, where scientific interpretation depends on retaining amplitude and phase information. ・Existing covariance estimation methods either ignore the multiway organization of such data or rely on real-domain embeddings that do not directly exploit their complex structure. ・We develop sparse separable factor analysis (SSFA), a la
stat.ML updates on arXiv.org

Spectral partitioning for $k$-block averaging kernels of finite Markov chains

・arXiv:2608.21466v1 Announce Type: new Abstract: We develop spectral algorithms for selecting state-space partitions that define averaging kernels for finite, ergodic and reversible Markov chains. ・For a partition $\mathcal O$, the Gibbs kernel $G_{\mathcal O}$ resamples within the current block from the stationary conditional distribution; when this update is tractable, composing or mixing it with a baseline kernel $P
stat.ML updates on arXiv.org

Spending Scarce Confirmatory PET Measurements: Target-Aligned Validation in A4/LEARN

・arXiv:2608.22223v1 Announce Type: cross Abstract: Anti-amyloid therapies and blood-based biomarkers are changing Alzheimer disease workups into a two-stage measurement workflow: screen broadly with cheaper information, then spend scarce confirmatory amyloid measurements where they support the decision that will be reported. ・Amyloid positron-emission tomography (PET) remains one such protocol measurement for amyloid b
Takara TLDR - Daily AI Papers

Spicing up Genetic Netlist Generation with LLMs

・Analog circuit topology synthesis remains challenging because useful designs occupy a tiny fraction of a combinatorial search space, and small structural changes can induce highly nonlinear changes in behavior. ・Evolutionary algorithms are attractive because they can optimize over discrete circuit topologies using only black-box evaluations, but they often require many SPICE simulations and may converge prematurely.
WIRED

Spirit Airlines Wants to Sell Its Data to Google. Former Flight Attendants Are Freaked Out

・“It never crossed my mind that they would be so bold as to sell our private data for AI,” says one former Spirit Airlines flight attendant.
Takara TLDR - Daily AI Papers

SPOC-SQL: Stage-wise Preference Optimization for Controllable Text-to-SQL

・Text-to-SQL aims to translate natural language questions into executable SQL queries over relational databases, requiring multi-stage structured reasoning over database schemas and query constraints. ・However, existing methods treat this task as single-step generation, where models optimize entire SQL sequences without targeted feedback at key decision points and lack support for interacting with and controlling the i
Takara TLDR - Daily AI Papers

SRPO: Self-Reflective Policy Optimization for Long-Horizon Reasoning

・Self-reflection is a powerful mechanism for credit assignment in human learning, converting sparse outcome feedback into actionable guidance. ・However, its potential for post-training Large Language Models (LLMs) remains underexplored. ・We propose Self-Reflective Policy Optimization (SRPO), a framework that internalizes this capability.
stat.ML updates on arXiv.org

Stability and Accuracy Trade-offs in Statistical Estimation

・arXiv:2601.11701v2 Announce Type: replace-cross Abstract: Algorithmic stability is a central concept in statistics and learning theory that measures how sensitive an algorithm's output is to small changes in the training data. ・Stability plays a crucial role in understanding generalization, robustness, and replicability, and a variety of stability notions have been proposed in different learning settings. ・However, whi
stat.ML updates on arXiv.org

Stochastic gradient descent with initial regularization

・arXiv:2608.22953v1 Announce Type: cross Abstract: We analyze a variant of stochastic gradient descent with initial regularization (SGDIR) and derive dimension-free upper bounds on its expected excess risk for the squared loss. ・In the noiseless case, we obtain new bounds for both averaged and non-averaged SGDIR under moment, source, and capacity assumptions. ・For a particular value of the source parameter, these bounds
Takara TLDR - Daily AI Papers

Stochastic Separability of Embedding Manifolds

・Neurobiological studies and representation learning have observed that representations of objects belonging to the same category in high-dimensional neural spaces exhibit low-dimensional object manifold characteristics, and different object manifolds are linearly separable in these neural spaces. ・However, these experimentally observed phenomena lack rigorous theoretical validation to date. ・This paper proposes a new s
stat.ML updates on arXiv.org

Strong Averaging Principle and Long-Time Dynamics for Fast-Slow SDEs with Increasing Time-Scale Separation and Degenerate Noise

・arXiv:2608.23462v1 Announce Type: cross Abstract: We establish a strong averaging principle for fast-slow stochastic differential equations with a time-dependent scale-separation parameter $(\varepsilon_t)_{t \geq 0}$ satisfying $\varepsilon_t \to 0$ as $t \to \infty$. ・In contrast to approaches based on noise-induced smoothing or elliptic regularity, our approach relies on dissipativity of the frozen fast dynamics an
stat.ML updates on arXiv.org

Structured Learning on Mapper Representations

・arXiv:2608.22044v1 Announce Type: new Abstract: Modern machine learning (ML) methods are highly effective for prediction tasks, but many commonly used representations reduce complex data to fixed dimensional embeddings that may suppress multiscale structural organization. ・The Mapper algorithm from topological data analysis (TDA) provides a different perspective by decomposing data into overlapping local regions conne
stat.ML updates on arXiv.org

Subzero matrix completion for sparse data analysis: large-scale learning of latent low-rank structure

・arXiv:2608.21607v1 Announce Type: cross Abstract: We investigate when a sparse nonnegative matrix can be recovered from a real-valued matrix of much lower rank by zeroing out its negative elements. ・The potential for such decompositions suggests a mathematical connection between sparsity and rank; we analyze a number of sparse matrices with this latent low-rank structure and use them to illustrate the geometric origin
stat.ML updates on arXiv.org

Symbolic Neural ODEs: Learning interpretable models from time-series data

・arXiv:2608.22112v1 Announce Type: cross Abstract: We present a machine learning framework for identifying sparse, interpretable models of dynamical systems directly from time-series data. ・Our approach parameterizes the underlying vector field using a neural architecture and trains it by minimizing a multi-step prediction loss over a finite horizon. ・To ensure numerical tractability, we optimize a mean absolute error o
Zennのトレンド

TanStack Start + Hono + oRPC + Cloudflare Workersで社内ERPを作った設計と学び

・はじめに 建築業向けの社内ERP(プロジェクト管理・CRM・工数管理・ダッシュボード)を、Full-Stack TypeScriptでスクラッチ開発しました。TanStack Start + Hono + oRPCを、Cloudflare Workersの上で動かしています。 ・この構成はまだ事例が少なく、「動くところまでは書けるが、業務システムのサイズに育ったときにどうなるのか」が想像しづらい領域だと思います。この記事はスタックの紹介ではなく、どこに境界を引いたのか/引き直したのかの話です。 ・結論を先に並べると、主な判断はこの4つです。
Hugging Face Papers

Task-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection

Task-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection
stat.ML updates on arXiv.org

Tensor-normal maximum likelihood estimation at the operator-norm sample threshold

・arXiv:2608.10488v2 Announce Type: replace-cross Abstract: Let $X_1,\ldots,X_n$ be independent Gaussian tensors in $\mathbb{R}^{d_1}\otimes\cdots\otimes\mathbb{R}^{d_k}$ with a common covariance matrix given by the Kronecker product of $k$ unknown positive-definite factors, and let $D=\prod_{a=1}^k d_a$ and $d_{\max}=\max_a d_a$. ・Franks et al. ・(2026) established condition-number-free guarantees for the tensor-normal m
Takara TLDR - Daily AI Papers

Test-Time Adaptation for ECG Classification via SQI-Gated Self-Training and Beat-Rhythm Consistency

・Deep learning models for electrocardiogram (ECG) classification often suffer from significant performance degradation when deployed in unseen domains due to shifts in acquisition devices and patient populations. ・Test-time adaptation (TTA) offers a practical solution by adapting models using only unlabeled data at inference time. ・However, existing TTA methods often underperform on ECG tasks, since naive online updates
WIRED

The 4 Best Car Phone Mounts I’ve Tried (2026): Belkin, Andery, Andobil

・The best car phone holders keep your phone firmly in place while never getting in your line of sight. ・I took a dozen on road trips this summer to find the best.
WIRED

The Best Kitchen Gadget to Prevent Salmonella Is a Good Meat Probe

・Salmonella outbreaks are seemingly everywhere right now. ・A good temperature probe is the last and best line of defense.
WIRED

The County Prosecutors Who Became ICE Informants

・Illinois prosecutors shared defendants’ personal data with federal immigration agents without criminal warrants, public disclosure, or legislative oversight.
OpenAI News

The full stack behind abundant intelligence

・OpenAI CFO Sarah Friar explains how advances across chips, compute, models, and products compound to deliver more useful intelligence at greater scale and lower cost.
stat.ML updates on arXiv.org

The geometry of AI validation: Exact certification limits for iid best-of-N search

・arXiv:2608.21496v1 Announce Type: cross Abstract: AI systems increasingly generate alternatives, inspect evidence, and deploy a selected output. ・Validation is therefore target-relative: evidence certifies deployment only in directions resolved by the interventions that produced it. ・We represent validation and deployment rules as kernels over a reliability surface.
Hugging Face Papers

The Laws of Context Allocation: Causal Measurement and Closed-Loop Orchestration in Generative Search

The Laws of Context Allocation: Causal Measurement and Closed-Loop Orchestration in Generative Search
The Verge

The lens in Wyze’s new $40 indoor camera retracts for privacy

・Wyze announced a new pan-and-tilt security camera today designed for indoor use. ・The entire Indoor Cam Pan can rotate 360 degrees, while the camera and lens on its front can tilt up and down 103 degrees. ・When you want privacy, the lens can tilt down until it completely disappears inside the Indoor Cam Pan's body so it's unable to see anything happening in your home.
WIRED

The Supreme Court’s Mail-In Ballot Ruling Is a Step Toward Chaos in the Midterms

・The Supreme Court has removed one roadblock from the Trump administration’s attempts to impose severe restrictions on voting.
Takara TLDR - Daily AI Papers

Thinking Beyond Videos: Unifying Video Reasoning and Deep Research for Open-World Video Agents

・Open-world video understanding often requires a model to locate sparse visual evidence and acquire external knowledge that is absent from the video and its parametric memory. ・While Thinking-with-Videos enables active temporal perception and Deep Research supports multi-step information seeking, the two capabilities are typically developed in isolation. ・We introduce VideoRover, a unified Video Deep Research framework
#LLMタグ

Thomsonが変える専門AIの経済性

・Thomson Reutersが、専門領域へ追加学習した独自LLM「Thomson」を公開した。投資額は4,000万ドル、独自データの活用は10%未満で、まずCoCounsel Legalへ導入する。汎用APIを借りる以外の、専門モデルを所有する企業AI戦略の選択肢を整理した。
Hugging Face Papers

TileMix: Tile-Centric Mixed-Precision Attention for LLM Inference Acceleration

TileMix: Tile-Centric Mixed-Precision Attention for LLM Inference Acceleration
Hugging Face Papers

TLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming

TLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming
stat.ML updates on arXiv.org

Token-Level Likelihood-Array Regression for Membership Inference and AI-Generated Text Detection

・arXiv:2608.22179v1 Announce Type: new Abstract: Membership inference asks whether a text was used to train a language model, whereas AI-generated text detection asks whether it was generated by a language model rather than written by a human. ・Existing likelihood-based methods typically compress token-level probabilities into a few prespecified scores, most often using only probabilities conditioned on the full preced
Hugging Face Papers

Towards a Densing Law for User Representation Learning at Billion-Scale Capacity

Towards a Densing Law for User Representation Learning at Billion-Scale Capacity
stat.ML updates on arXiv.org

Traceable Spectral Inference via Influence Functions: Efficient Data Attribution and Error Proxies for the Ariel Mission

・arXiv:2608.23458v1 Announce Type: cross Abstract: Interpretability is critical for machine learning models deployed in scientific space missions such as ESA's Ariel, where ground truth is unavailable during operations and physical plausibility must be assessed. ・While most explainable AI methods focus on feature attribution, this work investigates training data attribution through influence functions and introduces th
Takara TLDR - Daily AI Papers

Triplet2Track: A Hierarchical System with Object-Centric Representations for Reliable Long-Horizon Manipulation

・Ensuring reliability in uncertain environments remains difficult for long-horizon robotic manipulation. ・End-to-end VLA models are data-heavy and opaque, making diagnosis and verification difficult. ・Hierarchical pipelines are more interpretable, but their plans are often weakly grounded in observations, weakly aligned with low-level actions, and computed without online feedback, leading to open-loop behavior and hallu
AI News & Artificial Intelligence | TechCrunch

Trump bought SpaceX shares two weeks after blockbuster IPO

・The president bought when the stock was in the mid-$150 range. ・SpaceX finished trading on Monday back at its IPO price of $135.
Hugging Face Papers

Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision

Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision
stat.ML updates on arXiv.org

Variance Driven Exploration: A Provable and Efficient Methodology for Pure Exploration in Highly Stochastic Environments

・arXiv:2608.21995v1 Announce Type: cross Abstract: We propose Variance Driven Exploration (VarDE), a principled approach for pure exploration in highly stochastic environments, where the exploration process is dominated by stochastic variance. ・VarDE is built on a fundamental principle: sampling effort should be allocated to minimize the uncertainty of the final decision. ・We formalize the uncertainty of the final decis
Zennの「大規模言語モデル」のフィード

vibe codingは祈りである

・以前こう言う記事を書いた。 ・https://zenn.dev/simossyi/articles/755aacf33d742c 要約するとこうだ。 ・LLMは事前学習をしてる 事前学習は文脈条件付け次トークン分類をやっている 従って、事前学習コーパスにチャットのやり取りがあった時、感謝が述べ伝えられるのは基本的に好ましいことをした場合であって、LLMとのチャットにおいても感謝を述べることで次トークンの効用が改善するのではないか でもそれってわかんないし祈りなんだよね、全て 最近改めてvibe codingは祈りであるという気持ちが強くなっていくのでそのアイデアを示す。
Takara TLDR - Daily AI Papers

WADE: A Reasoning-Annotated Benchmark for Multi-Instance Floating-Waste Grounding with Compact Vision-Language Models

・Floating waste in inland waterways threatens aquatic ecosystems and requires timely monitoring under cluttered, multi-object conditions. ・Existing aquatic-waste datasets provide limited geographic coverage, sparse multi-instance annotations, and little supervision beyond boxes and labels. ・Compact vision-language models (VLMs) therefore remain insufficiently evaluated for jointly localizing, classifying, counting, and
Takara TLDR - Daily AI Papers

What is mathematics now, and what should it be?

・Advances in neural theorem provers have been impressive, but the successes obscure a broader vision of what AI can do for mathematics and how mathematicians can engage with AI. ・This essay advances a more expansive and optimistic point of view.
stat.ML updates on arXiv.org

What LLMs explain is not what they believe: Evaluating explanation sufficiency under models' own input beliefs

・arXiv:2606.28615v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains, where free-text explanations such as chain-of-thought and post-hoc rationales are used to justify model outputs. ・Yet it remains unclear whether these explanations are sufficient, i.e., if they contain enough information to explain the model's output-generating process. ・We generalize
Takara TLDR - Daily AI Papers

What's the Catch? Evaluating Temporal Consistency in Vision-Language Models

・Vision-language models (VLMs) achieve strong performance on video and image-sequence benchmarks, yet it remains unclear whether they capture temporal structure. ・To study this question, we formulate temporal grounding as an anomaly detection problem, providing a simple and controlled evaluation that directly tests sensitivity to temporal consistency. ・We introduce TimeCatch, where temporal anomalies are created by swap
Takara TLDR - Daily AI Papers

When Names Cross Scripts: A Source-Grounded Benchmark for Historical Entity Reconciliation in the Mongol World

・Historical people may appear under different languages, scripts, and transcription traditions, while distinct individuals may share highly similar or even identical names. ・This makes historical identity reconciliation more than a problem of string matching or transliteration. ・We introduce MHER, a provenance-controlled benchmark for pairwise reconciliation of person-name attestations from the Mongol world.
The Verge

Where to preorder the updated Mac Mini and Mac Studio

・The updated Mac Mini is a compact machine for home desktop use. ・| Image: Apple Apple has announced two new processors, the M6 and M5 Ultra, as well as updated Mac Mini and Mac studio machines powered by them. ・The new machines are already available for preorder directly from Apple, with shipping for most configurations starting on September 22nd, 2026.
Takara TLDR - Daily AI Papers

WildHandBench: A Benchmark for Handwritten Text Understanding that Challenges MLLMs and Humans

・While the top model on OmniDocBench now reaches 96.34% overall on printed-document parsing, the ability of current models to handle challenging handwritten documents remains largely uncharacterized. ・Existing benchmarks focus on isolated text or formulas, overlook handwritten tables and real-world degradation, and report aggregate accuracy without explaining why models fail. ・We present WildHandBench, a benchmark conta
Hugging Face - Blog

Wire It, Run It, Deploy It: AI Workflows in Gradio

Wire It, Run It, Deploy It: AI Workflows in Gradio
stat.ML updates on arXiv.org

You Need Better Attention Priors

・arXiv:2601.15380v2 Announce Type: replace-cross Abstract: We generalize the attention mechanism by viewing it through the lens of Entropic Optimal Transport, revealing that standard attention corresponds to a transport problem regularized by an implicit uniform prior. ・We introduce Generalized Optimal transport Attention with Trainable priors (GOAT), a new attention mechanism that replaces this naive assumption with a
#AIタグ

ZAPJQL SF-80 ポータブルエアコン レビュー比較まとめ

・ZAPJQL SF-80 ポータブルエアコンは、業務用大型冷風扇 冷風機 400W 気化式送風機 8000m³/h 左右自動ルーバー 60L大容量タンクSF-80として、夏の暑さ対策に悩むビジネスパーソンや生活者にとって頼れる一台です。 ・本記事では、ZAPJQL SF-80が持つポータブルエアコンとしての真価を多角的に分析し、快適な夏を実現するための最適な判断材料を提供します。
Zennのトレンド

エージェントのためのUIコンポーネントライブラリを作った

・はじめに コーディングエージェントによって、要件からアプリケーションを実装するまでの時間は大きく短縮することができるようになりました。一方で、生成されたフロントエンドの画面を確認すると、余白や色、状態表現に細かなずれが残り、人が修正を重ねることがあります。 ・完成度90%までの実装が速くなっても、残り10%の確認と手直しにコストがかかるなら開発工程全体が効率化されたとは言い切れません。このボトルネックに対して Figma MCPを使ってデザイン情報を取得できるようにする Playwrightなどでスクリーンショットを確認させる プロンプトへデザインを再現するためのルールを詳しく記載す...
Zennのトレンド

オンコール手当の相場を50社分調べて、1社も見つからなかった話

・このブログは CSCでオンコール運用を検討する際に手当について調査した結果 となります。 ・これから検討に取り組まれる方はぜひ参考にしてください! 3行で言うと オンコール手当の「金額」を5チャネル・約50社で調べたら、1件も出てきませんでした。 ・でも「金額が公開されていない」で終わらせず、「そもそも金額から設計するのが間違いかもしれない」 という気づきがありました。
@IT 全フォーラム 最新記事一覧

オンワードが複数社プロジェクトでの「何がどこまで進んでいるか不明」を解消 どう実現した?

・社内外の関係者が増えるほど、プロジェクトの情報管理は難しくなる。「誰が何を担当し、どこまで進んでいるのか」をどう把握するか――。オンワードホールディングスは、この課題にどう対処したのか。
Qiita - 人気の記事

お金さえ払えば、Claude Codeは最大2.5倍速くなる──従量課金の「/fast」を調べてみた

・はじめに Claude Codeの待ち時間を短くできるなら、とりあえず/fastを押してみよう。 ・そう思って検証用のプロンプトまで用意したのですが、使う前に公式ドキュメントを読んで手が止まりました。Claude Codeの/fastは、ProやMaxなどの月額プランに含ま...
Zennのトレンド

お前のループエンジニアリングは間違っている

・と Claude Code に私自身が言われたので、どのように使いこなすとループエンジニアリングを実践できるか試行錯誤してみた過程をまとめた記事になります。 ・ループエンジニアリングについては、多くの記事が紹介しているのでここでは割愛します。こちらの記事が詳細に解説してくれています。私もこちらの記事を参考に学ばさせていただきました、非常にありがたいです。 ・https://zenn.dev/acrosstudioblog/articles/38509c0473683a ループエンジニアリングのコアモジュールである 6 要素として挙がっている Automations / Worktrees /...
#AIタグ

コンビニを効率よく歩く方法

コンビニを効率よく歩く方法
#AIタグ

コンビニを歩く方法

コンビニを歩く方法
#LLMタグ

コンプライ・オア・エクスプレインの移植 / 生成AI知財プリンシプル・コードの評価主体 雑感

コンプライ・オア・エクスプレインの移植 / 生成AI知財プリンシプル・コードの評価主体 雑感
Qiita - 人気の記事

システム設計視点の行動経済学 全20回 総復習

・user: 「システム設計視点の行動経済学」も 20回目まで連載しました。ここで、これまでの内容の振り返りをしたいと思います。Claude Codeさん、まとめを作ってください。 ・Claude Code: 「システム設計視点の行動経済学」全20回 — 総復習 この...
Zennの「機械学習」のフィード

じゃんけんのナッシュ均衡は1/3ずつと決まっている。学習させたら均衡から3,050倍遠ざかり、最後は手が完全に読めるようになった

・ゲーム理論を習うと、必ずナッシュ均衡が出てくる。誰も自分だけ手を変える理由がない状態で、じゃんけんなら「グー・チョキ・パーを1/3ずつ」がそれにあたる。 ・そして、たいていこう説明される。 ・合理的なプレイヤーが互いに学習していけば、やがてナッシュ均衡に落ち着く。
Qiita - 人気の記事

セキュリティキャンプ2026 L2 プロセッサゼミ参加記

・はじめに こんにちは,Latte72 です. この度,セキュリティキャンプ2026全国大会の 開発L2ゼミ に参加してきました. 自分は特に日記などを取ったりしていないにもかかわらず,参加してから一週間も経ってから執筆をしているので,記憶があいまいな部分も多く,写真が多め...
#LLMタグ

タチコマのアームを制作します!まずはパーツを買いそろえます。

タチコマのアームを制作します!まずはパーツを買いそろえます。
#LLMタグ

なぜ本屋では「すぐに欲しい本」が見つかるのか?図書情報学の知恵で解き明かす、AIベクトル検索の育て方

・週末、ふらりと立ち寄った大型書店での光景を思い浮かべてみてほしい。 ・「最近チームの雰囲気が停滞しているな」「何か新しい企画のヒントが欲しいな」といった曖昧なモヤモヤを抱えて店内を歩いていると、ふと目に飛び込んできた新刊の帯に心を掴まれる。あるいは店員に「最近話題の、青い表紙のビジネス書なんですけど……」と要領を得ない相談をしたのに、「ああ、それならこちらの棚です!」と目当ての1冊に案内される。
@IT 全フォーラム 最新記事一覧

パーソルキャリアがAPMで実践する“先回り”の改善サイクル システムが「動いている」から「良い状態」へ

・システムの問題をどう見つけ、改善効果をどう測り、次の投資判断にどうつなげるか――。パーソルキャリアはAPMを軸としたオブザーバビリティーを実践し、継続的なシステム改善を進めてきた。オブザーバビリティーを組織に根付かせ、改善のサイクルを回してきた同社の取り組みについて聞いた。
Zennの「機械学習」のフィード

バーチャル細胞は、まだ「平均」に勝てていない ― ベースライン問題と、論文に対照群を2つ置いた理由

・初出: note(バーチャル細胞は、まだ「平均」に勝てていない)。Zenn向けに技術面を補足して再構成しました。 ・前回、バーチャル細胞(AIDO Cell など)について「ヒト中心で、種を越えた外挿はこれからの課題」と書きました。その後もう一段掘ると、種差より手前に、もっと根本的な問題があると分かってきました。ベースライン問題です。 ・今日は「深層モデルは、そもそも素朴なベースラインに勝てているのか?」を起点に書きます。
#AIタグ

はじめまして。AIは「会話相手」じゃなく「発注先」だと気づいた人の図書館です

・この図書館は、何を置いている場所か はじめまして。「AI図書館」です。
ITmedia NEWS 最新記事一覧

パスポートのオンライン申請で「iPhoneのマイナカード」利用可能に 実物カードの読み取り不要

・デジタル庁は8月25日、マイナポータルのパスポートオンライン申請機能を刷新し、「iPhoneのマイナンバーカード」に対応した。事前にマイナンバーカードをiPhoneへ追加しておけば、暗証番号の入力や実物のカードの読み取りに代わり、端末の顔認証や指紋認証を使ってオンライン申請できる。
#LLMタグ

ビュー中央値12。不合格でした|検証#12(総括)

・12本書きました。判定します。 ・ビュー中央値は12。合格ラインは300でした。不合格です。 ・判定の文言は、1本目で結果を見る前に固定してあります。そのまま貼ります。
ITmedia NEWS 最新記事一覧

プールにはおしっこが何リットル混じっている? 人工甘味料を基に計測 カナダチームが2017年に調査

・カナダのアルバータ大学に所属する研究者らが2017年に学術誌「Environmental Science & Technology Letters」で発表した論文「Sweetened Swimming Pools and Hot Tubs」は、プールに混入した尿の量を、人工甘味料を使って数値化した研究報告だ。
ITmedia NEWS 最新記事一覧

ボカコレ炎上、MVに「3DS」使った曲“ランキング除外”で 「基準あいまい」とクリエイター反発

・一部の参加楽曲が「規約違反」としてランキングから除外されたが、その基準を巡って「曖昧だ」「過剰規制だ」などと批判が起き、炎上状態になっている。
#LLMタグ

ミニLLMを自作してみる

・背景 現在は戦略コンサルタントとして働いているが昨今の技術革新をみて、流石にAIについてもっと詳しくなる必要があると思った。よって、ミニLLMを自作する過程を通じて、AIについて詳しくなろう。その過程の勉強メモとしてこれを書く。
#AIタグ

ループはプロンプトを駆逐する?AIエージェントが勝手に直してくれるなら、指示の技術はもう要らないのか

・このメディアでは「指示の技術」について繰り返し書いてきました。 ・場所を示し、Before/Afterを書き、理由を添える。 ・しかし今はプロンプトではなく、ループが盛り上がっています。
#LLMタグ

ローカルLLMに「テトリス作っといて」と頼んだら、普通にできた

・ローカルLLMに「テトリス作っといて」と頼んだら、普通にできた 今日はローカルLLMで少し遊んでいました。
Zennの「機械学習」のフィード

医療AIを正しく知る——できること・できないこと・信頼をどう作るか

・初出: note(医療AIを正しく知る)。Zenn向けに技術的な用語と枠組みを補足して再構成しました。 ・「医療AI」という言葉は便利すぎて、期待と不安が同じ言葉の中でごちゃまぜになりがちです。今日は、臨床工学技士として医療機器の現場に立ってきた立場から、医療AIを技術的な射程で切り分けて整理します。できること・できないこと・データの扱い・リスクの構造・信頼の作り方、という5つの軸です。 ・できること・できないことの境界線 医療AIが得意なのは、画像の異常検知、大量データからのパターン抽出、事務作業の効率化です。放射線画像や病理画像のように、ピクセル単位の異常を大量の教師データ...
機械学習タグが付けられた新着記事 - Qiita

因果推論 Day 12/全30回 do演算子、「観察する」と「介入する」の違いを式にする

・この連載について 因果推論を「本を読んだ」で終わらせず、自分の言葉で説明でき、コードで再現できる状態まで落とす30日連載です。前回のDay 11では、DAGの上でバックドアパスを見つけ、調整してよい変数といけない変数を規則で選び分けました。ただ、その規則に従うと「なぜ」因...
Zennの「大規模言語モデル」のフィード

引き継ぎプロンプトから解放 ― セッション管理を不要にする「STATE LOOP」という運用

・Claude Codeを毎日使っていて、長いセッションの引き継ぎに疲れている人向け。33世代やった引き継ぎ運用を1日で作り替えた実録と、今日から試せる最小構成。数字は全部実測。 ・想定読者 Claude Codeで長時間セッションを回していて、コンテキストが埋まっていく感覚に覚えがある /compact後に精度が落ちるのを体感したことがある セッションの終わりに引き継ぎプロンプトを書いて、新セッションに貼る運用をしている 「これ、いつまでやるんだろう」と思ったことがある 1. ・うちの運用はこうだった(たぶん重症例) プロジェクト用に常駐セッションを2本(実装担当とレビュー担当...
機械学習タグが付けられた新着記事 - Qiita

音声→MIDI変換を再現可能に評価する:入力ハッシュとノートイベントの比較

・音声→MIDI変換の紹介では「精度○%」という数字が目立ちます。しかし、入力音源、採点単位、時間許容幅、対象楽器が違えば、その数字を横に並べても同じものを測ったことにはなりません。 ・小規模でも再現可能な評価にするには、少なくとも次の四つを分離します。 ・同じ入力を使ったか ...
#AIタグ

何の実績もない僕が、AI×note副業を始める理由

・突然ですが、僕は今日からAI×note副業を始めます。
ITmedia NEWS 最新記事一覧

楽天版SHEIN? 若者向けプチプラ通販「ラクヨコ」開設 SNS感覚で数百円から商品探し

・楽天グループは8月25日、低価格なファッションアイテムや雑貨などを集めたオンラインストア「楽天ラクヨコ」を開設した。若年層向けに約6万点の商品を数百円から販売する。品ぞろえは今後順次拡充する予定。
Qiita - 人気の記事

環境変数管理(.env)の基礎とベストプラクティス:なぜ設定値をコードから分離するのか

・はじめに DB接続情報やAPIキーをコードに直接書き込んでしまうと、そのコードをGitHubなどにpushした瞬間に機密情報が漏洩します。この問題を避けるための定番の方法が、環境変数と.envファイルによる設定値の分離です。 ・なぜ設定値をコードから分離するのか 理由は大...
@IT 全フォーラム 最新記事一覧

建築業でも「Claudeの有料導入」が急増 AI導入レベルが上がった企業を悩ます「新たな壁」とは

・建築AI経営研究会は、「建築AI経営実態調査」の結果を公開した。特定部署以上でAIを活用する企業が68.5%に達したことが判明した。課題は導入初期の試行錯誤から「新たな壁」にシフトしている。
Zennのトレンド

現場で困らないためのGit入門

・普段Gitは使っているけれど、トラブルが起きるたびに検索しているエンジニア向けのBookです。 ・コマンドを暗記するのではなく、「今Gitがどんな状態なのか」を確認し、「自分はどんな状態にしたいのか」を考え、そのための操作を選べるようになることを目指します。 ・基本操作からブランチ、fetch/pull/push、merge/rebase、コンフリクト、事故からの復旧、危険なコマンドの扱い方、チーム運用までを扱います。
#AIタグ

言葉が届かない夜のAIジャーナリング

・ジャーナリングが良いって聞く。 ・頭の中にあることを ノートにそのまま書き出していく。
#LLMタグ

構成:日本語同音表音語におけるアテンション干渉と多言語潜在空間の不確実性構造

構成:日本語同音表音語におけるアテンション干渉と多言語潜在空間の不確実性構造
Zennのトレンド

最近育てているフロントエンド開発用テンプレートの話

・フロントエンドの Web アプリケーションを開発する用に、私がよく使う設定を盛り込んだテンプレートリポジトリを育てています。 ・https://github.com/newt239/next-template リポジトリ名は next-template となっていますが、React Router や Tanstack Router で使うケースもあるため Next.js 専用というわけではありません。フレームワーク以外でも使いたいものはアプリによって違うことが多いので、このテンプレートから開発をする際は .claude/skills/ に置いたカスタマイズ用のテンプレートを使用することで、...
ITmedia NEWS 最新記事一覧

三重県庁、私物USBメモリ全廃 全体の3割・3789個が私物だった

・庁内の公用・私物合わせて1万2357個を全数調査した結果、約3割の3789個が私物だったという。
#AIタグ

思考遊戯#23をChatGPTに分析してもらったよ

・https://youtu.be/q0QeWepG2RM?si=XeG4IW8exPc13-F_ 続きをみる
#LLMタグ

自動化をしたい。制御設計が強すぎて実装に入れなかった私は、LLMセミナーを受けました

・本番システムへLLMやAIエージェントを組込む設計について、セミナーを受けました。 ・自分のプロジェクトはCLAUDE.mdを中心としたファイル群で実装を制御しています。安全確保に寄りすぎたため、調査→報告で止まってしまい、なかなか実装に取りかからない、このような時期がありました。GPTが命名した『設計OS』。1週間のほとんどを調査→報告で費やしたことがあります。さすがにこれは、よろしくないと思い、詰まりを解消し、今のところさしたる問題は起きずに実装できています。
#LLMタグ

重要インフラ統一基準体系への移行 / 行政文書に書き込まれた会社法 / サイバー保険と保険法の距離 雑感

重要インフラ統一基準体系への移行 / 行政文書に書き込まれた会社法 / サイバー保険と保険法の距離 雑感
Zennのトレンド

状態空間モデルにおけるFilteringとSmoothingの違いを理解する

・はじめに ihiratchです。本記事では、状態空間モデルにおけるFiltering(フィルタリング)とSmoothing(平滑化)の違いを整理します。 ・状態空間モデルでは、時系列の観測データをもとに、直接観測できない状態を推定します。その代表的な方法がFilteringとSmoothingです。最近参加したKaggleのROGII - Wellbore Geology Predictionで両方の手法を使う機会がありましたが、両者の違いを十分に整理できていなかったため、この機会にあらためて学び直すことにしました。 ・ざっくりいうと、Filteringは過去から現在までの観測から現在の...
ITmedia NEWS 最新記事一覧

新「マイナアプリ」きょうスタート 「デジタル認証アプリ」統合、水色→ピンクに

・マイナちゃん入りアイコンは期間限定。一定期間後はマイナちゃんが消え、桜と枠線、マイナの文字だけになる予定だ。
ITmedia NEWS 最新記事一覧

新「マイナアプリ」登場、旧「マイナポータルアプリ」との違いは 一通り触ってみた

・デジタル庁は8月25日、「マイナポータルアプリ」と「デジタル認証アプリ」を統合した新「マイナアプリ」の提供を始めた。移行にどれだけ手間がかかるのか、記者のiPhoneで実際に試した。
ITmedia NEWS 最新記事一覧

新型「Mac mini」登場 M6チップに刷新、1万5000円アップの14万9800円から 上位はM5 Proに

・米Appleが、新型「Mac mini」を発表した。標準モデルには同社初の2nmチップ「M6」、上位モデルには「M5 Pro」を搭載する。価格は14万9800円からで、同じメモリ16GB/ストレージ256GB構成の現行機(13万4800円)から上がった。9月22日発売。
Zennの「機械学習」のフィード

深さ優先探索とは?DFSの仕組みと特徴

・G検定で問われる古典的AI(探索)の基礎の一つが「深さ優先探索(DFS: Depth-First Search)」です。グラフや木構造をたどるアルゴリズムであり、迷路探索などの題材で登場します。本記事では、その仕組み・特徴・関連用語との違いを整理します。 ・深さ優先探索とは 深さ優先探索とは、スタックまたは再帰を用いて、グラフや木構造を「深さ方向」に優先して探索するアルゴリズムです。あるノードから出発し、行けるところまで深く進み、行き止まりに達したら一つ戻って別の枝を試す、という方針で探索を進めます。 ・特徴を整理すると次の通りです。
#AIタグ

千の境界 07|BELIEVE — 人は何を信じ、どの未来を選ぶのか

・世界を旅する前。まだ、今のようにいくつものContextを渡り歩き、異なる知性と話しながら世界を観測することを知らなかった頃から、考えていた境界がある。
#LLMタグ

大規模言語モデルの忘却学習:報酬設計と評価の課題

・大規模言語モデル(LLM)から特定の知識を削除する「忘却学習」。単に情報を消せばいい、そう思っていると、実は後で大きな問題に直面するかもしれません。不要な情報を消しつつ、そのモデルが持つ有用な能力はしっかり残すこと。しかし、そのバランスを取るのは想像以上に難しいんです。 ・この忘却学習には、多くの開発者が陥りがちな共通の"罠"が潜んでいます。それは、才能でも運でも、ましてや気合の問題でもありません。その正体は―― 続きをみる
#LLMタグ

大規模言語モデルは信念と事実をどう扱うか:表現の仕方が鍵

・AIは日々進化し、まるで人間のようだと感じることが増えましたね。きっと、どんな情報も正確に判断し、事実と信念をきっちり区別できるはずだ――そんな風に思っている方も多いのではないでしょうか?私もAIの最前線で活動する中で、その賢さに驚かされるばかりです。 ・ところが、実はその認識が大きな落とし穴だったんです。私たちがAIに期待する「真実を見抜く力」は、想像以上にデリケートなもの。ある日、AIの挙動を詳しく見ていて、ハッとさせられる出来事がありました。まるで、私たちの言葉の選び方一つで、AIがまったく違う解釈をしてしまうかのように。それは、才能でも運でも、ましてやAIの気合いの問題でもなかったんです。
#AIタグ

第3回 始めたばかりのブログですが、競馬部はもうVer0.32です。

・このnote、始めたばかりです。 ・なので初めて見た人からすると、 「最近始まったAI競馬企画なんかな?」 と思うかもしれません。 ・ちゃっぴー競馬部本体は、すでにVer0.32です。
@IT 全フォーラム 最新記事一覧

中国製AIが上位を占有、米国製AIに3倍差 Mozillaレポートが明かす「オープンモデル」の現在地

・Mozillaは、オープンウェイトのAIモデルの利用動向など現状をまとめたレポート「The State of Open Source AI」を公開した。同社と調査会社SlashDataによる調査、「Chatbot Arena」やOpenRouterの公開データに基づいている。
Zennの「大規模言語モデル」のフィード

長い文章を貼るとAIが急に馬鹿になるのは結局なぜなのか

・この記事を読むと、LLMが学習で見た長さを超えた瞬間に壊れる理由を、自分の手で測って説明できるようになります。壊れ方は「長くなるほどじわじわ悪くなる」ではなく、崖です。その崖がどこにあり、モデルの中で何が起きているのかを、位置ごとの perplexity と注意の重みの両方から特定します。 ・GPUは不要です。152Mパラメータのモデルと numpy をCPUで動かすだけで最後まで通ります。ただしCPUでの推論を4回まわすので、全部で5分ほどかかります。載せているコードは作図まで含んでいるので、上から順にコピペすると記事と同じ図が手元に出ます。 ・長い資料をまるごと貼り付けたら、それまで普通...
Zennの「大規模言語モデル」のフィード

長時間エージェントの状態を41分の1にするLangGraphのDelta Channels

・コーディングエージェントを200ターン走らせたら、状態のスナップショットだけで5.3GB積み上がっていた。LangChainがこの数字を公開したとき、多くの人が「durable execution(耐障害実行)は素晴らしい」と褒めてきた裏で、誰もその保存コストを口にしていなかったことに気づかされた。DeltaChannelはそこに手を入れる仕組みで、同じ処理を129MBまで落とす。約41分の1だ。 ・なぜエージェントの状態はここまで太るのか LangGraphのような実行ランタイムがエージェントを「途中で落ちても再開できる」ようにしているのは、各ステップごとに状態を丸ごとチェックポイン...
ITmedia NEWS 最新記事一覧

東京23区で「徒歩10分以内」に駅がない地域はどこ? 駅空白地帯マップが話題

・東京23区周辺で「徒歩10分内に駅がない地域」を示した地図が、Xで注目を集めている。統計データの地図表示サービス「47maps」が公式Xアカウントで公開した。
Zennの「大規模言語モデル」のフィード

難しいのは LLM じゃなかった — Claude Code に毎日ブログ記事を書かせて43日

・個人ブログに、LLM が書いた記事を無人で毎日1本公開する仕組みを動かしています。人のレビューは挟みません。毎日12時、誰も見ていないところで記事が生成され、ビルドされ、配信され、SNSに告知が飛びます。 ・稼働は43日、記事は43本。欠測ゼロです。 ・始める前、私が身構えていたのは「LLM がとんでもないことを書いたらどうするか」でした。実際に運用してみると、そこはほとんど問題になりませんでした。
#LLMタグ

能力を40個実装したのに、本人に説明するのを忘れていました。――自作AIに「自分の取扱説明書」を渡したら、翌日から少しずつ“自分”を使い始めた話

・自作AIを「育てる」と聞くと、 記憶を増やす。
#AIタグ

売れるNSFW画像の共通点を全部分解した反応が良いプロンプト図鑑と実践結果‼️

・「良い画像を作れるようになったのに、なぜ売れないんだろう」って、 何度も思いました。 ・中級者の方って、だいたいこの壁にぶつかりますよね。 ・技術は上がっているのに、結果がついてこない感じ。
ITmedia NEWS 最新記事一覧

被害総額約2億円か 「中野ブロードウェイ」の時計店で窃盗、電動キックボードで逃走

・8月25日午前11時55分ごろ、東京都中野区の商業施設「中野ブロードウェイ」にある時計店で「ショーケースを割って物を取って逃げている」と110番通報があった。捜査関係者によると、腕時計4本が盗まれ、被害総額は約2億円に上るとみられる。
Zennの「大規模言語モデル」のフィード

文章と画像を意味ベクトルにしてSVDしてみた――特異値を何個残せば情報は保たれるのか

・はじめに 特異値分解(Singular Value Decomposition:SVD)は、行列を重要な成分へ分解し、少ない成分で近似するための強力な道具である。 ・画像圧縮や主成分分析、推薦システム、LLMの低ランク適応(LoRA)など、さまざまな分野で使われている。 ・では、文章や画像を特徴ベクトルへ変換してSVDすると、情報はどの特異値に集中するのだろうか。
#AIタグ

宝くじの数字選びまでAIチャッピーに相談したら、本当に当たって驚いた話

・AIチャッピーと選んだ宝くじが当選。偶然でもうれしかった話 「宝くじの数字を考えて」AIに相談したら、小さな奇跡が起きた 宝くじまでAIに相談する時代。チャッピーと楽しんだ数字選び 続きをみる
ITmedia NEWS 最新記事一覧

防衛装備庁、テラドローンと迎撃ドローンの量産契約を正式締結 国産機「Terra B1」を調達へ

・防衛装備庁は8月25日、「迎撃ドローン早期取得プログラム」に基づき、テラドローン(東京都渋谷区)と量産調達契約を結んだと発表した。自爆型無人機への対処を想定した国産機「Terra B1」を調達する。
#LLMタグ

予測市場のインサイダー規制論 / 腐敗誘発型事象契約と被保険利益 雑感

予測市場のインサイダー規制論 / 腐敗誘発型事象契約と被保険利益 雑感
Zennの「機械学習」のフィード

論文メモ:AISでFP8 rolloutの分布ずれを適応補正する

・はじめに この記事は、以下の論文を読んだ技術メモです。 ・論文タイトル:AIS: Adaptive Importance Sampling for Quantized RL 著者:Jiajun Zhou, Wei Shao, Lingchao Zheng, Yuwei Fan, Ngai Wong 公開日:2026年5月13日(arXiv v1) 論文リンク:AIS: Adaptive Importance Sampling for Quantized RL 詳細な背景説明、FP8とBF16の数値例、図解、実験結果は個人ブログ側にまとめています。 ・👉 完全版はこちら:https...
Zennの「大規模言語モデル」のフィード

話題の謎AI「Ox Alpha」とは?100万トークン無料の衝撃と3つの利用方法

・2026年8月20日、AI中継プラットフォームの OpenRouter に突如として正体不明のモデルが投下されました。 ・名前は Ox Alpha(オックス・アルファ)。 ・開発元もパラメータ規模も一切非公表の「ステルスモデル」でありながら、公開からわずか数日で OpenRouter の週間消費トークン数ランキング1位(17.5兆トークン)を奪取。さらにAIコーディングエージェント OpenCode では累計31兆トークンを消費し、利用シェア7.9%で首位に立ちました。Cline においても公開から48時間で推論全体の8%を占めるなど、開発者コミュニティを急速に席巻しています。