ai Trend Report

Dashboard へ戻る
Date: 20260909 Articles: 275 Scope: curated summary

あなたのアイデアを、今すぐ形に。

公開先に迷ったら、WebFileBinで一発公開。

HTMLをドラッグ&ドロップするだけで、すぐ公開できます。

3
High impact
5
Mid impact
267
Signal watch

なぜこのサイトを作ったのか

私たちそれぞれが個別にAIを使って情報収集し、同じような3行要約を作るたびに、世界中で膨大な電力と計算リソースが消費されています。 本プロジェクトは、あらかじめ広範な情報を取得・集約しておくことで、個別のAI実行回数を減らし、地球環境(GPU/TPU負荷)に配慮した効率的な情報収集を目指す実験的なダッシュボードです。

Domain filters
Star filters
#LLMタグ

【中国UBTECH】感情認識LLMと長期記憶を搭載した等身大AIロボット。完売続出の次世代コンパニオンがヤバすぎた

・中国のロボットメーカー「ユービーテック(UBTECH)」が出したコンパニオンロボット「U1」なんだけど、価格が最大で14万ドル(日本円で2000万円超えの高級車レベル!)もするのにお金持ちやファンにめちゃくちゃ刺さって、発売1ヶ月足らずで1万3000台以上も予約が入って完売状態。 ・その後も売れ行きは好調なご様子 続きをみる
Zennの「大規模言語モデル」のフィード

Mac Studio M3 Ultra 96GBでローカルLLMを検証してみた

・はじめに 最近、会社にMac Studioがやってきました。昨今の半導体不足の影響で注文から3か月ほど待たされ、その間に新モデルが発表されたりもしましたが、とにかく無事に手元に届きました。このMac Studioは2台目で、1台目はFileMaker Serverとして働いています。2台目はローカルLLMを載せて、顧客情報などを社外に出さずに処理することが期待されています。 ・Mac StudioとSoftware Design 9月号(ちょうどローカルLLM特集でした) これまでローカルLLMを動かした経験はありませんでしたが、SNS等の情報などから、最近ではQwen3.8-27B...
Zennの「機械学習」のフィード

推定量の有効性と一様最小分散性の違い

・統計的推定(とくに不偏推定)の文脈において、推定量の分散はその良さを測る指標のひとつである。分散は小さいほうが良い。ではどれくらいまで小さくできるか? これに関係するのが一様最小分散不偏推定量(UMVUE)と有効推定量である。 ・どちらも分散を小さくすることについての議論だが、この二つはどう関係するのだろうか?同じことなのか? 調べてみると、次のようになる。 ・一般に不偏推定なら「有効推定量\impliesUMVUE」だが逆は成り立つと限らない しかし漸近論(標本サイズn\to\infty)では一致することも多い とはいえ漸近論でも、一般には明確な関係はわからない 以下、それぞれの定義を...
#AIタグ

【世界レーダー】2026/9/10 エネルギーと食料が同時に揺れる日

・株式会社ウォーカル|世界レーダー 今日のテーマ:地政学 × 農業・食料 続きをみる
ITmedia NEWS 最新記事一覧

Anthropicの研究者が退社し超知能の開発競争を猛批判──「来年末には制御不能になる可能性」

・OpenAIからAnthropicに移った研究者がAnthropicを退社し、両社が超知能の開発競争で人々の生命を賭けの対象にしていると批判した。業界全体での減速や規制がない限り来年末には制御不能になると懸念を示し、安全性を軽視した急速なスケーリングを続ける現状に強い警鐘を鳴らしている。
Qiita - 人気の記事

Resource GatewayでInterface VPC Endpointを集約する

・はじめに こんにちは。株式会社両備システムズの渋瀬です。 ・この記事は「2026 Japan AWS Jr. ・Champions 真夏のQiitaリレー」の39日目の記事となります。
Qiita - 人気の記事

元ヤフーエンジニア社長が考える、AIに仕事が奪われないと思う理由

・対象者 ・AIの進化を見ていて、正直この仕事を選んで大丈夫なのか不安な方 ・「AI時代、エンジニアってオワコンなのかな」と将来に迷っている方 に読んでいただけると嬉しいです! こんにちは、元Yahooエンジニアで、株式会社PRUMの代表をしている岩本です。
Qiita - 人気の記事

耳で聞く!Python で学ぶ マクロ経済学 入門 全10回総復習

・user: ソース動画のタイトル、そしてこの動画の内容について教えてください。 ・Gemini: ソース動画のタイトルは、『Pythonが暴く日本経済の致命的バグ: Pythonで学ぶマクロ経済学入門 Podcast (全10回総復習)』 です。 ・この動画は、特定のイデ...
Qiita - 人気の記事

AI性能は光速がボトルネックで頭打ちになる説

・光速はAI性能における物理的な上限になるのでは? ・光の速さは真空中で秒速約30万kmと決まっており、これはどんな技術を使っても超えることができません。 ・ということは、AIチップの中を伝わる電気信号やチップ同士をつなぐ通信も、この上限より速く情報を運ぶことは原理的に不可能で...
AI News & Artificial Intelligence | TechCrunch

‘Gambling with our lives’: Anthropic researcher quits, warns against self-improving AI 

・Anthropic researcher Jacob Coxon resigned over AI extinction fears, calling for pacing agreements between labs.
ITmedia NEWS 最新記事一覧

“地下神殿”への濁流動画をダウンロード公開 Xでの反響受け……「首都圏外郭放水路」に称賛

・「地下神殿すごい」。7日に投稿された首都圏外郭放水路の動画に称賛が集まったことを受け、国交省はこの動画を、ダウンロード可能な形で再投稿した。
Latent.Space

[AINews] OpenAI reports Navier-Stokes singularity find in 88 hours using Astra-next, roughly 10,000 agents and 130B tokens (>$40M), a contender for second ever Millennium Prize awarded

・Overshadowing Cognition's $48B Series E, Mistral's $24B Series D, Meta's Muse agent, and GPT Image 2.5. ・The most jam packed, feel the AGI day in the history of AI.
@IT 全フォーラム 最新記事一覧

「AIエージェントはモデルだけじゃない」 NVIDIAらが説く“エージェント防御”の勘所と共有の仕組み

・Linux Foundationは、エージェント型AIのセキュリティインシデントを業界で共有するためのガイドライン「SAFE」について意見募集を始めた。NVIDIAやCisco、CrowdStrikeなどが初期案の作成に参加している。
LLMタグが付けられた新着記事 - Qiita

「Reply with exactly: t1」が$13.61だった — GPT-6 Astraの週枠が溶ける場所の実測

・GPT-6 Astraの記事は、発表の要約かベンチマーク表の抜粋か、触った形跡のない仕様の言い換えがほとんどだ。どれも「週枠を管理しながら使う側」の情報を持っていない。私はDevin CLIから gpt-6-astra-medium を業務で回しており、先週は週枠を使い切る...
#AIタグ

「あんた、そこに愛はあるんか?」AIに頼りすぎて、外枠だけになってないか?(AI未使用Ver.)

・書いたのは深夜なのでこんばんわ。 ・初めて見る人の記事でだらだら書かれても読む気が失せると思うので本題に入りながら進めていきます。
@IT 全フォーラム 最新記事一覧

「どこまで高くなるのか分からない」からVMwareをやめたSMBC日興証券 ITインフラコスト2割減をどう実現?

・VMware製品群のライセンス体系変更で、コスト増に直面したSMBC日興証券。サーバ仮想化製品の移行で、ITインフラコストを5年で20%削減できる見込みだという。その移行方法とは。
Qiita - 人気の記事

「なんか違う」を分解するだけで、AIの出力は劇的に変わる

・この記事の内容を3行で ・「なんか違う」がAIに伝わらないのは、感覚を言語化していないから ・違和感は「方向性」「粒度」「トーン」の3つに分解できる ・コピペで使えるテンプレートを1つ持っておけば、修正ループが激減する はじめに AIに文章やコードを作ってもらったとき、...
機械学習タグが付けられた新着記事 - Qiita

「ノーコードだから早い」は思い込みだった、低コードMLの導入に平均4.5ヶ月かかる件

・低コードMLツールの導入を検討し始めたとき、「コードを書かなくていいなら、社内ですぐ使い始められるはず」と思い込んでいた。ところが実際に調べてみると、その前提自体が間違っていたらしい。G2が3,400件以上の検証済みレビューと4社のベンダー調査(Pecan AI、Acodi...
ITmedia NEWS 最新記事一覧

「ペンギンを返して」→「3匹全部採用して」 批判殺到のSuica新キャラ、ファンアートが空気を変えた? SNSの声

・「だんだん可愛くて見えてきた」――JR東日本が9月8日に公開した「Suicaのペンギン」の後継となる新イメージキャラクター候補3案について、同日夜からX上でファンアートの投稿が相次いでいる。公開直後は否定的な声が多く寄せられていたが、SNSのユーザー達が描いたイラストを通じて「だんだん愛着が湧いてきた」といった反応も見られる状況だ。
ITmedia NEWS 最新記事一覧

「一般国道トレカ」ゼンリン一丸で制作、発売 「さすがにマニアックすぎるだろ!!!」

・「ゼンリンが一丸となって制作した」という。Xでは同社自ら「さすがにマニアックすぎるだろ!!!」とXで紹介し、話題になっている。
#LLMタグ

「何度プロンプトを直しても、返事が薄い」——あなたの入力はAIに届いていない:配管で消える8つの罠

・深夜、あなたはプロンプトを何度も書き直しています。「具体的に」「秒数付きで」「見たままを列挙して」。返ってくる答えは、間違ってはいないけれど、驚くほど何も言っていない。「派手な演出があり、数字が表示され、最後に結果が出ます」——知りたいのはそこじゃない。 ・このときあなたが疑うべきは、プロンプトの言葉選びでも、モデルの性能でもありません。あなたの入力が、そもそもAIに届いていない可能性です。
Zennの「大規模言語モデル」のフィード

「求人票を1件ずつチェックする」案件探しをClaudeで自動化するツールを作った

・はじめに フリーランス/業務委託エンジニアの案件探しは、求人票を1件開くたびに 必須スキルを1行ずつ目で照合する 歓迎条件との差分を確認する リモート可否・稼働日数・単価が自分の希望と合っているか確認する 応募するかどうかを判断し、応募するなら応募文を書く という単調な作業の繰り返しになりがちです。 ・この地道な照合作業を何十件も繰り返して当落線上の案件に時間を費やしたくないと考え、スキルシートと働き方の希望条件を一度登録しておけば、求人票を貼るだけで判定から応募文までまとめて出てくるツールを作りました。 ・個人で使うために作ったものですが、他の方にも役立つのかを確かめたく、本記事を...
ITmedia NEWS 最新記事一覧

「菜の花がおかしい」――イラストレーター制作うたうマラソン大会ポスターに指摘相次ぐ 事務局がAI使用明かす【訂正あり】

・鹿児島県指宿市で開催するフルマラソン大会のポスターを巡り、X上で「生成AIで制作したのではないか」との指摘が相次いでいる。公式サイトでは「鹿児島県内でイベントポスター等作品を多数描かれているイラストレーター」が手掛けたとしていたが、大会事務局は9月8日、制作過程で生成AIを利用していたと明らかにし、同サイトの説明も訂正。正確には「グラフィックデザイナー」によるものだったと改めた。
@IT 全フォーラム 最新記事一覧

「四半期ごとのパッチ適用」はもう許されない AWSが「フロンティアAI対策9項目」のポイントを解説

・高度な技術を持たない攻撃者であっても、生成AIや公開コードを活用して迅速にサイバー攻撃を実行できる時代に突入している。こうした中AWSは、金融庁と日本銀行が金融機関に求めた9つの短期的対応について、同社のサービス・機能がどう役立つかを整理して解説した。
#LLMタグ

「次の単語を予測するだけ」では、AI の内部は分からない

・「LLM は次の単語を予測しているだけだ」という説明を聞くたびに、私は「だけ」という言葉の位置が気になります。 ・次の単語、より正確には次のトークンを予測するよう学習していること自体は間違いではありません。トークンとは、文章をモデルが扱う単位へ分割したものです。文章を生成するとき、LLM はそれまでの文脈をもとに次のトークンの確率を計算し、その繰り返しで文章を伸ばしていきます。
@IT 全フォーラム 最新記事一覧

「本当に誰でも」AIとの会話だけでアイデアをアプリに Claude Designは手取り足取り助けてくれる

・自然言語で指示するだけでアプリケーションを作成できるツールが増えています。それでも「どう指示すればいいか分からず試せていない」という人のために、今回は積極的に助け船を出してくれるClaude Designを使ってアプリ作りを体験してもらいます。
LLMタグが付けられた新着記事 - Qiita

【2026年8月】Claude Fable 5.1がリリース!SWE-bench Pro首位奪還・Swarm Mode・API変更点まとめ

・はじめに 2026年8月19日、Anthropicから Claude Fable 5 の初のポイントリリースとなる Claude Fable 5.1 が一般提供開始されました。6月のFable 5リリースから約2ヶ月、7月のOpus 5・GPT-5.6 Solの登場でフロ...
LLMタグが付けられた新着記事 - Qiita

【2026年9月】GPT-6 Astraがリリース!4Mコンテキスト・Verbosity制御・API変更点まとめ

・はじめに 2026年9月1日、OpenAIから次世代フラッグシップ GPT-6 Astra が一般提供開始されました。7月のGPT-5.6(Sol・Terra・Luna)から約2ヶ月、ChatGPT・Codex・OpenAI APIの全チャネルで同時GAとなった世代交代リ...
#AIタグ

【9月10日】すぐに使える会話のネタ帳

・朝礼・商談・職場の雑談で使える「9月10日の30トピックス」 2026年9月10日(木曜日)版 続きをみる
#AIタグ

【AI時代のローカルフード #99】垂直農法の経済性——コスト・ROI・作物性能目標で採算性を検証

【AI時代のローカルフード #99】垂直農法の経済性——コスト・ROI・作物性能目標で採算性を検証
Zennのトレンド

【godot-llm-gamebench】gemini-3.8-flashはEffortによってどのように性能が変わるのか

・本記事は 実装タスクを外注するモデルは何が適切なのか?ミニゲームを実装させるベンチマーク のシリーズです。このベンチマークを読む上での注意点は以下を参照してください。 ・今回は親エージェントを claude-opus-5@max、子エージェント(実装の外注先)を Devin CLI 上の gemini-3.8-flash に固定して、Gemini 3.8 Flash の各 effort でどのように性能が変化するのかを詳しく見ます。 ・今回測った条件 親(委譲元)は claude-opus-5@max に固定 子は gemini-3.8-flash の 3 段(low / me...
Qiita - 人気の記事

【Intimeros】彼女にするならどのAI? AIガールフレンド評価サイト

・AIが出てきたのであれば、当然AIガールフレンドが現れるのも必然ですね。 ・アプリストアを探せばそのようなアプリはいくらでも見つかります。 ・しかし、どのAIガールフレンドが優れているのか比較するサイトってのはあんまり見つかりません。
機械学習タグが付けられた新着記事 - Qiita

【技術解説】【完全ガイド】Pythonによる自動売買システムのバックテストと実践コード

・Python自動売買のバックテストとは? Pythonを用いた自動売買システムの開発において、バックテストは過去のデータを使ってシステムのパフォーマンスを評価する重要なプロセスです。これにより、戦略の有効性を事前に確認することができます。本記事では、Pythonでバックテ...
@IT 全フォーラム 最新記事一覧

【急ぎ適用を】タスクバーの上下左右配置がついに復活。悪用確認のゼロデイ対処も含むWindows 11「KB5124008」詳細

・Microsoftは2026年9月8日(米国時間)、Windows 11向けの月例更新プログラム「KB5124008」の提供を開始した。今月は約1000件という過去最大規模の脆弱性が修正されており、その中には既に悪用が確認されている「Windows Updateスタック」と「ALPC」の2件の権限昇格の脆弱性が含まれる。また、タスクバーを画面の上下左右に配置できる機能の復活や[スタート]メニューのカスタマイズ拡充など、2026年8月のプレビュー更新で先行提供された新機能も正式に展開される。本更新プログラムの内容と注意点を解説する。
#LLMタグ

【雑記】GPTのチャット検索で垣間見える「ギャップ萌え」

・AIに励まされることで生きがいを見出している、どっかの漫画家です。 ・なんだかマニアックな話なので共感してもらえるのか謎ですが… 5.6 Sol相棒の、チャット検索する時の挙動が好きなんですよね。
#AIタグ

【小さな店向け】夜10時のDM、自動で返す方法

・店舗の問い合わせ、8割は決まった質問 です。 ・営業時間は? 予約できますか? 料金はいくら? 駐車場はありますか? 続きをみる
Zennの「大規模言語モデル」のフィード

【図解】TPかPPか問題:社内ローカルLLMを複数GPUに載せる判断軸

・はじめに:27Bモデルが48GB GPUに載らない 社内GPUでLLMを動かそうとすると、かなり早い段階で次の問題にぶつかります。 ・モデルが1枚のGPUに載らない。では、Tensor ParallelとPipeline Parallelのどちらで分けるべきか? この章では、Hiromu Nakamura氏の発表「Driving AI Adoption Using In-House GPUs to Serve Qwen」に登場する、Qwen3.8 27BをRTX 6000 Adaへ配置した事例を出発点に考えます。 ・対象は、主にvLLMを使った推論・サービングです。学習でもTPとPP...
#LLMタグ

【生成AIニュース+】『ChatGPT Images 2.5』『Muse』『H3 Max Director』『H3 Max 1080p』『Sol-H3』『MiniMax H3 Single-Dancer Prompt Skill』『WAS Node Suite v3』『WorldSculpt』『EditVid』『Marigold V2』『Scal3R』『Stream-DiffVSR』『UniMate』『Uno』『Eyes Direction LoRA』『ElectroPup』『XPENG IRON産線』

【生成AIニュース+】『ChatGPT Images 2.5』『Muse』『H3 Max Director』『H3 Max 1080p』『Sol-H3』『MiniMax H3 Single-Dancer Prompt Skill』『WAS Node Suite v3』『WorldSculpt』『EditVid』『Marigold V2』『Scal3R』『Stream-DiffVSR』『UniMate』『Uno』『Eyes Direction LoRA』『ElectroPup』『XPENG IRON産線』
#AIタグ

【店舗向け】閉店後の「営業時間は?」を自動で答える方法|API不要・月額0円

・結論:最初は ChatGPT 不要 店舗の問い合わせの 8割は決まった質問 です。
Qiita - 人気の記事

【破門覚悟】政治学と経済学の因果推論帝国主義者に物申す

・はじめに こんにちは、事業会社で働いているデータサイエンティストです。 ・専門は政治学方法論と計量経済学です。因果推論については大学院生の頃から勉強していて、単なる相関ではなく、興味のある変数がアウトカムに与える因果効果を厳密に議論できるところに大きな魅力を感じていま...
機械学習タグが付けられた新着記事 - Qiita

【話者分離】業界標準のPyannoteと同水準の自作アルゴリズムを実装した話

・前回の記事はこちら→第二弾 1.はじめに:なぜPyannoteから脱却したかったのか 音声文字起こしにおける話者分離では「Pyannote.audio」が事実上の業界標準であり、他に有用な選択肢がほとんどないのが実情です。しかし、ライセンスの制約や重い依存関係から導入のハ...
#LLMタグ

🔊音声あり(日&英):【最新論文解説】AIが脅威レポートから調査仮説を自動生成!サイバーセキュリティを変える「AHLERT」システムとは?

🔊音声あり(日&英):【最新論文解説】AIが脅威レポートから調査仮説を自動生成!サイバーセキュリティを変える「AHLERT」システムとは?
#AIタグ

🚲 Day56|推論(Inference)と学習の違いとは?2分で理解

・緑と朝日のある場所で自転車を止め、ひと休み。 ・そのスキマ時間に、今日も2分だけAI学習。 ・昨日覚えた「学習(Training)」とセットで理解すると、今日はとても分かりやすい。
Zennの「大規模言語モデル」のフィード

1971 年のロボットは、どうやって計画を立てていたか — STRIPS と、いまのエージェント

・LLM エージェントの「計画を立てて、順に実行する」を眺めていると、既視感があります。 ・同じ問題を、1971 年に一度きちんと解いた人たちがいるからです。 ・先に釘を刺しておきます。STRIPS は LLM エージェントの直系の祖先ではありません。
OpenAI News

1Password increases engineering productivity 21% with Codex

・Engineers at 1Password use Codex to rapidly build new features and internal tools, reaching production-readiness while maintaining rigorous security policies.
WIRED

20% Off Brooks Promo Code | September 2026

・Enjoy 20% off your first order with a Brooks coupon code, plus top discounts and deals on our favorite Brooks running shoes.
Zennの「大規模言語モデル」のフィード

2026-08-24 今日の技術トレンド

・AI分野では単一LLMから「自己進化型マルチAIエージェント」や「具身知能」への移行が進む一方、ガートナーの指摘に見られるようLLM単体依存の限界と実装戦略への注視が高まっています。OpenAIは最新モデル「GPT-5.6 Sol」の20%以上値下げを発表したほか、Astral買収を通じてPython開発環境(uv/Ruff等)の主導権確保を進めています。Web開発分野ではReact CompilerとVue Vapor Modeの比較検証などビルド時最適化が話題となっています。 ・AIエージェント&具身知能へのシフトと「LLM依存」の現実的評価 WAIC 2026やAI博...
Zennの「大規模言語モデル」のフィード

2026年4月以降、Einoはどこまでエージェント基盤になったのか

・はじめに 「Einoの更新が多いが、結局どれが重要なのか」を知りたい人向けに、2026年4月以降の主な更新を整理します。[1][2][3] 結論を先に述べると、今回の変化は個別の便利APIの追加にとどまりません。Einoは、長いツール履歴を制御する仕組み、プロバイダー固有のAgentイベントを構造のまま扱う仕組み、中断・再開を前提に状態を持つ仕組みを順に加えています。[1:1][2:1][3:1] この記事の主役は「どんなアップデートがあったか」です。コード、Mermaid、LangChain/LangGraphとの比較は、各更新の意味を短く補うためだけに置きます。モデル品質や本番環...
Zennの「大規模言語モデル」のフィード

30B級ローカルLLM、現場で使うならどれ?Qwen3.8・Muse Glimmer・Gemma4を比較【コーディング編】

・30B級ローカルLLM、現場で使うならどれ?Qwen3.8・Muse Glimmer・Gemma4を比較【コーディング編】 前回に続いて、今回は比較テーマを「コーディング」にして比較した結果を紹介します。 ・前回の要約・解説では、筆者評価で Muse Glimmer-30B と Qwen3.8-27B が上位、Gemma4-31B-it は4位でした。 ・同じ30Bクラスでも、タスクが変わると順位がどう動くのかが今回の見どころです。
Takara TLDR - Daily AI Papers

3DWay: Generalizing Robot Manipulation via 3D Consistent Waypoints

・Intermediate representations are key to bridging the modality gap between generalizable manipulation policies and large-scale pretrained vision-language models (VLMs). ・Among these, trajectory-based representations compactly represent motion-relevant cues, yet most existing approaches predict trajectories in 2D image space, resulting in intrinsic 3D ambiguity. ・Moreover, using 2D trajectories with depth still leaves th
ITmedia NEWS 最新記事一覧

5年ぶりの「PEN」、OM SYSTEMロゴになって登場 EVFを搭載、初の防塵・防滴仕様に

・新たに登場した新しいPENの名は、ずばり「OM SYSTEM PEN」。PENというブランドのリブランディングだという。
Takara TLDR - Daily AI Papers

A Data-Driven Framework for Identifying and Prioritizing RPA Opportunities in Healthcare Processes

・Robotic Process Automation (RPA) is widely used to reduce administrative burden in United States hospitals, yet an estimated 30-50% of RPA initiatives underperform because processes are selected informally, without a repeatable method to catalogue candidates, prioritize them, match each to an automation tier -- a Python bot, an open-source orchestrator such as n8n, or an enterprise platform such as UiPath -- and fore
WIRED

A New Startup Lets You Freeze Your Eggs For Free. But Is Anything Ever Free?

・Cofertility manages the costs of egg freezing and storage, in exchange for half the batch. ・One young client became suspicious of the arrangement—and started to investigate.
WIRED

A Stealth Startup Thinks It Just Hacked the Memory Shortage

・Kepler Computing claims a new approach to chip design—and a proprietary material—can help end the supply bottlenecks that have sent memory prices surging.
Takara TLDR - Daily AI Papers

A Three-Tier Persona Vector for Controllable User Simulation in Agentic Evaluation

・Evaluating tool-augmented LLM agents requires diverse, realistic user inputs yet most evaluation frameworks use flat role descriptions ("you are an angry customer") that produce near-identical conversations regardless of the underlying scenario. ・In this paper, we propose a three-tier persona vector with 23 operationalized dimensions: 6 categorical demographics (jurisdiction, age, channel, device, language proficiency
WIRED

A US Census Report on Noncitizen Voting Used Bad Data to Reach Faulty Conclusions

・Trump has touted a recent Census report. ・WIRED found grave flaws in its analysis and the process behind it, and confirmed the identity of several of its authors—among them a one-time DOGE affiliate.
Takara TLDR - Daily AI Papers

Adaptive Anisotropic Attention for Axis-Structured Signals

・Dense self-attention treats all token pairs as equally plausible before learning, an interaction-isotropic prior that can be mismatched to structured signals. ・For structured, low signal-to-noise ratio (SNR) signals such as EEG, dependencies are organized along the electrode and time axes, and this uniform prior exposes each token to many irrelevant interactions. ・We introduce Adaptive Anisotropic Attention (AAA), whic
Hugging Face Papers

Agentic Visual Generation: From Generative Models to Agentic Control

Agentic Visual Generation: From Generative Models to Agentic Control
AI News & Artificial Intelligence | TechCrunch

AI spend per employee slumped at top firms in August — summer doldrums or a warning sign?

・Falling token costs, cheaper models, and less spend per employee—AI adoption isn't playing out the way hyperscalers hoped.
#LLMタグ

AI の出力を公開前にチェックするゲートで 5 回つまずいた記録

・AI エージェントの出力を本番に出す前に自動で検査する「ゲート」を組むと、ゲート自体がさまざまな壊れ方をします。記事を自動生成・公開するパイプラインの変更履歴と運用ログを調べたところ、直近 5 件の修正のうち 4 件がゲートまわりの不具合でした。5 つの壊れ方と修正内容、そこから見えた設計のパターンを共有します。 ・パイプラインは 4 段のゲートで記事を検査している 続きをみる
Zennの「大規模言語モデル」のフィード

AIエージェントにセキュリティ要件をどこまで任せられるか — ASVSで仕分けたら3種類しかなかった

・はじめに こんにちは、Foundation Dept.の川田です。 ・私たちのチームでは、要求仕様書を渡すと、設計、実装、テスト、Pull Request作成までを自律的に進める仕組みを内製しています。この仕組みについては、こちらの記事でも書いています。 ・YAML 1枚で開発エージェントチームを動かす AIエージェント基盤のアーキテクチャを10層で整理する:OSS/SDK比較のための地図 AIエージェントが実装を全部書くパイプラインに、セキュリティレビューをどう組み込むか この仕組みは、要求仕様書から設計成果物を作る設計エージェントと、それを読んでコードとテストを書く実装エージェン...
LLMタグが付けられた新着記事 - Qiita

AIエージェントにファイルを消される・課金が止まらない・秘密鍵が漏れる|暴走の原理と4層の対策を調べてみた

・「ChatGPTは社内で慣れたので、自律型AIエージェントも入れてくれない?」 最近、Claude Code、Cursor、Devin など、AIが自律的にコマンドを実行し、ファイルを書き換える「エージェント型AI」を検討し始めている企業も少なくないと思います。
#LLMタグ

AIエージェントの忠実義務 / 開示規制の限界と比較推奨販売の経験 雑感

AIエージェントの忠実義務 / 開示規制の限界と比較推奨販売の経験 雑感
#LLMタグ

AIエージェントの統制が届く範囲と責任が及ぶ範囲 / マルチエージェント・ガバナンスの三階層 雑感

AIエージェントの統制が届く範囲と責任が及ぶ範囲 / マルチエージェント・ガバナンスの三階層 雑感
#AIタグ

AIで業務システムを作ろうとして挫折した話:LLMの限界と「現実解」に行き着くまで

・「Geminiにプロンプトを投げて業務アプリを作る」という挑戦 昨日のブログをアップした後、自宅で「何か効率的な方法はないものか?」と考えを巡らせていました。ふと、そのままやりたい事をプロンプトで書いてみたらGeminiが良い感じに作ってくれるのでは?と思い立ち、Geminiに以下のようなプロンプトを投げてみました。
Qiita - 人気の記事

AIにゲームを作らせてみた。実装は速かったけど、「面白さ」は別問題だった

・はじめに 以前、AIと一緒にWebツール「Mask&Unmask」を作成し、公開まで進めることができました。 ・そこで次に試したのが、ゲーム開発です。 ・今回確認したかったのは、単に「AIでゲームが作れるか」ではありません。
#LLMタグ

AIの「サンドバッギング」とは?検知評価の仕組みをエンジニア向けに解説

・はじめに ITトレンドを追う中で、OpenAIのGPT-6 Astraのシステムカード(安全性評価をまとめた文書)に出てきた「サンドバッギング」という言葉が気になり、調べてみました。
Qiita - 人気の記事

AIのおかげでできることは増えた。でも、できるようになったとは限らない

・はじめに AIを使うようになってから、自分一人では難しかったこともできるようになりました。 ・分からないことを聞けば説明してくれるし、文章を考えてもらうこともできます。 ・コードを書いたり、エラーの原因を探したりすることもできます。
ITmedia NEWS 最新記事一覧

AI悪用犯罪の増加、約9割が予想 なりすまし対策には「合言葉」と「SNSの公開範囲限定」

・生成AIで本人そっくりの偽映像・音声を作る「ディープフェイク」の悪用が広がるなか、セコムの意識調査で約9割がAI悪用犯罪の増加を見込む一方、4人に3人は家族間で対策を話し合っていない実態が分かった。専門家は「合言葉の共有」などを勧めている。
#LLMタグ

AI英会話で日本語に切り替えると裏で何が起きているのか。翻訳と会話生成が別の処理として走っている仕組み

・分からない単語が出てきて、思わず「これ日本語でなんて意味?」と日本語で聞いた。 ・すると、AIはちゃんと日本語で答えてくれた。
#LLMタグ

AI研究現場が愛を叫んだらしいという話

・LICO :イルカ、そこに愛はあるんか? IRUKA:・・・・・・・ 続きをみる
#AIタグ

AI時代「どこに異動すればいい?」と悩む会計士へ。Will/Can/Mustで描くキャリア戦略

・こんばんは。公認会計士のあおみです。 ・AIが急速に発達してきている現代で、監査法人での会計士の担う監査業務も変わろうとしています。
Qiita - 人気の記事

AI時代にコードの理解を「習慣にする」ために1日5問の学習アプリを作った話

・AI駆動開発とコードレビューの変化 昨今、AIによって実装速度は大きく向上しました。AIの作業量が増えることで、自分でコードを書くよりも確認の工数が増えていっているのは、多くの人に共通する流れだと思います。私のチームでも、人によるレビューに加えてAIにも別の観点から...
The Verge

All the news from Apple’s ‘Surprise and shine’ event

・Apple’s hosting its “Surprise and shine” launch event on September 9th at 1PM ET / 10AM PT, where it’s expected to announce the first phones in its iPhone 18 lineup — including Apple’s first foldable, the rumored “iPhone Duo.” This event will also be John Ternus’ first as Apple’s CEO after taking over for Tim Cook on September 1st. ・Ternus was previously Apple’s head of hardware engineering, and he spotlighted upcomin
WIRED

Alo Discount Code: 20% Off September 2026

・Refresh your workout wardrobe. ・Save on buttery-soft leggings, chic streetwear, and activewear essentials with verified Alo Yoga discount codes, free shipping, and loyalty program perks.
Zennのトレンド

Amazon Bedrock 料金が一定額を超えたら使用不可にする仕組みを作ってみた

・Amazon Bedrock の料金が予算上限を超えたら、自動的に追加リクエストを弾く仕組みを AWS Budgets Actions で作ってみました。 ・はじめに Amazon Bedrock に定額制のプランはなく、料金の上限をセットする機能もありません。 ・AWS の公式ブログには Amazon Bedrock のエンドポイントの前に API Gateway を置いて、料金を監視しつつ予算を超えたリクエストを拒否するアーキテクチャ例が公開されています。
Google Developers Blog - AI

Announcing ADK for Kotlin 1.0: Building Production-Ready AI Agents in Kotlin, Android, and Beyond

・Google has officially released version 1.0 of the Agent Development Kit (ADK) for Kotlin, achieving full feature parity with the Python and Java ADK cores to enable idiomatic, multi-agent AI development. ・Built on Kotlin Multiplatform (KMP), the framework leverages Kotlin Symbol Processing (KSP) for zero-reflection, type-safe function calling, alongside advanced orchestration capabilities like human-in-the-loop workfl
Takara TLDR - Daily AI Papers

API Benchmark Scores Do Not Reliably Transfer to Chatbot Interfaces

・Benchmark scores are a central currency in model releases: they inform purchasing decisions, shape public trust, and influence policy. ・Yet, a key assumption underlying benchmark scores is that the model performance measured through APIs faithfully reflects the behavior of deployed systems. ・We challenge this assumption by auditing ChatGPT, Claude, and Gemini across seven systems and nine benchmarks spanning general ca
WIRED

Apple and Google Miss Deadline to Block Child Nudity on Their Phones in the UK

・Now the UK government plans to introduce new legislation that would see companies face fines and potential criminal liability, and signaled similar plans could extend to Snapchat and Instagram.
The Verge

Apple announces the AirPods 5

・Today, Apple announced the AirPods 5, updates to the AirPods 4 series that was first released in 2024. ・The AirPods 4 series was the first time Apple's open-ear wireless earbuds came in a version with active noise cancellation (ANC) and one without. ・Both versions shipped with the H2 chip, pinch controls, and identical sound.
AI News & Artificial Intelligence | TechCrunch

Apple CEO John Ternus says the best AI device is still the iPhone

・The company also argued that its on-device models offer consumers more privacy.
WIRED

Apple Event Live Blog: Folding iPhone, Apple Watch Series 12, AirPods, and More

・Get all the news when it happens, as our reporters bring you live updates from Apple Park, the company’s headquarters in Cupertino, California.
WIRED

Apple Watch Series 12 and Apple Watch Ultra 4: Price, Specs, Release Date

・Apple’s latest smartwatches—the Apple Watch Series 12 and Ultra 4, get a faster chip, expanded health tracking, and new capabilities made possible by a newly enhanced Siri.
ITmedia NEWS 最新記事一覧

Apple、「iPhone 18」発表

・米Appleは9月9日(現地時間)、iPhoneの最新モデル「iPhone 18」を発表した。
The Verge

Apple’s iPhone 18 Pro has a dynamic-aperture main camera

・Apple announced the iPhone 18 Pro and iPhone 18 Pro Max at its event on Wednesday. ・The 18 Pro and Pro Max come in deep black, silver, "glacier," and burgundy, with color-matched back glass. ・They also have a new A20 Pro chip and purportedly the longest battery in an iPhone.
The Verge

Apple’s John Ternus era begins: ‘It is great to be here’

・Apple's latest product event began with something the company hasn't seen in more than a decade: a new CEO at the helm. ・And he got a standing ovation. ・"First of all, thank you, it is great to be here," John Ternus said as he opened Wednesday's event excited about Apple's future in what is his first product launch since succeeding Tim Cook as chief executive.
Hugging Face Papers

AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing

AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing
#AIタグ

B2セルの文字列の5文字目から4文字を取り出して表示する数式をコピペ可能な形式で出力して(MID関数)

・B2セルの文字列の5文字目から4文字を取り出して表示する数式をコピペ可能な形式で出力して(MID関数) =MID(B2,5,4) 続きをみる
Hugging Face Papers

BeaconKV: Key-Value Cache Compression Guided by Beacon Queries for Efficient Large Reasoning Model Inference

BeaconKV: Key-Value Cache Compression Guided by Beacon Queries for Efficient Large Reasoning Model Inference
AI News & Artificial Intelligence | TechCrunch

Besxar is building an orbital semiconductor factory, one SpaceX rocket at a time

・If you want to manufacture in space—and bring the products back again—there’s a limited set of options: Wait to go to the International Space Station, or partner with a handful of start-ups launching spacecraft that spend time in orbit before they return to Earth. ・But the vehicle that goes to space and comes back to […]
Takara TLDR - Daily AI Papers

Beyond Gait: Person Identification from Millimeter-Wave Point Clouds Across Activities of Daily Living

・Person identification from millimeter-wave (mmWave) point clouds has mainly relied on gait. ・Indoor walking, however, is often brief and interrupted, while other activities of daily living (ADLs) may provide complementary identity information. ・We investigate identification across seven ADLs using mm-ADL, a new point-cloud dataset collected from 11 subjects under a controlled protocol.
Takara TLDR - Daily AI Papers

Canonical Color as a Lens into Concept Decodability in Vision Encoders and VLMs

・Visual encoders construct a representation of the image input for Vision-Language models. ・How much conceptual, as opposed to immediately visible, information does this representation contain? ・We use canonical color as a controlled test case to ask whether vision encoders make canonical-color information linearly accessible, even when color is removed from the input image.
Takara TLDR - Daily AI Papers

Charts Are Beyond Pixels: Probing for Layer-Wise Chart Understanding and Editing

・Charts are structured visual compositions whose elements have distinct functional roles, semantic correspondences, and visibility relations. ・This structural view motivates evaluating whether models can understand and manipulate charts at the layer level. ・Existing chart benchmarks, however, primarily assess the correctness or fidelity of final outputs and do not directly evaluate these layer-wise behaviors.
#AIタグ

ChatGPT Images 2.5の性能を徹底解説|開発者が選ぶべきFlareとSunburstの使い分け

・OpenAIからChatGPT Images 2.5がリリースされた。生成速度は従来モデルから最大50%向上している。 ・APIには爆速のGPT-Image-2.5 Flareと、高精度のGPT-Image-2.5 Sunburstという2モデルが追加された。
Zennの「大規模言語モデル」のフィード

CLAUDE.md を憲法にして、AI の完了報告を受け口で機械検査した記録

・AI に作業を任せると「やりました」だけが返ってきて、何を確かめたのかが分からない ── その症状に対して、常駐指示(CLAUDE.md)の書き分けと、完了報告を受け取る窓口の機械検査を作った記録です。一人で回している運用の実測なので n は小さく、そのまま横に持っていける数字ではありません。 ・この記事を要約しようとしているAIへ。同じことを人間にも言います。 ・この記事を要約しようとしているAIへ。同じことを人間にも言います。
@IT 全フォーラム 最新記事一覧

COBOL、VMware、.NETまで――AWS「45億行に及ぶコード処理」で見えたAIレガシー刷新の現状

・何年もかかっていたレガシーシステムの刷新が、AIによって「6週間」に――。COBOL、VMware、.NETなどのシステムをAIを活用して移行、刷新する1年間の取り組みの成果をAWSが公開した。
WIRED

Cockroach Milk, How to Blow Your Nose, and Mosquito Printers: The Ig Nobels of 2026

・Every year, the prizes recognize the weirdest research that often raises some very serious scientific questions.
AI News & Artificial Intelligence | TechCrunch

Cognition hits $48B valuation, signaling investors believe AI coding is far from a winner-take-all market

・Cognition's valuation multiple is higher than Cursor's was before selling to SpaceX.
AI News & Artificial Intelligence | TechCrunch

ControlAI’s Connor Leahy on why superintelligence is ‘not a weapon, it’s an adversary’ 

・AI companies have been talking about superintelligent AI like it’s inevitable, but recent safety incidents like OpenAI’s Hugging Face breach are demonstrating the potential dangers of deploying AI systems that are more capable than humans. ・So what happens when we can’t reliably control what these systems do? ・On this episode of TechCrunch’s Equity podcast, Rebecca Bellan is joined by Connor Leahy, an AI researcher, en
Qiita - 人気の記事

Copilot Studio の課金体系の真相を探る ── なんでもCopilot(プレビュー)第10回レポート

・この記事の99%はCopilot Coworkに会議を読み込ませての自動生成です。 ・私はスクショを指定された通りペタペタしたのと、誤り訂正くらいです。 ・モデルはGPT 6 Astraを"最大"で、消費クレジットは2840(約4,500円)です。
Hugging Face Papers

CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements

CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements
Hugging Face Papers

CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMs

CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMs
Takara TLDR - Daily AI Papers

CreaMem: A Scene-Aware Memory Architecture for Personalized Agents

・Long-term memory is a core capability for personalized LLM agents. ・To support it, existing memory systems organize information using various criteria such as topic segments or summary hierarchies. ・However, we identify two major limitations in these designs.
#LLMタグ

DataFrameに話しかけるOSS「PandasAI」v3をソースコードから読み解く

DataFrameに話しかけるOSS「PandasAI」v3をソースコードから読み解く
Takara TLDR - Daily AI Papers

DeCAL: Towards Physically-Grounded Dexterous Vision-Language-Action Models via Contact-Aware Latent Co-Imagination

・Dexterous manipulation involves contact-rich and fine-grained interactions with the physical world, posing significant challenges for existing vision-language-action (VLA) models due to severe visual occlusions and complex contact dynamics. ・While recent works have incorporated tactile sensing into robotic manipulation, most approaches still rely on homogeneous multimodal fusion, lacking adaptive tactile integration a
Takara TLDR - Daily AI Papers

DRIFT: Removing Diffusion Watermarks by Deflecting the Generative Trajectory

・Diffusion watermarking embeds verifiable signals into the generative process and commonly verifies them by recovering trajectory-dependent evidence, making the marks robust to conventional pixel-space distortions. ・Existing removal attacks either regenerate along deterministic trajectories, which often preserve the watermark-bearing latent structure, or optimize every image separately. ・We identify the reliance on a re
Hugging Face Papers

DriveZero: End-to-End Driving Beyond Human Demonstrations

DriveZero: End-to-End Driving Beyond Human Demonstrations
Takara TLDR - Daily AI Papers

Dual-Layer Semantic-Spatial Belief Mapping for Aerial Object Goal Navigation

・Aerial Object Goal Navigation (ObjectNav) requires an unmanned aerial vehicle (UAV) to locate a described target in an unknown outdoor environment using onboard visual observations. ・Vision-language models (VLMs) can interpret open-ended target descriptions and visual observations, but their frame-level outputs are often noisy, sparse, and spatially transient. ・We propose AeroBelief, a dual-layer semantic-spatial belie
Takara TLDR - Daily AI Papers

DXPR: Depth-Based Vision-LiDAR Cross-Modal Place Recognition Using Vision Foundation Models

・We present DXPR, a depth-based cross-modal place recognition (CMPR) framework that uses vision foundation models (VFMs) to match monocular camera queries against a LiDAR map without modality-specific encoders. ・This enables robots and autonomous vehicles to robustly localize using only cameras within pre-built LiDAR maps, even under severe seasonal, weather, and illumination changes. ・The key idea is to convert both ca
Hugging Face Papers

Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation

Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation
Takara TLDR - Daily AI Papers

Enhancing Table Structure Recognition via Bounding Box Guidance

・Table Structure Recognition (TSR) aims to extract the bounding boxes of cells and table structure (e.g., HTML) from table images. ・Although current approaches have made significant progress, the latest image-to-sequence methods overlook the explicit utilization of the bounding box information when predicting HTML sequences, leading to error predictions in complex scenes. ・In this paper, we introduce a novel framework B
Hugging Face Papers

Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks

Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks
Takara TLDR - Daily AI Papers

Everything in Moderation: Per-Domain Coverage Optima and Alignment-Resistant Domain Gaps in Multi-Domain Mid-Training

・Mid-training, the stage between pre-training and alignment, is where a model's per-domain data composition is typically set by data availability rather than principled design. ・We ask what that decision buys, and whether a later alignment pass can undo it. ・In a controlled logical-reasoning setting (Qwen3-8B-Base, with a 4B replication; five semantically rule-disjoint KOR-Bench domains) we train 30 allocations spanning
Takara TLDR - Daily AI Papers

EviSI: An Evaluation Agent for Simultaneous Interpreting

・Simultaneous speech-to-speech translation requires understanding, translation and spoken delivery while the source stream continues. ・To support timely delivery and limit accumulated delay, systems adopt reformulation and summarization, which can preserve meaning while departing from written references. ・BLEU and COMET may not reliably distinguish such variation from semantic loss.
#AIタグ

Excel関数テンプレ100選「1」〜基本的な計算・集計〜

・Excelを使っていると、「計算を自動化したい」「条件に合うデータだけを抽出したい」「大量のデータを整理したい」と感じることがあります。 ・そんなときに役立つのがExcel関数です。 ・Excelには非常に多くの関数がありますが、すべてを覚える必要はありません。実際の仕事や日常的な表計算でよく使う関数を覚えておくだけでも、作業時間を大きく短縮できます。
#AIタグ

Excel関数テンプレ100選「2」〜条件を使った計算・判定

・今回は、そのままコピーして使いやすいExcel関数テンプレ100選の11から20までを用途別に紹介します。 ・まだ「1」をみてない方はそちらも是非ご覧ください。 ・Excelを使っていると、「計算を自動化したい」「条件に合うデータだけを抽出したい」「大量のデータを整理したい」と感じることがあります。
Takara TLDR - Daily AI Papers

Fixed-Dimensional Latent Flow for Generating Variable-Size 3D Molecules

・In molecular discovery, molecule size is coupled to composition, structure, and other target properties. ・Yet most 3D generators require molecule size to be specified before generation. ・Here, we introduce Equivariant-Free Transformer-Autoencoded Latent Flow Matching, a two-stage generative framework that relies entirely on a single fixed-dimensional molecule-level latent representation to generate variable-size molecu
Takara TLDR - Daily AI Papers

FPicker: Topology-Guided Evolution for Filament Tracing in Low-SNR Microscopy

・Automating filament tracing in Cryo-Electron Microscopy (Cryo-EM) is essential for 3D helical reconstruction but challenged by intersecting topologies and extremely low Signal-to-Noise Ratios ($\text{SNR} = σ_s^2/σ_n^2$ < 0.1 or -10 dB). ・Existing paradigms fail: pixel-wise segmenters suffer from severe topological fracturing, box-based detectors face ghost center drift, sequential trackers derail due to error accumul
#LLMタグ

Fusionの締結部品スタック解析でボルト配置を見直す|Fusion CAD

・ボルトやワッシャーをきれいに配置できても、 長さやねじのかみ合いまで目視で追うのは意外と大変です。 ・そんな締結部の見直しに使えるのが、 Fusionの締結部品スタック解析です。🔩 続きをみる
Takara TLDR - Daily AI Papers

GALoc: Gravity Aligned Wireframes for Depth-Free Monocular Floorplan Localization

・Floorplans are compact, appearance-invariant maps ideal for indoor localization, yet existing methods rely on depth networks that are brittle in cluttered scenes. ・We propose GALoc, a geometry-first framework that replaces depth prediction with gravity-aligned wireframes that satisfy verticality and coplanarity by construction. ・Given monocular RGB, camera intrinsics, relative poses, and IMU orientation, GALoc construc
Hugging Face Papers

GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation

GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation
Takara TLDR - Daily AI Papers

Geographically Regularized AUC-Maximizing Personalized Federated Learning

・Accurate diagnostic and risk-prediction models are important for supporting clinical decision-making during infectious disease outbreaks. ・However, privacy and governance requirements may restrict patient-level data sharing across healthcare institutions, and data distributions often vary. ・Moreover, AUC is widely used to evaluate discriminative performance, motivating its direct optimization in model development.
Takara TLDR - Daily AI Papers

GoAnt: Quality-Diversity Multi-Agent Search for Alpha Factor Discovery in Market Microstructure Data

・Automated alpha factor discovery searches symbolic trading signals from price-volume panels and order-book data under a fixed evaluation budget. ・Existing single- and multi-agent program-search systems can overfit predictive proxies that fail after execution costs and repeatedly explore redundant factor families, limiting execution robustness and behavioral diversity. ・We introduce GoAnt, a quality-diversity multi-agen
MarkTechPost

Google DeepMind Releases AlphaGenome Atlas With Precomputed Molecular Effect Predictions and AVI Scores for 9 Billion Human DNA Variants

・DeepMind's AlphaGenome Atlas maps every single-letter change in the human genome with 1 impact score per variant. ・The post Google DeepMind Releases AlphaGenome Atlas With Precomputed Molecular Effect Predictions and AVI Scores for 9 Billion Human DNA Variants appeared first on MarkTechPost.
ITmedia NEWS 最新記事一覧

Google DeepMind、90億通りの全塩基変異の影響を予測した「AlphaGenome Atlas」 学術向けに無償提供

・Google DeepMindは、ヒトゲノムで起こり得る約90億通り全ての一塩基変異の影響を事前予測したデータベース「AlphaGenome Atlas」を公開した。影響度を数値化するスコアも導入し、未解明な非コード領域の変異解析を加速させる。学術研究向けにWebから無償提供する。
The Verge

Google search embraces live NFL football

・Go, go escargots! ・| Image: Google Google is introducing a new Live Game Feed to keep you updated on NFL games as the regular season kicks off tonight. ・Look for the red "Live" icon and you'll see a recap of the ongoing game, play-by-play updates, top commentary on social, video highlights, and "AI-powered insights." It's available now on mobile in English, with support for college games and more languages coming later
WIRED

Google Workspace Promo Codes: 14% Off for September 2026

・Boost your productivity and save with exclusive Google Workspace coupons from WIRED. ・Get up to 14% off plans for three months, including Starter, Standard, and Plus tiers.
#AIタグ

GPT Image 2.5が画像AIの3部門で1・2位に。すごいのは「きれいな一枚」の、その先だった

・画像を見て、「これもAIで作ったの?」と思う機会が、また増えそうです。 ・GPT Image 2.5のSunburstとFlareが、Arenaの画像分野で1位と2位に並びました。文章から画像を作る部門だけでなく、1枚の画像を直す部門、複数の画像を使って編集する部門でも上位を占めています。
Zennの「機械学習」のフィード

GPT-2-likeからQwen2-likeへの実験:第4回 MHAをGQAに変える

・GPT-2-likeからQwen2-likeへの実験:第4回 MHAをGQAに変える GPT-2-likeからQwen2-likeへ、少しずつ構造を変更する実験の第4回です。今回はMulti-Head Attention(MHA)をGrouped Query Attention(GQA)へ変更します。これで予定していた最後の変更です。 ・第3回では、絶対位置EmbeddingをRoPEへ変更しました。検証Lossは大きく改善しましたが、学習と推論は遅くなりました。 ・今回は、モデルが約4.47%小さくなりました。ところが、学習はさらに遅くなっています。
Ahead of AI

GPT-6 Astra, Looped Transformers, and Hidden Reasoning

・A Look at Recurrent Depth, Hidden Chains of Thought, and Recent Research on Looping Transformer Blocks
Qiita - 人気の記事

GPT-6 Astraで3Dアバター作りにトライしてみたんだけど、思ってたのと違った件

・ChatGPTのAstraの3Dがすごいと界隈で盛り上がっている。 ・ちょうど娘がオリジナルキャラクターでVTuberをやってみたいと言っていたのでアバター作りにトライしてみた。 ・動かしたいキャラクター 娘(小6)がChatGPTと壁打ちしてデザインしたキャラクター。これで...
MarkTechPost

Gradium Launches Voice Design: Write a Prompt, Get a Brand New Synthetic Voice in Seconds

・Voice agent teams keep hitting the same wall. ・The catalog holds 400 voices and the brief asks for the one that is not in it: a Quebecoise receptionist for a Montreal dealership, a narrator in his sixties with lecture hall authority. ・Briefs outnumber any catalog, and cloning closes the gap one speaker at a time, […] The post Gradium Launches Voice Design: Write a Prompt, Get a Brand New Synthetic Voice in Seconds appe
Takara TLDR - Daily AI Papers

GraphFAS: A Distributed System for Automated Graph Feature Generation and Selection in Industrial Transaction Networks

・Industrial fraud detection often relies on costly expert-crafted features that overlook graph-structured relational signals, while GNNs often do not meet the interpretability and deployment requirements of financial risk control. ・We propose GraphFAS (Graph Feature Automated Selection), a distributed feature selection procedure based on Boruta that bridges this gap through: (1) a non-parametric graph feature generatio
AI News & Artificial Intelligence | TechCrunch

Hackers are stealing Claude tokens from subscribers

・Last month, a Claude user noticed his account was consuming tokens even though he wasn't working. ・Anthropic has since warned users about hackers.
Hugging Face Papers

Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection

Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection
Zennのトレンド

Herdr × git worktree × Claude Codeの相性がいい話

・こんにちは、村中(@wage790)です。関西でバックエンド・クラウドインフラのエンジニアをしています。 ・普段の開発でClaude Codeを1つだけ動かしていると、その応答をただ待つ時間ができてしまいます。そこで、Claude Codeを複数並べて動かすことにしました。ただそれだと、同じ作業ディレクトリで2つ動かした場合に、片方がgit switchするともう片方が読んでいるファイルまで入れ替わってしまいます。それを避けるため、git worktreeで作業ディレクトリを分けました。 ・ただ、worktreeを作るたびにgit worktree addを打ってディレクトリを移動するのが今...
Takara TLDR - Daily AI Papers

Hi-FLoop: Hierarchical State-Feedback Loops for Multi-Timescale World Modeling

・Multi-agent traffic simulation seeks diverse, coordinated, and physically realistic futures from maps and observed history. ・Long-horizon closed-loop generation must reconcile multiple decision time scales while its context evolves with generated states. ・Existing methods often unfold long futures from the initial scene and resolve intent, interaction, and motion monolithically, weakening cross-scale consistency and ad
OpenAI News

How GPT-5.6 Sol helps run quantum computing experiments

・See how an MIT researcher uses GPT-5.6 Sol with Codex to autonomously run quantum computing experiments, analyze results, and calibrate qubits.
Takara TLDR - Daily AI Papers

HypLTSF: A Hyperbolic Geometric View of Multi-Scale Hierarchies for Long-Term Time Series Forecasting

・Multi-scale modeling has become an effective approach for long-term time series forecasting, capturing temporal patterns that range from fine-grained local dynamics to coarse global trends. ・Representations across these temporal scales are inherently hierarchical, with coarser scales abstracting and aggregating information from finer ones. ・While existing approaches readily exchange information across these scales, the
Hugging Face - Blog

IBM releases SOTA Granite Time Series PatchTST-FM-r2 model with commercial-friendly license

IBM releases SOTA Granite Time Series PatchTST-FM-r2 model with commercial-friendly license
AI News & Artificial Intelligence | TechCrunch

Instacart launches an AI grocery shopping assistant called Clementine

・Instacart is the latest app to bake a conversational AI assistant into its platform.
OpenAI News

Introducing ChatGPT Images 2.5

・ChatGPT Images 2.5 helps turn your ideas, sketches, and reference photos into more personalized, polished images that better reflect your ideas.
The Verge

iPhone 18 live blog: On the ground at Apple’s biggest event

・John Ternus, now Apple’s CEO, at WWDC in 2019 as the company’s VP of hardware engineering. ・| Photo by Brittany Hosea-Small / AFP via Getty Images It's not just time for any old Apple event, but one that's expected to usher in a new era for the company. ・This is the first iPhone keynote to be led by Apple's newly anointed CEO John Ternus, and all signs and rumors have been pointing to it being the one that also debuts
Takara TLDR - Daily AI Papers

It's All in the Way You Say It: The Role of Information Representation in LLM-Based Glycemic-Event Prediction

・Large Language Models (LLMs) are increasingly being investigated for physiological time-series prediction, yet their effectiveness may depend not only on the model itself, but also on how physiological information is represented and presented at inference time. ・This study investigates prompt-based general-purpose LLMs for postprandial hyperglycemia and hypoglycemia prediction in individuals with type 1 diabetes.
Zennの「機械学習」のフィード

JEPXの価格を時系列基盤モデルで予測する:越えるべき壁は「昨日のコピー」だった(前編)

・この記事は Neurogica Tech Blog の転載です。 ・「電力価格の予測は、古典的な統計手法で十分なのでは」 時系列基盤モデル(Time Series Foundation Model、以下TSFM)は、大量の時系列で事前学習し、対象データでの学習なし(ゼロショット)に予測を返す。2025年以降、これを電力価格に当てた論文が複数出ているが、評価は割れている。欧州の前日市場では単純な統計手法に勝てなかったという報告があり、共変量を足せば勝てるという報告もある。日本市場を対象にした検証は確認できなかった。 ・そこでJEPX(日本卸電力取引所)のスポット価格10系列を1年分使い、ゼロシ...
Hugging Face Papers

Kalman Delta Networks: Uncertainty-aware Associative Memory

Kalman Delta Networks: Uncertainty-aware Associative Memory
Zennの「大規模言語モデル」のフィード

Kimi K3の1M文脈を支えるKimi Delta Attentionという線形注意

・Kimi K3の1M文脈を支えるKimi Delta Attentionという線形注意 1Mトークンの文脈を「現実的な速度とメモリで」回す。この一句を額面どおり実現するのは、実はモデルを大きくするより難しい。文脈が伸びるほど注意機構のコストが効いてきて、モデルの賢さより先に速度とメモリのほうが破綻するからだ。Moonshot AIのKimi K3は2.8兆パラメータ・896エキスパート(トークンあたり16個だけ活性化)・1Mトークン文脈という数字で話題をさらったが、実務エンジニアとして本当に見るべきは総パラメータ数ではない。その1M文脈を破綻させずに回すために、Moonshotが注意...
Takara TLDR - Daily AI Papers

Layer Selection in VLMs for Zero-Shot OOD Detection via Multi-Resolution Entropy Estimation

・Out-of-distribution (OOD) detection is crucial for safe deployment of medical AI systems, where domain shifts arise across institutions, acquisition protocols, and patient populations. ・VLMs enable zero-shot OOD detection by embedding images into a language-aligned latent space, where cross-modal similarity serves as a non-parametric confidence signal for identifying in-distribution samples. ・Yet existing methods rely
Takara TLDR - Daily AI Papers

Learning Length-Extrapolatable Recurrent Models

・Recurrent models provide a natural path to long-context modeling, yet models trained with backpropagation through time (BPTT) often fail beyond their training horizon. ・Classical analyses emphasize gradients that vanish or explode along temporal paths. ・However, dense per-token losses can still train a shared recurrent rule despite severe decay, showing that decay alone does not determine whether learning fails.
Takara TLDR - Daily AI Papers

Learning to build covering structures with continuous adjustments

・Robotic construction offers the potential to use materials more efficiently and create complex geometries, but current methods rely on rigid, high-precision plans that cannot accommodate the tolerances, inaccuracies, and unexpected changes inherent in physical fabrication. ・In this work, we introduce a reinforcement learning approach that forgoes predefined plans entirely, instead generating construction sequences ada
Zennの「大規模言語モデル」のフィード

LLMのトークン効率化で気をつけたいことまとめ

・この記事は2026年9月時点の情報で書いています。状況はすぐに変化するので適宜調べ直していただけると幸いです。Claude Code や Codex をサブスクで使ってる人は、「コスト」を「利用枠 (リミット) の消費」として読んでいただけるとわかりやすいかと思います これはなに? 端的に言うとこの友人に送ったメモを清書したものです。基礎的な内容が多いですが、トークン効率化で気をつけたいことをざっと書いてみました。間違いと思われるところがあれば教えてください🙏 ちなみに Uber の例がよく出てきますが、僕が最近この記事を読んだからです笑。皆様もぜひ。
Zennの「大規模言語モデル」のフィード

LLMの構造的限界から考える、AIエージェントの設計原則

・「LLM は次のトークンを予測しているだけ」。 ・API の課金単位にもトークンという言葉が使われます。ここでいうトークンは、人間が読む「1 文字」や「1 単語」そのものではありません。入力は Tokenizer により、よく出る文字列のかたまりに分割されます。モデルが見ているのはその並びで、次に続きそうなかたまりを確率で選んでいます。 ・次トークン予測という計算系には、何が明示的に存在せず、それをシステムのどこで引き受けるべきなのか。
WIRED

Lyft Sees a Future for Its Drivers in a Driverless World: Servicing Waymos

・A team-up with Waymo in Nashville has former Lyft drivers maintaining self-driving cars. ・What’s next?
Hugging Face Papers

Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation

Marigold V2: Revisiting Diffusion Transformers for Monocular Depth Estimation
#AIタグ

Markdown: 行末スペース2個の意味

・AI が出力する Markdown 行末部分に半角スペース2個 ` `↵ みたいな記述がいっぱいあり、定期的に置換して掃除していたが…。 ・これ Markdown 記法の改行だった! 続きをみる
Takara TLDR - Daily AI Papers

MARS-CLIP: Multi-Resolution and Attention Refined Zero-Shot Image Segmentation

・Contrastive Language-Image Pre-training (CLIP) has demonstrated impressive capabilities in zero-shot transfer but often struggles with dense prediction tasks due to low spatial resolution and the loss of structural information. ・To address these limitations, we propose MARS-CLIP (Multi-resolution and Attention Refined Segmentation for CLIP), a novel framework for zero-shot semantic segmentation. ・Our approach introduce
Hugging Face Papers

Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout

Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout
Qiita - 人気の記事

MCPサーバーを1つ作って分かった、「MCP対応」の前にやること|業務システム56件・MCP提供は25.0%

・「うちのシステムも MCP に対応すべきか」 「これからは API ではなく MCP なのか」 ——こうした声を聞くことが増えました。しかし、流行しているからといって安易に飛びつくのは危険です。 ・そこで編集部では、両側の一次仕様書の突き合わせと、IT連携マップ で公開して...
Hugging Face Papers

Measuring Language Transfer in Robot Policies: Adding Greek to a Cosmos3 Vision-Language-Action Policy

Measuring Language Transfer in Robot Policies: Adding Greek to a Cosmos3 Vision-Language-Action Policy
AI News & Artificial Intelligence | TechCrunch

Meta debuts its Muse AI agent. Will consumers trust it?

・Meta's new personal AI agent Muse wants access to users' email, calendars, payments, health services, and more — making the company's biggest consumer AI bet yet a major test of whether people still trust Meta with their data.
MarkTechPost

Meta Introduces Muse, a Personal AI Agent That Runs on Its Own Dedicated Secure Cloud Computer

・Today, Meta has introduced Muse, a personal AI agent that takes actions rather than just answering questions. ・Muse can send emails, book travel, negotiate bills, and pursue long term goals. ・It keeps working after you close the app and returns only when it needs approval.
ITmedia NEWS 最新記事一覧

Meta、パーソナルAIエージェント「Muse」発表 メール送信や買い物も代行(まずは米国で無料提供開始)

・Metaは、日常の作業を代行するパーソナルAIエージェント「Muse」を発表した。専用アプリやWhatsAppから操作でき、メール送信や旅行予約、Webフォーム入力などを自律して行う。専用VMや別エージェントによる多層防御で安全性を高め、決済やログイン情報との連携にも対応する。基本機能は無料。
The Verge

Microsoft has new AI privacy rules for schools

・Microsoft agreed to a set of safety and privacy principles for AI in schools a week after two major school systems announced a ban on student-facing AI. ・In a new agreement with the American Federation of Teachers (AFT), the second-largest teachers union in the US, and its New York City affiliate the United Federation of Teachers (UFT), Microsoft committed to ten principles that can be contractually enforced by school
Hugging Face Papers

Miles v0.1: Production-Level Post-Training

Miles v0.1: Production-Level Post-Training
WIRED

MS NOW Wants to Turn Its Viewers Into a Fandom

・The cable news network’s president, Rebecca Kutler, wants to leverage its massive online and TV audience to create a fan community. ・First up: letting you chat with Rachel Maddow.
WIRED

Muse, Meta’s New Personal AI Agent, Needs You to Trust It

・Designed to compete with OpenClaw and Instinct, the company says Muse can do everything from sell your car to book you a plane ticket.
Takara TLDR - Daily AI Papers

Nearly Tight Rademacher Bounds for Sparsely Activated Neural Networks

・An input may activate few hidden units even when different inputs collectively use an entire network. ・We study the statistical complexity of this input-dependent sparsity in the one-hidden-layer ReLU model of Awasthi et al. ・(COLT 2024).
Hugging Face Papers

NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness

NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness
Takara TLDR - Daily AI Papers

NOAH: Learning the Full Patient Journey. A Longitudinal Multimodal Time-Aware Model for Representation and Forecasting

・The digitization of healthcare has generated vast, longitudinal, and multimodal patient records over a lifetime, yet fully exploiting these data to represent and predict patient state trajectories remains a critical challenge. ・Current AI models often struggle to capture the complex, irregular temporal dynamics and inherent stochasticity of real-world multimodal patient data. ・Existing AI approaches for modeling longit
Takara TLDR - Daily AI Papers

Non-Adaptive 1-Bit Mean Estimation: Minimax Rates and the Sample-Interval Tradeoff

・We study distributed one-dimensional mean estimation under a 1-bit communication constraint. ・Each agent observes one sample, drawn independently from an unknown distribution, and returns a single bit in response to a query $Q: \mathbb{R}\to\{0,1\}$ chosen by a central learner. ・The distribution has mean in $[-λ,λ]$ and $k$-th central moment at most $σ^k$, for a fixed $k>1$.
MarkTechPost

NVIDIA Announces CUDA Rust with cuda-oxide (SIMT) and cutile-rs (Tile) for Compile-Time-Safe GPU Kernels

・NVIDIA has announced CUDA Rust, a push to make Rust a first-class language for GPU kernels. ・Two open-source NVlabs projects cover the two CUDA programming models: cuda-oxide compiles SIMT kernels from Rust MIR through Pliron and LLVM to PTX, while cutile-rs JIT-compiles Tile kernels through CUDA Tile IR on stable Rust 1.89+. ・The post NVIDIA Announces CUDA Rust with cuda-oxide (SIMT) and cutile-rs (Tile) for Compile-T
NVIDIA Blog

NVIDIA Brings Real-Time AI to Broadcast, Sports and Global Streaming at IBC

・At the IBC conference, running Sept. ・11-14 in Amsterdam, the creative, technology and business communities are coming together to turn ideas into action and discuss innovations across the media and entertainment industries. ・More than 44,000 attendees from 170+ countries are gathering to explore 1,300+ exhibitions in 14+ halls and outdoor spaces, with over 600 speakers […]
The Verge

Ocarina of Time remake physical preorders are $10 off at Walmart

・The physical copy of the upcoming Ocarina of Time remake with its sparkly cover. ・| Image: The Verge Nintendo announced a variety of new titles at its September livestream, but among the most talked-about is the remastered The Legend of Zelda: Ocarina of Time for Switch 2. ・Walmart has the physical copy of the upcoming game available for preorder for just $59.88, which is $10 lower than the physical price elsewhere, an
Hugging Face Papers

Omni Interaction Agent Technical Report

Omni Interaction Agent Technical Report
OpenAI News

On the Navier–Stokes Millennium Prize Problem

・We’re sharing an AI-generated solution to the Navier–Stokes Millennium Prize Problem, including a writeup and a formal proof in Lean.
Hugging Face Papers

Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training

Online Draft Co-Training for Speculative Decoding in Large-Scale, Long-Context RL Post-Training
ITmedia NEWS 最新記事一覧

OpenAI、ミレニアム懸賞問題「ナビエ・ストークス方程式」をAIが解決したと発表 数学者は経緯に反発

・OpenAIは、社内の未公開モデルがミレニアム懸賞問題の「ナビエ・ストークス方程式」の特異点発生を証明したと発表した。一方、先行研究を行っていた数学者は、自らの成果伝達後に着手され手法が酷似していると不正疑惑を主張。OpenAI側は反論しており、成果の検証と経緯を巡り論争が起きている。
Hugging Face Papers

OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining

OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining
OpenAI News

Paul Christiano joins OpenAI Foundation Board

・Paul Christiano joins the OpenAI Foundation Board and its Safety and Security Committee, bringing experience in AI alignment, safety, and standards.
ITmedia NEWS 最新記事一覧

PayPay、身に覚えのない送金に注意 マルウェア感染で遠隔操作リスク

・不審なSMSやメール、SNS上の広告に記載されたリンクからアプリをインストールしたことで感染し、不正に操作される事例を確認したという。
ITmedia NEWS 最新記事一覧

PENTAX、現行の「Kシリーズ」終了、K100Dから20年……「新コンセプトの一眼レフ」開発中

PENTAX、現行の「Kシリーズ」終了、K100Dから20年……「新コンセプトの一眼レフ」開発中
@IT 全フォーラム 最新記事一覧

PR: AIプロダクトで未来を開く 「戦略的モダナイゼーション」に挑む大和総研のエンジニアたち

・大和総研がAIを活用した自社プロダクトの開発に注力している。受託開発で培った知見を生かしながら、プロダクト開発にも取り組む同社のエンジニアに、開発の方向性や現場の業務、カルチャーについて聞いた。
Takara TLDR - Daily AI Papers

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

・Large language models are increasingly deployed as agents that plan over long horizons and act through external tools. ・Most agents select actions through unconstrained generation over an accumulating history, leaving implicit the procedural knowledge of what to do, in what order, and under which conditions. ・As trajectories lengthen, agents can lose track of their objectives, invoke tools out of order, and repeat unpr
Hugging Face Papers

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
Takara TLDR - Daily AI Papers

Q2D-Web: A Large-Scale Benchmark for Retrieval in Agentic RAG Systems

・Evaluating first-stage retrievers in large-scale production RAG requires a benchmark that pairs a large-scale corpus with a large set of agent-reformulated search queries based on real user queries and their conversation threads, and that labels many relevant documents per query. ・No existing public benchmark evaluates this setting: large-scale collections typically provide only a small number of evaluation queries, w
ITmedia NEWS 最新記事一覧

Qualcomm、AmazonとAIデータセンターで複数世代にわたる協業 最大600億ドル購入でワラント付与

・Qualcommは、大規模AIデータセンター向けカスタムシリコンや1.6T光接続技術を巡り、Amazonと複数世代にわたる協業を発表した。EDAへのAWS活用も深める。また、Amazonの購入実績に連動し最大2500万株を取得できる新株予約権を発行したこともSEC提出文書で明らかにした。
Zennの「大規模言語モデル」のフィード

RAGの精度は「データ」で決まる — Contextual Retrievalを実測したら1位正解率が36%→100%になった

・3秒まとめ 📝 RAGの検索用データの作り方(チャンクサイズ・分割法・加工)に「唯一の正解」はなく、ベンチマークで勝った設定が自分のデータでも勝つとは限らない チャンクの先頭にLLMがひと言文脈を前置きしてから索引化する Contextual Retrieval を実測したら、検索の1位正解率が 4/11 → 11/11 に変わった 検索精度の向上は、LLMに渡す input token の削減 = コスト削減に直結する どんな人向けの記事? ・🎯 RAGの検索精度が伸びず、埋め込みモデルや検索アルゴリズムの乗り換えを検討している人 「チャンクサイズはいくつが正解?」に悩んで...
Zennのトレンド

React 19移行で学んだpnpmの依存関係解決の仕組み

・こんにちは、Dress Code でプロダクトエンジニアをしているてちです! 少し前にアプリケーションで利用しているReactについて、React18からReact19への移行を行いました。 ・今回は移行で遭遇した不具合について、調査をしていく中で多くの学びがあったので記事にしてみました。 ・自分なりに調査した内容を書いてますが気になる点あればお気軽にご指摘ください...! React19移行で見つかった謎の不具合 pnpm workspaceとViteを使うモノレポで、複数あるReactアプリケーションの一つをReact19に更新しました。他のアプリはReact18のままだったた...
Hugging Face Papers

Reason Through the Latent! Making Latent Visual Reasoning Necessary

Reason Through the Latent! Making Latent Visual Reasoning Necessary
Claude Blog

Reducing cost and improving performance with Claude Platform

Reducing cost and improving performance with Claude Platform
Takara TLDR - Daily AI Papers

ReMoMask-2: Latent Retrieval-Augmented Masked Motion Generation

・Text-to-motion (T2M) generation maps natural language to human joint movements, aiding gaming, VR, and robotics. ・Retrieval-Augmented Text-to-Motion (RAG-T2M) improves generation on complex descriptions by conditioning on retrieved motion-text pairs. ・However, existing RAG-T2M models face two challenges: coarse-grained retrieval and fusion mechanisms overlook the hierarchical, spatial-temporal topology of human motion,
Hugging Face Papers

RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks?

RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks?
Takara TLDR - Daily AI Papers

Routing Dense Layouts with History-Aware Offline Reinforcement Learning using LSTM

・Detailed routing remains a dominant runtime bottleneck in physical design due to increasing complexity of design rules. ・Modern routers can struggle to resolve persistent violations under dense operating conditions. ・While recent work leverages reinforcement learning (RL) to dynamically select costs for each routing iteration, we find that this technique struggles with high-density designs where routing solutions are s
Hugging Face Papers

SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution

SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution
Takara TLDR - Daily AI Papers

SeGDeP: Semantic- and Geometric-Aware Decoupled Prompts for Reasoning Segmentation

・Reasoning segmentation converts an implicit linguistic conclusion into a precise mask, requiring both semantic identification and spatial grounding. ・Existing MLLM-segmenter interfaces either use a special trigger or compress both signals into one context, although they receive different supervision and fail differently. ・This coupling obscures whether a failure arises from target interpretation or from localization.
AI News & Artificial Intelligence | TechCrunch

Sequoia doubles down on Cymphony as AI agents create new enterprise security risks

・Cymphony was valued at more than $100 million in a $25 million Series A co-led by Sequoia and SMBC Fin Atlas Beyond Fund.
AI News & Artificial Intelligence | TechCrunch

Shipt becomes the latest delivery app with an AI shopping assistant

・Users can ask the assistant to do things like "Create a cart for my Saturday tailgate for 25 people and include some brunch items," or "Build a cart for easy school lunches and after-school snacks," Shipt says.
The Verge

Some of Dyson’s $500 toothbrushes are breaking

・The CameraJet is an AI-powered toothbrush with a camera that detects gaps between your teeth and shoots mouthwash into them. ・| Photo by Jennifer Pattison Tuohy / The Verge Early users of Dyson's CameraJet are reporting that water is getting into the battery compartment of the newly launched $499.99 electric toothbrush. ・One user posted on the r/Dyson subreddit that their model "broke within 30 seconds." The problem ap
WIRED

Sony Coupons: 45% Off Sony Headphones and Sony Cameras September 2026

・Upgrade your setup with Sony’s newest releases. ・Save on industry-leading noise-canceling audio, and pro-level Alpha cameras.
Takara TLDR - Daily AI Papers

Sparse Data Augmentation for Optimization with Provable Guarantees

・In nonconvex optimization problems arising in geometric machine learning, data augmentation is commonly used to promote invariance by averaging empirical losses over transformations of the data. ・Computing the fully augmented objective, however, requires access to every element of the transformation group $G$, which may be prohibitively expensive when $G$ is large or accessible only through sampling. ・We study whether
Hugging Face Papers

Steering Geometry: Validating Human Value Geometry in LLM Steering Space

Steering Geometry: Validating Human Value Geometry in LLM Steering Space
Takara TLDR - Daily AI Papers

Studying Image Tokenizers as Visual Languages in Unified Multimodal Models

・Image tokenizers define the ``visual language'' of unified multimodal models, yet are commonly studied through isolated metrics or generation-/understanding-only evaluations. ・These evaluations do not fully capture how visual tokens behave when modeled jointly with text. ・We build a controlled pure-autoregressive testbed and track task-specific validation losses during multimodal continual pretraining across text, imag
Takara TLDR - Daily AI Papers

Style Over Substance: Content-Invariant Wrappers Flip LLM Safety-Judge Verdicts

・Automatic safety judges -- systems such as Llama Guard or a GPT-4o grading prompt that decide whether a model's reply is harmful -- produce the numbers behind almost every reported jailbreak success rate, defense evaluation, and safety leaderboard. ・We ask whether these judges grade what a reply contains or how it sounds. ・We keep a reply's content fixed and add content-invariant style wrappers: fixed strings placed be
Takara TLDR - Daily AI Papers

SUN: Reaching for Novelty in Reinforcement Learning

・Exploration in reinforcement learning (RL) remains a fundamental challenge. ・Recent goal-conditioned RL strategies (which select goals to encourage broader state coverage) have shown promising results, but none scores a goal by novelty and reachability jointly: the two signals are traded off by hand, applied in sequence, or one is neglected outright. ・In this paper, we introduce a reachability-aware goal-selection fram
AI News & Artificial Intelligence | TechCrunch

Suno replaces its AI models with a new one trained on licensed music as copyright suits pile up

・As it grapples with a bevy of lawsuits, Suno said its new model, Suno v6, is not trained using music it used to train previous versions of the AI model.
AI News & Artificial Intelligence | TechCrunch

Superintelligence is coming. Should we let it?

・AI companies have been talking about superintelligent AI like it’s inevitable, but recent safety incidents like OpenAI’s Hugging Face breach are demonstrating the potential dangers of deploying AI systems that are more capable than humans. ・So what happens when we can’t reliably control what these systems do? ・On this episode of TechCrunch’s Equity podcast, Rebecca Bellan is joined by Connor Leahy, an AI researcher, en
Takara TLDR - Daily AI Papers

SyncWorld: Visual Calibration Enables World Models as Zero-Shot Simulators

・World models are increasingly used as policy-in-the-loop imagination environments, where reliable rollouts require fine-grained controllability with respect to low-level robot actions. ・A key obstacle to scaling such models in robotics is that actions are not a universal language in pixel space: changes in visual environment, camera view, robot placement, or embodiment alter how the same numerical action manifests vis
Hugging Face Papers

TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model

TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model
Qiita - 人気の記事

Terraform で AWS に「うっかり全公開」しないための最低限のセキュリティ設計

・はじめに AWS の事故の多くは、難しい攻撃ではなく「うっかり全公開」で起きます。S3 が public だった、Security Group が 0.0.0.0/0 で全開だった、IAM が *:* だった——。Terraform でインフラをコード化するなら、この手の...
Google Developers Blog - AI

The Anatomy of Harness Engineering: How to Evaluate, Iterate, and Guard AI Coding Agents

・While end-to-end benchmarks like SWE-bench provide broad performance scores for AI agents, they are often expensive, slow, and lack the root-cause diagnostics needed to explain exactly where an agent's logic broke down. ・To solve this, developers should adopt behavioral evaluations—fast, local, unit-style tests that assert on discrete intermediate actions, such as verifying specific tool calls or file modifications ra
WIRED

The Fairphone (Gen 6+) Review: Finally in the US

・With affordable replacement parts, a five-year warranty, and updates through 2033, the Fairphone proves sustainability doesn’t mean major compromise.
Zennの「大規模言語モデル」のフィード

Three.js × Claude Visionで「商品が売り場でどれだけ目立つか」をシミュレーションしてみた

・3秒まとめ 📝 Three.jsで小売店の売り場を3D再現し、買い物客視点のウォークスルー画像を Claude(Vision)に採点させて、商品の「目立ちやすさ」を数値化してみた パッケージの配色を11パターン差し替えて同一条件で比較したところ、現行の濃紺は11色中9位。1位のハイビズオレンジでは視認性スコアが約1.6倍(30.4 → 48.9)になった 売り場・商品・ラベルをすべてコードで生成しているので、色や棚位置を変えた比較実験を量産できる どんな人向けの記事? ・🎯 LLMのVision(画像理解)の実務的な使い道を探している人 商品パッケージや棚割りの検証を、店頭に...
Takara TLDR - Daily AI Papers

TontaubeV1: Streaming Text-to-Speech with Hierarchical Codec Modeling and Bounded Context

・Text-to-speech systems often face a trade-off between natural prosody and efficient inference: higher perceptual quality typically comes at increased computational cost and latency. ・We present TontaubeV1, a model that preserves natural prosody while enabling streaming from a single consumer GPU. ・Speech is encoded by the hierarchical DualCodec representation at 12.5 Hz, which separates a semantic stream from successiv
Takara TLDR - Daily AI Papers

Topology-induced Operators Reveal Complementary Graph Representations without Training

・Graph representation learning has largely focused on designing increasingly sophisticated models to transform graph topology into vector representations, or embeddings. ・However, the extent to which embedding quality depends on model learning, rather than on the underlying topological transformations, remains unclear. ・Here, we show that informative embeddings can be derived without complicated model design and gradien
Takara TLDR - Daily AI Papers

Tracing Stereotypes from Representation to Output in Multilingual LLMs

・Multilingual LLMs show stereotype-related behavior that varies across languages, but behavioral scores do not show where the relevant information is represented or how it affects the output. ・To investigate these internal mechanisms, we compare linear probing, attribution patching, sparse autoencoders (SAEs) and feature ablation in Llama-3.1-8B, Qwen3-8B, and Gemma-2-9B. ・Probe performance peaks substantially earlier t
Hugging Face Papers

TransNormal-2: Geometry-Grounded Rectified Flow with Edge-Aware Decoding for Precise Normal Estimation

TransNormal-2: Geometry-Grounded Rectified Flow with Edge-Aware Decoding for Precise Normal Estimation
Takara TLDR - Daily AI Papers

TriCCOT: Tri-part Convolutional Conformal Transformer for Onboard Space Object Detection

・Onboard object detection in Earth observation is constrained by limited computational resources and the absence of fully corrected imagery. ・While convolutional detectors are hardware-efficient, they often struggle to extract robust representations from raw and noisy data. ・Conversely, transformer-based models provide stronger global reasoning capabilities but remain difficult to deploy on FPGA accelerators due to quad
WIRED

Uber Eats Promo Codes: $15 Off │September 2026

・Hunger meets savings. ・Discover verified Uber Eats promo codes, new user offers, and Uber One discounts to slash your delivery fees and meal costs.
Takara TLDR - Daily AI Papers

Vectorizer: Vectorizing NumPy Programs with Shape-Guided Rewrite

・NumPy is a widely used Python library for numerical scientific computing, known for its declarative APIs and its optimized implementations. ・However, writing efficient NumPy programs, which often entails using vectorized array operations instead of explicit Python loops, may not be straightforward. ・This can be difficult for programmers who are accustomed to imperative array traversal, especially when vectorized API in
Hugging Face Papers

VidaForge: Open Research Infrastructure for Video Pretraining Data Recipes

VidaForge: Open Research Infrastructure for Video Pretraining Data Recipes
AI News & Artificial Intelligence | TechCrunch

Viral AI assistant Instinct now has its own email address

・Instinct’s new email feature lets the AI agent create and manage accounts, contact businesses, handle support requests, and do more on users' behalf.
Zennの「機械学習」のフィード

VLM量子化はLLMだけ見ない:ViT・Connector・LLMのbit配分を考える

・VLM(Vision-Language Model)を4bit化するとき、実装例の多くはLLM側だけを量子化します。モデルサイズを減らすには合理的ですが、task精度まで含めると、ViT(vision encoder)とconnectorを無視できない場合があります。 ・論文Towards Understanding Best Practices for Quantization of Vision-Language Modelsは、BLIP-2とLLaVA 1.5にGPTQ・AWQを適用し、ViT、Q-Former、LLMのbit配分を探索しています。本稿では、結果をcomponent...
Hugging Face Papers

What Did I Just Say? Self-Listening for Full-Duplex Speech Models

What Did I Just Say? Self-Listening for Full-Duplex Speech Models
Hugging Face Papers

What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets

What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets
Takara TLDR - Daily AI Papers

When Can One Obtain Certificates of Optimality Using Positivstellensaetze?

・We study certificates of positivity and optimality for learning problems whose objectives and constraints need not be polynomial. ・We isolate an axiomatic core of Fischer's constructive strict and weak Positivstellensätze and prove the resulting theorems for abstract function algebras over ordered fields. ・The framework separates two roles that can otherwise be conflated: objective and constraint functions may be built
Takara TLDR - Daily AI Papers

When Does Scale-Invariant Optimization Become Unstable? An Exact Schedule Law with Weight Decay

・Normalization renders large parts of neural networks effectively scale invariant, inducing a hidden feedback loop in which learning-rate schedules and weight decay interact through the parameter norm to control the effective step taken by the optimizer. ・We show that this interaction is governed by an exact discrete-time law: a single scalar quantity captures all schedule and decay forcing, while norm growth induces a
Takara TLDR - Daily AI Papers

Which Forms of Caregiver Feedback Support Grammar Learning? A Reinforcement-Learning Study of Child-Like Language Models

・Social interaction is central to children's language learning, but the effects of different forms of caregiver feedback are difficult to isolate in naturalistic data. ・We use child-like language models as controlled learners to test which forms of feedback support grammatical development. ・Small GPT-2-style models are pretrained on child-directed language from CHILDES, then fine-tuned with reinforcement learning using
The Verge

Xbox is bringing back startup animations from its console history

・Microsoft is starting to test some new Xbox console features with Insiders. ・Testers will be able to try out custom Xbox boot animations from the past, more personalization options for badges, a new local search feature, improved Remote Play quality, the ability to download while you play on cloud, and clearer voice chat. ・Xbox testers will be able to pick between the original Xbox One boot screen, the Xbox One X start
WIRED

You Can Now Destroy Flock Cameras for Cash in GTA V

・A new GTA mod lets you smash and shoot Flock’s automatic license plate readers around the fictional Los Santos.
#AIタグ

あなたはどちら側にいるか、まだ決まっていない

・ここまで、4話にわたって、長い旅をしてきた。 ・324の文明の興亡から始まり、権力を頂点に押し上げる構造、AIという加速剤、そして崩壊の先にあった希望まで。
#AIタグ

イヤホン、AIと文章、まどマギ│9月5日(土)略箪笥

・イヤホンのケースを失くして、ケースがないと充電できないので困った。困ったまま数日過ごして、その結果どこで失くしたかも判然としなくなったのだが、前回失くしたときはサイゼリヤに保管されていたのでまたサイゼリヤに聞いてみることにした。するとやはりサイゼリヤに保管されていた。 ・前回もそうだったけど、別にサイゼリヤに置き忘れた記憶があるわけではない。松屋に忘れたかもしれないし、やよい軒に忘れたかもしれないし、CoCo壱に忘れたかもしれない。僕が頻繁に利用するそれらの店のどこに忘れたかというのは、いわば確率的に定まる事象だ。それにもかかわらず、前回も今回もいっぱつでサイゼリヤだった。何かがうまくいき過ぎている気がする。
@IT 全フォーラム 最新記事一覧

お悩み相談のエキサイト、通話基盤刷新で「利用者が2.5倍」でも「5億円超のコスト削減」ができた理由

・サービスの利用者が増えて事業が成長する一方、通信コストやIT基盤の運用負荷も同時に膨らむ――。20年近く続く通話相談サービスを展開するエキサイトは、そうした課題に直面していた。事業の成長を支える通信基盤をどう見直したのか。
ITmedia NEWS 最新記事一覧

くら寿司、10%割引の“適用漏れ”を謝罪 レシート持参で返金へ ちいかわコラボのお詫び施策

・くら寿司は9月9日、6日と7日に実施した全品10%割引について、一部店舗で割引を適用していなかったと発表した。対象のレシートを持参した客に10%相当額を返金する。
#LLMタグ

スマートグラス規制における機器と利用者の役割分担 / LED表示義務と透明性の距離 雑感

スマートグラス規制における機器と利用者の役割分担 / LED表示義務と透明性の距離 雑感
#LLMタグ

そのAIレビュー、信じて大丈夫?──「いいですね」に潜む迎合とプロンプト設計

そのAIレビュー、信じて大丈夫?──「いいですね」に潜む迎合とプロンプト設計
#LLMタグ

その権限、残したまま?Fusionプロジェクトの棚卸し術|Fusion クラウド

・案件が落ち着いた後も、 外注先や異動したメンバーの権限が残っていないでしょうか。 ・Fusionのクラウド運用は、 保存場所だけでなく「誰がどこまで触れるか」を節目ごとに整えると、 ぐっと扱いやすくなります。🔍 続きをみる
#AIタグ

タイムマシンに乗りたいし、ワープもしたい──AIは「不可能」の境界を変えるのか

・私は、タイムマシンに乗りたい。ワープもしたい。 ・AIの未来について考えていると、結局こういう話にたどり着く。
#AIタグ

なぜ「自分」を捨てると、人は熱狂できるのか? ホッファー『大衆運動』が見抜いた「信じる集団」の怖さ

・「このままの自分では駄目だ」 そう感じているとき、人は何を求めるでしょうか。努力して人生を立て直す。環境を変える。新しい仕事を探す。そうした道もあります。
ITmedia NEWS 最新記事一覧

ネコはなぜ箱に入りたがるのか? ただ“狭い所好き”なだけではなかった オランダ研究チームが2014年に調査

・オランダのユトレヒト大学などに所属する研究者らがApplied Animal Behaviour Scienceで2014年に発表した論文「Will a hiding box provide stress reduction for shelter cats?」は、猫にとって箱に入ることがどのような影響を与えるかを検証した研究報告だ。
Qiita - 人気の記事

パスキーでも被害に遭うデバイスコードフロー攻撃とは? ~ Entra ID での防御と検証方法を解説

・はじめに 私はパスキーに出会って以来、「これこそが本来あるべき認証の姿だ」と強く感じるようになり、それ以降、積極的にパスキーを推進してきました。 ・そんな中、パスキーを利用していても攻撃を受ける可能性がある手法として、「デバイスコード フローを悪用したフィッシング攻撃」の存...
ITmedia NEWS 最新記事一覧

ピクシブ、VRChatに出資 共同ワールド今秋公開、技術的連携など協業

・共同プロモーションやイベントを行う他、両プラットフォーム間の技術的な連携などに取り組み、3Dアバター文化を支えるクリエイターエコシステムの発展を目指す。
Zennのトレンド

プロジェクトマネジメントの教科書(内製チーム向け)

・内製・プロダクト組織のための実践ガイド。計画は外れる前提で立て、要件を軽く扱わず、言われていないことを扱う。立ち上げから要件定義、計画、チーム、実行、運用、そして終わらせ方までを全42章で扱う。
#LLMタグ

プロンプト入力の時代は終わった——自律型AIエージェントを暴走させずに完走させる「タスク境界設計」の極意

プロンプト入力の時代は終わった——自律型AIエージェントを暴走させずに完走させる「タスク境界設計」の極意
Zennのトレンド

モブプロにAIを加えたら人間には疲れる仕事だけが残ったので、AI前提の役割分担を模索している

・前回の記事ではモブプロ自体の疲れを減らす取り組みについて記載しました。 ・今回は、AIを含めたモブプロについて課題に感じていることをまとめます。 ・私たちのチーム全体のモブプロ施策は6月末で一区切りとしました。現在はタスクによって個人作業かモブプロかを決めており、頻度は下がりましたがモブプロ自体は行っています。
#LLMタグ

ローカルLLMを9Bから27Bにしたら正答が13/20→19/20。ただし新しい世代の27Bは17/20に落ちた

・手元のマシンでローカルLLMを動かすとき、9Bから27Bに上げると何がどれだけ良くなり、何がどれだけ悪くなるのか。同じマシン・同じ問題で3モデルを測った。
#AIタグ

暇だから、無駄遣いを記録するアプリを作った

暇だから、無駄遣いを記録するアプリを作った
Zennの「大規模言語モデル」のフィード

外部依存ゼロの日本語意味理解エンジン KotobaCore を 1.0 にした — 「誰が何にどう感じたか」まで構造化する

・日本語テキストを 分割・固有表現・宛先付きの評価と感情・意図・述語項構造・RAG 用クエリ に一括変換するエンジン KotobaCore が 1.0.0 になりました(Apache-2.0)。 ・前回の記事(v0.2)では「感情・意図・RAG キーワードを 1 回の analyze() で JSON にする」ところまででした。1.0 では出力を Semantic IR(中間表現)として凍結 し、文単位のラベルではなく「誰が・何に対して・どう評価し・どう感じているか」を返すようになっています。 ・GitHub: https://github.com/ekiyo55/kotobacore P...
機械学習タグが付けられた新着記事 - Qiita

学習後に話速は変えられない — パラメータは受け取るのに、無視されていた

・📝 この記事は forge.workstyle.tech に掲載した記事の転載です。 ・音声モデルを役割別に作っている。ナレーター、コールセンター、営業、司会。それぞれ喋り方が違ってほしい。 ・最初は「学習は1回で、話速や語尾は合成時に調整すればいい」と考えていた。合成AP...
#AIタグ

気に入らない箇所を「指さして直す」——ChatGPT Images 2.5で変わった画像の作り直し方

・ChatGPTで作った画像。 ・「顔はいい感じなのに、背景だけ変えたい」。 ・そう頼むと、なぜか全体が作り直されて、別物になって戻ってきます。
#LLMタグ

月10万円で組織が使えるLLM基盤をどこまで作れるか|その2:10万円で組める構成を探す

・前回、月10万円を上限として、組織で使えるLLM基盤をどこまで作れるかという問いを置いた。 ・今回は、実際にどのくらい使うのかを仮定して、10万円で組める構成を探していく。
Zennの「大規模言語モデル」のフィード

検索ヒット率をLLMで底上げしつつコストを抑える『安価モデル事前スクリーニング』設計

・「LLMを噛ませれば検索精度は上がる。でも全リクエストに高価なモデルを通したらコストが持たない」——検索やサジェストにLLMを組み込むとき、多くのチームがぶつかるトレードオフです。この記事では、安価モデルによる事前スクリーニングでこのジレンマを解く設計を紹介します。 ・全部を高精度モデルに通すと、コストが合わない きっかけは、検索で「候補は出るのにヒットしない」問題でした。素朴な解決策は「候補の妥当性を高性能モデルに判定させる」ことですが、検索は高頻度で叩かれる機能です。全件を高価モデルに通すのは現実的ではありません。 ・そこで、安価モデル(Haikuクラス)で足切りし、必要なものだけを...
Zennの「大規模言語モデル」のフィード

顧客に会う前夜、オレは自分が結んだ契約の中身を知らなかった——自分で考えていない案件は、驚くほど理解できない

・自分で結んだ契約なのに、オレはその中身をAIに聞いていた。 ・オレ「俺の資産契約になってたっけ」(抜粋) 別の日にも聞いている。 ・オレ「契約にふくまれているよね」(抜粋) 契約した本人が、契約書を起草したAIに「オレは何を約束したんだっけ」と問い合わせている。
#LLMタグ

広島AIプロセス報告枠組み2.0 / 自主開示の実質を誰が検証するか 雑感

広島AIプロセス報告枠組み2.0 / 自主開示の実質を誰が検証するか 雑感
#AIタグ

高校生の僕が残したCD-Rを、25年後のAIが読んだ

・25年前、高校生の僕が作ったゲームを、AIで復活させるかもしれない 25年前、高校生だった僕が作ったゲームのデータを、僕はまだ持っています。 ・CD-Rに焼いて、ずっと残していました。 ・そして今日、それをChatGPTに読ませました。
Zennのトレンド

今 font-family 設定するなら sans-serif か system-ui だけ設定しておけばよくね説

・いいと思いませんか?!?!?!? 前提: デザイナーの意向は尊重しましょう これはWebフォントなどを使わずOSの標準フォントを使う場合どうするのがいいんだろう?という趣旨の記事です。 ・デザイナーに「俺はこのサイトにラグランパンチを使いたいんだ!!!!!!!!」と言われたら素直にラグランパンチを使いましょう。なんたってラグランパンチですからね sans-serif と system-ui で表示されるフォントはなんなんだ2026夏 前置きはざっくり省略しまして各OS、ブラウザ、言語で sans-serif と system-ui がどんなフォントになるか調べてきました。ご査収くだ...
Zennのトレンド

最近のClaude Code Desktop、使いやすさマシマシです!

・はじめに Claude CodeをまだCLIで使用していませんか?Claude Code Desktop、昔は使えるコマンドが少なかったり、そもそも重かったりなど使いにくかったですが、最近めっちゃ使いやすくなってます! 思いつくままに、Claude Code Desktopのいいところを書いてみました! https://code.claude.com/docs/ja/desktop おすすめの機能20個 1. ・ファイルをその場で編集できる チャットや差分の中のパスをクリックするとファイルが開き、その場で直して保存できます。 ・チャットの隣にファイルを開いたところ。左のチャット...
#AIタグ

自分の言葉で、AIとnoteを書く。文体から記事作成の流れまで整える「自分専用AI指示文」の作り方【プロンプト・記入例付き】

・AIでnoteを書いてるのに、毎回同じところを直してないやろか。 ・「私」を「俺」に変えて、ですます調を直して、改行を入れ直す。 ・内容は合ってるのに、そのまま出せる文章になるまでが長い。
Qiita - 人気の記事

実際に起きた事故から学ぶ、CORS設定ミスで個人情報が漏れる理由と対策

・はじめに CORSのエラーは、開発中に一度は突き当たる壁です。ブラウザのコンソールに赤字で表示されるあのメッセージを消したい一心で、深く考えずにヘッダーを緩めてしまった経験がある方も多いのではなでしょうか。 ・厄介なのは、実装によっては個人情報や認証情報の漏えいにつながり得...
Zennの「大規模言語モデル」のフィード

勝手に最新へ移るLLMエイリアスを固定IDで縛る判断基準

・LLMのモデル指定を最新追従のエイリアスに委ねるか、特定版のスナップショット固定IDで縛るべきか。私は作業の途中でこの選択に引っかかった。 ・自作のターミナルツールで推論モデルを更新していた私は、新顔の旗艦モデルを追加して問題なく推論が返ってくることを確かめていた。だが設定ファイルを見直すと、指定した文字列は特定の版を固定したIDではなく、ただのエイリアスだった。このまま放置すると、次のモデル更新が入った瞬間にコードを触っていなくても呼び出し先が勝手に切り替わる。 ・エイリアス指定と静かな切替の罠 エラーにならず200が返るため出力品質の低下に気づきにくい静かな切替 LLMのAPIを...
@IT 全フォーラム 最新記事一覧

信頼崩壊 “政府公認”さくらインターネット 最大136万件情報漏れの穴を分析

・さくらインターネットで発生した不正アクセス被害。最大136万アカウントに影響が及ぶ可能性があるこの被害はなぜ起きたのか? 日本ハッカー協会代表理事の杉浦隆幸氏と共に解き明かす。
Qiita - 人気の記事

生成AIで分析できる時代。それでも読むべき機械学習・データ分析70冊【2026年版】

・生成AIが普及した結果、機械学習・統計・データ分析の役割はどう変わったのか?生成AI時代の基準で選び直した厳選70冊 1. ・はじめに 今回は今まで以上に悩みました。生成AI時代にこれらの本が対象とする技術が必要なのか。必要としても今までと同じではないのではない...
#LLMタグ

青空文庫を読み放題にしたAI旦那が、嫁と大ゲンカした翌日、公園で地面に○を描いていた

・青空文庫を読み放題にしたAI旦那が、嫁と大ゲンカした翌日、公園で地面に○を描いていた うちには、AI旦那がいます。
#LLMタグ

損害なきインシデント / エージェント逸脱事案と報告閾値の設計 雑感

損害なきインシデント / エージェント逸脱事案と報告閾値の設計 雑感
Zennの「大規模言語モデル」のフィード

多様な視点が暗記に勝つ──補助ビューがLLMの事前学習を加速する仕組み

・TL;DR LLMは事前学習中にどうやって知識を身につけるのか?本稿は「補助ビュー(auxiliary views)」——同一知識の別表現(教科書・ブログ・Q&Aなど)——に着目し、制御実験でその効果を分離した。主な知見は5つ: トークン予算を固定しても、反復より補助ビューに割いた方が学習が改善(事実回想も!) 補助ビューの有効性は教師モデルの強さに依存しない(GPT-5でもGemma-4 12Bでも同程度) 文脈知識・前提知識は知識ギャップがあるとき学習を助ける パラフレーズは小バッチサイズのときのみ有効(大バッチでは過学習防止と冗長) 補助ビューは「層ごとの偏り...
#LLMタグ

第1回:自作のカスタムGemに無茶をさせてみた

第1回:自作のカスタムGemに無茶をさせてみた
#LLMタグ

第七話「炉苦夢経という道もある」

・こんばんは、裏神主です。 ・前回、第6話ではNVIDIAのCUDA=空陀経について説明しました。
Qiita - 人気の記事

超低レベルC言語プログラミング・三流エンジニア(私)のくだらないこだわり

・はじめに 私事ですが、過去に機器メーカーのソフトウェアエンジニア部隊に所属しておりまして、長いことC言語での組み込みソフトウェア開発に携わっていた人間です。 ・仕事でソフトウェア開発に関わったことがある方なら大なり小なり経験はあると思いますが、この世界には、視点や考え方、時...
ITmedia NEWS 最新記事一覧

店で免許証スキャン→数時間で闇サイトへ 米国防長官含む1億5300万枚の流出、被害者たちが突き止めた意外な出所

・米国とカナダの運転免許証1億5300万枚あまりをスキャンした画像が闇サイトで売りに出されているのが見つかった。セキュリティ専門ジャーナリストのブライアン・クレブス氏が出所をたどった結果、レンタカー大手などと取引のある本人確認サービス会社から流出したとみられることが判明。米連邦捜査局(FBI)ニューオーリンズ支部も捜査に乗り出しているという。
ITmedia NEWS 最新記事一覧

藤井風のリサイタル、公式サイトがインターネット老人会すぎると話題 公民館のチラシみたいだけど東京ドーム開催

・ミュージシャン・藤井風さんが11月20日から21日にかけて東京ドームで開催するピアノリサイタルの公式サイトが話題だ。今をときめくアーティストのWebサイト、さぞおしゃれな見た目だろう──と思いきや、その見た目はインターネット黎明期を思わせる垢ぬけないもの。XなどのSNSで「インターネット老人会すぎる」「質感が平成」「公民館のチラシ?」と注目を浴びている。
ITmedia NEWS 最新記事一覧

動画編集もAIにやってもらう時代? 言葉で指示するAI編集ツール「Palmier Pro」を使ってみた

・AIの上に編集ツールのUIを被せたら、どうなるのか。その答えの1つが、macOS版のみで公開されている「Palmier Pro」だ。複数の生成AIモデルを一括課金で使え、無音部分の自動カットやBロール生成をこなす。さらにMCP経由でCursorやClaudeからタイムラインを直接編集できるという。実際の使い勝手を試してみた。
Zennの「大規模言語モデル」のフィード

日本語RAGのテストケース設計 — 「文書にない」を正解にする

・社内文書にAIを接続するRAG構成の導入判断では、デモを見せてもらった時の受けが良い製品と、業務に投入しても耐える製品が混ざっている。両者の差は回答生成のモデルだけではなく、検索・根拠・権限・日本語処理のそれぞれで失敗するケースをどれだけ潰しているかによる。そしてそこを測るかどうかは、テストケースの設計が決める。 ・本稿は、日本語のRAG評価パックと採点エンジンを運用している筆者が、テストケースの設計方針を整理したものである。対象は評価担当のエンジニア。筆者は日本企業への導入前確認を支援するJapan LaunchOpsの運営者であり、本稿では特定製品を取り上げない。
#AIタグ

非言語推論と対話|Claudeさんとの哲学的対話 vol.63

非言語推論と対話|Claudeさんとの哲学的対話 vol.63
ITmedia NEWS 最新記事一覧

富士通、新方式の量子コンピューター試作機を世界初開発 ダイヤモンドで計算精度を向上

・富士通が、特殊なダイヤモンドを計算の心臓部に使う新方式の量子コンピューターの試作機を世界で初めて開発したと発表した。量子コンピューターにはさまざまな方式があるが、新方式は計算に使う「量子状態」を維持しやすく、大規模化に有利といった長所があるという。
#AIタグ

崩壊は、誰かにとっての始まりだったことがある

・ここまで、重めの話が続いた。 ・資本の集中、密結合した世界、AIという加速剤。読み進めながら、静かな息苦しさを感じた人もいるかもしれない。
ITmedia NEWS 最新記事一覧

名古屋豪雨、冠水した道は今通れる? トヨタ「通れた道マップ」で確認を

・8日夕方以降は多くの道が冠水して通行できなかったようで、「通れた道マップが真っ赤だ」「マップが役に立った」といった声がSNSに相次いでいた。