<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<oembed>
  <author_name>NTTCom</author_name>
  <author_url>https://blog.hatena.ne.jp/NTTCom/</author_url>
  <blog_title>NTT docomo Business Engineers' Blog</blog_title>
  <blog_url>https://engineers.ntt.com/</blog_url>
  <categories>
    <anon>テクノロジー</anon>
    <anon>AI</anon>
    <anon>OSS</anon>
    <anon>アドベントカレンダー</anon>
  </categories>
  <description>この記事は、NTT docomo Business Advent Calendar 2025 19日目の記事です。 こんにちは、イノベーションセンターの鈴ヶ嶺です。普段はAIアクセラレータの検証に関する業務に従事しています。 本記事では、まずTenstorrentのAIアクセラレータアーキテクチャを紹介し、その特徴について説明します。次に、複数の演算を1つのkernelに統合するfused kernelによる最適化に注目し、標準正規乱数(randn)を例にTenstorrentのアクセラレータにおける具体的な実装方法と性能評価を共有します。その結果、従来の演算の組み合わせの標準正規乱数の実装と…</description>
  <height>190</height>
  <html>&lt;iframe src=&quot;https://hatenablog-parts.com/embed?url=https%3A%2F%2Fengineers.ntt.com%2Fentry%2F20251219-tenstorrent%2Fentry&quot; title=&quot;Tenstorrentにおけるfused kernel実装と性能評価 - NTT docomo Business Engineers&amp;#39; Blog&quot; class=&quot;embed-card embed-blogcard&quot; scrolling=&quot;no&quot; frameborder=&quot;0&quot; style=&quot;display: block; width: 100%; height: 190px; max-width: 500px; margin: 10px 0px;&quot;&gt;&lt;/iframe&gt;</html>
  <image_url>https://cdn.blog.st-hatena.com/files/26006613764871753/17179246901334487731</image_url>
  <provider_name>Hatena Blog</provider_name>
  <provider_url>https://hatena.blog</provider_url>
  <published>2025-12-19 20:12:53</published>
  <title>Tenstorrentにおけるfused kernel実装と性能評価</title>
  <type>rich</type>
  <url>https://engineers.ntt.com/entry/20251219-tenstorrent/entry</url>
  <version>1.0</version>
  <width>100%</width>
</oembed>
