<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<oembed>
  <author_name>wk0</author_name>
  <author_url>https://blog.hatena.ne.jp/wk0/</author_url>
  <blog_title>SB Intuitions TECH BLOG</blog_title>
  <blog_url>https://www.sbintuitions.co.jp/blog/</blog_url>
  <categories>
  </categories>
  <description>Responsible AIチームの綿岡晃輝、Evaluationチームの高山隼矢です。 本記事では、日本語マルチターンにおける安全性評価ベンチマーク「JMT-Safety」をご紹介します。 なお、本研究は早稲田大学 河原研究室（河原研）との共同研究です。 原論文はこちら 著者：五十里渚（早稲田大学）(若手奨励賞受賞)、福田創（早稲田大学）、高山隼矢（SB Intuitions）、綿岡晃輝（SB Intuitions）、河原大輔（早稲田大学） データはこちら 要約 ① 実対話に即したリスクを捉えるため、日本語で複数ターンの対話を評価する12,409件のベンチマーク JMT-Safety を構築…</description>
  <height>190</height>
  <html>&lt;iframe src=&quot;https://hatenablog-parts.com/embed?url=https%3A%2F%2Fwww.sbintuitions.co.jp%2Fblog%2Fentry%2F2026%2F07%2F29%2F100000&quot; title=&quot;JMT-Safety：日本語マルチターン安全性評価ベンチマークを公開 - SB Intuitions TECH BLOG&quot; class=&quot;embed-card embed-blogcard&quot; scrolling=&quot;no&quot; frameborder=&quot;0&quot; style=&quot;display: block; width: 100%; height: 190px; max-width: 500px; margin: 10px 0px;&quot;&gt;&lt;/iframe&gt;</html>
  <image_url>https://cdn-ak.f.st-hatena.com/images/fotolife/w/wk0/20260714/20260714105622.png</image_url>
  <provider_name>Hatena Blog</provider_name>
  <provider_url>https://hatena.blog</provider_url>
  <published>2026-07-29 10:00:00</published>
  <title>JMT-Safety：日本語マルチターン安全性評価ベンチマークを公開</title>
  <type>rich</type>
  <url>https://www.sbintuitions.co.jp/blog/entry/2026/07/29/100000</url>
  <version>1.0</version>
  <width>100%</width>
</oembed>
