AIAny
Icon for item

Joyo Kanji Yomi Benchmark

Provides kanji-level evaluation data for Japanese TTS: disambiguated sentence contexts targeting 4,378 kanji-reading pairs (2,136 Jōyō kanji) with 13,095 native-speaker–verified sentences and katakana-marked ground-truth readings for kanji-level error metrics.

Introduction

Oops! Something went wrong

[next-mdx-remote-client] error compiling MDX: Expected a closing tag for `<>` (6:125-6:127) before the end of `paragraph` 4 | - Coverage and granularity: covers all 2,136 Jōyō kanji and 4,378 kanji-reading pairs with three sentence contexts per reading (13,095 sentences), so you can evaluate per-reading behaviour rather than only word- or sentence-level quality — useful for pinpointing specific polyphony errors. 5 | - Native verification and disambiguation: sentences and annotations were reviewed by 35 native Japanese speakers through a multi-stage process, and ambiguous kanji-reading pairs that cannot be uniquely disambiguated by context were excluded — improving label reliability for evaluation. > 6 | - Evaluation-ready format: each sample includes a full-sentence katakana transcription with the target reading delimited by <> for automatic extraction; an accompanying evaluation toolkit handles TTS synthesis, ASR transcription, alignment, and metric computation, streamlining kanji-level experiments. | ^ 7 | - Focused on TTS/ASR pipelines: designed to measure pronunciation selection in synthesized speech (kanji→phoneme mapping under sentential context), not as a general-purpose language modeling corpus. 8 | More information: https://mdxjs.com/docs/troubleshooting-mdx

Information

  • Websitehuggingface.co
  • Organizationssbintuitions
  • Published date2026/06/11

Categories

More Items

Evaluates whether video models reason according to physical laws by treating generated videos as visible reasoning traces and using a three-stage Perception–Formulation–Deduction protocol. Includes Orchard (400 mechanics videos), chain-of-frames prompting on annotated first frames, and a hybrid MLLM-plus-objective scoring suite for stage-resolved diagnostics.

Hugging Face

Provides intermediate pretraining checkpoints for the Aether-7B-5Attn base model to enable reproducible training-dynamics research. Includes three raw checkpoints (110k, 115k, 162k steps) packaged with model.safetensors, config, and tokenizer; uses a custom aether_v2_7way architecture requiring the aether_pkg loader.

Hugging Face

Provides 2,056 penetration-free cloth simulation trajectories (240 frames each, 493,440 frames, ~33 GB) across human garments, robotic manipulation, and object-collision scenarios. Includes per-vertex positions, per-frame displacements, mesh topology and collision fields under CC BY 4.0 — useful for training and evaluating learning-based cloth simulators.