September 3: BAAI releases 3 embodied datasets in quick succession
human judgment becomes a bottleneck for robot collaboration
This week scanned 86 HF orgs · 50 GitHub orgs · 71 blogs · 125 X accounts
Agent-related GitHub repositories continued expanding as of 2026-09-05 [P0], BAAI released multi-robot and embodied planning data on 2026-09-03 [P0], and NVIDIA plus LAION turned code Agent trajectories into the main battleground [P1]. The strongest data demand signal in this scan: code Agent trajectories.
Key Findings
This week's 5 high commercial value findings
As of 2026-09-05, NousResearch/hermes-agent has 241,719 stars, up 22,688 from 2026-07-23. openai/codex has 121,647 stars, up 20,903 from 2026-07-23. anthropics/skills has 174,335 stars, up 10,853 from 2026-07-23. anthropics/claude-code has 144,124 stars, up 5,374 from 2026-07-23. anthropics/claude-plugins-official has 35,926 stars, up 3,416 from 2026-07-23.
BAAI/Discoverse-L was released on 2026-09-03, with 129 downloads. BAAI/MobileVLA-CoT was released on 2026-09-03, with 56 downloads. BAAI/Orchestra-Bench was released on 2026-09-03, with 186 downloads, containing 12,000 samples and aimed at three-robot collaborative planning. BAAI/ToolPrivBench was released on 2026-08-31, with 53 downloads, focusing on how agents choose between "standard tools" and "high-privilege tools".
NVIDIA/Open-SWE-Traces currently has 21,253 downloads, up from 8,274 on 2026-07-23, a 156.9% increase. This repository was released on 2026-04-16, and the v1.2 update added agent trajectories generated by Qwen3.8-27B. On the LAION side, multiple terminal_bench- and swebench-related datasets were added this scan, including laion/terminal_bench_2_a3_rl_DCAgent_exp_rpt_unitsyn_python_v3_10_8B_20260829_192924, laion/terminal_bench_2_a3_rl_DCAgent_selfinstruct_naive_sandboxes_2_verified_70_8B_20260f49e33da, laion/swebench_verified_random_100_folders_a3_rl_DCAgent_selfinstruct_naive_sandboxes_2_c2a08420, and others.
allenai/BenchMIRT was released on 2026-09-01, with 327 downloads. allenai/BenchMIRT-item-statistics has 352 downloads. allenai/BenchMIRT-model-statistics has 342 downloads. allenai/BenchMIRT-eval-data has 39 downloads. Meanwhile, the paper "Subspace Inference Enables Efficient Active Reward Learning from Preferences" was published on 2026-09-03, "Patterning in Practice: Debiasing Reward Models with Susceptibilities" was published on 2026-09-01, and "Small Language Models as Judges for Rubric-Based Reinforcement Learning" came into community view on 2026-08-30.
google/WaxalNLP was released on 2026-01-19, with 25,857 downloads and 274 likes, and is a large-scale multilingual speech corpus for African languages. The paper "TalkFa: A Unified Benchmark for Farsi Dialogue Generation and Understanding" was published on 2026-09-01, targeting Persian dialogue tasks for more than 120 million users. The paper "MemeBridge: A Dataset for Benchmarking and Mitigating the Bidirectional Cultural Gap in Meme Interpretation" was published on 2026-08-31, focusing on cross-cultural meme understanding.
Demand Signals
Infer training data demands from model releases
Download Movers
Datasets with the largest download changes this week
| Dataset | Downloads | Weekly Growth |
|---|---|---|
| lerobot/community_dataset_v3 | 32,028 | +526.3% |
| allenai/asta-bench-submissions | 215 | +186.7% |
| nvidia/Open-SWE-Traces | 21,253 | +156.9% |
| allenai/asta-summary-citation-counts | 1,095 | -22.8% |
Want to discuss this issue?
Auto-generated by AI Dataset Radar · Updated weekly
AI Dataset Radar →