Monitors and detects risky behavior in enterprise AI agents via high-fidelity telemetry, security benchmarking, and a two-tier detector. Comprises ADR Sensor, ADR-Bench, and ADR Detector; deployed in production at Uber and validated on public benchmarks.
Generates English text matching pre-1931 style — a 13B language model trained on ~260B tokens of pre-1931 English, useful for historical-language generation and stylistic research. An instruction-tuned variant exists for interactive tasks.
Synthetic Korean-language persona dataset for training and evaluating conversational and generative models — 1M records (≈7M persona entries) with 26 fields aligned to South Korea’s demographic distributions. Built with NeMo Data Designer and released under CC BY 4.0.
Provides a large-scale multimodal embodied dataset (vision, depth, hand/arm kinematics, tactile) captured with an exoskeleton glove and egocentric sensors; organized as clip-level Zarr volumes for manipulation, imitation learning, and vision–action research. Includes both high-precision glove measurements and natural bare-hand clips; sizable storage required.
A GGUF-format preview checkpoint derived from Qwen3.6-27B — a multimodal, image-text-to-text reasoning model fine-tuned for more structured reasoning and consistent answer style; packaged for local inference and compatible with engines like vLLM/SGLang/llama.cpp.
Aggregates 750k+ Harbor-compatible agentic tasks from 100+ public sources (Parquet shards preserved). Includes tasks with and without verifiers for RL evaluation or SFT/datagen workflows, enabling reproducible trace generation.
Curated multimodal training corpus for spatial intelligence: ~8.16M QA-style samples paired with ~2.72M unique images (≈1.1 TB). Provides JSONL annotations, a 1,000-sample preview, and 52 independent image archives — used to train SenseNova-SI models.
Instruction‑tuning dataset of 8,706 Claude Opus 4.6/4.7–generated examples where each assistant turn begins with a synthetic <think> block to emulate chain‑of‑thought. Provided as four splits (full/instruct/roleplay/code), ~17M tokens total, Apache‑2.0, not manually reviewed.
A Chinese public-transit route-planning dataset for training and benchmarking LLMs that generate structured transit routes from origin–destination pairs. Releases include a large CPT corpus, SFT train/test splits, and a 30K real-world benchmark; anonymized and real testsets are provided for privacy-aware, fair evaluation.
Provides tools and samples to build context management, enrichment, and retrieval solutions on Google Cloud Knowledge Catalog — an AI-oriented data catalog that builds a dynamic knowledge graph for structured and unstructured data, suitable for RAG and agent workflows.
Large-scale synthetic video dataset of physically simulated multi-object interaction scenes for training and evaluating models on physical reasoning, depth and optical-flow estimation, instance segmentation, and physics-grounded captioning. Provides RGB + lossless depth, per-frame instance masks, per-object physics annotations (NPZ), VLM-grounded captions, and USD scene files — useful for world-model and simulation-to-real work; commercial use permitted.
Provides 10k–100k Indonesian-language cooking recipes in Parquet format, including dish names, ingredients and instructions — suitable for text-generation, recipe parsing, and culinary data analysis. Check the dataset card for license and field details.