Team Ai
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01DataDrivenConstruction /cwicr-construction-rates CWICR — Construction Works, Items, Costs & Resources A multilingual, machine-readable database of national construction rate books for 30 countries / language locales. Each rate is fully decomposed into its work composition and resource breakdown (labour, machinery, materials), with unit prices, hierarchical classification, and physical parameters preserved in the source language. This dataset is the tabular source-of-truth behind the cwicr-vector-db-bgem3-v3 Qdrant snapshots. Use… See the full description on the dataset page: https://huggingface.co/datasets/DataDrivenConstruction/cwicr-construction-rates.tabular10M<n<100M3 likes856 downloads5mo agoHugging Face02srii2829 /cwicr-construction-rates CWICR — Construction Works, Items, Costs & Resources A multilingual, machine-readable database of national construction rate books for 30 countries / language locales. Each rate is fully decomposed into its work composition and resource breakdown (labour, machinery, materials), with unit prices, hierarchical classification, and physical parameters preserved in the source language. This dataset is the tabular source-of-truth behind the cwicr-vector-db-bgem3-v3 Qdrant snapshots.… See the full description on the dataset page: https://huggingface.co/datasets/srii2829/cwicr-construction-rates.tabular10M<n<100M0 likes139 downloads23d agoHugging Face03cwiz /soarm_pickplace_31025This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so100_follower", "total_episodes": 50, "total_frames": 5788, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 500, "fps": 30, "splits": { "train": "0:50" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/cwiz/soarm_pickplace_31025.tabularrobotics1K<n<10K0 likes57 downloads1y agoHugging Face04cwiz /soarm_pickplace_29925This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so100_follower", "total_episodes": 116, "total_frames": 13074, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 500, "fps": 30, "splits": { "train": "0:116" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/cwiz/soarm_pickplace_29925.tabularrobotics10K<n<100K0 likes51 downloads1y agoHugging Face05razvanalex /cwi-2018-en-newstabular10K<n<100K0 likes33 downloads3y agoHugging Face06cwinkler /green_patents Green patents dataset num_rows: 9145 features: [title, label] label: 0, 1 The dataset contains patent titles that are labeled as 1 (="green") and 0 (="not green"). "green" patents titles were gathered by searching for CPC class "Y02" with Google Patents (query: "status:APPLICATION type:PATENT (Y02) country:EP,US", 05/01/2023). "not green" patents titles are derived from the HUPD dataset (random choice of 5000 titles). We could not find any patents in HUPD assigned to any CPC class… See the full description on the dataset page: https://huggingface.co/datasets/cwinkler/green_patents.tabulartext-classification1K<n<10K1 likes20 downloads4y agoHugging Face07razvanalex /cwi-2018-estabular10K<n<100K0 likes18 downloads3y agoHugging Face08razvanalex /cwi-2018-en-wikinewstabular1K<n<10K0 likes12 downloads3y agoHugging Face09razvanalex /cwi-2018-detabular1K<n<10K0 likes8 downloads3y agoHugging Face10razvanalex /cwi-2018-en-wikipediatabular1K<n<10K0 likes5 downloads3y agoHugging Face11razvanalex /cwi-2018-frtabular1K<n<10K0 likes5 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.